Jobs in United States

Senior Infrastructure Platform Engineer in United States

1,941 active opportunities · Updated October 2026

Explore current senior infrastructure platform engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

C
📍 United States· Remote
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a highly skilled Staff Data Engineer, Observability Engineering to join the Enterprise Observability Platform organization and help advance the next generation of observability, infrastructure, and security data capabilities. The Staff Data Engineer, Observability Engineering will play a critical role in designing, building, and operating scalable data pipelines and data products that power enterprise observability, operational intelligence, and security analytics across the organization. The Staff Data Engineer, Observability Engineering is a senior individual contributor responsible for developing and optimizing Databricks-based data engineering solutions that ingest, transform, govern, and deliver high-volume telemetry, infrastructure, application, and security data. This role combines deep hands-on technical execution with ownership of engineering excellence, operational reliability, performance optimization, and data platform best practices. Working closely with Observability Engineering, Security Engineering, Infrastructure Engineering, and Data Platform teams, the Staff Data Engineer, Observability Engineering will contribute to the evolution of the enterprise observability lakehouse by building resilient ingestion frameworks, establishing data quality standards, enhancing governance controls, and driving efficient, scalable data processing patterns. The id

PythonSQLAzure
N
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -40%

Company Description Novo is a venture-backed fintech that simplifies banking for small businesses. Since our Fall 2018 beta launch, we've expanded offerings from free checking accounts and debit cards to lending products, business tools (invoices, bookkeeping, and more), and integrations.Today , Novo is a powerfully simple banking platform that serves over 250,000 small businesses. In addition to providing smart tools built for entrepreneurs to better run and grow their businesses, Novo has processed billions in transactions in partnership with a number of established banking partners. Our vision is to be the go-to platform serving small businesses – from launch to everyday – so business owners can focus on growing, while Novo provides seamless money movement, money storage, and access to capital. We'd like to look back 5–10 years from now and know that we helped new generations of small businesses succeed because of the work we did at Novo. Novo raised $170 million in venture capital and is backed by leading investors, including Stripes, Valar Ventures, Crosslink Capital, and Notable Capital (formerly GGV). Learn more at https://www.novo.co . Role Description Novo is seeking a highly experienced and motivated product manager to own and scale our core banking experience — the accounts, money movement, cards, and program infrastructure that power banking for over 250,000 small businesses. This role is critical to ensuring the reliability, usability, and depth of the products at the heart of Novo, supporting our rapidly expanding customer base and business operations. The ideal candidate is a hands-on product leader who can set product direction for complex, regulated financial products, partner closely with Engineering, Design, Risk, and Compliance, and work hand-in-hand with banking partners to ship reliable, secure, high-quality experiences within a dynamic, regulated fintech environment. Responsibilities Product Strategy & Leadership Own the strategy, roadmap,

V
📍 United States· Full-time
✓ Quality checkedCompany trend -90%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Our Data and Analytics team is currently looking for a Senior Data Scientist to join us! You’ll be responsible for laying the foundation for a best-in-class business analytics function. You’ll partner closely with our business stakeholders to ensure that our analytics stack and processes meet the business needs today with an eye towards the future. What you’ll do as a Senior Data Scientist at Vanta: Build and maintain trusted product data assets using dbt, Snowflake, and modern analytics infrastructure Leverage AI-powered analytics tools and data agents (e.g., Snowflake Cortex) to accelerate insight generation, automate repeatable analysis, and scale decision-making Define and evolve measurement frameworks for product health, customer lifecycle, and AI-powered product experiences Partner closely with Product, Engineering, Design, and Customer Success to influence product strategy through data Help define Vanta’s analytics strategy and AI measurement practices as our product and data platform evolve Lead executive analytics reviews, translating complex analyses into clear recommendations that drive company decisions How to be successful in this role: 4+ years of experience working with data as a Data Scientist, Product Analyst, or Analytics Engineer in an applied business setting Strong foundation in SQL, Python (or R), statistics, and machine learning Experience designing and evaluating experiments, predictive models, and other statistical analyses to inform product decisions Experience building scalable data assets, metrics, and analytical frameworks on modern cloud data platforms (e.g., Snowflake, dbt) Deep experience with da

PythonSQLRestMachine Learning
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 ClickUp is seeking a talented and experienced Senior Search Engineer to join our team and help revolutionize our search capabilities. As a key member of our engineering team, you'll be responsible for optimizing and enhancing our search functionality, which is a critical feature of our platform and underpins our AI efforts. Your impact at ClickUp As a Senior Engineer on the Search squad, you'll be responsible for designing, optimizing, and scaling our search infrastructure. Our platform is built as backend services on top of Postgres and OpenSearch, focused on real-time search ingestion and serving. You'll work hands-on with these systems to ensure users can instantly find exactly what they need across their workspace. Your expertise will directly contribute to making ClickUp the most intuitive and responsive productivity platform available. Core Responsibilities Design and implement robust search solutions that scale with our rapidly growing user base Improve search relevance, accuracy, and speed to deliver the most relevant results to users at blazing fast speeds Improve our real-time indexing pipelines to ensure search results remain up-to-date Create measurement frameworks to evaluate and improve search quality Build and enhance vector search capabilities to power next-generation search experiences Collaborate with AI, backend, and product teams to integrate search into new features Troubleshoot complex search-related issues at scale Design and implement robust search solutions that scale with our rapidly growing user base Improve search relevance, accuracy, and speed to deliver the most relevant r

TypeScriptAWSMachine LearningAI
G
📍 United States· Full-time
✓ High-confidence listingCompany trend -100%

From $154K/yr

Quick readStrong listing-quality and freshness signals

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s

PythonKubernetesAISwift
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $221.4K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. WHY DATA SCIENCE & ANALYTICS? The Data Science & Analytics organization’s mission is to increase our speed, frequency and acumen of making decisions at scale by instilling a data-influenced approach to building products. We cover a wide area of the data spectrum including analytical data engineering, product analytics, experimentation, causal inference, statistical modeling and machine learning. Aligned and partnering with product verticals, we use this extensive toolbelt to discover new opportunities and unmet use cases, influence and shape the product roadmap and prioritization, build data products and measure impact on our community of players and creators. WHY CREATOR SERVICES? At Roblox, the Creator Services team enables unbounded creation through reliable core services and novel AI applications. As a Senior Data Scientist focused on Machine Intelligence , you will bridge the gap between high-tier engineering infrastructure and cutting-edge ML applications. You will be the primary strategic partner to product and engineering leadership, transforming unstructured data into actionable business insights and user-facing products. This is a "zero-to-one" environment. You will be tas

AWSGitMachine LearningAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $293.8K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Enterprise Security Engineer, you will advance Roblox’s Enterprise Security strategy by shaping and evolving security architecture in alignment with business objectives. You will lead the design, deployment, and governance of security solutions that safeguard Roblox’s corporate infrastructure while enabling scalable, secure operations. Partnering cross-functionally with Corporate Engineering and Trust & Safety, you will translate organizational priorities into resilient security capabilities that balance risk, compliance, and productivity. You will join the Platform, Enterprise, and Application Security group, reporting directly to the Senior Manager of Enterprise Security Engineering. You'll partner with security professionals across the InfoSec team, and work cross-functionally with teams throughout Roblox to drive security initiatives that scale with our business. You will: Define and maintain enterprise-wide security standards and principles that guide how security is implemented across business workflows, ensuring consistency, scalability, and alignment with organizational risk posture. Lead and drive initiatives across core security domains, including Endpoint Secur

AWSGitAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. A majority of our users interact with our products in languages other than English, and our products must work seamlessly across languages, regions, and cultures. The Internationalization team builds the infrastructure that enables OpenAI products to ship globally by default. We develop the systems that power localization, international product launches, and high-quality global user experiences across all OpenAI products. About the Role As a Senior Software Engineer on the Internationalization team, you will build the systems that power localization and international product launches at OpenAI. You’ll work on the platform that manages product content, translation workflows, and localization infrastructure across our products. This role sits at the intersection of AI systems, developer platforms, and product infrastructure. In this role, you will Build and scale OpenAI’s localization, content, and experimentation platform used across OpenAI product teams, including open-source components: Develop AI-powered translation pipelines combined with human-in-the-loop review workflows. Design systems that reliably deliver localized product content across web and mobile apps. Build tools that enable linguists and localization teams to review and improve translations. Develop developer tooling that simplifies localization and internationalization workflows. Build and maintain internationalization libraries used across OpenAI products: Design systems that correctly handle numbers, currencies, dates, and pluralization across locales. Improve support for multilingual interfaces and right-to-left languages. Partner with product teams to improve the international readiness of new features. You might thrive in this role if you Have strong software engineering experience building backend or full-stack systems. Have familiarity with Java, React, MySQL, and cloud infrastructure p

JavaReactSQLMySQL
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Security Engineering is the engineering function inside the Plaid security org that focuses on developing the industry-leading security systems and infrastructure. Security Engineering owns most of Plaid’s security-related infrastructure: secure data storage, key management systems, internal identity platform, internal authentication systems, internal permission management, and internal authorization service. We develop solutions across data encryption, key management, access control, and data loss prevention to protect sensitive consumer data. We believe in the Zero Trust security model and are always looking for ways to improve our authentication and access control platforms. About the role: You will develop security capabilities to secure Plaid infrastructure and sensitive data access. You will lead the team’s strategic planning in collaboration with the manager and other senior engineers. You will own, maintain, and build Plaid’s security infrastructure and services like IAM Gateway, Key Management System and Network Firewall. You will consult with product engineers to ensure Plaid services meet security standards. You will help educate and support other engineering teams to improve security in

G
📍 United States· Full-time
✓ High-confidence listingCompany trend -100%

From $128K/yr

Quick readStrong listing-quality and freshness signals

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team… GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the industry, powering the object, block, and file storage platforms that underpin hosting, applications, internal infrastructure, and next-generation AI/HPC workloads. If you're passionate about distributed systems, large-scale storage architecture, and solving complex reliability challenges, you'll work on infrastructure that few engineers ever experience. At GoDaddy, Ceph isn't a side project — it's a critical platform. Our environment spans 80+ production clusters, 20,000+ OSDs, and approximately 300 PB of raw storage capacity, supporting tens of billions of objects across multiple continents. The scale demands deep technical expertise in storage architecture, automation, observability, and performance engineering. As a Senior Site Reliability Engineer, you'll be a key technical owner of the platform, responsible for maintaining reliability, driving operational excellence, and influencing the future evolution of our storage ecosystem. You'll tackle challenging production problems, develop automation that operates at massive scale, contribute to architectural decisions, and collaborate with some of the industry's most experienced Ceph engineers. This is an opportunity to have direct impact on a storage platform that serves millions of customers worldwide. What You'll Get to Do… Own the reliability, performance, scalability, and capacity of large-scale production Ceph environments supporting object, block, and file storage wor

PythonKubernetesLinuxAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $295.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer on the Sharing team, you will lead large, multi-team initiatives with long-term technical vision and group-level impact. You will be expected to define and drive the platform-wide strategy powering how millions of users capture, share, and discover content on Roblox. In this role, you will set engineering standards, mentor senior engineers, and serve as a key architect of our long-term technical direction. You Will Drive Strategy & Execution: Own the outcome of complex, business-critical programs spanning several teams, often lasting years. Innovate at Scale: Develop and drive a multi-year technical vision for content creation and sharing, anticipating scale, technology, and business evolution. Elevate Reliability: Lead high-severity incident response across groups; drive durable systemic solutions that improve reliability and velocity. Architect Foundations: Regularly improve shared infrastructure and foundational systems, introducing frameworks that uplift development speed across the organization. Align Teams: Aligns multiple teams on shared technical direction, producing detailed design docs, phased roadmaps, and planning models that balance short an

JavaAWSGitAI
D
📍 Massachusetts, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $296K/yr

Quick readStrong listing-quality and freshness signals

Datadog’s Cloud Observability group is one of the core data retrieval and processing groups powering our foundational product, Infrastructure Monitoring. The group’s scope includes integration with all major hyperscalers (AWS, Azure, GCP, OCI), as well as both regional and GPU-specific cloud providers. As Director, you will own engineering for all clouds, generating more than 10 million metric points per second, managing ~40 engineers through a team of Engineering Managers. You’ll partner with Senior Directors and product leadership to shape the roadmap, not just execute against it, managing the growth of one of Datadog’s foundational teams. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You'll Do: Own engineering for all of Cloud Observability Manage ~40 engineers through a layer of Engineering Managers; this is a manager-of-managers role Shape the roadmap alongside product leadership rather than simply executing against it — push back on, iterate on, and help author the strategy for your area Drive AI adoption across the engineering org, from tooling and workflows to product features and team practices Navigate cross-team dependencies across the Agent, Telemetry Onboarding, Integrations, Action Platform, and Infrastructure Monitoring. Build and retain engineering talent in NYC, Boston, and Paris, mentor Engineering Managers toward Director readiness, and participate in the on-call rotation Who You Are: You have directly managed Engineering Managers, not just individual contributors You have deep experience with one or more cloud providers, ideally with experience operating large-scale systems in the cloud. You have a solid understanding of cloud economics, as well as how to balance performance and cos

AWSAzureGCPAI
N
📍 Remote, United States· Remote
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

NVIDIA DGX Cloud is an AI Factory designed to power the next generation of AI and industrial-scale breakthroughs. As a Principal Engineer for Security Architecture, within our Security Engineering organization, you will own a core security domain of the AI factory: the architecture, the paved road that delivers it, and much of the code underneath. You will hold the security design bar across DGX Cloud from inside the teams doing the building, and this is a founding seat on a new team. Security Engineering is a new organization at DGX Cloud, accountable for the security outcome of the platform, and Security Architecture is the function inside it that holds the design bar. Security here is fleet horizontal and stack vertical, so your work will cross every DGX Cloud engineering organization: you will embed with the teams building GPU clusters, control planes, and services, join their designs as a participant rather than an approver, and leave behind systems in which an entire class of risk is no longer possible. There is no architecture review board here and no approval queue. You are a senior IC with deep security domain knowledge, and the security bar holds because you helped set it and then helped ship it. What You Will Be Doing: Own a Security Domain End to End: Take architectural ownership of a core domain of DGX Cloud security, from the design through the system running in production. That could be tenant and GPU workload isolation, workload identity, infrastructure and network, supply-chain provenance, hardened baselines and patching, or deploy-time policy and admission control. Embed with the Teams Building It: Join the design early, write the code, and help land it. The posture is not "you did this wrong." It is "here are the considerations we need to meet, I will help, let's go to work." Build Paved Roads, Not

KubernetesLinuxArtificial IntelligenceAI
V
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -90%
Quick readStrong listing-quality and freshness signals

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's App Primitives team owns the shared building blocks that span data structures to UI: comments, notifications, event logs, feature flags, and task management. We own the general infrastructure; product teams own the logic built on top. Getting this layer right unlocks velocity for every team at Vanta. As the Engineering Manager, App Primitives at Vanta, you'll lead the team building the shared product primitives that every Vanta product team depends on, owning the infrastructure layer that makes collaboration, communication, and core workflows possible across the entire platform. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow the App Primitives team, owning hiring, team health, delivery, and the development of senior engineers and technical leads Own the strategy and roadmap for Vanta's shared product primitives: comments infrastructure, notifications platform, event log, feature flag system (Statsig), and task management Define and steward the interface model between App Primitives systems and product teams, ensuring product teams can build on top of shared infrastructure quickly and safely, without owning the underlying systems themselves Partner closely with product engineering leaders across Vanta to surface developer ne

RestAIRustHR
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $295.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Join Roblox as an Engineering Manager of Application Security and lead a team responsible for improving the security of our products, services, and development ecosystem. In this role, you will drive security across the software lifecycle, partnering with engineering teams to identify risks, improve secure development practices, and build scalable solutions that protect Roblox at scale. You will balance hands-on security work with longer-term investments in automation, tooling, and developer enablement. You will work closely with engineering, infrastructure, and security teams to reduce risk while enabling teams to move quickly and safely. This role reports to the Senior Manager of Application Security and is based in San Mateo with a hybrid schedule. You Have: 8+ years of experience in Information Security 2+ years of experience managing engineers Strong background in Application Security or Product Security Experience driving security programs across the software development lifecycle Solid understanding of common vulnerabilities (e.g., OWASP Top 10) and secure coding practices Experience working closely with engineering teams in modern environments (cloud, microservices, CI/CD) Prov

AWSCI/CDGitMicroservices
🔔

Get new senior infrastructure platform engineer jobs in United States by email

Daily job updates · Unsubscribe anytime