About the Team OpenAI’s Infrastructure organization builds the systems that power frontier AI workloads at global scale. As compute demand accelerates, our ability to rapidly convert infrastructure investments into usable production capacity has become mission critical. The CPU / Storage / PoP / WAN team is responsible for the end-to-end infrastructure layers required to bring compute online: server and cluster activation, storage platforms, Points of Presence (PoPs), backbone connectivity, and global network expansion. We operate across first-party facilities, colocation environments, and strategic cloud partners to ensure OpenAI can scale reliably and quickly. About the Role We are seeking a highly technical Program Manager to lead execution across CPU, Storage, PoP, and WAN infrastructure programs that directly unlock OpenAI’s next generation compute capacity. In this role, you will own complex cross-functional programs spanning compute cluster activation, storage deployment, PoP bring-up, and backbone expansion. You will coordinate hardware readiness, site readiness, network pathing, storage availability, vendor execution, and engineering dependencies required to turn contracted infrastructure into live training and inference capacity. This role requires strong technical fluency across hardware systems, network infrastructure, storage architecture, and deployment execution. You should be comfortable operating from rack-level implementation details through executive-level capacity planning discussions. This role is based in San Francisco, CA, with travel as needed. Key Responsibilities Lead end-to-end execution of CPU / GPU cluster activation programs across OpenAI’s global infrastructure footprint Drive readiness to convert contracted compute capacity into schedulable production clusters Own deployment programs for new PoPs, backbone nodes, WAN expansion, and interconnection initiatives Build integrated schedules spanning procurement, logistics, installation, st
Jobs in United States
Lead Systems Engineer in United States
2,434 active opportunities · Updated October 2026
Showing
15 jobs
Explore current lead systems engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
At Datadog, we're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale, enabling seamless collaboration and problem-solving among Dev, Ops, and Security teams globally for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. APM at Datadog is on its way to redefine how users interact with their telemetry. We are integrating intelligence directly into troubleshooting workflows to help engineers find root causes faster, navigate complex distributed systems seamlessly, and optimize application performance with minimal cognitive load. APM provides deep visibility from end-user interactions to backend services and we are now expanding this foundation with new AI-driven insights, guidance, and automation. As a Product Manager II for APM, you will work with world-class engineers, designers, and partner product teams to shape the future of Distributed Tracing, Performance Analysis, and Intelligent Troubleshooting. You will help build advanced capabilities that scale to thousands of customers and make sophisticated observability workflows accessible to every engineer, from experts to beginners. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What you will do: Develop a deep understanding of APM customers, their performance challenges, telemetry workflows, and competitors Lead conversations with design partners and strategic customers to uncover real-world performance issues, validate product assumptions, and guide solutions from early prototypes through General Availability Define and deliver the next generation of APM features with engineering and design, especially agentic on
At Datadog, we’re on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale, enabling seamless collaboration and problem-solving among Dev, Ops, and Security teams globally for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Observability Data Platform (ODP) is the backbone of everything Datadog delivers – powering how data is ingested, stored, routed, and surfaced across every product at planet scale. As a Senior Product Manager for ODP, you will work with world-class engineers and cross-functional partners to shape how the platform is deployed, controlled, and operated. You will define product direction across the control plane and data layer, translate complex infrastructure trade-offs into clear roadmap decisions, and help customers get the most from their observability investment – regardless of architecture, topology, or scale. At Datadog, we place value in our office culture – the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You Will Do: Develop a deep understanding of the Observability Data Platform customers – platform engineers, SREs, and product managers that own the product verticals – their infrastructure challenges, deployment topologies, and cost-to-serve trade-offs. Define product direction across multiple ODP surfaces, including the control plane and data layer, by articulating clear problem statements and desired outcomes, and partnering with engineering on technical approach and sequencing Lead conversations with design partners and strategic customers to understand real-world platform pain points, validate product assumptions, and guide solutions from early prototypes through General Availability Develop a co
From $224K/yr
Datadog's Technical Solutions (TS) organization is one of the largest organizations in the company — spanning Sales Engineers, Technical Account Managers, Enterprise Customer Success Managers, Technical Support Engineers, and Solutions Architects who work with prospects and customers across every stage of their journey with Datadog. Technical Solutions Operations (TSO) exists to make that organization faster, smarter, and more scalable. We build the systems, analytics & programs that give TS teams more leverage — and we measure our success by the business outcomes we drive, not the projects we complete. The programs this team runs touch every function in TS, and the operating model you build will define how that scales. We're looking for a Director of Technical Program Management to lead the TSO Program Management team. This role sits at the intersection of strategy and execution: you'll own the programs that shape how TS operates at scale, lead a team of technical program managers, and serve as a peer to the Directors and VPs who run the teams you support. Your counterparts are leaders overseeing hundreds of customer-facing technical professionals, and your role is to successfully interface with each organization with proactive solutions on how your team can help them be more effective. This is a rare opportunity to lead a function where the output isn't a deliverable, it's organizational capability. If you want to build something that compounds across an entire organization, this is the role. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do Run the PMO as a business impact function. Own the PMO operating model end-to-end, including intake, prioritization, scoping, execution, and impact measurement — with every program directly linked to measurable bu
$226K – $285K/yr
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As a Supply Chain Program Manager, you will own material readiness and supply chain execution for critical hardware programs spanning custom silicon, systems, memory, storage, networking, and rack infrastructure. You will work cross-functionally with Engineering, Strategic Sourcing, Manufacturing Operations, Finance, Planning, Quality, and external suppliers to develop and execute scalable supply strategies that support aggressive product development and deployment timelines. This role requires deep understanding of hardware supply chains, material planning, NPI execution, supplier management, and operational scaling in constrained and rapidly evolving environments. In this role you will: Material Readiness & Supply Planning - Own end-to-end material readiness across NPI and production phases, including building the necessary framework and processes for enablement. Drive supply planning and execution for long lead-time and constrained commodities including ASICs, HBM, DDR, SSDs, networking, optics, power, thermal, and mechanicals. Build and manage material readiness plans aligned to proto/pre-EVT, EVT, DVT, PVT, and mass production schedules. Monitor supply health, lead times, inventory positions, allocation risk, and capacity constraints. Drive shortage management, allocation mitigation, and recovery planning. Coordinate supply commits, forecast alignment, and supply continuity planning with suppliers and manufacturing partners. Cross-Functional Program Ma
About the Team The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust. About the Role As a Research Program Manager, you will lead initiatives that test the safety and robustness of OpenAI’s models through creative experimentation and structured evaluation. You’ll coordinate efforts across research and engineering teams to transform ambiguous risks into concrete research programs and influence future model development and deployment. We’re looking for people who are technically savvy, comfortable with ambiguity, and excited about shaping the future of safe AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead programs that explore unexpected model behaviors and identify failure modes. Translate vague or emergent risk signals into clear priorities and actionable research plans. Design and run creative evaluations, experiments, and red-teaming campaigns. Collaborate with research, product, and deployment teams to integrate findings into model training and deployment cycles. Develop repeatable systems for tracking model performance and understanding emerging behavior patterns. You might thrive in this role if you: Have strong experience in technical program management, with excellent organizational and communication skills. Are familiar with large language models, prompt engineering, or model evaluation techniques. Are comfortable managing fast-paced, high-uncertainty projects and shaping them from the ground up. Are creative and resourceful in devising new methods for testing model behavior and performance. Can effectively coordinate across technical and non-technical stakeholders to drive alignment and execution. About OpenAI OpenAI is an AI resear
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron is seeking a Facilities Water Services UPW and Wastewater Coordinator to support wafer manufacturing by coordinating daily work, maintenance, vendors, contractors, and projects for ultrapure water, process water, reclaim, and wastewater systems. This role partners with Operations, Engineering, Maintenance, Construction, Procurement, EHS, vendors, and leadership to plan, implement, document, and align work with production needs. Success requires strong organization, technical understanding, communication, attention to detail, and the ability to lead priorities across operations, maintenance, projects, and production schedules. Ideal candidates are proficient with SAP or other CMMS tools, Microsoft Office, trackers, dashboards, drawings, and documentation systems. They coordinate schedules, track action items, communicate status, support scope development, identify gaps, and improve safety, reliability, documentation, cost control, and execution quality. This role helps maintain critical facility systems while demonstrating Micron’s core values of People, Innovation, Tenacity, Collaboration, and Customer Focus. Responsibilities: Coo
About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.
About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. Working alongside leading cloud providers, engineering firms, construction partners, utilities, and equipment manufacturers, we are delivering hyperscale AI campuses that enable the next generation of frontier AI models. The Strategic Sourcing team develops and executes the commercial strategies that ensure our infrastructure programs have reliable access to the equipment, materials, and strategic partners needed to deliver at unprecedented scale. We partner closely with Infrastructure Delivery, Capacity Planning, Design Engineering, Hardware Operations, Finance, Legal, and our external suppliers to build a resilient global supply network capable of supporting Industrial Compute's long-term growth. As we continue expanding globally, strategic sourcing becomes a critical competitive advantage, ensuring our infrastructure programs remain cost-effective, resilient, and capable of executing against aggressive deployment timelines. About the Role We are seeking a Strategic Sourcing Manager, Data Center Infrastructure: Owner Furnished Equipment to lead sourcing strategy for the critical infrastructure systems that power Industrial Compute campuses. This role will develop commercial strategies, negotiate strategic supplier agreements, and manage relationships across engineering, construction, manufacturing, and infrastructure partners responsible for delivering mission-critical facilities. You will work closely with Infrastructure Delivery, Capacity Planning, Engineering, Finance, Construction, and external suppliers to ensure Industrial Compute has the capacity, supplier relationships, and commercial frameworks required to support rapid global expansion. The ideal candidate has experience sourcing major infrastructure systems for hyperscale data centers, mission-critical facilities, industrial construction, semiconductor manufacturing, energy infrastr
About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. Safely delivering increasingly capable AI systems requires scalable technical safeguards, clear ownership of emerging risks, rigorous deployment readiness, and close coordination across research, engineering, product, operations, legal, policy, and external partners. Our Technical Program Managers lead complex, high-stakes initiatives that turn safety commitments into deployed systems and measurable outcomes. We work across model development, infrastructure, product, and operational response to help ensure our technology is deployed responsibly and cannot be used to cause serious real-world harm. About the Role We’re seeking Technical Program Managers to drive complex product, platform, and safety initiatives across ChatGPT, API, enterprise, and related deployment environments. These roles operate at the intersection of technical strategy and execution: you will turn safety and product priorities into actionable plans, influence architectural and operational decisions, and deliver durable capabilities across model, infrastructure, application, and platform layers. Depending on the role, you may enable sensitive or high-impact model deployments, integrate safeguards into cloud and API platforms, prevent violent misuse and other serious harms, improve detection and enforcement systems, create platform solutions for safety or establish new programs as risks evolve. You will partner deeply with engineers, researchers, product managers, and operational teams while communicating technical tradeoffs and program decisions to senior leadership. You bring technical fluency, product judgment, and a strong execution record. You’re comfortable navigating ambiguity, advocating for users and developers, balancing safety with model usefulness, and leading cross-functional work with urgency, rigor, and empathy. Specific focus areas and scope will vary by opening and level. Thi
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role As a Solutions Consultant, Public Sector at Ramp, you’ll serve as the primary technical and financial advisor for government agencies, educational institutions, and nonprofits evaluating Ramp. You will guide prospective customers through technical discovery, design tailored solutions that align with compliance and security requirements, and deliver compelling demos that highlight Ramp’s ability to modernize financial workflows. Partnering closely with Sales, Product, Engineering, and Compliance, you’ll act as the bridge between complex customer needs and Ramp’s platform capabilities. Your role is equal parts storyteller, architect, and trusted advisor—helping public sector organizations unlock efficiencies while navigating procurement, compliance, and multi-stakeholder environments. What You'll Do Lead Technical Discovery & Demos : Conduct deep technical and workflow discovery with public sector prospects; deliver customized demos showcasing Ramp’s value for finance, procurement, and compliance teams. Design Public Sector Solution
From $116K/yr
We are seeking a motivated and experienced Technical Program Manager II to join Datadog’s Technical Solutions organization. This organization is made up of 1,200+ customer-facing technical experts around the world — including Sales Engineers, Technical Account Managers, Support Engineers, and Solution Architects — who work with prospects and customers throughout their journey with Datadog to deliver outstanding experiences and drive growth through product adoption. As a Technical Program Manager II, you will lead and support a variety of technical programs that help these customer-facing teams work more effectively. Depending on the needs of the organization, this could include programs related to internal tooling, process improvement, knowledge and content systems, or other cross-functional initiatives. You’ll partner closely with team leads, subject matter experts, and other stakeholders to bring structure and clarity to moderately complex, cross-team programs, and you’ll act as a multiplier for the teams you support by driving programs from planning through delivery. Candidates who have previously led projects for customer-facing technical teams are especially encouraged to apply. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do Drive delivery of cross-team programs that support one or more Technical Solutions teams — coordinating stakeholders, dependencies, and timelines to keep work on track from kickoff through delivery. Partner with team leads, subject matter experts, and cross-functional stakeholders to translate program goals into clear, actionable plans, and document scope, timeline, and quality expectations along the way. Establish and maintain program management fundamentals — project trackers, status updates, risk logs — so stakeholders always
From $162K/yr
This is a senior individual contributor role for someone who wants to actively shape how Engineering, one of the most important parts of how Datadog develops its people. You'll sit at the center of Datadog's biggest talent bets for Engineering: how we build leaders, define career paths and org design, evolve performance, move talent internally, and plan succession for our most critical roles. You’ll own this work end to end, from the first framing conversation with senior leaders through to delivering a program running at scale. AI is changing how Engineering builds software, and it is changing how People builds the programs that support Engineering too. This role sits at the centre of both: understanding how AI is reshaping engineering roles, skills and structures, and building AI-powered solutions within People to keep pace. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Design and lead complex talent programmes for Engineering, spanning leadership capability, career architecture, org design, performance, internal mobility and succession for critical roles. Partner directly with PBPs and senior leaders to turn ambiguous problems into clear programme goals, design principles and success measures. Help Engineering and People understand and respond to how AI is reshaping roles, skills and ways of working, and translate that shift into practical talent and org design choices. Stay hands-on from concept through to adoption: this is a build and run role, not a strategy and handover role. Work across Enablement, Learning, People Analytics and People Systems so what you build scales and embeds into core people processes. Equip PBPs with frameworks, tools and executive-ready narratives that support real adoption in the business. Operate in ambigu
About the Team OpenAI’s mission is to ensure that general-purpose artificial intelligence benefits all of humanity. The Payments team works across product, engineering, design, and finance to build the financial infrastructure that makes OpenAI’s products accessible to consumers and enterprises around the world. As AI introduces new ways for people and organizations to work, the team is defining how to support and monetize emerging forms of product usage, from usage-based pricing to agentic work. We’re building the foundational systems that help OpenAI products deliver clear, reliable, and scalable payment experiences while ensuring that this powerful technology is deployed responsibly. About the Role In this role, you’ll lead design for one of OpenAI’s most foundational product areas: the payments and monetization infrastructure that supports our consumer and enterprise products. You’ll partner closely with product, engineering, and cross-functional teams to shape how customers understand, manage, and pay for entirely new kinds of AI usage. Your work will extend beyond traditional checkout and billing. You’ll help define the systems, frameworks, and experiences behind durable pay-as-you-go models, Codex usage, and agentic workflows, translating complex business and technical requirements into intuitive experiences. As a product designer in a highly ambiguous and rapidly evolving space, you’ll influence both product strategy and the underlying infrastructure that OpenAI products depend on. This role is based in our San Francisco HQ. We offer relocation assistance to new employees. In this role, you will: Lead the design direction for foundational payments, billing, and monetization experiences across OpenAI’s consumer and enterprise products. Design and ship high-quality, end-to-end product experiences, from early systems and interaction concepts to high-fidelity prototypes and production-ready designs. Shape the infrastructure and product frameworks that support em
About the Team The GTM Data Science team partners with Go-to-Market, Technical Success, Product, Engineering, RevOps, and Strategic Finance to build the shared intelligence layer for OpenAI's B2B business. The team turns product usage, customer behavior, revenue, field activity, and customer feedback into rigorous insight products that help leaders and field teams understand where customers are succeeding, where adoption is blocked, and what actions will accelerate durable growth. We are building systems that make customer intelligence proactive: surfacing risk, expansion potential, product gaps, and repeatable playbooks before they show up as escalations or missed opportunities. About the Role As the Applied Data Science & Insights Lead for GTM Intelligence Solutions and Technical Success, you will be a hands-on technical leader responsible for shaping how OpenAI measures, understands, and improves customer adoption across our B2B products. You will build AI/ML-powered intelligence products that connect account health, product usage, customer lifecycle, support tier, qualitative sentiment, commercial context, and field actions into a practical operating system for GTM and Technical Success. This role will build the data science foundation for Technical Success: defining the metrics, models, operating insights, and decision systems that help the team scale customer adoption and expansion with rigor. You will also be expected to build and lead a small mighty team over time: setting direction, hiring and developing talent, creating operating cadences, and holding a high bar for technical rigor and business impact. You will lead the development of models, metrics, and decision systems that recommend what GTM and Technical Success teams should do next, explain why, and measure whether those interventions worked. Your work will help customers move from pilots to production, deepen usage across products, identify high-value use cases, reduce churn risk, and create a f
Other cities to consider
More places hiring for this role
Get new lead systems engineer jobs in United States by email
Daily job updates · Unsubscribe anytime