Jobs in United States

Lead Systems Engineer in United States

2,434 active opportunities · Updated October 2026

Explore current lead systems engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $204K/yr

Quick readStrong listing-quality and freshness signals

The opportunity Datadog’s Infrastructure products help engineers understand and operate the systems their applications depend on. Our customers work in complex environments like Kubernetes and serverless, where infrastructure changes constantly, information is dense, and decisions about reliability, performance, and cost are closely connected. We’re looking for a Staff Product Designer to join Modern Compute, with an initial focus on Containers Autoscaling. Autoscaling helps engineering teams make better decisions about how their applications and infrastructure use resources. Designing these experiences requires making deeply technical systems understandable, helping customers act with confidence, and fitting into the tools and workflows they already use. The team is rethinking how workload and cluster autoscaling come together as a more coherent product experience. This includes how customers get started, understand recommendations, evaluate value, and safely apply changes across their environments. The work also connects to other parts of Datadog, including observability, Cloud Cost Management, permissions, and AI-assisted workflows. As a Staff Product Designer, you will help define that direction and lead the work from early problem framing through shipped product. You will partner closely with product and engineering, bring a high level of interaction and visual craft to complex workflows, and help raise the quality of design across Modern Compute. At Datadog, we place value in our office culture, the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to help our Datadogs find a work-life rhythm that works for them. What you’ll do Lead end-to-end product design for Modern Compute, initially focused on our Autoscaling product. Help define the product direction for an area that is still evolving, from early framing and exploration through detailed design and delivery. Design clear, trustwort

KubernetesGitAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse, misalignment, and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. About the Role We are hiring a Product Manager to focus on risk related to multimodal models. In this role, you will drive initiatives which ensure that OpenAI’s audio, image, and video deployments are safe, impactful, and aligned with user needs and technical innovation. You will clarify strategic priorities, develop safety-focused product roadmaps, and collaborate closely with AI researchers, software engineers, policy experts, and cross-functional partners. This role suits a proactive, technically skilled product manager adept at adversarial thinking and excited to tackle challenging, ambiguous problems through structured analysis and collaborative decision-making. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Partner closely with AI research, engineering, data science, policy teams, and other stakeholders to embed safety throughout the development and deployment of multimodal AI models - such as GPT-Live and ChatGPT Images - as well as multimodal capabilities in frontier AI models. Develop comprehensive frameworks for understanding and mitigating deployment safety risks, drawing on data analysis, expert consultation, and adversarial assessments. Define strate

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. About the Role The Safety Measurement Product Manager owns OpenAI's approach to measuring harm and safeguard efficacy in production, including driving the strategy for our suite of safety measurement platforms and products used across the company. You will partner closely with our safety research and engineering teams to determine what we measure, where we measure it, and how we measure it, feeding those insights directly into critical leadership decisions and back into our safety work. You will also represent the company's topline safety metric as well as prioritize incoming requests from partner teams to expand our safety measurement platform to more use cases. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Partner closely with data science, research, engineering, policy teams, and other stakeholders to craft a vision for understanding safety outcomes and prevalence on our platforms. Define strategic priorities and product roadmaps focused on improving safety measurement approaches will scaling our measurement platform to more use cases, products, and cross-functional team needs. Establish repeatable processes to integrate cutting-edge AI safety research into OpenAI’s safety m

AWSRestAIGo
G
📍 Austin, Texas, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the job Lead the team that proves our AI silicon performs reliably before it reaches customers. You will build and lead Graphcore's characterisation capability for next generation silicon and system platforms. Your work will help ensure our products perform consistently across real world conditions. You will define bring up and characterisation strategies, lead technical execution, and shape the lab infrastructure needed for success. You'll work across silicon, hardware, manufacturing, architecture and product teams to solve complex engineering challenges. This is a hands on leadership role with the opportunity to influence both product design and how Graphcore validates future AI systems. The team and culture This is a newly formed team within Manufacturing Operations. You'll have the opportunity to establish how the team works while building strong partnerships across engineering and operations. Day to day, you'll work closely with architecture, silicon, hardware, production test and product teams. Decisions are driven by data, technical evidence and close collaboration across disciplines. We value ownership and clear communication. You'll be trusted to lead technical direction, remove blockers and help teams make progress with confidence. What we're looking for Essential Proven track record of delivering complex technical projects as an individual contributor, manager, or project manager, with the ability to work independently and drive execution. Strong expertise in silicon digital device design, bring-up, characterisation, and silicon process technologies, with an understanding of their impact on transistor- and system-level performance. In-depth knowledge of high-performance processors, system-on-chip (SoC) architectures, and high-speed digital interfaces such as PCIe, Ethernet, and DDR. Experience with measurement automation, data analysis, and scripting/coding to develop automated test and analysis workflows, with familiarity of ATE systems and t

GitAIGoRust
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%
Quick readStrong listing-quality and freshness signals

Datadog's integrations are the connective tissue between our platform and the technologies our customers run in the real world. As a Sr. PM on the Agent Integrations team, you will own the vision, prioritization, and execution for 100+ integrations that run directly inside the Datadog Agent from foundational infrastructure (MySQL, Kafka, Kubernetes) to the rapidly growing landscape of self-hosted AI and on-premise enterprise technologies. This is a high-impact, breadth-first role at the intersection of infrastructure observability and the frontier of AI-native workloads. At Datadog, we place value in our office culture; the relationships it builds, the creativity it brings, and the collaboration of being together. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own the Agent Integrations roadmap. Determine which new integrations to build and which existing ones to improve, balancing customer demand, business impact, and engineering capacity across a catalog of 100+ technologies. Drive the expanding AI integration surface. Lead product strategy for self-hosted AI workloads, including LLM inference frameworks (e.g., Hugging Face TGI, BentoML), AI agents, MCP servers, and model orchestration tools, so Datadog customers can monitor every layer of their AI stack. Expand on-prem and hybrid coverage. Prioritize and execute new integrations for on-prem technologies including storage systems, HPC schedulers, network devices, and legacy enterprise platforms where customers run critical workloads. Build observability for ERP systems. Define and drive Datadog's strategy for monitoring enterprise ERP platforms (SAP, Oracle EBS/Fusion, Microsoft Dynamics) covering performance, job execution health, and integration layer telemetry so enterprise customers can observe their ERP stack alongside the rest of their infrastructure. Analyze adoption and customer feedback at scale. Use data from multiple sources to

SQLPostgreSQLMySQLMongoDB
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Stargate and 3P Engineering teams are responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems. We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute. Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy. As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads. About the Role We are seeking a Technical Program Manager, Token-as-a-Service (TaaS) to lead delivery of external compute capacity that directly serves OpenAI model workloads. In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale. You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput. This role sits at the intersection of infrastructure execution, systems readiness, and business impact. Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners. This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally. This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week. Relocation assistance is available. Key Responsibilities Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply. Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments. Build

AWSRestAIRust
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $244K/yr

Quick readStrong listing-quality and freshness signals

We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—allowing for seamless collaboration and problem-solving among Dev, Ops and Security teams globally for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Team: As organizations rapidly invest in AI applications and build out AI labs, telemetry volumes are growing exponentially and costs are becoming unpredictable. From LLM interactions to agentic workflows, AI systems generate unpredictable streams of logs, driving up costs and making it harder to maintain efficient observability. These challenges are critical for organizations in regulated industries with strict data residency requirements, where data must remain within controlled environments. Datadog’s Bring Your Own Cloud (BYOC) team is reimagining what observability and security look like at petabyte scale in the AI era. The Opportunity: The Group Product Manager - Bring Your Own Cloud (BYOC) role is responsible for defining and bringing to market the next generation of telemetry analytics and insights capabilities in an AI-first environment. This role is highly technical and creative in nature as you will envision novel ways to enable customers to cost-effectively explore, analyze and report over petabytes of data through a welcoming and easy-to-use interface. You will partner with various teams to take advantage of BitsAI capabilities and surface critical insights on volume usage and retention for popular use cases. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a team

SQLAWSAzureRest
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $280K/yr

Quick readStrong listing-quality and freshness signals

Datadog is seeking a Director of Product Management to lead our AI Observability portfolio and shape how organizations build, monitor, and scale AI systems in production. This role leads LLM Observability and helps define the next wave of innovation across GPU Monitoring, Distributed AI Monitoring, and emerging research-oriented tooling such as Model Lab. You will set the vision and strategy for this rapidly growing area, expanding established products while incubating new capabilities that deliver deep visibility into AI infrastructure, model performance, and distributed AI environments. As AI becomes core to modern applications, this team plays a critical role in ensuring customers can deploy and scale AI with confidence. We’re looking for a builder-minded product leader with strong technical depth and hands-on curiosity - someone who has built or worked closely with AI-powered products and understands the realities of production AI. You will lead a team of product managers and partner closely with engineering and design to advance Datadog’s leadership in AI observability. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the vision and strategy for AI-driven products, ensuring alignment with overall company goals and customer needs. This will include managing our embed program to enhance the capabilities of existing products as well as developing dedicated and independent AI products. Lead and mentor a team of product managers, helping them grow and advance their careers while ensuring the delivery of high-quality, AI-powered features. Collaborate with cross-functional teams including engineering, data science, marketing, and sales to deliver AI product solutions that meet customer needs and business objectives. Identify new opportunities for

Machine LearningAIGoRust
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team The Search research team focuses on building the systems that help AI systems find, retrieve, and use information from the world. We aim to make answers more useful and grounded for more than a billion ChatGPT users. About the Role We’re looking for a Technical Program Manager to lead a broad portfolio of research and engineering programs that power search. You’ll partner closely with researchers, engineers, and product leaders to turn ambitious goals into clear plans, resolve dependencies, and move complex technical work forward. This role combines technical depth, product judgment, and hands-on execution. You’ll work across retrieval, indexing, and model improvements, while collaborating with policy, legal, and external data partners. You’ll help teams make informed tradeoffs and build practical ways of working that support a fast-moving research environment. This role is based in San Francisco, CA. In this role, you will: Lead programs across model training, retrieval, large-scale indexing, and search infrastructure. Translate evolving goals into prioritized workstreams with clear owners, milestones, dependencies, and resource needs. Partner with research, engineering, and product leads to define requirements and make tradeoffs across scope, quality, performance, timelines, and cost. Establish program success metrics and use them to guide priorities and track improvements in coverage, answer quality, responsiveness, and trust. Identify technical and cross-functional risks early, drive blockers to resolution, and communicate progress and decisions clearly to teams and leadership. Coordinate with product, policy, legal, and external partners on data access, use, and presentation, helping teams resolve decisions that span technical and non-technical domains. Manage dependencies with data providers and build repeatable processes that help research and engineering teams execute effectively as the search effort grows. You might thrive in this role if you

AWSRestAIGo
C
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Coder is looking for a Senior Staff Product Manager to lead our product management function and raise the bar for how we build. You will develop PMs, strengthen product execution, and help turn strategy into measurable outcomes for customers and teams. You will manage a small team of product managers while partnering closely with Engineering, Design, GTM, and executive leadership. You will coach people, improve systems, and keep product decisions grounded in real customer evidence. What you’ll do here Coach and develop product managers at every stage, helping strong executors grow into strategic leaders. Set a high bar for product discovery, planning, execution, and measurement. Partner with Engineering, Design, Customer Success, Sales, and Marketing to align priorities and deliver outcomes. Create clarity around product priorities, tradeoffs, and investment decisions. Build a culture of ownership, accountability, authority, and autonomy across the product organization. Lead customer discovery and ensure product decisions are grounded in objective evidence, not assumptions. Improve how the product organization operates, including planning, roadmap development, product reviews, and decision-making frameworks. Serve as a senior product leader and trusted partner to executive leadership. What we’re looking for Significant product management experience, including experience leading or managing product managers. A strong track record of coaching and developing PM talent. Excellent customer discovery and validation skills, including interviews that uncover objective insights and real customer needs. Sound judgment and pragmatism. You know when to optimize for outcomes, when to improve process, and how to balance strategy with execution. A strong sense of ownership and accountability, with the ability to build those qualities across a team. Clear communication skills and the ability to influence across functions, levels, and priorities. Experience operating in fast-moving

AWSAIGoRust
M
📍 United States· Full-time
✓ High-confidence listingCompany trend -93.7%

From $1.3M/yr

Quick readStrong listing-quality and freshness signals

The Opportunity MongoDB’s partner ecosystem — systems integrators, ISVs, and the major cloud providers — is a strategic growth engine for the business. We’re looking for a Senior Manager, Sales Plays and Offerings to design the joint go-to-market motions that turn partner relationships into pipeline and revenue. This is a highly cross-functional, strategic role: you’ll build the sales plays and joint offerings themselves, partner with Enablement to get the field and our partners ready to sell them, instrument how they perform, and work with the Programs lead to make sure incentives reward the behavior we want to see. You will report directly to the VP of Partner Strategic Operations and act as a connective layer between Partnerships, Sales, Enablement, and Programs — translating ecosystem strategy into repeatable, measurable, field-ready motions. We are looking to speak to candidates who are based anywhere in the US for our hybrid working model. What You'll Do Build sales plays and joint offerings Design and package partner sales plays and joint solution offerings with priority ISVs, SIs, and cloud partners — defining the joint value proposition, target segment, competitive positioning, and playbook for how field and partner sellers execute it Partner with Product Marketing, Solutions Engineering, and partner counterparts to validate technical integration stories and translate them into a compelling, sellable narrative Prioritize which plays to build and scale based on market opportunity, partner readiness, and alignment to MongoDB’s strategic pillars (e.g., AI, migrations, industry verticals) Own the lifecycle of each play from concept through launch, iteration, and eventual retirement or refresh Drive partner and field enablement Partner closely with the Enablement team to translate each sales play into field- and partner-facing assets: pitch decks, battlecards, demo scripts, certification conte

MongoDBAWSAzureGCP
M
📍 United States· Full-time
✓ High-confidence listingCompany trend -93.7%

From $92K/yr

Quick readStrong listing-quality and freshness signals

We are hiring a Senior Technical Product Marketing Manager to lead positioning and messaging and to grow adoption of MongoDB Search and Vector Search as foundational components of our platform – the retrieval layer powering the next generation of grounded AI applications and agents. This is a high-impact role for a marketer who thinks like a builder. As developers architect increasingly sophisticated systems – RAG pipelines, agentic workflows, multi-modal search experiences – retrieval has moved from an implementation detail to a core design decision. You’ll join a high-performing, globally distributed team and partner closely with Marketing, Builder Relations, Product Management, Engineering, Partners, and Sales to develop, measure, and achieve cross-functional goals. The role requires technical depth in information retrieval — lexical and vector search, hybrid approaches, embeddings, re-ranking, agentic retrieval loops, and the tradeoffs that matter in production systems — paired with the product marketing instincts to turn that depth into crisp, differentiated messaging for distinct user and buyer personas. Hands-on experience building or shipping AI-enabled products is a strong advantage. Individuals with prior experience in technical sales, developer relations, or technical marketing are encouraged to apply. This person is a voracious consumer of AI research and pays close attention to shifting patterns in application architectures and development, including agentic systems. This individual is confident in communicating with technical practitioners and non-technical decision makers in one-to-few and one-to-many engagements for internal and external audiences. We are looking to speak to candidates who are based in the US for our hybrid working model. What You’ll Do Drive Strategy & Execution: Act as a strategic partner for high-impact initiatives that align with MongoDB’s long-term business goals in collaboration with Marketing, Developer Relations, Product

MongoDBAWSAzureAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. Through strategic partnerships and self-built campuses, we are scaling one of the world's fastest-growing AI infrastructure platforms. The Supply Chain organization ensures critical infrastructure components—from compute systems and networking equipment to integrated rack solutions—are sourced, manufactured, qualified, and delivered with the speed and reliability required to support frontier AI development. We partner closely with Hardware Engineering, Manufacturing Quality Engineering, Infrastructure Delivery, Hardware Operations, Finance, and suppliers worldwide to build a resilient, scalable supply chain capable of supporting rapid infrastructure expansion. As Industrial Compute continues to grow, Supply Chain serves as the operational bridge between engineering innovation and large-scale infrastructure deployment. About the Role We are seeking a Supply Chain Manager to lead strategic execution across sourcing, supplier operations, manufacturing quality, and infrastructure delivery for OpenAI's AI infrastructure portfolio. This role will oversee a multidisciplinary team responsible for strategic sourcing, manufacturing quality engineering, and technical program management while partnering closely with engineering, finance, hardware operations, and deployment teams. You will drive supplier strategy, manufacturing readiness, production planning, quality performance, and operational execution across the full hardware lifecycle. Success requires balancing long-term supplier strategy with day-to-day execution. You'll establish scalable operating mechanisms, strengthen supplier partnerships, manage complex cross-functional programs, and ensure OpenAI can rapidly deploy AI infrastructure without compromising quality, cost, or reliability. This is a people leadership role responsible for developing a high-performing organization while driving operati

AWSRestAIGo
P
📍 US; Remote, United States· Full-time· Remote
✓ High-confidence listingCompany trend -86.3%

From $189.3K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . The Conversion Visibility Modeling team enables a performant ads marketplace and helps prove value to advertisers by connecting Pinterest onsite activity with conversions that happen offsite (both digital and physical) in a privacy-preserving way. As a Machine Learning Engineering Manager on this team, you will lead a hybrid team of ML engineers and backend software engineers to build end-to-end identity and conversion visibility solutions across modeling, serving, and data infrastructure, so advertisers retain accurate, privacy-aware performance visibility as signals fragment and degrade. You will set the technical direction for high-impact ML systems that feed ranking, bidding, measurement, and reporting across Pinterest’s ads stack. What you’ll do: Attract, hire, develop, and lead a hybrid team of ML engineers and backend softwar

AWSGitRestMachine Learning
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Industrial Compute team is responsible for building the physical infrastructure that powers OpenAI’s largest-scale AI systems. We design, deploy, and operate next-generation compute infrastructure across a rapidly expanding global footprint, combining OpenAI-owned infrastructure with strategic cloud and infrastructure partners to support frontier AI workloads. As our infrastructure footprint grows, operational excellence across third-party providers becomes increasingly critical. Our team ensures external infrastructure partners consistently deliver the reliability, performance, and operational maturity required to support OpenAI’s rapidly expanding compute environment. About the Role We are seeking a Hardware Technical Program Manager, Infrastructure Partner Operations to lead operational delivery across OpenAI’s third-party infrastructure partners, including major cloud service providers and strategic compute vendors. In this role, you will serve as the primary operational program manager for external infrastructure partners, driving accountability for service delivery, operational readiness, incident management, performance reporting, and continuous operational improvement. You will work closely with partner engineering and operations teams while coordinating internally across Hardware Engineering, Infrastructure Operations, Capacity Planning, Networking, Supply Chain, Deployment, Reliability Engineering, and executive leadership. Success in this role requires someone who understands how hyperscale infrastructure organizations operate, can establish strong operational governance with external partners, and is comfortable driving complex technical programs without direct ownership of the underlying infrastructure. Key Responsibilities Own operational engagement with third-party infrastructure providers, ensuring consistent execution against operational commitments, service-level agreements (SLAs), and performance expectations. Develop operationa

AWSAzureGCPRest
🔔

Get new lead systems engineer jobs in United States by email

Daily job updates · Unsubscribe anytime