Jobs in United States

Deployment Lead in United States

636 active opportunities · Updated October 2026

Explore current deployment lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

C-
📍 New York, New York, United States· Full-time
✓ High-confidence listing

$180K – $220K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We're looking for Software Engineers to help build the next generation of CLEAR's identity platform. Beyond verifying identity, we're creating a secure, networked digital identity that enables seamless experiences across travel, enterprise, healthcare, financial services, and beyond. As a Software Engineer, you'll own complex technical problems from design through deployment, partnering closely with Product, Design, Security, and Operations to deliver reliable, scalable solutions. We're looking for engineers with a strong builder mindset who thrive in ambiguity, take ownership, and enjoy turning ideas into production systems. Level and team matching (open roles across the three pillars that make up Technology at CLEAR: Core Identity, CLEAR1 , and CLEAR Travel ) will occur towards the end of our interview process. Tech stack overview: Python / Java / React / Typescript What you’ll do: Design, build, test, and deploy scalable applications that power CLEAR's identity platform. Own projects end-to-end from technical discovery and architecture through implementation, rollout, and operational support. Partner closely with Product, Design, Data, Security, and Operations to translate business problems into simple, scalable technical solutions. Drive engineering excellence by improving system reliability, performance, testing, observability, and developer experience. Contribute to architectural decisions and continuously improve the scalability, security, and maintainability of our platform. Partner with teammates through thoughtful code reviews

TypeScriptPythonJavaReact
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team pAGI Infra team builds and operates the systems that make large-scale model training and evaluation reliable, efficient, and easy to run. Our work spans distributed training infrastructure, inference and grading platforms, compute scheduling, and research tooling. We partner closely with researchers and engineering teams to turn new research needs into dependable infrastructure, improve GPU efficiency, and shorten the path from an experiment to a validated model. About the Role We’re looking for an AI Systems Engineer to help scale the infrastructure behind our training and evaluation workflows. You’ll own projects from identifying bottlenecks and designing solutions through deployment and operation. The work combines distributed systems engineering, performance optimization, and close collaboration with researchers. You might build a shared grading service, improve resource allocation across workloads, or bring a new training stack into production — directly improving how quickly and reliably research moves forward. In this role, you will: Build and operate infrastructure for large-scale training and evaluation, improving reliability, throughput, and resource efficiency. Develop shared inference and grading platforms with automated capacity management, health monitoring, and visibility into performance. Improve compute scheduling and resource allocation to reduce idle GPU time and help workloads recover quickly from failures. Diagnose bottlenecks across training, inference, and orchestration, and work across teams to improve end-to-end performance. Build self-service tools, automated validation, and observability that help researchers launch experiments, diagnose issues, and compare results with less manual intervention. You might thrive in this role if you: Are excited about the potential of personal AGI and want to build the infrastructure that enables it. Have strong software engineering fundamentals and experience building or operating large-scal

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team The Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to train nuanced safety behaviors, how to make the model robust to bad actors, how to address privacy and security risks, and how to make the model trustworthy in safety-critical situations. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. About the Role We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities. In this role, you will: Research and implement methods for safety training, reinforcement learning, and adversarial robustness. Develop evaluations, identify model failure modes, and use findings to improve training. Work with research, engineering, security, and policy partners to support safe, reliable deployment. You might thrive in this role if you: Bring 4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness. Have a degree in computer science, machine learning, or a related field, and strong deep learning research or engineering skills. Have experience improving model safety for deployment and enjoy collaborative research. Are motivated by OpenAI’s mission and the responsible use of AI in safety-critical settings. Security Requirements Active TS/SCI clearance or equivalent. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefi

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team The Developer Experience team at OpenAI has a singular focus: empowering developers globally. Our mission is to provide every developer and startup on the planet with the most delightful and seamless experience to integrate AI into their applications and products. We ensure developers have the tools, resources, and support they need to unlock AI’s full potential. We create inspiring demos, developer tools, sample applications, and technical content that show developers how to build with Codex and frontier models like GPT-6 Astra, GPT-Live, and GPT-Image-2.5 to create powerful agents and AI-native applications. We collaborate closely with product, engineering, research, and GTM teams to ensure the developer journey, from onboarding with Codex to first API call to production deployment, is seamless, effective, and delightful. About the Role As a Developer Experience Engineer, Cyber, you will create technical content, developer tools, and sample applications that help developers and security teams secure the software they build and depend on. You will turn OpenAI’s cybersecurity capabilities, including Codex Security, into clear, runnable examples and useful learning experiences. You will work directly with developers, security practitioners, and technical founders to understand their needs and show how AI can improve secure development and defensive security workflows. We’re looking for a hands-on engineer who combines strong software skills, creativity, developer empathy, and curiosity about cybersecurity, and who can own projects from the first prototype through release and ongoing improvement. In this role, you will: Build demos, sample applications, and developer tools that demonstrate practical ways to use Codex, OpenAI’s APIs, and Cyber products for secure software development and defensive workflows. Create tutorials, blog posts, videos, and code samples that help developers understand security findings, validate and prioritize vulnerabilities, r

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI's Research Team is at the forefront of AI research, pushing the limits of what AI can achieve. Our team is dedicated to developing advanced AI systems that are powerful, safe, and beneficial for everyone. About the Role The Research IP Partnerships team is in need of Technical Program Managers (TPMs) to streamline the integration of our applied research with external strategic partners. This role is critical for synthesizing research from cross functional teams, enabling model deployment, and ensuring new technologies are effectively adopted. You will act as the connection that enables our partners to deploy the most advanced AI models. Your primary focus will be to increase our research velocity and ensure that our deployments are successful and collaborative with our partners. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Build and share a deep understanding of frontier AI model development Partner with internal and external teams to drive deployment of the latest OpenAI technologies Manage critical inquiries from both technical and non-technical partners Design and implement simple, scalable processes that solve complex problems Deliver high-profile pipeline and tooling projects on tight deadlines Work across research and engineering to align goals, streamline communication, and support business priorities You might thrive in this role if you: Have experience in a strategic partnerships and technical program management role Can right-size process to align stakeholders while ensuring speed of delivery (action-oriented) Are fantastic at building cross-functional relationships and having empathy for the many roles involved in deploying research Can design and build tools (via code / no-code / AI) to facilitate internal processes Are a great communicator across written, presentation, and visual forms. Are engaged and c

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Industrial Compute team is building and productizing infrastructure capabilities that help organizations deploy and operate advanced AI systems at scale. The team works across AI hardware, systems engineering, physical infrastructure, and customer delivery to turn emerging technologies into reliable, repeatable infrastructure solutions. Our work sits at the intersection of technical strategy, product development, engineering, and deployment. We partner closely with customers and internal engineering teams to solve complex infrastructure challenges spanning compute, power, cooling, controls, and facility efficiency. About the Role We are seeking a senior, hands-on Data Center Infrastructure Architect to develop and optimize the physical infrastructure required for large-scale AI deployments. This is a broad technical role spanning data center architecture, electrical and mechanical systems, high-density compute, controls, telemetry, and digital modeling. You will use simulation, operational data, and digital-twin approaches to evaluate infrastructure designs, identify system-level constraints, and improve efficiency, reliability, cost, and speed of deployment. The ideal candidate can move fluidly between first-principles analysis, facility and equipment design, computational modeling, engineering review, and real-world implementation. You should be comfortable working across disciplines rather than operating solely within electrical, mechanical, or software boundaries. Key Responsibilities Define system-level architectures for high-density AI data centers across power, cooling, IT equipment, controls, and facility infrastructure. Develop digital twins and other computational models that represent the behavior of data center systems under changing workloads, environmental conditions, equipment configurations, and failure scenarios. Use design and operational data to identify constraints, improve PUE and related efficiency metrics, and optimize

PythonAWSGitRest
B
📍 Berkeley, United States
✓ Quality checkedCompany trend +515.8%

Associate and Mid-Level Software Engineers Company: The Boeing Company The Boeing Company is looking for an Associate or Mid-Level Software Engineer to join our team in Berkeley, MO. Position Responsibilities: Lab Environment Provisioning: Design, provision, and maintain scalable lab environments on-premises and on cloud platforms (Azure, AWS) using Infrastructure as Code (IaC) tools such as Terraform and Ansible. Automated Integration Testing Orchestration: Develop and manage automated integration testing workflows that aggregate inputs from multiple program segments, ensuring comprehensive test coverage and timely feedback. Deployment Management: Manage and optimize automated deployment pipelines and mechanisms for both physical and virtual systems, ensuring reliable and repeatable software delivery. Data-Driven Feedback & Reporting: Orchestrate processes to collect, analyze, and deliver comprehensive feedback on product stability, performance, and integration issues to segment development teams, enabling continuous improvement. Collaboration & Communication: Work closely with cross-functional teams including development, QA, security, and physical lab team to align deployment strategies, testing requirements, and environment configurations. Security & Compliance: Integrate security best practices into provisioning, deploying, and testing processes, supporting compliance with relevant standards and frameworks. This position is expected to be 100% onsite. The selected candidate will be required to work onsite in Berkeley, MO. Basic Qualifications (Required Skills/ Experien

PythonAWSAzureDocker
R
📍 Foster City, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -85.9%
Quick readStrong listing-quality and freshness signals

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our FDE team and work directly with some of the world's largest organizations to turn their most ambitious ideas into production applications on Replit. As a Forward Deployed Engineer, you'll partner closely with customers to understand their technical and business needs, architect solutions, and build the integrations and applications required to deploy Replit successfully within complex enterprise environments. This is a deeply technical and hands-on role. You won’t just advise customers on what to build—you’ll build alongside them. You’ll take applications from initial idea and prototype through deployment and production, navigate complex enterprise environments, and solve the technical challenges that emerge when AI-powered software development meets real-world infrastructure, data, security, and organizational constraints. You will: Build with Customers: Embed with strategic enterprise customers to design and build high-impact applications and AI-powered workflows on Replit, taking projects from initial concept through production deployment. Architect Enterprise Solutions: Design secure, scalable architectures that connect Replit with customers’ existing systems, data, APIs, identity providers, and infrastructure. Integrate with the customer's stack: SSO, data warehouses, internal APIs, and SaaS systems. Own Technical Deployments: Serve as the technical owner for complex enterprise implementations, identifying blockers, debugging issues, and driving projects through to successful production adoption. Bridge Customers and Product: Develop a deep understanding of how enterprises use Replit and translate field insights, technical constraints, and recurring customer needs into actionable feedback

AISEMWarehouseHR
N
📍 Remote, United States· Remote
✓ Quality checkedCompany trend -8%

NVIDIA is looking for an experienced software engineer with infrastructure experience to become a senior member of the Cloud Foundations Automation - Development Team. We build and manage the automation ecosystem supporting NVIDIA's GPU Cloud and NVIDIA SuperPod deployments. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most hard-working and dedicated people on the planet working for us. If you're creative and autonomous, we want to hear from you! What you'll be doing: Developing software to enable efficient network design, deployment and day 2 management. Building product focused software solutions, used by internal and external customers. Helping us as we transform our workflows and organization into a centrally orchestrated configuration management framework, operating at scale across geographies. Owning and driving integrations with various service APIs such as Cloud Service Providers, to automate creation of environments and auto populate data sources in turn. Building on open source software, designing and implementing data structures and UI interfaces to automate processes from equipment purchase to device config generation to deployment to operations. Streamlining deployment mechanisms and life cycle operations Developing modern service architectures around streaming data and event pipelines. Working with infrastructure domain experts on true, zero touch deployment solutions and utilizing best of breed high performance computing management solutions. Be a proactive problem solver, looking out for new opportunities to improve our services and customer experience. Communicate readily with your peers across the organization, b

PythonKubernetesAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA is hiring an NCX Senior Engineer who is passionate about NVIDIA Cloud Partner (NCP) infrastructure operations to join our DSX team. This role involves working closely with strategic NVIDIA Cloud Partners to build and improve the operational capabilities essential for running large-scale NVIDIA accelerated infrastructure reliably in production. Your role involves guiding partners beyond the initial cluster deployment and validation phase into advanced Day 2 operations. These operations cover ongoing infrastructure health, observability, lifecycle management, quick remediation, performance validation, and operational readiness. You will engage directly with partner engineering and operations teams to develop consistent approaches that support NVIDIA workloads and the broader external customer environments of the partners. This is a highly technical, hands-on role at the intersection of NVIDIA accelerated computing, cloud infrastructure, distributed systems, and production operations. What you'll be doing: Lead NCP Day 2 operational readiness efforts. Collaborate directly with NVIDIA Cloud Partners to set up the systems, procedures, automation, and operational methods necessary to consistently manage NVIDIA accelerated infrastructure following initial deployment and activation. Build continuous infrastructure validation. Develop and implement methods to continuously validate GPU, CPU, storage, and network health. Do this across large-scale AI clusters to identify degraded infrastructure before it impacts critical training or inference workloads. Establish observability and operational telemetry. Help NCPs implement comprehensive telemetry, monitoring, alerting, dashboards, and operational signals across compute, GPU, InfiniBand/RoCE networking, storage, Kubernetes, and AI workloads. Devel

PythonKubernetesLinuxArtificial Intelligence
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA is seeking a Senior Firmware Engineer to join our CSP Engagements team, focusing on system software for Datacenter products such as GB200. This role combines deep technical expertise in embedded firmware development with customer-facing responsibilities to enable cloud service providers with next-generation computing platforms. You will work at the intersection of hardware and software, driving technical solutions from concept through deployment. What you will be doing: Design and develop firmware solutions for manageability and observability of data center servers. Actively participate in hardware bring-up activities, OOB firmware development, protocol stacks (Redfish, PLDM, MCTP, NSM) and hardware-software co-design for Cloud Service Provider deployments. Debug and troubleshoot NVIDIA GPU firmware issues, power management, performance, and thermal control problems for data center deployments, providing active support to CSPs. Partner directly with CSPs to deliver technical solutions, co-develop & co-debug features and optimizations, and provide support during new product introductions. Perform advanced system debugging, root cause analysis, and performance optimization for large-scale data center environments. Collaborate with AE, FAE, and Solution Architect teams to deliver integrated customer solutions and technical documentation. What we need to see: Deep expertise in data center server architectures, HPC systems, and hardware-software co-design. Deep expertise in embedded firmware, server management controllers, and hardware bring-up with proven track record of shipping production BMC solutions Strong knowledge of DMTF protocols (Redfish, IPMI, PLDM, MCTP, SPDM), telemetry frameworks, and out-of-band management architectures Expert-level skills in C/C&

Artificial IntelligenceAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team Industrial Compute is building the world's most advanced AI infrastructure. Working alongside our capital partners, engineering teams, and construction organizations, we design and deliver large-scale, mission-critical compute campuses that power the next generation of AI. Our Design organization brings together engineering, construction, and digital design to ensure facilities are coordinated, constructible, and optimized from the earliest planning phases through deployment. About the Role We are seeking a BIM Designer & Coordinator to support the planning and design of large-scale industrial and mission-critical facilities. In this role, you will develop, coordinate, and maintain multidisciplinary BIM models across Civil, Electrical, Mechanical, Architectural, and Structural disciplines, enabling early design validation, constructability reviews, equipment planning, and cross-functional coordination. You will partner closely with internal engineering teams, external design consultants, contractors, and project stakeholders to produce coordinated BIM deliverables that improve design quality, reduce project risk, and support efficient execution across Industrial Compute's global infrastructure portfolio. Key Responsibilities: Develop and maintain conceptual and schematic BIM models for large-scale industrial and MEP-intensive facilities. Coordinate BIM models across Civil, Electrical, Mechanical, Architectural, and Structural disciplines. Create and maintain federated models used for design reviews, spatial coordination, constructability analysis, and clash detection. Model major building systems including equipment layouts, utility corridors, electrical rooms, mechanical rooms, structural framing, site infrastructure, and architectural constraints. Coordinate equipment clearances, maintenance access, routing zones, shafts, risers, utility entrances, and major MEP pathways. Translate engineering sketches, basis-of-design documents, equipment lists

PythonAWSGitRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Health team, within OpenAI’s broader Personal AGI organization, has a mission to ensure AGI improves health for all humanity. Improving human health will be one of the defining impacts of AGI. Hundreds of millions of people already turn to ChatGPT for questions about their health and millions of clinicians use it weekly to support care delivery. Increasingly capable models create an opportunity to make high-quality medical intelligence more accessible across patients and clinicians—raising the floor of human health—and accelerate the new capabilities and scientific advances that raise the ceiling of human health. Our job is to make those benefits real. We work across the full model stack—pretraining, midtraining, reinforcement learning, post-training, evaluations, harnessing, and deployment—and connect that research to the patients, clinicians, and real-world outcomes we aim to improve. About the Role We’re looking for an exceptional, hands-on researcher who wants to build frontier health capabilities and turn them into impact at scale. This is a role for someone who can take an important, underdefined problem from 0→1: identify the right bet, build what’s needed to test it, and drive it all the way to a measurable improvement in the models and products we actually ship. We’re especially excited about two kinds of people: researchers with the technical depth to move the frontier in pretraining, reinforcement learning (RL) / post-training, or evals; and researchers with real depth in developing frontier biomedical AI capabilities. Prior experience in healthcare is helpful but not required. Research excellence, velocity, ownership, and alignment with the mission are most important to us. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own a high-leverage research direction end to end—from deciding which problem matters and h

Machine LearningArtificial IntelligenceAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

Join the engineering teams that bring OpenAI’s ideas safely to the world! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role As OpenAI continues to grow, we are looking for experienced, problem-solving engineers to ensure our systems scale. Our success depends on our ability to quickly iterate on products while also ensuring that they are performant and reliable. You will work in a deeply iterative, collaborative, fast-paced environment to bring our technology to millions of users around the world, and ensure it’s delivered with safety and reliability in mind. Successful candidates will play a crucial role in ensuring the reliability, scalability, and performance of our systems as we continue to expand. As a reliability expert, you will be at the forefront of maintaining and enhancing the stability, scalability, and performance of our rapidly evolving infrastructure. You will work closely with cross-functional teams, including software engineers, product managers, and data scientists, to build and maintain resilient systems that can handle our growing user base and workload. In this role, you will: Design and implement solutions to ensure the scalability of our infrastructure to meet rapidly increasing demands. Build and maintain the load, chaos and synthetic-testing software leveraged by development teams to make the systems they design and operate more reliable. Build and maintain automation tools to streamline repetitive tasks and improve system reliability. Build and maintain the platform for CPU, storage, GPU, and network lifecycle management to drive efficiency, accountability and dynamic optimization of our resources. Implement fault-tolerant and resilient design

AWSKubernetesRestMicroservices
V
📍 United States· Full-time
✓ Quality checkedCompany trend -88.6%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta’s new Enterprise Resilience team is being formed to support the next era of growth by powering services that are reliable, scalable, and resilient by design. As our customer base expands and our systems scale, we need a dedicated group focused on partnering closely with product engineering teams to build and operate robust distributed systems across all of Vanta’s environments, including our new FedRAMP deployment. In this role, you’ll help define the foundations of reliability at Vanta including shaping best practices, building core infrastructure, and guiding teams as they design services that perform consistently for customers. This team will have a broad and deep impact across product engineering. Your work will influence how every Vanta engineer builds, deploys, monitors, and maintains their services, whether for our commercial environment or regulated customers with more stringent requirements. You’ll develop tools and frameworks that make it easier to detect and remediate issues, improve operational readiness, and support feature development that meets the needs of increasingly large and complex enterprise customers. Vanta engineers design and develop new product functionality and infrastructure using modern frameworks and tooling, including TypeScript, React, Node.js, MongoDB, GitHub Actions, and AWS services such as Fargate and ECS. If you're excited to help define a new function, raise the reliability bar across an entire engineering organization, and build systems that scale with Vanta’s growth, we’d love to meet you. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll

TypeScriptReactNode.jsMongoDB
🔔

Get new deployment lead jobs in United States by email

Daily job updates · Unsubscribe anytime