Jobs in United States

Deployment Lead in United States

636 active opportunities · Updated October 2026

Explore current deployment lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI's Enterprise team builds AI-powered enterprise products and shared platform capabilities that help organizations put advanced AI to work securely and at scale. Our work spans enterprise workflows, agent experiences, integrations, identity, administration, security, governance, and deployment. About the Role As a Technical Program Manager on Enterprise, you will lead the technical strategy and execution behind the products and shared capabilities that make ChatGPT, Codex, and future OpenAI products useful, secure, and scalable for organizations. You will translate customer needs, competitive dynamics, and product priorities into actionable plans, influence architectural direction, and deliver durable capabilities across application, platform, and infrastructure layers. The role requires deep technical fluency, strong product judgment, and the ability to move between hands-on execution and broader enterprise strategy. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Drive technical strategy and execution for enterprise product and AI workflow initiatives, from design through implementation, launch, customer rollout, and iteration. Partner with engineering teams to influence architectural direction, interface definitions, and implementation tradeoffs across full-stack products, APIs, integrations, and shared platform systems. Translate enterprise customer requirements into actionable product priorities across AI-powered workflows, agent experiences, integrations, permissions, data access, evaluations, identity, security, governance, and deployment readiness. Represent the needs of enterprise buyers, IT administrators, security teams, business leaders, developers, and end users in product and technical decisions. Identify adoption barriers, competitive gaps, and opportunities to make OpenAI products easier for organizations

AWSRestAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten’s Inference Stack team builds the distributed runtime that powers large-scale LLM inference across our platform. We operate at the intersection of distributed systems, model performance, infrastructure, and developer experience. We enable customers to deploy and operate cutting-edge LLM models with industry-leading performance, scalability, reliability, and ease of use. As a Software Engineer on the Inference Stack team, you’ll work across the stack - from the developer experience customers use to deploy models, the libraries used for features like tool calling and reasoning, all the way down to the systems we use to orchestrate deployments in Kubernetes and route traffic efficiently. This is an ideal role for engineers who enjoy owning systems in production, solving hard integration problems, and making complex infrastructure simple and reliable for users. EXAMPLE INITIATIVES Blog Posts https://www.baseten.co/blog/nvidia-dynamo-day-baseten-inference-stack/ https://www.baseten.co/blog/how-baseten-achieved-2x-faster-inference-with-nvidia-dynamo/ https://www.baseten.co/blog/how-baseten-multi-cloud-capacity-management-mcm-powers-cloud-self-hosted-and-hybr/#comparing-deployment-options-cloud-vs-self-hosted-vs-hybrid RESPONSIBILITIES Develop infrastructure and orchestration systems for deploying and managing large-scale distributed LLM inference Work across the stack, from customer-facing features to low-le

KubernetesCI/CDRestMachine Learning
C
📍 United States· Full-time
✓ Quality checkedCompany trend -94.7%

We are looking for a Senior Forward Deployed Engineer to join the Customer Solutions team. You will be the technical authority embedded with our most complex customers, guiding them through deployment, architecture, onboarding, and the adoption of agentic development workflows. You bring deep hands-on experience from prior roles and use that depth to advise, design repeatable patterns, and drive customer outcomes end to end. You operate autonomously, own the technical success of your customers, and bring their experience back to shape how Coder builds and delivers. This is not an execution-only role. You are equally comfortable doing deep technical work with a customer and stepping back to design the repeatable pattern behind it. You are energized by ambiguity, motivated by customer outcomes, and capable of influencing organizational change alongside the technical work. This position is required to sit in the Eastern Time Zone. What You'll Do Serve as the primary technical authority for post-sales customers, guiding deployment architecture, environment design, and adoption of Coder across both human and AI development workflows Own onboarding engagements end to end, ensuring customers move from contract to productive adoption with speed and confidence Lead Get Well engagements where architecture decisions, rollout patterns, or organizational dynamics are limiting customer health or growth Help customers implement the technical and organizational changes required to adopt agentic development practices at scale Design and document repeatable delivery patterns across onboarding, architecture, and adoption that can scale across customer segments Design and recommend reference architectures tailored to each customer's cloud environment, security posture, and organizational constraints Translate customer environment complexity into clear guidance on networking, ingress, identity, and infrastructure patterns Anticipate technical and operational risks, escalate to the right

AWSAzureKubernetesCI/CD
SF
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $200K/yr

Quick readStrong listing-quality and freshness signals

About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role At Stitch Fix, we are at the forefront of innovation, creating cutting-edge solutions that blend fashion, technology, and data science. Our data science team combines machine learning with expert human judgment to generate innovative recommendations and insights that transform the way our clients discover what they love. We believe in a curiosity-driven data science culture where members are empowered to deliver impact through end-to-end model development. The diversity of the problems that we work on and the data-rich environment of our business make it possible, even essential, to bring the tools of multiple disciplines to bear on our hardest problems. We are looking for an experienced Styling Algorithms Team Manager to lead a group of talented machine learning engineers and data scientists. In this role, you will shape the future of fashion technology by driving the development and deployment of our styling algorithms, which empower our human stylists to delight clients by nailing their fit and style. This includes ML-, AI-, and product-driven feature curation and testing for our proprietary styling platform, as well as client-facing AI personalization experiences, such as Stitch Fix Vision, our virtual try-on. Responsibilities: Champion bold AI and ML interventions to improve our styling experiences, enabling our stylists to have a multiplicative impact on their client connection points. Likewise, actively shape the product roadmap for direct client-facing styling experiences, expand

PythonRestMachine LearningAI
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $272K/yr

Quick readStrong listing-quality and freshness signals

We’re looking for a Senior Staff Software Engineer with deep experience in GenAI/ML to join Datadog’s Application Performance Monitoring (APM) team. APM is a product which provides deep visibility into applications, enabling users to identify performance bottlenecks, troubleshoot issues, and optimize services. With distributed tracing, profiling, out-of-the-box dashboards, and seamless correlation with other telemetry data, Datadog APM provides some of the deepest and most structured visibility into the health and performance of applications. This context sets us up for an opportunity to be the world leaders in agentic investigations and incident troubleshooting. You’ll act as a technical leader within the APM group, focused on agentic workflows. You’ll lead efforts to design, train, evaluate, and deploy GenAI/ML models at scale. We’re looking for a product-minded ML engineer with strong technical expertise, excellent communication skills, and a track record of driving impactful initiatives end to end. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Serve as the technical owner for GenAI initiatives within APM, leading design, development, and deployment of ML/AI-powered features across multiple teams. Guide long-term strategy and technical direction for GenAI workflows across APM and related products. Build and benchmark GenAI/ML models using state-of-the-art techniques. Contribute to Datadog’s broader senior engineering community through thought leadership and collaboration on company-wide initiatives. Collaborate with cross-functional teams to build automated investigation and triaging tools. Influence product direction by bringing a strong product mindset to your work, always advocating for the end user. Guide teams through ambiguity, sc

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $154K/yr

Quick readStrong listing-quality and freshness signals

Datadog’s Implementation Services team helps customers implement and deploy Datadog quickly and successfully. Our team of architects leads the discovery, design, build, and launch of the Datadog platform to help customers accelerate time to value and get the most out of their investment. As a Senior Services Architect focused on Security and Cloud SIEM, you will help customers design, implement, and operationalize Datadog’s security capabilities across cloud, infrastructure, application, and log data sources. You will lead structured, outcome-driven professional services engagements delivered through a day-based professional services delivery model, partnering directly with customers through co-development working sessions, architecture workshops, implementation planning, and operational handoff. This role is ideal for someone who combines customer-facing consulting experience with strong cybersecurity knowledge, hands-on Cloud SIEM implementation skills, and an understanding of security control frameworks such as NIST 800-53, the NIST Cybersecurity Framework, CIS Controls, MITRE ATT&CK, SOC 2, PCI, HIPAA, ISO 27001, or similar standards. At Datadog, we place value in our office culture — the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Design and guide execution of Datadog implementations, focusing on Security and Cloud SIEM deployment, including discovery, requirements gathering, technical architecture, deployment planning, and launch. Partner with customers to map security requirements, controls, and monitoring objectives to Datadog capabilities, including frameworks such as NIST 800-53, NIST Cybersecurity Framework, CIS Controls, MITRE ATT&CK, SOC 2, PCI, HIPAA, ISO 27001, or similar standards. Advise customers on security data strategy, including log source prioritization, parsing, normalizat

KubernetesAIGoRust
M
📍 Boise, ID - Main Site, United States· Full-time
✓ Quality checkedCompany trend -75%

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Position Overview: As a Staff/Principal Equipment Engineer within PROCT ADT, you will provide technical leadership to improve equipment performance, capability, and global manufacturing outcomes. You will lead complex, multi-site initiatives to resolve equipment challenges, enable technology node readiness, and drive yield and defectivity improvements. This role requires deep expertise, strong ownership, and the ability to influence across teams and suppliers to deliver measurable impact in cost, performance, and scalability. Responsibilities Lead global ownership of equipment performance across toolsets, driving improvements in stability, availability, matching, and efficiency. Direct complex, multi-fab root cause analysis for equipment-driven yield, defectivity, and reliability issues using data-driven and physics-based approaches. Provide technical leadership in tool hardware and chamber behavior to expand process capability and improve performance limits. Partner with Technology Development, integration, and process teams to enable node readiness, volume ramp, and disciplined change control. Influence equipment supplier strategy by leading technical engagements, driving design improvements, and aligning roadmaps to business needs. Drive global standardization through development, validation, and deployment of Best Known Methods (BKMs) across sites. Lead initiatives to improve cost of ownership, reduce variation,

AIRecruitment
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Role As a Sales Manager, Energy, you will build and lead a team of Account Directors focused on strategic growth across utilities, oil and gas, renewables, power generation, and energy services. The team will partner with complex organizations modernizing operations, improving reliability, accelerating the energy transition, and adopting enterprise AI responsibly at scale. You’ll help the team navigate regulated enterprise sales cycles, deepen relationships with business, technology, operations, engineering, security, and risk leaders, and drive adoption of OpenAI’s platform across safety-conscious, asset-intensive organizations. Key Responsibilities Recruit, develop, and lead a high-performing team of Energy Account Directors. Create a strong coaching culture through deal reviews, account strategy sessions, ride-alongs, and structured 1:1s. Define the Energy GTM strategy, including subsector segmentation, account prioritization, partner strategy, executive engagement, and territory planning. Drive disciplined pipeline generation, forecast accuracy, and operational rigor. Guide multi-stakeholder opportunities involving operations, engineering, digital, data, security, legal, risk, procurement, and executive leadership. Help customers translate AI and API capabilities into measurable outcomes across asset and field operations, grid and generation planning, engineering knowledge, customer service, commercial workflows, and enterprise productivity. Partner with Product, Solutions Architecture, Technical Success, Legal, Security, Finance, and policy experts to support responsible deployment. Provide structured feedback on customer requirements, integration blockers, reliability and governance needs, and emerging industry trends. What We’re Looking For 15+ years of enterprise sales, GTM, or sales leadership experience. Proven experience building and scaling enterprise sales teams responsible for complex strategic accounts and large revenue targets. Deep underst

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Enablement team translates OpenAI’s product innovation into real-world customer impact by equipping customer-facing teams with the knowledge, practice, and frameworks they need to succeed. Within GTM onboarding, we establish a shared foundation during the first 30 days and build deeper role readiness for technical GTM teams as they progress through ramp. About the Role OpenAI is hiring a GTM Technical Onboarding Program Manager to define and deliver the role-specific onboarding experience for technical customer-facing GTM roles. Foundational onboarding establishes a common baseline during the first 30 days. This role owns the deeper development that follows, with a particular focus on days 30–90, when new hires need to translate product and deployment knowledge into practical, job-specific capability. You will partner with technical GTM leaders, Product, and subject-matter experts to turn complex product and deployment concepts into role-specific curricula, hands-on practice, assessments, certifications, and manager reinforcement. Success means technical GTM new hires build the product fluency, scoping ability, solution judgment, and customer-facing confidence required for their roles. Managers have clear readiness signals, learning experiences stay current as products evolve, and role-specific ramp becomes repeatable across teams and regions. In this role, you will: Own the design and delivery of day 30–90 role-specific onboarding for technical GTM roles Define learning outcomes and readiness standards in partnership with technical GTM leaders Build curricula and learning paths that translate complex product and deployment concepts into job-specific behavior Design and facilitate hands-on experiences such as labs, demos, simulations, technical discovery practice, and customer scenarios Create assessments, certifications, and manager-ready signals of proficiency and deployment readiness Partner with Product, technical subject-matter experts, and G

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. Through strategic partnerships and self-built campuses, we are scaling one of the world's fastest-growing AI infrastructure platforms. The Supply Chain organization ensures critical infrastructure components—from compute systems and networking equipment to integrated rack solutions—are sourced, manufactured, qualified, and delivered with the speed and reliability required to support frontier AI development. We partner closely with Hardware Engineering, Manufacturing Quality Engineering, Infrastructure Delivery, Hardware Operations, Finance, and suppliers worldwide to build a resilient, scalable supply chain capable of supporting rapid infrastructure expansion. As Industrial Compute continues to grow, Supply Chain serves as the operational bridge between engineering innovation and large-scale infrastructure deployment. About the Role We are seeking a Supply Chain Manager to lead strategic execution across sourcing, supplier operations, manufacturing quality, and infrastructure delivery for OpenAI's AI infrastructure portfolio. This role will oversee a multidisciplinary team responsible for strategic sourcing, manufacturing quality engineering, and technical program management while partnering closely with engineering, finance, hardware operations, and deployment teams. You will drive supplier strategy, manufacturing readiness, production planning, quality performance, and operational execution across the full hardware lifecycle. Success requires balancing long-term supplier strategy with day-to-day execution. You'll establish scalable operating mechanisms, strengthen supplier partnerships, manage complex cross-functional programs, and ensure OpenAI can rapidly deploy AI infrastructure without compromising quality, cost, or reliability. This is a people leadership role responsible for developing a high-performing organization while driving operati

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

Overview: The Data Acquisition team within the Foundations organization at OpenAI is responsible for all aspects of data collection to support our model training operations. Our team manages web crawling and GPTBot services and works closely with Data Processing, Architecture, and Scaling teams. We are looking for a skilled Software Engineer to join our Data Acquisition team. Responsibilities: Own and lead engineering projects in the area of data acquisition including web crawling, data ingestion, and search. Collaborate with other sub-teams, such as Data Processing, Architecture, and Scaling, to ensure smooth data flow and system operability. Work closely with the legal team to handle any compliance or data privacy-related matters. Develop and deploy highly scalable distributed systems capable of handling petabytes of data. Architect and implement algorithms for data indexing and search capabilities. Build and maintain backend services for data storage, including work with key-value databases and synchronization. Deploy solutions in a Kubernetes Infrastructure-as-Code environment and perform routine system checks. Conduct and analyze experiments on data to provide insights into system performance. Qualifications: BS/MS/PhD in Computer Science or a related field. 4+ years of industry experience in software development. Experience with large web crawlers a plus Strong expertise in large stateful distributed systems and data processing. Proficiency in Kubernetes, and Infrastructure-as-Code concepts. Willingness and enthusiasm for trying new approaches and technologies. Ability to handle multiple tasks and adapt to changing priorities. Strong communication skills, both written and verbal. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an

AWSKubernetesRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The ChatGPT Model Flywheel team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement. Team Focus Areas Model Experimentation: Enable rapid, safe model validation for ChatGPT and Codex products through experiment automation and lifecycle management. Model Deployment: Ensure safe, scalable deployment of model capabilities with robust rollout and operational tooling. Automate capacity management and incorporate platform-wide health monitors. Model Measurement: Build comprehensive evaluation and measurement systems for model quality, from user signals to launch scorecards. Improve end-to-end feedback loops for continual model improvement. Key Partnerships Collaborate cross-functionally with teams including Model Measurement DS, Research, Codex, Fleet, Inference, and API. In this role, you will: Elevate and consolidate ChatGPT’s harness, context management, and system prompt frameworks. Drive expansion and improvement of multi-tier model experiences. Support and scale self-serve experiment capabilities and automated guardrails. Lead model rollout automation, capacity management, and health monitoring. Shape end-to-end measurement systems (evals, grader signals, user feedback, etc.). You might thrive in this role if you have: Proven experience leading engineering teams in complex, cross-functional environments. Demonstrated success shipping production systems at scale (ideally for AI or large backend services). Deep understanding of model-driven product development, deployment lifecycle, and measurement tooling. Excellent communication and collaboration skills—experience interfacing directly with engineering, research, and product stakeholders. Prior involvement with large language models, distributed infrastructure, or experimentation platforms is a plus. Why Work With Us Tackle highly impactful technical challenges at the cutting edg

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The compute infrastructure team runs the GPU fleet and large-scale compute clusters that serve the models backing ChatGPT and the API, while also supporting training workloads for our next generation models. We operate a large, modern GPU fleet and provide a unified platform for other OpenAI teams to seamlessly run production Applied AI and Research training workloads. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role You will be part of an engineer-first TPM team as a Technical Program Manager for Compute Infrastructure who owns the end-to-end delivery of large-scale GPU clusters, partnering with engineers to bring clusters online across external providers and partners. You’ll run a broad, parallel portfolio spanning hardware, networking, power, and cooling—driving execution, risk management, and crisp alignment from working teams through leadership to deliver production-ready capacity at scale. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead end-to-end delivery of both New Compute SKUs and large-scale GPU clusters across an external partner ecosystem while supporting capacity planning for training and inference. Ability to contextually drive multi-threaded bring-up programs spanning hardware, networking, power, and cooling—owning plans, dependencies, and critical paths. Interface with chip providers to derisk long-term onboarding to new hardware platforms by working across kernels, comms, hardware, and scheduling engineering teams. Build and operationalize program mechanisms (roadmaps, milestones, risk registers, runbooks) that make delivery predictable at massive scale. Partner with engineering to improve cluster turn-up reliability, repeatability, and automation

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role you will Own technical delivery across multiple deployments from first prototype to stable production Build full-stack systems that deliver customer value and sharpen how we learn Embed closely with customer teams, understand their needs, and guide adoption of what you build Scope work, sequence delivery, and remove blockers early Make trade-offs between scope, speed, and quality; adjust plans to protect delivery Contribute directly in the code when progress or clarity depends on it Codify working patterns into tools, playbooks, or building blocks that others can use Share field feedback that helps Research and Product understand where the models succeed and where they can improve Keep teams moving through clarity and follow-through You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work Have scoped and delivered complex systems in fast-moving or ambiguous environments Write and review production-grade code across frontend and backend using Python, JavaScript,

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s Strategic Finance organization provides the financial insight and guidance that support the company’s long-term strategy and ambitious growth. We partner across Product, Partnerships, Engineering, and GTM to allocate resources to the highest-impact opportunities while preserving sustainable unit economics. The Applications Technology Strategic Finance team works closely with Product and Engineering to scale core offerings and bring new products to market. We lead revenue forecasting and performance management across the product portfolio, and partner closely with the Infrastructure Engineering team to forecast infrastructure and compute demand and to manage margin targets and headcount planning with rigor. About the Role We are hiring a Director of Product Finance, Infrastructure and Compute Demand to help drive strategic decision-making across our product organization. This person will serve as the primary finance partner to the AGI Deployment (Applications) Infrastructure Engineering organization and play a critical role in connecting product usage, compute demand, infrastructure costs, and margin outcomes across the portfolio. You will help build the frameworks that translate technical decisions into economic outcomes, shape planning across product and infrastructure teams, and ensure OpenAI is scaling its products with rigor and efficiency. You will also manage a small team, providing mentorship and leadership while remaining hands-on with analysis and execution. This role is ideally based in our San Francisco HQ, but we are open to NYC and Seattle. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Serve as the primary finance partner to the AGI Deployment (Applications) Infrastructure Engineering organization. Own the consolidated compute and infrastructure demand forecast for the products we bring to market, partnering closely with capacity planning and

AWSRestAIGo
🔔

Get new deployment lead jobs in United States by email

Daily job updates · Unsubscribe anytime