Jobs in United States

Ai Deployment Manager in United States

5,082 active opportunities · Updated October 2026

Explore current ai deployment manager jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

W
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend +8.1%
Quick readStrong listing-quality and freshness signals

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Join WRITER's security team as a staff detection and response engineer and help protect the AI infrastructure that's transforming how the world works. You'll build sophisticated detection systems that identify attacks targeting our AI platform, training data, and model deployments while creating automated response capabilities that scale with our explosive growth. This isn't just traditional security work – you're defending cutting-edge AI/AGI systems against adversaries who are evolving their tactics as fast as AI itself advances. This role combines hands-on security engineering with strategic thinking to stay ahead of novel threats that don't exist in textbooks yet. You'll be the operational arm of our security function, translating threat intelligence into real-time detections, coordinating incident response across multiple teams, and hunting for sophisticated attacks across GPU clusters and distributed training environments. If you're excited by the challen

PythonRestAIGo
O
📍 Washington, District of Columbia, United States· Full-time
✓ Quality checkedCompany trend -82%

About the team The OpenAI for Government team is a dynamic, mission-driven group leveraging frontier AI to transform how governments achieve their missions. Our team works to empower public servants with secure, compliant AI tools (e.g., ChatGPT Enterprise in custom configurations) and mission-aligned deployments that meet government technical requirements with strong reliability and safety. As part of the OpenAI for Government team you will work across U.S. defense enterprises, helping to accelerate adoption through hands-on support, tailored training, and early insights—so civil servants can spend less time on red tape and more on meaningful work. If you're passionate about responsibly deploying frontier AI to uplift public institutions and transform how the government serves the American people, we want you on our team. About the Role The Strategic Delivery Lead (SDL) for DoW will identify and deliver on some of OpenAI’s most complex and high-impact deployments for the U.S. national security enterprise, with a focus on CDAO-led efforts including GenAI.mil , Advana (WDP), and other programs that accelerate DoW adoption of data, analytics, and artificial intelligence across the DoW enterprise. The role is highly cross-functional and customer-facing: You’ll partner with Research, Engineering, Go-to-Market, CDAO, the Military Departments and Services supported by CDAO, and other external stakeholders to align on scope, unblock delivery, and communicate progress at every level, from technical teams to C-suite. You will own and drive execution of technical delivery workstreams led by internal teams (Forward Deployed Engineers, Researchers, Solutions Architects, Enablement), ensuring clear problem-framing, crisp milestone definitions, proactive risk identification, rapid issue resolution, and effective translation between business objectives and technical solutions, while consistently communicating clear progress and demonstrating tangible value to national security cus

AWSGitRestAI
W
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend +8.1%

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Join WRITER's security team as a staff detection and response engineer and help protect the AI infrastructure that's transforming how the world works. You'll build sophisticated detection systems that identify attacks targeting our AI platform, training data, and model deployments while creating automated response capabilities that scale with our explosive growth. This isn't just traditional security work – you're defending cutting-edge AI/AGI systems against adversaries who are evolving their tactics as fast as AI itself advances. This role combines hands-on security engineering with strategic thinking to stay ahead of novel threats that don't exist in textbooks yet. You'll be the operational arm of our security function, translating threat intelligence into real-time detections, coordinating incident response across multiple teams, and hunting for sophisticated attacks across GPU clusters and distributed training environments. If you're excited by the challen

PythonRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Enterprise Go-To-Market organization helps the world’s largest companies adopt and scale our AI platform across their business—from ChatGPT Enterprise to our developer platform and APIs. We partner with organizations across Financial Services, Life Sciences, and Retail to build new AI-powered customer experiences, transform operations, and reimagine how work gets done in highly regulated and technically complex environments. As adoption accelerates through pilots, experimentation, and developer-led use cases, the GTM team turns that momentum into durable, enterprise-wide deployments. Account Associates sit at the front of this motion—where technical curiosity turns into executive engagement, and early product signals become enterprise-wide deployments. About the Role This role is designed as a launchpad for future enterprise sellers at OpenAI. This is a highly strategic role with significant exposure to enterprise sales, product, and customer strategy. You’ll develop core skills in discovery, stakeholder mapping, deal shaping, and executive communication while building a deep understanding of how leading companies adopt AI. We are looking for individuals with 4-6+ years of professional experience who are motivated to grow into world-class enterprise sellers over time. While many individuals in this role go on to become Account Directors or take on broader GTM roles, progression is based on performance and business needs rather than a fixed timeline. For the right person, this is one of the most accelerated paths to break into sales at OpenAI and build a long-term career in AI go-to-market. In this role, you will identify, shape, and advance complex enterprise opportunities by turning product usage, pilots, and inbound demand into qualified, high-conviction enterprise opportunities for our Large Enterprise and Strategic Accounts teams. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week (Monday thro

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Enterprise Go-To-Market organization helps the world’s largest companies adopt and scale our AI platform across their business—from ChatGPT Enterprise to our developer platform and APIs. We partner with organizations across Financial Services, Life Sciences, and Retail in LATAM to build new AI-powered customer experiences, transform operations, and reimagine how work gets done in highly regulated and technically complex environments. As adoption accelerates through pilots, experimentation, and developer-led use cases, the GTM team turns that momentum into durable, enterprise-wide deployments. Account Associates sit at the front of this motion—where technical curiosity turns into executive engagement, and early product signals become enterprise-wide deployments. About the Role This role is designed as a launchpad for future enterprise sellers at OpenAI. This is a highly strategic role with significant exposure to enterprise sales, product, and customer strategy. You’ll develop core skills in discovery, stakeholder mapping, deal shaping, and executive communication while building a deep understanding of how leading companies adopt AI. We are looking for individuals with 4-6+ years of professional experience who are motivated to grow into world-class enterprise sellers over time. While many individuals in this role go on to become Account Directors or take on broader GTM roles, progression is based on performance and business needs rather than a fixed timeline. For the right person, this is one of the most accelerated paths to break into sales at OpenAI and build a long-term career in AI go-to-market. In this role, you will identify, shape, and advance complex enterprise opportunities by turning product usage, pilots, and inbound demand into qualified, high-conviction enterprise opportunities for our Large Enterprise and Strategic Accounts teams based in LATAM. This role is based in San Francisco. We use a hybrid work model of 3 days in the offi

AWSRestAIGo
O
📍 Washington, District of Columbia, United States· Full-time
✓ Quality checkedCompany trend -82%

About the team The OpenAI for Government team is a dynamic, mission-driven group leveraging frontier AI to transform how governments achieve their missions. Our team works to empower public servants with secure, compliant AI tools (e.g., ChatGPT Enterprise, ChatGPT Gov) and mission-aligned deployments that meet government technical requirements with strong reliability and safety. About the role Forward Deployed Engineers (FDEs) lead complex deployments of frontier models in production. You will embed with our most strategic government and public sector customers—where model performance matters, delivery is urgent, and ambiguity is the default. You’ll map their problems, structure delivery, and ship fast. This includes scoping, sequencing, and building full-stack solutions that create measurable value, while driving clarity across internal and external teams. You will work directly with defense, intelligence, and federal stakeholders as their technical thought partner, guiding adoption, maximizing mission impact, and ensuring successful deployments at scale. Along the way, you’ll identify reusable patterns, codify best practices, and share field signal that influences OpenAI’s roadmap. This role is based in Washington DC, Seattle or San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required, including on-site work with customers. In this role you will Own technical delivery across multiple government deployments, from first prototype to stable production. Deeply embed with public sector customers to design and build novel applications powered by OpenAI models. Enable successful deployments across customer environments by delivering observable systems spanning infrastructure through applications. Prototype and build full-stack systems using Python, JavaScript, or comparable stacks that deliver real mission impact. Proactively guide customers on maximizing business and operational value from

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Software Engineering team is responsible for designing and building the scalable, performant, and secure backend systems that power our products—from early prototypes to large-scale deployments. We collaborate closely with product, hardware, and full-stack teams to ensure our infrastructure enables fast iteration while setting a strong foundation for long-term growth. About the Role As a Backend Engineer , you will design and build services, APIs, and infrastructure that support evolving product needs. You’ll apply a deep understanding of backend systems and maintain enough end-to-end context—from hardware to cloud—to guide technical decisions that best serve the product and team. We’re looking for engineers who thrive in fast-paced, collaborative environments and care deeply about building robust systems that scale. This role is based in San Francisco, CA . We use a hybrid work model of four days in the office per week and offer relocation assistance to new employees. In this role, you will: Architect, build, and maintain high-performance, secure backend systems. Design APIs, data models, and infrastructure to support evolving product needs. Balance near-term development velocity with long-term maintainability and scalability. Collaborate with cross-functional teams to ensure cohesive, end-to-end solutions. You might thrive in this role if you: Have 7+ years of professional software engineering experience, with a focus on backend systems. Have a proven track record of building and scaling systems from early stage to large scale. Are proficient with Python and Go, and familiar with a range of server-side technologies. Have a strong grasp of system design, performance optimization, and security best practices. Can reason about full-stack tradeoffs from hardware through cloud infrastructure. (Nice to have) Have experience with distributed systems and cloud architectures. (Nice to have) Bring a background in instrumentation, analytics, and performanc

PythonAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with research, software, and external hardware partners to shape the next generation of AI systems, from silicon through full-scale deployments. Our team focuses on understanding and optimizing performance across the full system stack—ensuring that architectural decisions are grounded in rigorous, quantitative analysis of real-world workloads. About the Role We are seeking a Performance Modeling Lead to build and lead a small, high-impact team responsible for answering forward-looking architectural questions across AI infrastructure systems. You will develop modeling frameworks and methodologies to evaluate system-level tradeoffs and guide key design decisions. Your work will directly influence reference architectures, vendor designs, and long-term infrastructure strategy. This role sits at the intersection of AI workloads, system architecture, and quantitative modeling, and requires strong technical judgment, ownership, and the ability to translate complex analysis into clear, actionable guidance. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Build and own a performance modeling framework/toolchain to evaluate AI systems across multiple levels of abstraction. Analyze and quantify architectural tradeoffs across compute, memory, networking, storage, and system topology. Develop performance models to guide decisions on: scale-up vs. scale-out architectures interconnect and network design memory hierarchy and system balance. Translate modeling outputs into clear recommendations for internal teams and external hardware vendors. Influence reference designs and vendor roadmaps through data-driven insights. Partner closely with machine learning, systems, and hardware teams to understand workload characte

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team The GTM Enablement team helps OpenAI’s customer-facing organizations turn rapidly evolving AI capabilities into consistent, high-quality customer outcomes. We build the onboarding, learning experiences, playbooks, and knowledge systems that help teams develop technical depth, stay current, and confidently guide customers through successful AI adoption. About the Role We’re hiring a Field Enablement Lead, Technical Success to design and scale enablement for our rapidly growing technical customer-facing teams. You will own programs spanning onboarding, continuous skill development, technical pitches and demos, and subject-matter-expert knowledge sharing. Working closely with Technical Success leaders, Product Enablement, and technical SMEs, you will turn complex product knowledge and field experience into practical systems that improve readiness, consistency, and the quality of customer interactions and deployments. In this role, you will: Design, implement, and scale comprehensive enablement programs aligned to Technical Success onboarding, role-based skill development, and ongoing readiness needs. Redefine and operate the Subject Matter Expert (SME) program, creating clear pathways for technical experts to share knowledge and raise technical depth and consistency across GTM. Own and proactively maintain a versioned repository of technical pitches, demos, playbooks, and launch-ready assets. Partner with Technical Success leadership, Product Enablement, and product SMEs to identify skill gaps and deliver targeted learning interventions. Capture, vet, organize, and make field- and SME-generated technical content easy to discover, trust, and reuse. You might thrive in this role if you have: 5+ years of experience in technical enablement, solutions engineering, solutions architecture, technical success, or a related role. A proven track record of designing and scaling technical enablement programs in high-growth SaaS or technology environments. Exceptional

AWSRestAIGo
R
📍 Foster City, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -85.9%
Quick readStrong listing-quality and freshness signals

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: As a New Grad Software Engineer, you'll join a team of exceptional builders working on products that are reshaping how the world creates software. You'll have the opportunity to work on everything from our AI-powered development platform to the distributed systems that enable real-time collaboration for millions of developers. This is a chance to define your career while defining the future of software development. You'll work on problems that matter, with the autonomy to drive solutions and the support to grow into a technical leader. What you will build: Product features that delight users and make it possible for anybody to create software AI coding agent that understands intent and generates production-ready applications Cloud infrastructure that provides instant, powerful development environments at global scale Platform features that enable one click deployments and scale to millions of users Required skills and experience: Recent graduate (2027) with a degree in Computer Science, Computer Engineering, or related field Strong programming skills in a modern language (JavaScript/TypeScript, Python, Go, Rust) Full-stack capabilities with experience in React, Node.js, and database technologies Growth orientation - eager to learn new technologies and take on increasing responsibility Collaborative spirit - you work well in cross-functional teams and value diverse perspectives What we value : Problem-solving mindset: Ability to approach complex operational challenges systematically and devise effective solutions Self-directed and autonomous: Capable of working independently while collaborating effectively with cross-functional teams Strong communication skills: Ability to explain complex technical conce

JavaScriptTypeScriptPythonJava
O
25 days ago
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Role As a Field CTO (Strategic Pursuits) , you will serve as a strategic bridge between our customers, go-to-market (GTM) teams, and product organization. You will partner closely with Sales, Technical Success, and Product to shape high-impact deals, guide customer architecture decisions, and influence our product roadmap based on real-world adoption and feedback. This is a highly cross-functional, externally facing leadership role for someone who combines deep technical expertise with strong business acumen and customer empathy. In this role, you will: Customer & Deal Strategy Partner with Sales, Product and Technical Success teams to support complex, high-value deals as a technical and strategic advisor. Translate customer business needs into scalable technical solutions and architectures. Engage with senior customer stakeholders (CTO/CIO/VP-level) to drive alignment on vision, roadmap, and adoption. Lead technical strategy discussions during key deal stages, including discovery, solution design, and executive presentations. Architecture & Implementation Guidance Guide customers on best practices for deploying and scaling AI-driven solutions in production. Provide architectural oversight across use cases such as LLM applications, integrations, data pipelines, and security. Act as a trusted advisor to ensure long-term success, not just short-term wins. Product & Feedback Loop Bring structured customer insights back to Product and Engineering teams to inform roadmap and prioritization. Identify gaps, opportunities, and emerging patterns from customer deployments. Influence product direction based on real-world usage, scalability needs, and enterprise requirements. GTM Strategy & Thought Leadership Help shape GTM strategies by identifying repeatable patterns across industries and customer segments. Develop scalable frameworks, reference architectures, and playbooks for broader field teams. Represent the company externally through customer en

AWSRestAIGo
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by building high-performance, scalable and reliable machine learning systems? Do you want to help define and build the next generation of AI platforms powering advanced NLP applications? We are looking for Members of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will work closely with many teams to deploy optimized NLP models to production in low latency, high throughput, and high availability environments. You will also get the opportunity to interface with customers and create customized deployments to meet their specific needs. You may be a good fit if you have: 5+ years of engineering experience running production infrastructure at a large scale Experience designing large, highly available distributed systems with Kubernetes, and GPU workloads on those clusters Experience with Kubernetes dev and production coding and support Experience with GCP, Azure, AWS, OCI, multi-cloud on-prem / hybrid serving Experienc

AWSAzureGCPKubernetes
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by leading the design of high-performance, scalable and reliable machine learning systems? Do you want to set technical direction and help shape the next generation of AI platforms powering advanced NLP applications? We are looking for a Lead Member of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team. You may be a good fit if you have: 8+ years of engineering experience running production infrastructure at a large scale, with a track record of technical leadership Demonstrated experience leading the architecture

AWSAzureGCPKubernetes
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hiring a Developer Productivity Engineer to help scale the engineering systems, safeguards, and developer workflows that enable our teams to move quickly without compromising reliability or performance. This role sits at the intersection of developer experience, CI/CD infrastructure, release engineering, production readiness, and inference systems reliability. You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack. About the Role We’re looking for an autonomous, high-ownership engineer who cares deeply about making other engineers faster, safer, and more confident. A major focus of this role will be improving the tooling and infrastructure around deploy gates for inference engine images. These systems help ensure that every image released to production and research is correct, numerically sound, free of regressions, and performant across key metrics like time-to-first-token (TTFT) and time-between-tokens (TBT). You’ll help harden the systems that catch issues before they reach production, reduce noise from flaky or infrastructure-related test failures, and improve automation around triage, ownership, debugging, and escalation when failures occur. You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack. This is not generic internal tools work. The systems you build directly impact OpenAI’s ability to support new model launches, safely ship inference optimizations to the world, onboard new infrastructure providers, and operate one of the largest and most p

PythonAWSCI/CDRest
O
📍 Washington, District of Columbia, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role Our technologies support some of the most important and impactful work in the world, including our strategic and high-impact customers in the public sector. As a Forward Deployed Security Engineer (FDSecE) you will be responsible for securing these novel applications of OpenAI’s technology. We’re looking for motivated, tenacious, and curious people who will work closely with engineering teams to ensure our infrastructure deployments are highly secure against our adversaries. As an FDSecE, you will embed directly throughout the lifecycle, working on-site and being hands-on to ensure the overall security of these deployments from design to production and through ongoing operations. This role is preferred to be based in Washington DC but may consider remote work. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. Travel to and working from customer sites is required for this role. In this role, you will: Deeply embed with our most strategic public sector customers to implement and maintain robust security controls. Be a design and technical thought partner by leveraging security expertise on protective controls including access controls, authentication, encryption, network, and system security. Collaborate closely with teammates, cross-functional teams, customers, and service providers to achieve security and compliance goals. Ensure continuity of critical security and monitoring c

PythonAWSAzureKubernetes
🔔

Get new ai deployment manager jobs in United States by email

Daily job updates · Unsubscribe anytime