Jobs in United States

Ai Deployment Engineer in San Francisco

1,456 active opportunities · Updated October 2026

Explore current ai deployment engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

25/100

cooling · 8 related jobs

Hiring trend

-40%

Job postings compared with the previous 30 days

Remote options

37.5%

Share of matching jobs listed as remote

B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. Product at Baseten Product at Baseten is a nascent function. Our company today has a strong engineering culture, is heavily customer-obsessed, and moves fast. We're building the product function now, and you'd be one of the people who defines it. You'll work directly with our founders and with some of the best systems and AI engineers and you'll set the standard for what product looks like here. PMs at Baseten don't sit above engineers - you earn ownership by being technical, finding the truth in front of customers, building great cross-functional relationships, and just shipping great product experiences. The role Once a model is deployed, keeping it fast, reliable, and economical at scale is where production inference is won or lost. You'll own the surface that makes that happen: how deployments autoscale, how traffic is routed, how the system fails over, and how workloads scale across clusters and regions. You'll own these as products end to end - both how they work under the hood and how customers configure and observe them - and you'll help set and define the roadmap that infrastructure and product teams alike can build towards. This space is largely still evolving - think Cloud Infrastructure in mid-2000s. Your job is to make it 10x easier to reliably scale and serve AI models in production and set the market standard. Impact and outcomes you'll drive You will own how workloads scale and where they land — autosca

KubernetesRestMachine LearningAI
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by leading the design of high-performance, scalable and reliable machine learning systems? Do you want to set technical direction and help shape the next generation of AI platforms powering advanced NLP applications? We are looking for a Lead Member of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team. You may be a good fit if you have: 8+ years of engineering experience running production infrastructure at a large scale, with a track record of technical leadership Demonstrated experience leading the architecture

AWSAzureGCPKubernetes
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Employee Technology & Experience (ETX) team is responsible for delivering a world-class internal technology experience that enables employees to do their best work. We support and operate the employee-facing systems that keep the company moving quickly and efficiently. ETX spans support, logistics, AV, identity, endpoints, SaaS administration, automation, enterprise tooling, and internal infrastructure operations. We partner closely with Security, Engineering, Workplace, Finance, People, and other teams to keep OpenAI’s internal technology reliable, scalable, and moving at the pace of the company. About the Role We are hiring a Program Manager to help scale how IT operates across OpenAI. This role will lead complex cross-functional programs that improve operational maturity, streamline how teams work together, and turn high-impact initiatives into durable operational capabilities. You will work across IT, Security, Engineering, Workplace, and other functions to drive alignment, remove friction, and help build the operational foundation needed to support OpenAI’s rapid growth. You’ll be responsible for: Lead cross-functional operational programs that improve scalability, consistency, and operational maturity. Drive operational excellence initiatives across IT Support, employee lifecycle operations, meeting room and calendaring services, onsite support, vending, and research support environments. Build operating models, readiness plans, escalation paths, governance cadences, and success metrics for complex operational programs. Partner with technical teams to ensure new deployments, infrastructure investments, and internal platforms are operationally ready and sustainably supported at scale. Drive high-priority operational programs supporting company growth, including infrastructure expansion, operational integrations, and other emerging initiatives. Improve operational visibility, stakeholder alignment, and coordination across long-running cros

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%

$261K – $305K/yr

Quick readStrong listing-quality and freshness signals

About the team OpenAI’s Enterprise Go-To-Market organization helps the world’s largest companies adopt and scale our AI platform across their business—from ChatGPT Enterprise to our developer platform and APIs. We partner with organizations across Financial Services, Life Sciences, and Retail to build new AI-powered customer experiences, transform operations, and reimagine how work gets done in highly regulated and technically complex environments. As adoption accelerates through pilots, experimentation, and developer-led use cases, the GTM team turns that momentum into durable, enterprise-wide deployments. Sales Development sits at the front of this motion—where technical curiosity becomes executive engagement and OpenAI’s enterprise relationships begin. About the role We are hiring a Sales Development Manager, AMER to lead a team of SDRs supporting our Large Enterprise business across North America. Reporting into Sales Development leadership, this role will be responsible for coaching and developing a high-performing team focused on generating qualified pipeline and creating strategic enterprise opportunities. This is a foundational front-line leadership role within OpenAI’s enterprise GTM organization. You will help define how we engage prospective customers, convert product and developer interest into enterprise conversations, and scale repeatable outbound and inbound motions across some of the world’s largest organizations. You will partner closely with Sales, Marketing, Solutions Engineering, Partnerships, and Strategy & Operations to drive pipeline generation and help shape the future of OpenAI’s enterprise growth engine. In this role, you will: Lead, coach, and develop a team of SDRs supporting Large Enterprise, Verticals, and Strategic Accounts across AMER Drive high-quality pipeline generation through outbound prospecting, inbound qualification, account-based motions, and product-led signals Establish strong operational rigor around activity metrics,

AWSRestAIGo
O
24 days ago
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Role As a Field CTO (Strategic Pursuits) , you will serve as a strategic bridge between our customers, go-to-market (GTM) teams, and product organization. You will partner closely with Sales, Technical Success, and Product to shape high-impact deals, guide customer architecture decisions, and influence our product roadmap based on real-world adoption and feedback. This is a highly cross-functional, externally facing leadership role for someone who combines deep technical expertise with strong business acumen and customer empathy. In this role, you will: Customer & Deal Strategy Partner with Sales, Product and Technical Success teams to support complex, high-value deals as a technical and strategic advisor. Translate customer business needs into scalable technical solutions and architectures. Engage with senior customer stakeholders (CTO/CIO/VP-level) to drive alignment on vision, roadmap, and adoption. Lead technical strategy discussions during key deal stages, including discovery, solution design, and executive presentations. Architecture & Implementation Guidance Guide customers on best practices for deploying and scaling AI-driven solutions in production. Provide architectural oversight across use cases such as LLM applications, integrations, data pipelines, and security. Act as a trusted advisor to ensure long-term success, not just short-term wins. Product & Feedback Loop Bring structured customer insights back to Product and Engineering teams to inform roadmap and prioritization. Identify gaps, opportunities, and emerging patterns from customer deployments. Influence product direction based on real-world usage, scalability needs, and enterprise requirements. GTM Strategy & Thought Leadership Help shape GTM strategies by identifying repeatable patterns across industries and customer segments. Develop scalable frameworks, reference architectures, and playbooks for broader field teams. Represent the company externally through customer en

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. Our Go-to-Market team helps organizations understand, adopt, and deploy OpenAI’s technology to solve meaningful business challenges and create lasting value. The Technology team works with leading software, internet, cloud, infrastructure, cybersecurity, semiconductor, and digital-native companies as they build new AI-powered products, transform internal operations, and rethink how they serve their customers. We partner with executives, product leaders, engineers, and go-to-market teams to help organizations integrate OpenAI’s capabilities into their products and businesses responsibly and at scale. The team collaborates closely with Solutions Engineering, Customer Success, Product, Research, Partnerships, Marketing, and Operations to turn customer priorities into successful, durable deployments. About the Role We are looking for an experienced Account Director, Tech to help build and grow OpenAI’s business across the technology industry. You will own relationships with a portfolio of strategic technology companies, helping executive, product, and technical leaders understand how OpenAI’s products can accelerate innovation, improve productivity, and create differentiated customer experiences. You will be responsible for developing account strategies, creating qualified pipeline, navigating complex enterprise sales cycles, and expanding adoption across products, teams, and use cases. This role requires a combination of enterprise sales leadership, technical fluency, commercial judgment, and the ability to operate credibly with both business and engineering stakeholders. You should be comfortable engaging with customers that have sophisticated technical environments, rapidly evolving AI strategies, and high expectations for product performance, security, reliability, and scale. Success in this role will be measured by revenue growth, depth of customer adoption,

AWSGitRestMachine Learning
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The AI Deployment Management (ADM) team enables organizations to turn OpenAI products into real, sustained impact through world-class services execution. Our mission is to help customers successfully adopt and operationalize AI across their organizations. We partner with enterprises to translate the potential of OpenAI’s technology into durable capability - through structured training, technical enablement, and services. By helping customers move from experimentation to production, the ADM team accelerates time-to-value, deepens product adoption, and helps make OpenAI indispensable to how organizations work. About the Role The AI Deployment Manager role is a specialist post-sales enablement role focused on delivering high-impact enablement and adoption services across OpenAI’s product suite. This role is responsible for designing and delivering enablement experiences that support a repeatable adoption framework, driving sustained activation, expanding breadth and depth of usage, and measurable business value across OpenAI’s product suite, including ChatGPT Enterprise and Agents. This role blends strong product fluency, instructional design, and customer advisory. You will lead live workshops, deliver services, and design adoption interventions for audiences ranging from everyday business users to technical practitioners and executive leaders, helping customers understand not just what OpenAI’s products can do, but how to apply them effectively in real world workflows. Success in this role means accelerating customer confidence, increasing product adoption, supporting successful launches of new product capabilities, and helping customers translate product features into tangible outcomes across teams and business functions. You will own outcomes related to activation and sustained usage by shaping how enablement drives measurable customer impact. This role is based in our San Francisco office. We use a hybrid work model of 3 days in the office per week

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Role The AI Deployment Manager (ADM) - Pilots is a customer-facing role responsible for leading structured, time-bound enterprise AI pilots from initial scoping through final executive readout. This role is focused on helping customers evaluate OpenAI’s products in real-world contexts, identify high-value use cases, and generate clear, decision-ready signals tied to business value. You will design and lead pilot engagements that drive activation, sustained usage, and measurable impact across ChatGPT Enterprise, Codex, and adjacent workflows. This includes partnering with customer stakeholders to define success criteria, guiding users from experimentation to real adoption, and translating pilot outcomes into clear recommendations that support expansion or purchase decisions. This role requires strong judgment, the ability to operate in ambiguity, and a consistent focus on connecting technical capabilities to business outcomes. You will regularly engage both executive stakeholders and working teams, adapting your approach to meet customers where they are and move them forward. In this role, you will: Own the design and execution of enterprise AI pilots, including scoping, cohort definition, and success criteria aligned to a clear commercial decision. Identify and prioritize a small set of high-impact use cases that can generate credible signal within a 30–45 day pilot. Drive activation and sustained engagement across pilot cohorts through targeted enablement, office hours, and workflow-level coaching. Monitor pilot performance and adapt in real time, diagnosing gaps in engagement, use case traction, or stakeholder alignment. Translate pilot signals into clear, executive-ready recommendations, including whether and how the customer should expand. Navigate customer constraints such as security, data access, and competing tools while maintaining pilot momentum. Partner closely with ADs, SEs, and customer stakeholders to align on scope, risks, and next steps. Ca

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the role As a Pricing Strategist focused on GTM, you will help shape pricing strategy for B2B enterprise customers across our product portfolio. Working within Finance and partnering with GTM, Sales, and Product, you will focus on pricing analytics, price performance, and the effectiveness of credit programs while contributing to broader commercial strategy. In this role, you will Build regular pricing-performance reviews, analyze discounting and concessions, and turn findings into better pricing guidance. Develop commercial strategies and pricing programs for customer segments such as education, startups, and government. Define objectives and success measures for credit programs and promotions, and evaluate their return on investment. Translate segment goals and product pricing strategy into scalable pricing frameworks and commercial structures. Partner on strategic enterprise deals to align commercial proposals with sound deal economics and pricing strategy. You might thrive in this role if you Approach problems from first principles, identify root causes, and develop clear options for stakeholders. Work effectively through ambiguity and differing perspectives. Understand how sales organizations operate and collaborate well across GTM strategy, Finance, and Product. Bring relevant experience from consulting, strategy and operations, or pricing, particularly work involving analytical teams. Experience with pricing analytics, pricing frameworks, discounting guardrails, or new-product pricing is helpful, but a strictly pricing-specific background is not required. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we mu

Artificial IntelligenceAIFinance
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the role We are hiring a Manager, Revenue Accounting focused on enterprise revenue to provide hands-on accounting support within Controllership. You will help the team meet the demands of a growing enterprise business through journal entries, reconciliations, close support, and audit support. In this role, you will Prepare and book journal entries and perform reconciliations for enterprise revenue accounting activity. Support the monthly accounting close. Respond to audit requests and prepare supporting documentation. Work with large datasets to support enterprise revenue accounting execution and investigate accounting activity. You may thrive in this role if you Bring hands-on accounting execution skills across journal entries, reconciliations, close activities, and audit support. Are comfortable working with large volumes of data. Have SQL experience or otherwise demonstrate strong data-analysis capability. Lack of SQL experience is not an automatic disqualifier. Enjoy contributing to a team supporting a rapidly growing business. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be admini

SQLArtificial IntelligenceAIAccounting
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

By applying to this role, you will be considered for Research Scientist roles across all teams at OpenAI. About the Role As a Research Scientist here, you will develop innovative machine learning techniques and advance the research agenda of the team you work on, while also collaborating with peers across the organization. We are looking for people who want to discover simple, generalizable ideas that work well even at large scale, and form part of a broader research vision that unifies the entire company. We expect you to: Have a track record of coming up with new ideas or improving upon existing ideas in machine learning, demonstrated by accomplishments such as first author publications or projects Possess the ability to own and pursue a research agenda, including choosing impactful research problems and autonomously carrying out long-running projects Be excited about OpenAI’s approach to research Nice to have: Interested in and thoughtful about the impacts of AI technology Past experience in creating high-performance implementations of deep learning algorithms About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About The Team The Data Understanding team is responsible for creating the high quality datasets and their quantized representation for OpenAI. This includes synthesizing data, building VQ representations, and processing, filtering, deduplication, quality control, and tokenization so it can be used effectively in big model training runs. About The Role We're looking to advance how OpenAI builds and understands pretraining data at scale. You'll treat data quality and curation as core research problems: developing new methods to select, combine, and transform data; creating datasets that improve model capabilities; and designing rigorous experiments to understand how data choices and interventions affect model learning and downstream behavior. You'll work closely with frontier models and web-scale data to build evidence for which approaches work and why, then translate successful research into scalable data processing pipelines We Expect You To Have a strong track record of new or improved ML ideas, through publications, projects, or applied research. Own and drive a research agenda, from choosing the right problems to carrying long-running work through to impact. Be excited by OpenAI’s empirical, collaborative approach to research. Nice To Have Thoughtfulness about AI’s impact, including privacy, provenance, and data quality. Experience building high-performance deep learning or large-scale data processing systems. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About The Team The Data Understanding team is responsible for creating the high quality datasets and their quantized representation for OpenAI. This includes synthesizing multimodal data, building VQ representations, and processing, filtering, deduplication, quality control, and tokenization so it can be used effectively in big model training runs. About The Role We’re looking to advance how OpenAI prepares, curates, synthesizes and understands multimodal data at scale. You’ll work on research and production problems like synthesizing multimodal content (images, audio, and video) and their supervisions, improving noisy data pipelines, building better quality filters, using models to automate data prep, and measuring whether changes in the dataset improve model performance. We Expect You To Have a strong track record of new or improved ML ideas, through publications, projects, or applied research. Own and drive a research agenda, from choosing the right multimodal data problems to carrying long-running work through to impact. Be excited by OpenAI’s empirical, collaborative approach to research. Nice To Have Experience with multimodal learning, audio, vision, video, synthetic data, or data-centric ML. Thoughtfulness about AI’s impact, including privacy, provenance, and data quality. Experience building high-performance deep learning or large-scale data processing systems. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%

$342K – $445K/yr

Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking a Technical Lead to lead deployment and operations for OpenAI’s Silicon & Systems team. This person will become the Directly-Responsible Individual responsible for bringing OpenAI’s custom silicon and associated systems into data center environments, ensuring successful deployment, bring-up, validation, operational readiness, and ongoing reliability at scale. This role sits at the intersection of silicon, systems, infrastructure, data center operations, and software. You will lead a team focused on taking new hardware platforms from lab validation into production data center deployment. You will be responsible for building the operational processes, technical workflows, tooling, and cross-functional alignment required to deploy and operate custom AI hardware reliably in OpenAI’s supercomputing infrastructure. The ideal candidate is both a strong leader and a deeply technical operator. You should be comfortable staying close to the technical details of hardware bring-up, fleet deployment, debugging, system validation, data center integration, and production operations. This role requires strong execution, excellent cross-functional judgment, and the ability to drive clarity in ambiguous, fast-moving environments. In this role, you will: Lead a team responsible for deployment and operations of OpenAI’s custom silicon and systems in data center environments Own the path from hardware bring-up and validation through production deployment, operati

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional experiences to developers, consumers, and businesses alike. The ChatGPT infrastructure team is responsible for ensuring that our products can serve rapidly growing demand with the performance, reliability, and quality our users expect. This work sits at the intersection of product demand, model deployment, inference, research, fleet, and capacity. The team translates changing product and model needs into clear capacity decisions and safe, scalable launches. About the Role We are seeking a Technical Program Manager to lead the operating system for Chat capacity and model deployment. You will connect demand forecasting and capacity allocation with model readiness, rollout planning, launch coordination, and post-deployment learning. You will also own mode deployment beyond capacity by working with cross functional teams across research, post-training, inference and product to own mainline model deployment. You will bring structure to constrained-capacity decisions, improve the tooling and mechanisms teams use to prioritize demand, and help new models reach users safely and efficiently. Success requires technical depth, sound judgment under ambiguity, and crisp execution across product, research, infrastructure, and operations teams. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own cross-functional programs for Chat capacity forecasting, allocation, headroom planning, and constrained-capacity operations. Build durable intake, prioritization, and decision mechanisms that connect product demand and model requirements to available serving capacity. Partner

AWSRestAIRust
🔔

Get new ai deployment engineer jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime