Jobs in United States

Codex Deployment Engineer in San Francisco

181 active opportunities · Updated October 2026

Explore current codex deployment engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team OpenAI for Financial Services is part of OpenAI's Verticals organization, which focuses on accelerating the economy and knowledge work. We build AI products for financial institutions and the professionals who power them, from investment bankers and research analysts to investors and other financial services teams. We combine OpenAI's models, financial data, and enterprise knowledge to help professionals research companies, analyze markets, and produce high-quality work in the tools they use every day. We're a small, entrepreneurial team working closely with customers and partners across research, product, design, engineering, and go-to-market to bring new capabilities from idea to production. About the Role We're looking for full-stack engineers to build new, AI-native products on top of ChatGPT Work and Codex. This is zero-to-one work: you'll help define how financial professionals research, analyze, and make decisions alongside AI. You'll own the experience across the stack, work directly with customers to understand their workflows, and collaborate across OpenAI to turn new model capabilities into products that professionals can trust. Your work will shape how some of the world's largest financial institutions adopt AI and how financial knowledge work gets done. In this role, you will: Build a new financial services app within ChatGPT Work and Codex, creating intuitive AI-native experiences for company research, financial analysis, document review, and professional work products. Develop new product experiences around enterprise memory that learn from an organization's knowledge, workflows, and context, and adapt to how its teams work. Develop the APIs, services, and integrations required to connect user experiences with financial data providers, enterprise systems, and OpenAI's models. Work closely with product and design to turn ambiguous customer problems into polished, useful, and reliable products. Work directly with financial institutions to

TypeScriptPythonReactAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team OpenAI for Financial Services is part of OpenAI's Verticals organization, which focuses on accelerating the economy and knowledge work. We build AI products for financial institutions and the professionals who power them, from investment bankers and research analysts to investors and other financial services teams. We combine OpenAI's models, financial data, and enterprise knowledge to help professionals research companies, analyze markets, and produce high-quality work in the tools they use every day. We're a small, entrepreneurial team working closely with customers and partners across research, product, design, engineering, and go-to-market to bring new capabilities from idea to production. About the Role We're looking for backend engineers to build the systems that make advanced AI useful, reliable, and trustworthy in financial services. You'll build the data systems, agentic workflows, and enterprise integrations behind our products. You'll also help bring them into production at some of the world's largest financial institutions. This is a product-minded engineering role with significant ownership and zero-to-one building. You'll shape new products from the ground up, work directly with customers to understand their workflows, and collaborate across OpenAI to turn new model capabilities into products that professionals can trust with high-stakes work. In this role, you will: Design and build backend systems that power AI-native financial workflows across ChatGPT Work and Codex. Build infrastructure to ingest, index, retrieve, and serve financial data, company filings, market information, and firm-specific knowledge at scale. Develop integrations with financial data providers, enterprise knowledge systems, and customer environments, including the authentication, authorization, and entitlements required to use them securely. Build the systems that let models and agents use the right tools and data, preserve source provenance, and produce accurate,

TypeScriptPythonAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

AI Systems Engineer - Codex Core Agents About The Team The Codex Core Agents team builds the agent harness that turns model capability into real-world action. We own the systems around the model: prompting and interpreting model outputs, executing actions safely in real environments, and feeding production experience back into better models and better agent behavior. This team sits close to research and works across the stack: harness, model interaction, inference, sandboxed execution, orchestration, evals, production reliability, and the performance envelope around tokens, latency, cost, capacity, and quality. The harness is open source and increasingly part of how models are trained and evaluated, making this one of the highest-leverage layers in Codex. About The Role We’re looking for engineers to build the AI systems that make Codex agents dependable in production. The ideal candidate is an agent-systems builder: hands-on across low-level systems and ML workflows, able to debug Codex behavior end to end across the harness, model behavior, inference/runtime stack, GPU fleet, and product surface. You’ll work with research, infrastructure, and product to design agent harness capabilities, run experiments and ablations across the model + system prompt + harness stack, build frameworks for assessing production agent performance, and turn messy failures into durable improvements. What You’ll Do Design and build the core agent harness and execution loop that lets Codex agents interpret model outputs, use tools, execute code, and complete long-horizon tasks safely. Build sandboxing, isolation, orchestration, state, and workflow infrastructure for agents operating in real development environments. Develop evaluation, experimentation, and debugging systems that distinguish harness issues, model behavior, inference/runtime issues, and product failures. Run ablations across prompts, model-facing interfaces, context construction, tool-use strategies, and harness behavior to

PythonAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team Codex is OpenAI’s first-party developer product focused on agentic software engineering. We’re building tools that help engineers design, write, test, and ship code faster—safely and at scale. We partner tightly with research and product to translate model advances into tangible developer productivity. About the Role As a Data Scientist on Codex, you will measure and accelerate product-market fit for AI developer tools. You’ll define what “developer productivity” means for our product, run experiments on new coding models and UX, and pinpoint where the model helps or hurts across languages and tasks. Your insights will directly shape how an entire industry builds software. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will Embed with the Codex product team to discover opportunities that improve developer outcomes and growth Design and interpret A/B tests and staged rollouts of new coding models and product features Define and operationalize metrics such as suggestion acceptance, edit distance, compile/test pass rates, task completion, latency, and session productivity Build dashboards and analyses that help the team self-serve answers to product questions (by language, framework, repo size, task type) Diagnose failure modes and partner with Research on targeted improvements (model quality signals, user feedback, evals) You might thrive in this role if you have 5+ years in a quantitative role at a developer-facing or high-growth product Fluency in SQL and Python; comfort with experiment design and causal inference Experience defining product metrics tied to user value Ability to communicate clearly with PM, Eng, and Design—and to influence product direction You could be an especially great fit if you have Strong programming background; ability to prototype, run simulations, and reason about code quality Familiarity with IDE/extensi

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team The Codex Research team creates the frontier agents OpenAI ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that can operate computers, collaborate with people and other agents, and expand what people and organizations can imagine, attempt, and achieve. We define what the next generation of agents should be able to do, build the training signal that teaches those abilities, and run the experiments that make them real. Our work spans coding, tool use, computer use, multi-agent coordination, long-horizon execution, factuality, instruction following, calibrated reasoning, and taste. Our team is where new model capabilities get made. We build the data, environments, graders, training methods, and feedback loops that shape what OpenAI's next agents can do, then carry those capabilities through major training runs and into the products people use. About the Role As a member of the Codex Research team, you will improve the capabilities, reliability, and product fit of OpenAI's agentic models. You might own a research direction, build the infrastructure that makes large training runs faster and more trustworthy, create evals that reveal where models fail, or drive a capability from an idea through experimentation, integration, and launch. This role is intentionally broad. The strongest candidates are not defined by one method or subfield; they are people who can take an ambiguous capability problem and make progress across research, engineering, data, evals, and product. You should be excited to work on models that act in the world: writing and debugging code, using tools, calling functions, operating computers, collaborating with other agents, and completing valuable work on behalf of users. You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measu

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for engineers to build the infrastructure that powers Codex agents in production. This role focuses on the systems that let models safely execute code, interact with tools, complete long-running tasks, and operate reliably and efficiently at scale. You’ll design and operate the infrastructure behind sandboxed execution, orchestration, stateful workflows, app-server and SDK boundaries, and model rollouts. You’ll work at the intersection of distributed systems, developer tooling, and AI, building primitives that make Codex faster, safer, more reliable, and easier for the rest of the organization to build on. What You’ll Do Design and build execution environments for AI agents, including sandboxing, isolation, and reproducibility. Develop systems for agent orchestration across multi-step, tool-using workflows. Build infrastructure for running, testing, and debugging code generated by models. Create state and memory systems that allow agents to persist context across long-running tasks. Optimize tokens, latency, reliability, and cost across Codex’s production fleet. Support model rollouts, capacity planning, and the core tradeoffs between quality, speed, and economics to manage a fleet of frontier agents at scale. Build shared platform capabilities that unblock product teams, partner teams, and open source Codex. Yo

AWSCI/CDRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for applied AI engineers to help bring Codex agents from impressive demos to dependable tools. This role is about improving agent performance on real software engineering tasks and closing the gap between research capability and real-world usefulness. You’ll work closely with research, infrastructure, and product to ensure agents are not just powerful, but useful, steerable, and reliable in practice. The job is not only to improve model behavior in isolation, but to turn those improvements into measurable gains in solve rate, usefulness, and economic value for users. What You’ll Do Design and iterate on agent behaviors across real-world coding tasks and long-horizon workflows. Work closely with research to develop and run evals to measure agent performance, regressions, failure modes, and edge cases. Improve performance through prompting, tool-use strategies, context construction, and model-facing experimentation. Analyze failures in production and systematically improve robustness and reliability. Build feedback loops and data systems that get better real-task data into evaluation and research. Work with product teams to shape user-facing agent experiences and the interfaces the agent depends on. Help define what “good” looks like for agents completing complex tasks end-to-end. You Might Be a Good Fit If You Ha

PythonAWSRestMachine Learning
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s mission is to ensure that general-purpose artificial intelligence benefits all of humanity. The Codex team plays a key role in this by changing how software gets built and who can build it. The team is a fast-moving group within OpenAI that works across research, engineering, product, and design to build an AI software engineer. One that you can pair with, delegate to, or even ask to take on future tasks proactively. Design plays a critical role here. To succeed in our mission, it’s crucial that we make coding agents intuitive and accessible. We’re hiring a product designer to create products that are easy to use, beautiful, and push the boundaries of what’s possible. As an early team member, you’ll have a huge part in shaping our product direction and design culture. About the Role As a product designer on Codex, you’ll collaborate closely with a small team of product designers and cross-functional partners to shape the future of coding. You’ll partner closely with world-class engineers and researchers to bring cutting-edge capabilities into the hands of developers. Much of the work is 0–1, requiring you to balance a high-bar for craft, fast iteration, and strong product intuition. You’ll work on the end-to-end design of new features and improvements across a range of Codex products. You’ll help chart our course into the future as we evolve our technology, product, and design paradigms. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Contribute to the overall design and product direction of OpenAI products. Design and ship high-quality products and improvements, from early concepts to high-fidelity prototypes and visuals. Partner closely with engineering, product management, research, and design peers to define both long-term strategy and short-term tactics. Engage in user research to better understand our users (sof

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.2%
Quick readStrong listing-quality and freshness signals

About the Team The Applied AI Engineering team is responsible for helping customers turn frontier AI capabilities into real products, workflows, and business impact. We act as trusted technical partners across solution design, architecture, implementation, evaluation, and adoption, working alongside customers to build and scale effective AI applications with OpenAI’s technologies. The Codex Applied AI Engineering team focuses on helping organizations transform how software is built with AI. We partner directly with engineering teams and technical leaders to integrate Codex into their software development lifecycle — from identifying high-impact use cases and designing AI-enabled workflows to implementation, evaluation, and scaled adoption. Our work helps ensure AI-powered software development is effective, reliable, secure, and deeply integrated into how engineering organizations operate. About the Role We are seeking an experienced technical leader to join as Manager, Applied AI Engineering (Codex) , leading a team of Applied AI Engineers responsible for driving successful Codex adoption across strategic customers. Your team will work hands-on with customer engineering organizations to design and build AI-enabled development workflows, solve complex implementation challenges, and establish scalable patterns for AI-powered software development. As a manager, you will shape how these technical engagements operate at scale — setting strategy, coaching engineers, determining where the team can have the greatest impact, and ensuring consistently strong execution across customers. You will serve as both a people leader and senior technical advisor, partnering closely with Sales, Product, Research, and Engineering to translate customer needs and real-world usage into better technical approaches, reusable patterns, and product insights. Success in this role will be measured by meaningful and sustained Codex adoption, successful customer outcomes, and the creation of repeat

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team : We build core first-party app experiences in ChatGPT and Codex, define the primitives for high-quality app third-party app experiences, and collaborate with best-in-class partners across consumer and enterprise categories to bring delightful experiences to our customers. About the Role: We are hiring a Product Manager to shape and scale the app ecosystem across ChatGPT and Codex. This person will own both 1P product experiences and partner-led launches, translating user needs, model capabilities, platform constraints, developer and partner requirements, and enterprise controls into products that feel delightful, reliable, and safe. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Develop the strategy and roadmap for ChatGPT and Codex app ecosystem experiences across consumer and enterprise use cases. Build and ship high-quality first-party app experiences that demonstrate the best of what apps can do inside ChatGPT and Codex. Collaborate with best-in-class partners to create app experiences that solve real user and business workflows. Define the product foundations and quality standards needed for a trusted app ecosystem. Lead cross-functional execution across engineering, design, research, partnerships, GTM, legal, privacy, security, support, and data/evals. Use customer, user, partner, and model-behavior insights to prioritize the roadmap and improve post-launch performance. You might thrive in this role if you: Have built and scaled consumer or enterprise product experiences with strong product taste and measurable user impact. Have built app platforms, marketplaces, partner ecosystems. Are technically fluent enough to reason about MCPs, APIs, SDKs, and model/product constraints. Can move fluidly between strategy, product, partner judgment, and operational execution. Communicate crisply in writing and bring clarity to ambiguous, fast-moving product areas. Care deeply about user trust, safe

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team The Coding team is reimagining how software is built in the AI era. We build tools and workflows that help software engineers work faster, tackle more ambitious projects, and spend less time on repetitive tasks. AI has already transformed how code is written, but software engineering extends far beyond coding. Our mission is to apply AI across the entire software development lifecycle (SDLC) — from design and implementation to code review, testing, debugging, issue remediation, maintenance, documentation, and user support. The team is also responsible for developer-facing Codex experiences including the Codex IDE Extension and the terminal interface, which are used daily by developers ranging from individual open-source contributors to some of the world’s largest engineering organizations. The team also works closely with the open-source software community, building tools that help maintainers and contributors manage increasingly complex projects. We believe AI can make open-source development more sustainable by reducing the operational burden of reviewing contributions, triaging issues, maintaining quality, and supporting growing communities. By building the future of software development, we're helping advance OpenAI's mission of ensuring that the benefits of AI reach people around the world. About the Role We’re hiring a Full Stack Software Engineer to help invent the next generation of AI-powered software development workflows. “Full stack” in this role means much more than traditional frontend and backend development. You'll own complete product experiences, spanning user interfaces, workflow orchestration, agent and prompt design, backend systems, and cloud infrastructure. This is a highly product-oriented role. You'll work directly on the workflows developers use every day, identifying bottlenecks and rethinking how software gets built in a world where AI agents are active participants in the development process. The features you ship will inf

TypeScriptAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team The Codex Web Layer team provides the web-based systems and user experiences for Codex across the entire stack, from the Electron-like application framework that powers the application, to the user-facing in-app browser. About the Role In this role, you will be responsible for designing and implementing infrastructure and features end-to-end for the Codex desktop client application. You will help define what it means to be a hybrid agentic/interactive web browser. The team embodies “full stack” development from the lowest-level OS integration to the highest-level interaction design. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role you will: Partner closely with product and design to conceive, design, and build features for Codex web browsing features on macOS and Windows. This role will focus mostly on the backend C++ layer and Chromium, but many features cross the full stack including some TypeScript. Partner with the wider Codex team to deliver a high-performance, stable, and secure application platform for client development. This includes API design and implementation (mostly in C++) and the infrastructure that supports deploying it (in Python, TypeScript, and agentic skills). Work with a small, experienced team of engineers on this critical and rapidly growing product. You might thrive in this role if you: Have significant experience building technically complex features end-to-end. Are a strong C++ developer, especially with experience in browser environments like Chromium and Electron. Since this role is more backend focused, general knowledge of web development and TypeScript is helpful but not required. Thrive in a fast-paced, ambiguous environment. Communicate clearly and concisely across many different roles in the organization. Are self-directed, identifying important work and executing it end-to-end. About OpenAI OpenAI is an AI

TypeScriptPythonAWSRest
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.2%
Quick readStrong listing-quality and freshness signals

About the Team The Plugin Ecosystem team builds the platform and product experiences that let people extend ChatGPT and Codex. We work on plugins, skills, connectors, interactive apps, and open standards like the Model Context Protocol (MCP). We make plugins easy to discover, install, and use, ensure they’re invoked at the right time, and help people find new ways to get value from them. We want anyone to be able to turn a useful workflow into a plugin, share it, and have other people use it. A plugin can package instructions and skills with connections to the tools and data it needs. Our work spans creation and publishing, reliable execution across our products, clear permissions and approvals, and the controls admins need to bring plugins to their organizations. We work closely with research to improve plugin quality as models evolve. About the Role We’re looking for product-minded engineers to build the systems behind plugins and improve how models use them. Depending on your focus, you may scale generalist infrastructure and identity-related integrations across products, or improve plugin quality at the intersection of backend engineering and applied AI or work on the product experience itself to drive plugin usage. You’ll work across teams and own problems from diagnosis and design through implementation and release. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design and ship APIs, SDKs, and services that developers use to extend ChatGPT and Codex. Build intuitive experiences that help users discover, install, and use plugins to get more done. Make plugins easier to create, test, publish, update, and share. Improve when and how models use plugins, from choosing the right plugin to completing a task. Work with Research to diagnose failures and measure improvements as models evolve. Improve plugin reliability and interaction quality acros

Artificial IntelligenceAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.2%

$266K – $295K/yr

Quick readStrong listing-quality and freshness signals

About the Role As a member of the Platform Partnerships team, you will lead some of the most important integrations connecting the world into OpenAI’s flagship products, including Codex and ChatGPT. This role has enormous impact, as OpenAI users and their agents increasingly interact with third-party services. You will set partnership strategy, craft mutually beneficial commercial and technical agreements, and work closely with product and engineering teams to bring new platform capabilities to market. You will help decide what gets built, which partners to prioritize and how a new generation of AI-native services come to life. This role sits at the intersection of product, engineering, partnerships, and ecosystem strategy. The right person is someone who can run deals end-to-end, navigate ambiguity, push product thinking when needed, and move fast with partners and internal teams. In this role, you will Lead full lifecycle partnerships, from identifying and prioritizing partners to negotiating terms, closing complex agreements, and driving partner success. Work directly with Product and Engineering to shape how connectors, agents, and third-party workflows function inside ChatGPT and Codex. Develop ecosystem relationships to drive deeper engagement across distribution, co-marketing, and product innovation. Translate partner commercial needs, technical constraints, and market dynamics into program and product recommendations, prioritization calls, and go-to-market plans Bring structure to ambiguous opportunities and help the team decide where to go deep, move fast, or say no. Drive aggressive timelines by clearing blockers and aligning internal and external workstreams. You might thrive in this role if you 10+ years of experience in product partnerships, product management, or developer relations. Demonstrated success structuring and negotiating complex commercial and technical deals, with strong negotiation skills with a focus on creative solutions. Excellent produ

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -79.2%

About the Team About the Team: Our team brings OpenAI’s most capable technology to the world through our products. We've released ChatGPT, GPT5.6, and Codex. We empower consumers and developers alike to use and access our state-of-the-art AI models, allowing them to do things that they’ve never been able to before. Across all product lines, we ensure that these powerful tools are used responsibly. This is a key part of OpenAI’s path towards safely deploying broadly beneficial Artificial General Intelligence (AGI). Safety is more important to us than unfettered growth. The Youth Well-Being team focuses on helping young people have safe, age-appropriate, and positive experiences with AI. The team builds features such as age prediction, parental controls, and access to crisis support resources when people need help. Their work brings together product, policy, safety, research, and operations to keep improving how ChatGPT supports teens and families. About the Role We’re looking for an experienced Product Manager for OpenAI’s Youth product efforts. This role will shape how teenagers benefit from AI across OpenAI’s products. This is a high-impact role with the opportunity to improve daily life for millions of people. You’ll work closely with engineering, design, research, policy, legal, communications, finance, and operations teams. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Create age specific product experiences for <18 users Identify user needs and turn them into high-impact product opportunities, owning products end-to-end, from concept through launch and iteration. Build experiences that balance usefulness, simplicity, trust, safety, and accessibility. Define success metrics and use data and research to guide roadmap decisions. Partner cross-functionally to execute quickly and effectively. You might thrive in this role if you: Have 10+ years of product management experience leading large-scale consum

Artificial IntelligenceAIFinance
🔔

Get new codex deployment engineer jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime