About the Team The ChatGPT Search Product Infrastructure team builds the foundational systems that power search experiences across ChatGPT. We develop the product infrastructure that connects models with search systems and other sources of real-time information, enabling ChatGPT to deliver timely, relevant, and trustworthy answers to users around the world. Our work sits at the intersection of product engineering, AI, and large-scale infrastructure. We build shared platforms and abstractions that enable product teams to independently develop, evaluate, and launch new search-powered experiences. These platforms provide the guardrails, testing capabilities, observability, and rollout controls needed to prevent reliability, scalability, quality, and latency regressions while supporting rapid product iteration. The team partners closely with: Post-Training on model launches, experimentation, and prompt optimization Search product verticals on new user experiences Inference on GPU efficiencies Indexing and Retrieval on the systems that identify and deliver relevant information Capacity/Fleet team to ensure optimal regionalized provisioning of GPUs and CPUs About the Role We are looking for an Engineering Manager to lead the team responsible for ChatGPT’s Search Product Infrastructure. You will set the technical and organizational direction for the systems that bring search capabilities into ChatGPT. You will guide architectural decisions across search orchestration, model and prompt integration, serving infrastructure, experimentation, observability, evaluation, and product integrations. You will balance immediate launch and product needs with the long-term reliability, scalability, latency, and maintainability of the platform. A central responsibility of this role is creating leverage for Search product verticals. You will lead the development of extensible platforms that allow those teams to independently build, test, and launch features without requiring ongoing invol
Jobs in United States
Applied Ai Engineer in United States
441 active opportunities · Updated October 2026
Showing
15 jobs
Explore current applied ai engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The Core Services organization builds and runs the mission-critical online services that product teams rely on in production. We own foundational distributed systems and platform capabilities that enable reliable execution, high-performance services, and large-scale file/data needs across our products. This team is distinct from developer infrastructure and data infrastructure—our focus is production service foundations and core runtime services. About the Role We’re hiring an Engineering Manager, Core Services to help lead teams responsible for highly reliable, high-scale distributed systems that sit on the critical path for OpenAI products. Your team will own foundational production systems that OpenAI’s product engineering teams build on. You’ll collaborate closely with product and infrastructure partners to ship reliable services quickly, and help scale systems and teams as OpenAI grows. You’ll partner closely with senior engineering leaders to scale the org, mature operations, and drive major platform initiatives. This role requires strong technical ability. You’ll be responsible for: Managing and growing a high-performing team of infrastructure engineers. Leading teams building and operating large, critical production platforms, including cluster reliability, scaling, and rollout safety. Building and operating mission-critical distributed systems with strong operational rigor (SLOs, incident response, capacity planning, reliability). Setting technical direction for platform foundations such as workflow/orchestration capabilities, large-scale file/blob/storage services, and core service foundations. Partnering with a broad set of stakeholders, including product engineering, adjacent infrastructure teams, and (where relevant) finance/cost partners. Coaching, mentoring, and developing engineers and emerging leaders. You might thrive in this role if you: Have significant experience leading teams that run mission-critical infrastructure in production
About the Team The Legal team is building the next generation of AI-powered products and experiences for the legal industry. We are exploring how advanced AI systems can transform legal workflows, improve access to information, and enable legal professionals and organizations to work more effectively. As a founding member of the Legal engineering team, you will help define the technical foundation for this new product area from the earliest stages. You’ll operate at the intersection of AI, product, and real-world legal workflows—identifying opportunities, building prototypes, and turning emerging ideas into scalable products that can create meaningful impact. We operate with a startup-like mindset inside OpenAI: small teams, rapid iteration cycles, and a willingness to explore bold ideas, learn quickly, and adapt based on user feedback. Our goal is to build products that meaningfully improve how legal professionals work while leveraging OpenAI’s cutting-edge models and infrastructure. About the Role As a Founding Full-Stack Software Engineer on the Legal team, you will help imagine, build, and scale new AI-powered products for the legal industry. You’ll work across the stack to design intuitive user experiences, build robust backend systems, and create the foundations for products used by legal professionals and organizations around the world. You’ll have significant ownership from the earliest stages—working closely with product, design, research, and go-to-market partners to understand customer needs, shape product direction, and deliver high-impact solutions. This includes rapidly prototyping new concepts, building production-quality applications on top of OpenAI’s platforms, and developing new technical approaches when existing systems are not sufficient. We’re looking for engineers who thrive in ambiguity, have strong product instincts, and enjoy building from 0→1. You should be comfortable moving quickly, making thoughtful technical decisions, and taking owner
About the Team The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions. Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After joining OpenAI, the team began its next chapter: bringing deep product expertise, customer intuition, and mature platform infrastructure into the product development system used by every OpenAI team. Today, teams across ChatGPT, Codex, model measurement, consumer monetization, business subscriptions, developer products, and shared infrastructure rely on Statsig to safely introduce new capabilities, measure impact, and roll changes forward or back with confidence. We are at a defining moment as adoption accelerates and the platform becomes a company-wide standard. About the Role We are looking for an Engineering Manager, Statsig Product to lead the product engineering organization responsible for Statsig’s post-acquisition journey at OpenAI. You will define how experimentation, rollout, configuration, and analytics become a simple, reliable, and trusted part of how every OpenAI product team ships. You will set strategy across multiple product and platform workstreams, build the organization and leadership structure needed for the next phase, and establish the operating model for a platform that serves teams across the company. The right leader can operate across product strategy, technical architecture, organizational design, developer experience, reliability, and executive alignment. You will help preserve what made Statsig strong while integrating it deeply into how OpenAI launches, measures, learns, and makes product decisions. In this role, you will: Build, lead,
About the Team The ChatGPT Model Flywheel team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement. Team Focus Areas Model Experimentation: Enable rapid, safe model validation for ChatGPT and Codex products through experiment automation and lifecycle management. Model Deployment: Ensure safe, scalable deployment of model capabilities with robust rollout and operational tooling. Automate capacity management and incorporate platform-wide health monitors. Model Measurement: Build comprehensive evaluation and measurement systems for model quality, from user signals to launch scorecards. Improve end-to-end feedback loops for continual model improvement. Key Partnerships Collaborate cross-functionally with teams including Model Measurement DS, Research, Codex, Fleet, Inference, and API. In this role, you will: Elevate and consolidate ChatGPT’s harness, context management, and system prompt frameworks. Drive expansion and improvement of multi-tier model experiences. Support and scale self-serve experiment capabilities and automated guardrails. Lead model rollout automation, capacity management, and health monitoring. Shape end-to-end measurement systems (evals, grader signals, user feedback, etc.). You might thrive in this role if you have: Proven experience leading engineering teams in complex, cross-functional environments. Demonstrated success shipping production systems at scale (ideally for AI or large backend services). Deep understanding of model-driven product development, deployment lifecycle, and measurement tooling. Excellent communication and collaboration skills—experience interfacing directly with engineering, research, and product stakeholders. Prior involvement with large language models, distributed infrastructure, or experimentation platforms is a plus. Why Work With Us Tackle highly impactful technical challenges at the cutting edg
About the role We’re looking for an engineering manager to lead a team building software systems that detect and prevent harmful misuse of frontier AI models—before incidents occur. This is a builder’s role: you’ll lead engineers shipping production services, detection pipelines, and mitigation mechanisms that protect frontier model integrity and reduce high-severity misuse risk. While this work intersects with frontier model development, security and risk, we’re explicitly seeking someone with a software engineering foundation who is comfortable building reliable systems that can operate at billions of users scale. In this role you will: Lead a team of software engineers building detection + mitigation systems for frontier model misuse, with an emphasis on model IP protection / distillation detection and emerging risk surfaces from autonomous agents. Set the technical roadmap and execution strategy: prioritize, design, ship, iterate, measure impact. Build production systems: services, pipelines, tooling, instrumentation, and automation that scale with frontier model usage. Partner deeply with Research and Product to translate evolving model capabilities into concrete tests, signals, and mitigations that can be deployed at scale. Drive strong engineering fundamentals: architecture, reliability, monitoring, performance, and operational excellence. Hire and grow an exceptional team across backend, data systems, and applied ML engineering domains as needed. Anticipate what breaks at scale as agentic workflows become more capable. You might thrive in this role if you: Experience building systems in adversarial, fast-evolving environments Are comfortable with ambiguity and novelty Have experience adjacent to security (e.g., abuse prevention, fraud, integrity, platform defense, auth/identity, malware/spam, adversarial environments) Communicate clearly and build trust quickly with senior stakeholders—pragmatic, collaborative, and calm under scrutiny. Significant experience
About the Team The Online Data team builds and operates the core online database and indexing services for OpenAI’s production AI applications, including supporting the explosive growth of ChatGPT, the #1 AI app in the world, and Codex, the fastest growing agentic development toolset in the world. Our mission is to ensure the reliability, correctness, and scalability of our online data stack and to curate a comprehensive portfolio of services that matches the relentless ambition of OpenAI, enabling our product and research teams to build 0-100 without getting bogged down in the minutiae of multi-region, multi-cloud, exabyte-scale data infrastructure. About the Role We are seeking an Engineering Manager to lead our Online Data Systems team, responsible for our in-house database and indexing technology. This role is about shepherding a team of world-class engineers tasked with building and operating hyperscale data storage and retrieval technology. You’ll be overseeing the delivery of extremely challenging engineering work in areas like distributed query execution, multi-region federation, self-orchestrating and self-healing services, low-level performance optimization, and more. There are few companies in the world building this kind of technology in-house at this scale where you’ll still be getting in on the ground floor. Instead of being a cog in the machine spending months chasing small optimizations, you’ll play a major part of shaping our future. In this role, you will: Build, lead, and grow high-performing infrastructure engineering teams. Drive the evolution of OpenAI’s in-house online data technologies, our core, hyper-scale database systems, indexing technologies, and vector search. Anchor delivery around measurable reliability goals (SLOs, etc) to ensure system performance and resiliency is above reproach. Champion pragmatic use of agent technology to amplify execution velocity. Reduce operational toil and incident frequency through better abstractions, gua
About the Team The Artifacts team is building the AI-native creation layer for documents, spreadsheets, slide decks, dashboards, reports, analyses, and new forms of interactive work products. We are rethinking what creation looks like when models can move from an ambiguous user goal to a polished, editable artifact with strong structure, taste, correctness, and speed. This is a high-agency team working across product, infrastructure, and research. We partner closely with model training teams to shape how frontier models create artifacts, and with ChatGPT product teams to turn those capabilities into experiences that millions of people can use. The work spans full-stack product engineering, model integration, rendering and editing systems, collaboration, storage, evaluation loops, and production reliability. Our ambition is to build the premier product experience for AI-generated artifacts: starting with familiar work products like slides, sheets, and docs, then expanding into new artifact types that are only possible in an AI-native world. About the Role As Engineering Manager, Artifacts, you will lead and grow the engineering team responsible for building this product and technical foundation. You will manage a team of full-stack and infrastructure-oriented engineers, set technical direction, and stay hands-on enough to shape architecture and debug hard problems. This role sits at the intersection of product engineering, research, and infrastructure. You will partner with researchers on how models are trained and evaluated for artifact creation, with product and design on the user experience. This is a strong fit for a technical manager who wants to build and ship, not only coordinate. The team has a fast trajectory, so you will help define both the product surface and the team that builds it. In this role, you will: Lead, manage, and grow a team building AI-native artifact creation experiences across documents, spreadsheets, slide decks, and emerging artifact form
About the Team The Integrity team at OpenAI is dedicated to ensuring that our cutting-edge technology is not only revolutionary, but also secure from a myriad of adversarial threats. We strive to maintain the integrity of our platforms as they scale. The Integrity team is at the front lines of defending against misuse in all its forms: content abuse, scaled attacks, and other actions that could undermine the user experience or harm our operational stability. About the Role As a Machine Learning Engineer in OpenAI's Integrity team, you will have the opportunity to work with some of the brightest minds in AI. You’ll work on state-of-the-art models and classifiers, experiment with new architecture and approaches, and push forward our abilities in content and user understanding. You’ll help turn research breakthroughs into tangible solutions that improve the trust and safety of our platform. If you're excited about training LLMs and building ML models, this role is your chance to make a significant mark. In this role, you will: Innovate and Deploy: Design and deploy advanced machine learning models that solve real-world problems. Bring OpenAI's research from concept to implementation, creating AI-driven applications with a direct impact. Collaborate with the Best: Work closely with researchers, software engineers, and product managers to understand complex business challenges and deliver AI-powered solutions. Be part of a dynamic team where ideas flow freely and creativity thrives. Optimize and Scale: Implement scalable data pipelines, optimize models for performance and accuracy, and ensure they are production-ready. Contribute to projects that require cutting-edge technology and innovative approaches. Learn and Lead: Stay ahead of the curve by engaging with the latest developments in machine learning and AI. Take part in code reviews, share knowledge, and lead by example to maintain high-quality engineering practices. Make a Difference: Monitor and maintain deployed m
About the Team The Premium team owns some of the highest-leverage customer-facing levers in ChatGPT’s consumer revenue business, spanning the paid customer journey: helping users understand the value of paid plans, convert with confidence, and continue finding lasting value in their subscription. Our work is highly cross-functional, partnering with Product, Data Science, Design, FinEng, Finance, Legal, Support, and Marketing to improve free-to-paid conversion, renewal, customer lifetime value, and revenue while keeping the experience trustworthy, scalable, and low-friction. In This Role, You Will: Lead and scale an engineering team responsible for some of ChatGPT’s most important subscription and monetization experiences. Own the technical execution for Premium customer experiences across plan merchandising, paywalls, upgrade flows, checkout UX, plan management, renewals, downgrades, and cancellation. Partner with Product and Data Science to run high-quality experiments across upgrade, trial, renewal, downgrade, and cancellation flows. Improve key subscription metrics including conversion, renewal, churn, ARPU, and lifetime value. Build reliable customer-facing Premium experiences for purchase, plan management, renewal, downgrade, cancellation, and access-related states at scale. Partner closely with FinEng and other platform teams to evolve the billing, payments, and entitlement capabilities that power Premium experiences. Collaborate closely with Product, Design, Data Science, Finance, Legal, Support, and Marketing on monetization strategy and execution. You Might Thrive in This Role If You: Have 5+ years of engineering management experience, Have strong technical expertise in backend, frontend, or full-stack development, with experience building growth-oriented features. Have a track record of improving conversion, retention, or monetization through experimentation and data-driven product engineering. Are experienced with subscription products, plan merchandising
About the Team The Applied team brings OpenAI’s technology to the world through products used by hundreds of millions of people and by developers and businesses building on our APIs. We work across research, engineering, product, policy, safety, and operations to deploy frontier AI systems responsibly and safely. The Trust & Safety Data Engineering team builds the data foundations that help OpenAI understand, detect, investigate, and mitigate abuse and safety risks across our products. We partner with Integrity, Investigations, Safety Systems, Product Policy, Privacy, Data Science, Engineering, and Data Platform to create reliable, privacy-safe datasets and pipelines for fraud and abuse detection, enforcement workflows, safety measurement, ML feature generation, launch readiness, and transparency reporting. About the Role We are hiring a Technical Lead Manager to lead and grow the Trust & Safety Data Engineering team. This is a hands-on leadership role for someone who can set strategy, shape data architecture, align senior stakeholders, coach engineers, and drive execution on high-impact data systems. You will help turn fragmented launch and incident support into durable, reusable, privacy-safe data foundations that Trust & Safety teams can rely on. The systems your team builds will help OpenAI detect risk, investigate abuse, power operational workflows, develop and evaluate safety models, measure interventions, support product launches, and report accurately on platform integrity. In This Role, You Will Lead and grow a high-performing Trust & Safety Data Engineering team. Define the roadmap and technical strategy for Trust & Safety data systems. Build canonical, privacy-safe datasets and pipelines for abuse detection, fraud detection, risk signals, enforcement, scaled review, transparency reporting, and safety monitoring. Create reusable foundations for Trust & Safety model development, including features, labels, training data, backtesting,
From $109K/yr
The Agent Research and Tooling team, part of MongoDB's AI Builder Experience organization, owns the platform layer around agents: how teams author, distribute, evaluate, monitor, and improve agent skills and agent behavior. We are hiring a software engineer to build and maintain the tooling, evaluation systems, and quality gates behind MongoDB's agent skills. This is a software engineering role at the intersection of developer tooling, applied AI, and software quality. You will take loosely defined agent and tooling problems, break them into workable plans, and ship durable internal systems: command-line tools, reusable libraries, evaluation harnesses, and CI workflows. This role is open to remote work in the US or can be based out of any of our US offices. What you'll do Build and maintain agent skills and the infrastructure to validate, evaluate, publish, and maintain them Design evaluation datasets and workflows that compare agent behavior against a baseline and produce actionable quality signals Build agent metrics and observability: skill selection and routing, success and failure outcomes, tool calls, latency, and token usage Design safety and quality gates for agent-authored content: rule packs, static analysis, confidence thresholds, structured verdicts, and bounded suppression Create CLIs, libraries, and MCP integrations that other repositories adopt and that run in local development and CI Integrate tooling into GitHub Actions and other CI workflows, including secrets, annotations, exit codes, and artifacts Build code-generation quality checks, such as anti-pattern catalogs and linting for AI-generated MongoDB code Investigate real failures such as nondeterministic results, false positives, and unsafe generated guidance, and turn them into reusable improvements Collaborate with engineers, security partners, and product teams; communicate trade-offs, risks, and ownership across teams Examples of the problems you'll solve How can tests verify an agent tool's
From $399.4K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. About the role: Roblox Studio is the creation engine behind millions of immersive 3D games built by creators around the world. We are entering a major platform transition, evolving Studio into an AI-Native IDE where intelligent systems plan, act, validate, and iterate alongside creators to amplify their output. We are seeking a Technical Director, Applied AI to own the technical direction of the agentic AI systems embedded deeply into Roblox Studio. This is a hands-on individual-contributor role operating at the intersection of platform engineering, AI systems, and developer experience, where you will set the technical direction, design the architecture, and write the code. You will: Design and build the agentic AI systems at the core of the AI-Native IDE (planning, execution, validation, evaluation) so they are production-grade and reliable. Decide how Studio uses coding models, agents, retrieval, tool calling, and context management across the Assistant and the broader creator workflow. Own the quality bar and evaluation strategy , standing up quantitative and qualitative eval pipelines, including human-in-the-loop, so we ship with confidence. Optimize these systems for latency, throughpu
From $220K/yr
The Applied AI team designs and builds algorithmically driven features in the Datadog app. We work across a range of applications, primarily focusing on analysis on streaming data such as anomaly detection, error outliers and faulty deployment analysis. As an Applied Scientist you will work on building models and algorithms for machine learning powered features within the Datadog platform. You will work closely with our engineering and product partners to explore, build, scale and deliver these features that we incubate within the Applied AI team. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Design solutions for our different use cases. You will research and benchmark relevant algorithms to find the best fit for our use-cases Leverage machine learning algorithms and statistical techniques to build new scalable product features Develop, deploy and monitor new and existing features to production Participate in our journal club by reading and presenting the latest academic research papers to the team Explore, analyze and tell the story behind high volumes of data flowing through Datadog systems Maintain and monitor the models, services and infrastructure owned by your team Participate in your team’s on-call rotation Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering, Machine Learning or related scientific field or equivalent experience You have experience working with high-scale systems and datasets including building models, applying machine learning to real business problems, and writing production data pipelines You can explain complex ideas and algorithms to non-technical audiences You care about code simplicity and performa
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As a senior member of Baseten's Platform Team, you will own the systems that let every engineer at Baseten prove their code works before it reaches production. Our product runs mission-critical AI inference for customers who measure downtime in dollars per second, which means our internal bar for correctness, performance, and failure tolerance has to be exceptional. Your focus is the full testing stack: fast and reliable unit test tooling, integration harnesses that spin up realistic environments on demand, load and performance testing for GPU-backed inference workloads, and resilience testing that deliberately breaks things so our customers never have to find out what happens when a node dies mid-request. This is a builder role with org-wide leverage. You won't be writing tests for other teams — you'll be building the frameworks, harnesses, and feedback loops that make writing good tests the path of least resistance, and you'll set the standards for what "well-tested" means at Baseten. RESPONSIBILITIES Own Baseten's testing strategy end to end — define the standards, the tiers, and the tooling that engineering teams build against. Build and maintain unit, integration, load and performance testing frameworks Design end to end test infrastructure that provisions realistic dependencies
Other cities to consider
More places hiring for this role
Get new applied ai engineer jobs in United States by email
Daily job updates · Unsubscribe anytime