About the Team The SaaS and Software Governance team sits within Corporate IT and helps OpenAI scale from startup-speed tooling to mature enterprise architecture. The team owns practical governance for software, SaaS, integrations, APIs, identity, data access, and agent-enabled workflows, with a mandate to improve security, reduce software sprawl, and help business teams move faster through better foundations. About the Team OpenAI is scaling from an emerging, high-velocity startup into a mature enterprise operating model. The IT Software Architect will help shape the software, SaaS, Data integration, and agent-enabled architecture that lets the company move quickly while improving security, compliance, data quality, and customer, partner, and employee experience. This role is not a traditional ivory-tower architecture function. It is a hands-on governance and enablement role that partners with business teams, IT operations, Security, Procurement, Business Platforms, Applied teams and Data Engineering to guide software decisions, reduce unmanaged sprawl, and build reusable enterprise foundations. Why This Role Matters OpenAI’s software footprint is expanding rapidly across SaaS, internally built tools, agents, integrations, APIs, third-party platforms, and application systems that OpenAI. The company needs a stronger tools architecture layer that can help teams make good decisions early, avoid duplicate tools, govern sensitive data and identities, and identify where OpenAI should build instead of buy. The person in this role will help turn software governance from an approval checkpoint into an enterprise capability: a system that improves speed, reliability, security, and business outcomes. What You'll Do Own the target architecture for enterprise software, SaaS, integrations, APIs, and agent-enabled business systems across Corporate IT. Drive deprecation and consolidate targets for enterprise software. Build lightweight governance patterns that guide teams before
Jobs in United States
Agent Ai Engineer in San Francisco
287 active opportunities · Updated October 2026
Showing
15 jobs
Explore current agent ai engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI’s mission is to ensure that general-purpose artificial intelligence benefits all of humanity. The Identity team builds the foundational systems that enable people, organizations, devices, and agents to securely access OpenAI products. The team works across consumer and enterprise experiences, including account structures, sign-in, authentication, recovery, privacy, permissions, administration, and agent identity. As AI systems become increasingly capable of acting on behalf of people and organizations, we are defining new models for authorization and trust across human-to-agent and agent-to-agent interactions. About the Role In this role, you’ll lead design for foundational identity experiences across OpenAI’s products. You’ll make permissions, access, risk, and administration feel clear, safe, and genuinely usable—whether someone is securing a personal account, administering access across an enterprise, signing in to another product with ChatGPT, or authorizing an agent to perform sensitive work. Your work will span established identity challenges and emerging interaction models without established design conventions. You’ll help define how users understand what an agent can access, who or what it is acting on behalf of, and when confirmation or stronger authentication should appear. Working closely with product, engineering, security, privacy, and design, you’ll translate complex policies and technical systems into coherent, trustworthy experiences. This role is based in our San Francisco HQ. We offer relocation assistance to new employees. In this role, you will: Lead the design direction for identity, account security, permissions, and administrative experiences across OpenAI’s consumer and enterprise products. Design and ship high-quality experiences spanning sign-in, authentication, account recovery, device accounts, privacy, access controls, governance, and remediation. Define mental models and interaction patterns for human-to-agent and
About the Team The Account & Platform Integrity Operations team protects OpenAI’s ecosystem by ensuring that users, developers, partners, and organizations can access and build on our platform safely and responsibly. The team works across account integrity, fraud prevention, abuse detection, and platform risk to prevent bad actors from exploiting OpenAI’s products and developer surfaces. This role will sit within the Account & Platform Integrity Operations team while partnering closely with teams across the Ecosystem organization to ensure the platform can scale rapidly without compromising trust, quality, or safety. About the Role - Platform Operations Program Manager As OpenAI's developer ecosystem expands, we're building the operational foundation to support the next generation of plugins, MCP apps, agent skills, and integrations. As a Platform Operations Program Manager, you'll partner across Product, Engineering, Policy, Legal, Security, and Developer Experience to design the systems and operating models that enable ecosystem growth while maintaining quality, trust, and an exceptional developer experience. This role requires strong operational judgment, a deep understanding of developer platforms and APIs, empathy for the developer experience, and the ability to translate evolving product, platform, and policy requirements into scalable operational systems. Location: San Francisco, CA (Hybrid - 3 days in office) What You'll Do Own the operational processes that help developers bring apps, plugins, and integrations to market, from submission and review through launch, appeals, and ongoing monitoring. Manage external review partners and policy operations workflows, helping teams apply standards consistently while identifying areas where guidance, process, or quality expectations need to improve. Partner with Product, Engineering, Policy, Legal, Security, Developer Experience, and Go-to-Market teams to turn platform goals and requirements into clear, effec
About the Team: The OpenAI API team builds the foundation that enables every developer to harness OpenAI’s models safely, reliably, and at scale. Our mission is to make it effortless for any developer to build transformative products with OpenAI’s models. We’re responsible for the infrastructure and product layers that allow millions of developers to integrate our models, fine-tune behavior, manage data, and deliver experiences to their users. We collaborate deeply across product, research, and engineering teams to drive innovation at the model layer and then directly translate that into value for customers. About the Role: As the Agents Product Manager for the API team, you'll be at the forefront of defining and guiding the future of how developers build agentic applications on top of our AI models. You’ll set clear priorities and drive impactful improvements to model capabilities, balancing user needs, safety considerations, and technical innovation. This role is perfect for a proactive, technically adept PM who thrives on solving challenging, ambiguous problems through structured product thinking and close collaboration with customers, engineers, and researchers. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Deeply understand problems faced by agent builders and identify opportunities where our products and models can make building agents faster, more intuitive, more reliable, and more powerful. Define strategic priorities and roadmap for improving agentic infrastructure for API users, focusing on user outcomes and emerging capabilities. Partner with research and engineering teams at a technical level to translate those priorities into developer products and features (SDKs, APIs, and more). Deliver quickly while maintaining a high bar for product quality and user experience. You might thrive in this role if you: Have 5+ years of product management or related industry experience. Proven track record of b
About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.
About the Team The GTM Intelligence Solutions team builds the data and decision systems that help customer-facing teams take the right action at the right time. We combine product telemetry, commercial data, customer context, and field activity to identify account health and opportunity, recommend actions and use cases, deliver intelligence through field-facing products and agent workflows, and measure what happens next. We’re looking for a Data Scientist to help build the next generation of GTM intelligence at OpenAI. You will own a flexible portfolio of high-impact decision data products and work closely with Technical Success and other GTM teams to ensure the work drives better decisions. About the Role As a Data Scientist on GTM Intelligence Solutions, you will define and build the intelligence systems that help customer-facing teams prioritize accounts, identify risks and opportunities, choose interventions, and understand what worked. You will set the roadmap and methodology, build canonical features, ship reliable production workflows, monitor quality and adoption, and improve the systems using field feedback and business outcomes. This role combines hands-on technical depth with strong product and business judgment. You should be as comfortable writing production Python and advanced SQL, defining durable data contracts, and operating decision products as you are evaluating a ranking approach or designing an experiment. You will personally ship reliable first versions and partner with Analytics Engineering and Data Engineering when work requires shared infrastructure or additional scale. In This Role, You Will Set the roadmap and methodology for GTM intelligence and decision products, using deep stakeholder discovery to probe beyond stated requests, uncover the underlying decisions, workflows, constraints, and measures of success, and translate them into measurable systems. Own the full lifecycle of intelligence products, including feature definition, methodo
About the Team OpenAI Finance ensures the organization is positioned for long-term success as we pursue our mission. The Revenue team plays a critical role in enabling OpenAI to scale commercial offerings by overseeing billing operations, deal desk, revenue systems, revenue accounting, and controllership. We work cross-functionally with Product, Engineering, Go-To-Market, Tax, Legal, and Technical Accounting to support new monetization strategies, improve operational efficiency, and maintain financial integrity as the business grows. About the Role This senior leader will own key elements of Ads revenue accounting from technical assessment through operational execution. The role will guide accounting for products, pricing, contracts, incentives, credits, refunds, makegoods, international expansion, and new go-to-market motions. It will establish governance and translate approved accounting positions into launch, billing, data, close, reconciliation, and control requirements. Success requires deep technical revenue expertise, strong business partnership, and the ability to build durable 0-to-1 processes in a fast-changing environment. Advertising is a critical and growing monetization vector for OpenAI, and this role will help shape the financial foundations that enable Ads to scale responsibly and transparently. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead technical accounting assessments for Ads products and commercial arrangements, including performance obligations, variable consideration, allocation, principal-versus-agent, collectibility, contract modifications, refunds, incentives, credits, makegoods, and revenue presentation. Own and continuously evolve Ads revenue accounting policies and operating guidance as product behavior, pricing, contracting, incentive programs, and billing models change. Establish governance for new pro
About the Team The Intelligence and Investigations team seeks to rapidly identify and mitigate abuse and strategic risks to ensure a safe online ecosystem. We are dedicated to identifying emerging abuse trends, analyzing risks, and working with our internal and external partners to implement effective mitigation strategies to protect against misuse. Our efforts contribute to OpenAI's overarching goal of developing AI that benefits humanity. The Strategic Intelligence & Analysis (SIA) team provides safety intelligence for OpenAI’s products by monitoring, analyzing, and forecasting real-world abuse, geopolitical risks, and strategic threats. Our work informs safety mitigations, product decisions, and partnerships, ensuring OpenAI’s tools are deployed securely and responsibly across critical sectors. About the Role As an Agentic Risk Analyst, you will shape OpenAI’s operating picture for current agentic risk across products and platforms. You will bring a strategic, system-level perspective to current risks, connecting individual incidents, technical findings, abuse patterns, and external developments to relevant workstreams, mitigations, owners, dependencies, and residual gaps. You will analyze how risks emerge through autonomy, multi-step task execution, tool use, memory, retrieval, connectors, computer-use capabilities, and multi-agent workflows, with a particular focus on both adversarial misuse and unintended system behavior. By synthesizing signals from investigations, evaluations, red teaming, security reviews, product launches, external research, and real-world incidents, you will maintain a current view of material risks and evolving threat patterns. Your work will help turn complex and often ambiguous signals into coordinated decisions and measurable follow-through across product, safety, security, policy, and governance teams. You will work closely with investigators, engineers, product, policy, safety, and security teams, and measurement and forecasting
$192K – $259.8K/yr
Drata is building the trust layer between great companies - automating compliance, managing risk, and helping organizations prove trust continuously as they scale. We're Dratanauts: a global crew of 600+ professionals united by a culture that rewards integrity, ownership, and raising the bar, no matter where in the world we're working from. Why Join the Drata Team? At Drata, you're not maintaining legacy compliance software - you're building the agentic AI platform defining what trust looks like for the next generation of companies. Here's what makes the work itself worth showing up for: Problems without a playbook: You'll work at the edge of AI and security, building agentic governance, continuous compliance, and real-time trust verification to solve problems that don't have an established answer yet. You're writing it as you go. Real ownership, not just process: Our values center on owning outcomes and raising the bar, not checking boxes. You're expected to have opinions and back them. A seat at the table: Your perspective is unique and valued. Open debate and diverse viewpoints are built into how decisions actually get made here, at every level. Growth at rocketship speed: Drata is scaling fast, which means scope grows fast too. High performers get more ownership, visibility, and experience. A crew, not just coworkers: Dratanauts consistently describe a "come as you are" culture with sharp, curious people—the kind of team that makes hard problems genuinely fun to solve. See what they say here and follow us on LinkedIn for company news, employee stories, and career updates. Job Summary: Drata's AI Platform team builds the production infrastructure that powers AI features across our compliance platform — from MCP servers that make Drata's data available to AI agents, to LLM workflow orchestration that automates SOC 2, TPRM, and policy analysis. You'll own the systems that sit between our AI models and our customers: tool definitions that agents actually understand,
From $130K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. This is NOT a new grad role! This role is for candidates with 0-2 YOE and can start full time right away. About the Role As an engineer at Notion, you’ll help shape core user experiences and accelerate how people discover value in Notion. You'll tackle meaningful challenges with increasing autonomy, crafting code that millions of users will experience. You'll take ownership of projects that matter, make critical technical decisions, and contribute your unique perspective to our product vision. Working alongside passionate experts across design, product, and data, you'll help shape the future of how people work. We're looking for an Early Career AI Engineer to join as a strategic partner in shaping Notion's AI vision. You'll work on cutting-edge AI-powered features, leveraging LLMs, embeddings, and other AI technologies to make Notion more intelligent and capable. This role may be aligned to one of multiple AI-focused teams at Notion. Depending on team match and business needs, you could work on: AI product engineering: building model-powered features end-to-end (UX, APIs, retrieval, orchestration, quality, and reliability) Model &
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As an AI engineer at WRITER, you'll be at the forefront of shaping how enterprises harness superintelligence. This isn't just about theory; you'll be building tangible AI solutions that power the future of work for hundreds of the world's leading companies. Your work will directly impact the performance, scalability, and ethical alignment of our cutting-edge LLMs and AI agents, enabling businesses to unlock unprecedented levels of productivity and innovation with AI that is truly grounded in their data. This role can be hybrid in our San Francisco, New York City, or Seattle hubs. You'll report to the head of AI engineering. 🦸🏻♀️ What you'll do Architect, develop, and deploy high-performance, scalable AI applications into production environments, ensuring robust integrations with our end-to-end platform. Drive the development of intelligent agents and AI-powered features, translating complex research into practical, impactful solutions for our customers. Coll
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As an AI engineer at WRITER, you'll be at the forefront of shaping how enterprises harness superintelligence. This isn't just about theory; you'll be building tangible AI solutions that power the future of work for hundreds of the world's leading companies. Your work will directly impact the performance, scalability, and ethical alignment of our cutting-edge LLMs and AI agents, enabling businesses to unlock unprecedented levels of productivity and innovation with AI that is truly grounded in their data. This role can be hybrid in our San Francisco, New York City, or Seattle hubs. You'll report to the head of AI engineering. 🦸🏻♀️ What you'll do Architect, develop, and deploy high-performance, scalable AI applications into production environments, ensuring robust integrations with our end-to-end platform. Drive the development of intelligent agents and AI-powered features, translating complex research into practical, impactful solutions for our customers. Coll
$155K – $400K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Senior Software Engineer on Sentry’s AI/ML team, you’ll be responsible for building the evaluation infrastructure that measures the accuracy, reliability, and real-world performance of our AI systems. This role is critical to ensuring that our debugging agents and AI-powered features behave correctly, safely, and predictably as they scale. You’ll design datasets, benchmarks, and test harnesses that turn ambiguous AI behavior into measurable signals, helping the team ship AI with confidence. In this role you will Design and build robust evaluation frameworks to measure accuracy, reliability, regressions, and edge cases in AI systems Create and curate high-quality datasets, golden test cases, and benchmarks grounded in real production data Build automated test harnesses and metrics pipelines to continuously evaluate models, prompts, and agentic workflows Partner closely with applied AI engineers and product leaders to define what “good” looks like and translate it into measurable criteria Own the evaluation lifecycle for major AI initiatives, from early experimentation through production monitoring You’ll love this job if you Care deeply about correctness, rigor, and measurement in AI systems Enjoy turning fuzzy product goals and model behavior into concrete tests and metrics Like building foundational infrastructure that unlocks faster iteration and higher confidence for the entire AI team Thrive in cross-functional environments and enjoy influencing model design through better evaluation Qualifications Minimum 5+ years of professional experience with a Bachelor’s degree in computer science, machine learni
From $152K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role We’re looking for an AI Applications Engineer to help drive Notion’s business transformation efforts. In this role, you’ll be a strategic partner to our internal stakeholders (primarily GTM, Finance and People teams) and deliver and scale creative AI-driven solutions to multi-faceted problems with measurable business impact. You’ll also build reusable components, evaluation patterns, and operational guardrails that make AI delivery repeatable across teams. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You’ll Achieve Work with stakeholders to discover opportunities from ambiguous problem statements, translate them into scoped solutions, and drive iterative releases from idea to adoption Build and ship end‑to‑end AI solutions—from problem framing through data readiness, modeling, evaluation, and production rollout Establish evaluation and production-readiness patterns (met
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As a deployment engineer at WRITER, you'll be at the forefront of expanding human capacity through superintelligence. This isn't just a technical role; it's a deeply influential one where you'll partner directly with our leading enterprise customers. Your expertise will be crucial in uncovering their unique business challenges and architecting AI-powered solutions that leverage our powerful platform and enterprise-grade LLMs. You'll transform complex needs into tangible, high-impact applications, creating champions and driving tangible business results. Your builder's mentality and passion for bringing cutting-edge AI into the hands of real users will directly shape the future of work for some of the world's largest companies. This is a critical role that directly impacts our customers' success and product evolution. You'll contribute significantly to WRITER's mission, working with a dynamic team to push the boundaries of what's possible with generative AI. Thi
Other cities to consider
More places hiring for this role
Get new agent ai engineer jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime