About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry provides developer-first observability to over 4 million developers worldwide. The Events Analytics Platform (EAP) team is at the heart of that mission: it powers how all of Sentry's event data, such as errors, transactions, spans, profiles, replays, and metrics, is stored, queried, and analyzed. It also powers Sentry's latest AI push, Seer. The EAP team makes it possible for developers to efficiently search and debug across massive volumes of data, providing the context needed to understand and fix issues quickly. This team is also a cornerstone of Sentry's long-term strategy to become a context assembly and telemetry platform that unifies different signals so developers can see the complete picture. As an engineering manager on the EAP team, you will lead a group of engineers building and scaling one of Sentry's most critical data platforms. You will be responsible for driving architectural evolution, ensuring system stability, and mentoring a talented team. This is a highly visible leadership role with direct ties to Sentry's long-term product and platform strategy. What you'll do Grow and develop a team of engineers with high expectations for ownership and impact Set the technical and strategic direction for the team, balancing short-term stability with long-term architectural evolution Drive development of core EAP features, including support for complex analytical queries, dynamic routing logic across fidelity levels, storage and compute separation, and modern patterns for analytical storage Ensure EAP can support the workload demands of AI agents and MCP servers that unlock new AI capabilities fo
Jobs in United States
Agent Ai Engineer in San Francisco
287 active opportunities · Updated October 2026
Showing
15 jobs
Explore current agent ai engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
$220K – $450K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry provides developer-first observability to over 4 million developers worldwide. The Events Analytics Platform (EAP) team is at the heart of that mission: it powers how all of Sentry's event data, such as errors, transactions, spans, profiles, replays, and metrics, is stored, queried, and analyzed. It also powers Sentry's latest AI push, Seer. The EAP team makes it possible for developers to efficiently search and debug across massive volumes of data, providing the context needed to understand and fix issues quickly. This team is also a cornerstone of Sentry's long-term strategy to become a context assembly and telemetry platform that unifies different signals so developers can see the complete picture. As an engineering manager on the EAP team, you will lead a group of engineers building and scaling one of Sentry's most critical data platforms. You will be responsible for driving architectural evolution, ensuring system stability, and mentoring a talented team. This is a highly visible leadership role with direct ties to Sentry's long-term product and platform strategy. What you'll do Grow and develop a team of engineers with high expectations for ownership and impact Set the technical and strategic direction for the team, balancing short-term stability with long-term architectural evolution Drive development of core EAP features, including support for complex analytical queries, dynamic routing logic across fidelity levels, storage and compute separation, and modern patterns for analytical storage Ensure EAP can support the workload demands of AI agents and MCP servers that unlock new AI capabilities fo
About the Team The Developer Experience team at OpenAI has a singular focus: empowering developers globally. Our mission is to provide every developer and startup on the planet with the most delightful and seamless experience to integrate AI into their applications and products. We ensure developers have the tools, resources, and support they need to unlock AI’s full potential. We create inspiring demos, developer tools, sample applications, and technical content that show developers how to build with Codex and frontier models like GPT-5.6, GPT-Live, and GPT-Image-2 to create powerful agents and AI-native applications. We collaborate closely with product, engineering, research, and GTM teams to ensure the developer journey, from onboarding with Codex to first API call to production deployment, is seamless, effective, and delightful. About the Role As a Developer Experience Engineer, you will create compelling technical content, developer tools, and sample applications designed to inspire developers and enable them to succeed with Codex and OpenAI’s APIs and products for developers. You will engage with developers and technical founders, demonstrating best practices and building innovative applications powered by frontier models, multimodal capabilities, and tools like Codex. We’re looking for people who combine strong technical skills, creativity, and a passion for engaging with and empowering developers. In this role, you will: Develop demos and sample applications that showcase best practices for building with Codex, frontier models, multimodal capabilities, and agents. Create high-quality technical content—including tutorials, blog posts, videos, and code samples—to educate and inspire the developer community about our models, APIs, and Codex. Actively engage with and foster a vibrant local and global developer ecosystem around OpenAI’s platform and products. Represent OpenAI at developer events and online, serving as a knowledgeable and approachable advocate for
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why this role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why this role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei
About PostHog Product development used to mean manually writing code, running analysis, diagnosing bugs, and rolling out changes using dozens of tools. PostHog is the only platform that acts like a co-pilot for you (and your AI agents) to do it all – autonomously. We started with open-source product analytics, launched out of Y Combinator's W20 cohort . We've since shipped more than a dozen products , including: PostHog Code , the only AI devtool that understands your product, not just your codebase. A built-in data warehouse , so users can query product and customer data together using custom SQL insights. PostHog AI , an AI-powered analyst that answers product questions, helps users find useful session recordings, and writes custom SQL queries. We are: Product-led . More than 450,000 organizations have installed PostHog, mostly driven by word-of-mouth. We have intensely strong product-market fit. Default alive . Revenue is growing incredibly quickly, and we're very efficient. We raise money to push ambition and grow faster, not to keep the lights on. Well-funded. We've raised more than $180m from some of the world's top investors. We're set up for a long, ambitious journey. We're focused on building an awesome product for end users, hiring exceptional teammates, shipping fast, and being as weird as possible . Things we care about Transparency: Everyone can read about our roadmap, how we pay (or even let go of) people, our strategy, and how we work, in our public company handbook . Internally, we share revenue, notes and slides from board meetings, and fundraising plans, so everyone has the context they need to make good decisions. Autonomy: We don’t tell anyone what to do. Everyone chooses what to work on next based on what's going to have the biggest impact on our customers, and what they find interesting and motivating to work on. Engineers lead product teams and make product decisions . Teams are flexible and easy to change when needed. Shipping fast: Why not n
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry's issue platform processes billions of events every day to help millions of developers find and fix bugs. The hard part is deciding which events point to a problem, which belong together, and what a developer needs to know to investigate. As a Senior Software Engineer on the Issue Detection team, you'll design, build, and operate the systems that make those decisions. You'll work on real-time processing pipelines, monitors, and analysis systems that detect problems and turn them into issues. The work combines distributed systems with product engineering. Choices about detection accuracy and processing latency affect which problems developers see and how soon they can act. You'll help shape how developers monitor their applications and how Sentry groups related events into issues. You'll also build the context developers and AI agents need to investigate what went wrong. Keeping these systems reliable and fast as Sentry grows is part of the job, alongside making the issues they produce more useful. In this role you will Build and scale features on a product surface handling billions of events daily, where both query latency and correctness are immediately visible to users. Own the design and delivery of substantial projects end to end, scoping alongside product and design, making the technical calls within your scope, shipping, and instrumenting what you ship so the team can measure it. You will contribute to meaningful technical product decisions : grouping quality, search performance, migrations and backfills against enormous datasets, and making the surface work well for both humans and agents. Champi
From $165.4K/yr
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. What you'll do Identity & access Advance our identity posture: SSO coverage, phishing-resistant MFA rollout, SCIM lifecycle automation, and least-privilege access across the SaaS and cloud estate. Build the detections and guardrails that catch account takeover, MFA fatigue attacks, and session token theft before they turn into incidents. Endpoint & device lifecycle Write and ship device policy as code — configuration profiles, remediation scripts, and enforcement rules across macOS and Windows — with staged rollout and rollback built in from day one. Maintain and improve our EDR stack's detection and response coverage across the fleet. SaaS posture Reduce SaaS risk at scale through SSPM tooling and automation , including detection of risky OAuth grants, shadow IT, and configuration drift across our critical SaaS applications. Own security configuration for the SaaS tools hundreds of Flexporters use daily (Google Workspace, Slack, and similar), and keep pace as we add AI agents and MCP integrations to that surface. Automation & enablement Automate the parts of corporate security that don't need a human — device provisioning, access reviews, vendor securi
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity The Customer Journey team at Postman is responsible for shaping how users experience the product from their first interaction through to becoming active, engaged customers. We own the critical paths between sign-up and purchase including new user onboarding and the moments where users discover and adopt key features. As Postman expands what it means to be a customer, from enterprise organizations to teams of engineers to individual developers working with AI agents, this team plays a central role in defining and evolving those experiences. We’re looking for a Senior Engineer who is excited to build and improve product journeys end-to-end. You’ll work at the intersection of product, design, and engineering to create intuitive, high-quality experiences that help users find value in Postman—whether they’re exploring on their own or collaborating as part of a team. What You’ll Do Own and improve key product journeys from sign-up through onboarding, feature discovery, and purchase Build and refine user experiences that help developers and teams quickly understand and realize the value of Pos
About the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme level of severity. Our work involves: Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping misuse safeguards, alignment tools, and security measures on track to adequately address extreme threats that might arise in the future. Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness framework , and partnering with other staff to achieve these targets. This is urgent, fast-paced work that has far-reaching implications for the company and for society. About the role The stakes of securing OpenAI increases as our internal coding and research becomes increasingly driven by autonomous AI agents. Compromising these agents could allow a cyber threat actor to compromise many other parts of the company. In this role, you would lead Preparedness work defending the security of our internal AI agents against insiders, Advanced Persistent Threats (APTs), or powerful AI agents. We’re looking for a strong hands-on technical executor with experience working directly with advanced cyber threat actors. In this role, you will: Develop and maintain threat models via which advanced attackers could compromise our coding assistants and automated security systems. Identify security investments that are especially critical to make in advance; for example, prioritizing by implementation lead-times, costs, and benefit. Partner with Security, Infrastructure, Research, Legal, and Preparedness to align on implementation plans and tradeoffs. Lead technical execution directly when needed, including prototyping controls, writing and reviewing software, and coordinating engineers across teams. Work with penetration testers to close gaps in defenses. You might thrive in this role if you: Are an exceptional hands-on technical executor. Have worked with advanced
$196K – $230K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good” looks like not only for model output quality, but for the end-to-end product experience where people discover, set goals, delegate work, review results, and build trust over time with AI. This role sits at the intersection of research craft and evaluation operations: you’ll run studies that uncover user mental models, expectations, and failure/recovery behaviors, then translate those insights into reusable rubrics, workflows, and measurement approaches that product, design, engineering, and data science can apply consistently. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: Define what “good” looks like (frameworks & rubrics): Establish clear, reusable evaluation criteria that reflect real user expectations—helpfulness,
From $166K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Own how our Foundations business teams design, automate, and operate their most important workflows. Notion is growing quickly, and the processes that run our business — intake, triage, execution, approvals, escalations, reporting — need to scale without adding headcount or losing auditability. This role exists to redesign those processes end to end and to embed AI directly into them, so the work moves faster and the system stays trustworthy. What's different at Notion: you'll build on Notion as customer zero. You'll design multi-system workflows spanning Notion, Slack, email, support tooling, and internal data, then ship the custom agents, skills, and automations that run them — partnering with Finance, Legal, People, Ops, Data, and Engineering to drive real adoption, not just documentation. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: Redesign cross-f
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Product at Brex The Product team is at the forefront of Brex's mission to empower employees anywhere to make better financial decisions. With a deep understanding of the business, we identify and scope out the most impactful opportunities for Brex to tackle. We are responsible for aligning cross-functional teams — such as Engineering, Legal, Compliance, and Design — on key decisions. We set strategy and drive products from inception to launch, enabling Brex to grow rapidly and help our customers reach their full potential. Unlocking AI Agents for the Finance Admin Companies buy Brex to run their spend, and the person on the hook for that is the finance admin — a controller, an accounting manager, a finance ops lead. Their job comes down to two things they can never quite get ahead of: making sure money is being spent compliantly and documented in time to close the books, and knowing where the company's money is actually going. Today t
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About this role This is a chance to help shape the next chapter of WRITER's product at a moment when generative AI is redefining how work gets done. As senior staff product manager, you'll report directly to Wil Pong, VP of Product, and operate as a force-multiplier across the product organization — setting direction where it's unclear, aligning stakeholders, and making sure our most critical initiatives ship with excellence. You'll own ambiguous, high-impact problem spaces that span multiple product areas and teams, and you'll be equal parts strategist and operator. One moment you're shaping a multi-quarter roadmap; the next you're rolling up your sleeves to unblock execution. We're looking for someone who thrives at the intersection of customer empathy, technical depth, and business judgment — and who is energized by turning complex, evolving AI capabilities into products customers love. We are open to hiring for this role at multiple levels, with compensation ranges scaled
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About this role This is a chance to help shape the next chapter of WRITER's product at a moment when generative AI is redefining how work gets done. As senior staff product manager, you'll report directly to Wil Pong, VP of Product, and operate as a force-multiplier across the product organization — setting direction where it's unclear, aligning stakeholders, and making sure our most critical initiatives ship with excellence. You'll own ambiguous, high-impact problem spaces that span multiple product areas and teams, and you'll be equal parts strategist and operator. One moment you're shaping a multi-quarter roadmap; the next you're rolling up your sleeves to unblock execution. We're looking for someone who thrives at the intersection of customer empathy, technical depth, and business judgment — and who is energized by turning complex, evolving AI capabilities into products customers love. 🦸🏻♀️ Your responsibilities Drive high-impact product initiatives end-to-end, from di
Other cities to consider
More places hiring for this role
Get new agent ai engineer jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime