About the team The Agent Enablement AI Deployment Engineering (ADE) team works across engineering, product, design, partnerships, and strategic customers to grow an open ecosystem of agent-enabled sites and services. We help partners adopt the OpenAI tech stack related to identity, permissioning, agent-auth primitives so users can safely connect ChatGPT and Codex to the tools, services, and workflows they already use. Our team also works with external partners on defining the standards for agent access, marketplace offerings as well as other agent enablement initiatives to ensure users of ChatGPT and Codex go from intent to task completion seamlessly. About the role We are looking for an AI Deployment Engineer to help strategic partners design, build, validate, launch, and operate agent enablement integrations across web applications, connectors, APIs, CLIs, MCP servers, and developer tools. This is a hands-on, partner-facing product engineering role for someone who can contribute to the platform itself, lead sophisticated technical engagements, and turn ambiguous identity and agent-workflow requirements into secure, production-ready integrations. You will work across partner product and engineering teams and OpenAI’s product, engineering, design, partnerships, legal, policy, security, support, and go-to-market teams. You will identify high-value user journeys, choose the right integration path, prototype and review architectures, write code, run evaluations and dogfood, trace failures end to end, guide launch and rollout, and support post-launch iteration. The best person for this role moves fluidly between full-stack code, OAuth/OIDC and identity systems, product judgment, project leadership, and clear communication with engineers and executives. This role is a fit for a product-minded engineer who wants to stay close to users and partners while going deep on authentication, permissions, reliability, safety, and developer experience. The principle objective is to
Jobs in United States
Agent Ai Engineer in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current agent ai engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity (Summer 2026 AI Internship - Applications Open Now) We're seeking an AI Engineer Intern to work alongside our AI team on large-scale AI and Agentic systems from data pipeline to production deployment. This role is scoped for someone with foundational experience who wants to deepen it: you'll own discrete pieces of real systems under the mentorship of senior engineers, not shadow work or isolated coursework-style projects. What You'll Do You’ll work directly with the AI team, taking responsibility for well-scoped pieces of real systems, with mentorship from senior engineers. Benchmarks & Evaluation Contribute to APIFlow-Bench , our open-source benchmark for real API-development work: design and review benchmark tasks and their mock API environments, extend the evaluation harness and task-generation pipeline in Python, and help maintain the public multi-model leaderboard with statistical confidence intervals. Help build a new action-level AI safety benchmark: instead of grading what a model says, it scores what an agent actually does inside a simulated enterprise API environment. You’ll work on scenari
From $276K/yr
Team description At Datadog, AI agents are becoming first-class consumers of observability, security, and software delivery data — from third-party coding agents like Claude Code, Cursor, and Copilot, to our own Bits SRE, Bits Assistant, and Bits Dev Agent. The Agentic Interfaces team owns the platform that connects these agents to Datadog: the MCP Server, the tools and retrieval surfaces agents call into, and — critically — the evaluation systems that tell us whether an agent's experience on Datadog data is actually getting better over time. This role is about that last piece. We're hiring a Staff Applied Scientist to define what "good" means for an Agentic interface at Datadog and to build the measurement systems that make it true. "Good" isn't one number — it spans answer quality, tool-selection accuracy, retrieval relevance, latency, token cost, and end-to-end agent success on real customer workflows. You'll design the evals, build the datasets, define the metrics, and partner with the AI engineers on the team to land the platform that lets every product group at Datadog ship integrations that are demonstrably better release over release. The space is full of open research questions. How do you evaluate an agent end-to-end when the trajectory is non-deterministic? How do you score tool selection when the tool catalog has hundreds of entries and grows weekly? How do you build a measurement system that catches regressions across first-party and third-party agents at once, without each team writing their own harness? If those are the problems you want to spend your time on, come build this with us. Datadog values people from all walks of life. We understand not everyone will meet all the above qualifications on day one. That's okay. If you’re passionate about technology and want to grow your skills, we encourage you to apply. What You’ll Do: Own the evaluation strategy for Datadog's AI agent integrations. Define the metrics — offline and online, quali
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview: We are seeking a highly skilled Staff AI Engineer - Multi-Agent Frameworks to join our AI Platform team. In this role, you will play a pivotal part in building a cutting-edge platform that empowers our users to create and deploy sophisticated intelligent agents, with a key focus on enabling collaborative and multi-agentic behaviors . This is a backend-focused role that requires deep expertise in AI, large language models (LLMs), and orchestration software. Key Responsibilities: Design, develop, and maintain a robust platform to enable users to create and manage AI agents and their interactions. Integrate and work with multiple LLMs, ensuring seamless orchestration and scalability for both individual and coordinated agent operations. Leverage orchestration frameworks like LangGraph and others to build complex workflows and pipelines that support diverse agent functionalities, including frameworks for multi-agent coordination . Develop and implement evaluation frameworks for testing AI agents in challenging and complex scenarios, focusing on individual performance and system-level dynamics. Stay at the forefront of AI advancements, incorporating the latest research and technologies into our platform to enhance agent capabilities and collaboration. Collaborate with cross-functional teams, including product managers, designers, and frontend engineers, to deliver a seamless user experience for building and deploying intelligent systems. Address challenging AI privacy scenarios, ensuring compliance with data protection regulations and best practices within agent-based applications. Contribute
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview: We are seeking a skilled and experienced Senior AI Engineer - Multi-Agent Frameworks to join our AI Platform team. In this role, you will play a pivotal part in building a cutting-edge platform that empowers our users to create and deploy sophisticated intelligent agents, with a key focus on enabling collaborative and multi-agentic behaviors . This is a backend-focused role that requires deep expertise in AI, large language models (LLMs), and orchestration software. Key Responsibilities: Design, develop, and maintain a robust platform to enable users to create and manage AI agents and their interactions. Integrate and work with multiple LLMs, ensuring seamless orchestration and scalability for both individual and coordinated agent operations. Leverage orchestration frameworks like LangGraph and others to build complex workflows and pipelines that support diverse agent functionalities, including frameworks for multi-agent coordination . Develop and implement evaluation frameworks for testing AI agents in challenging and complex scenarios, focusing on individual performance and system-level dynamics. Stay at the forefront of AI advancements, incorporating the latest research and technologies into our platform to enhance agent capabilities and collaboration. Collaborate with cross-functional teams, including product managers, designers, and frontend engineers, to deliver a seamless user experience for building and deploying intelligent systems. Address challenging AI privacy scenarios, ensuring compliance with data protection regulations and best practices within agent-based applications. C
About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for applied AI engineers to help bring Codex agents from impressive demos to dependable tools. This role is about improving agent performance on real software engineering tasks and closing the gap between research capability and real-world usefulness. You’ll work closely with research, infrastructure, and product to ensure agents are not just powerful, but useful, steerable, and reliable in practice. The job is not only to improve model behavior in isolation, but to turn those improvements into measurable gains in solve rate, usefulness, and economic value for users. What You’ll Do Design and iterate on agent behaviors across real-world coding tasks and long-horizon workflows. Work closely with research to develop and run evals to measure agent performance, regressions, failure modes, and edge cases. Improve performance through prompting, tool-use strategies, context construction, and model-facing experimentation. Analyze failures in production and systematically improve robustness and reliability. Build feedback loops and data systems that get better real-task data into evaluation and research. Work with product teams to shape user-facing agent experiences and the interfaces the agent depends on. Help define what “good” looks like for agents completing complex tasks end-to-end. You Might Be a Good Fit If You Ha
From $152.8K/yr
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As Engineering Manager for the Agent Execution group , you'll lead more than 10 engineers who build foundations for AI agents to run safely and effectively across GitLab. You'll give the group clarity, remove obstacles, and help teams move quickly while protecting what matters. You'll contribute to design reviews on sandboxing and agent observability, work alongside staff engineers, and use modern AI coding tools and agent harnesses in your work. What you’ll do Lead the Agent Tools, Agent Observability, and Runner Execution teams across frontend, backend, and AI engineering, with each team anchored by a s
About the Team GTM Growth Engineering builds AI-native products that help OpenAI's go-to-market and B2B marketing organizations scale with greater speed, intelligence, and operational effectiveness. We apply OpenAI models to real business workflows and build the systems that make those applications useful and dependable: customer context, agent behavior, feedback, evaluation, experimentation, and appropriate human oversight. Our work brings together software engineering, applied AI, product, data, and GTM operations. We measure success through the quality of customer engagement, pipeline, conversion, and the effectiveness of our sales and marketing teams. About the Role We're looking for an Applied AI Engineer to build production systems that help AI-powered go-to-market workflows improve over time. You will connect agent behavior, customer and operator feedback, evaluation, experimentation, and business outcomes to make these systems more effective, reliable, and responsive to evolving customer needs. This is a deeply technical, cross-functional role with end-to-end ownership of the agent improvement loop: understand production behavior, identify failure modes, improve how the system decides or acts, and validate the resulting impact. You will partner with Engineering, Product, Data Science, Sales, and B2B Marketing to turn real-world signals into safer, more effective agent behavior and measurable improvements in customer engagement, conversion, qualified pipeline, and team productivity. In this role, you will: Own the production improvement loop across agent behavior, customer and operator feedback, evaluation, experimentation, and verified business outcomes. Instrument agent workflows so model interactions, tool use, decisions, failures, human edits, and downstream outcomes can be understood in context. Define meaningful quality standards, representative evaluation datasets, regression coverage, and production monitoring for real GTM workflows. Investigate why a
We are investing in agentic AI and need a Senior AI Engineer to lead the design and delivery of these systems. This is a foundational hire: you will own both the agent-facing workstreams — pipelines, orchestration, conversational interfaces — and the underlying context layer that makes them reliable, including memory management, knowledge graph integration, and retrieval infrastructure. You will work closely with data engineers, project leads, and client stakeholders, and play a key role in shaping how Lynx builds and ships AI solutions at scale. What This Involves: Lead the architecture and delivery of agentic AI systems end-to-end: agents, orchestration, tool use, and multi-step reasoning workflows. Own the context layer: design and implement memory architectures (episodic, semantic, working memory) and integrate GraphRAG and knowledge graph retrieval into agentic pipelines. Build robust RAG systems — including vector retrieval, graph traversal, and hybrid search — and ensure retrieval quality through evaluation frameworks. Translate client requirements into technical designs, presenting approaches and trade-offs to both technical and non-technical stakeholders. Define standards and reusable patterns for agentic AI development that other engineers at Lynx can build on. Set up observability, evaluation, and monitoring pipelines to ensure AI systems perform correctly in production. Requirements: 5–8 years of software or ML engineering experience, with at least 2–3 years building LLM-based or agentic AI systems in production. Deep hands-on experience with agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI, or similar) and LLM APIs (OpenAI, Anthropic, etc.). Strong understanding of agent design patterns: ReAct, planning loops, tool use, multi-agent coordination, and memory architectures. Practical experience with GraphRAG or knowledge graph-based retrieval (e.g., Neo4j, Microsoft GraphRAG) and vector databases (Pinecone, Weaviate, Qdrant, etc.). Proficiency in
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Are you the person on your team who builds the agent everyone else ends up using? We're looking for an AI Engineer to join our Training Product team and do that at Baseten. You'll build AI-driven product features for the customers training and post-training frontier models on our platform, and you'll raise the ceiling on how Baseten itself uses AI internally, turning manual workflows into agentic ones that make every other team faster. You'll work directly with our research engineers to scope and build products, taking ideas from a research loop that already works internally to something customers can run themselves. This is a hands-on role with real autonomy. You'll pick the problems worth solving, build the harnesses, execution flows, and guardrails that make AI systems reliable, and own the results. If you've been shipping agents and want that to be the job, let's talk. EXAMPLE INITIATIVES: Take a look at these blog posts written by members of our team: Baseten Training: an autoresearch substrate Introducing Baseten Loops Harnesses are everything. Here's how to optimize yours. Building with NVIDIA Nemotron 3 Ultra and LangChain Deep Agents Code on Baseten RESPONSIBILITIES: Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to GA. Design the harnesses, execution flows, and guardrails that make AI systems reliable in production. Build internal autom
From $152.8K/yr
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role The DAP Repository Flows team owns two areas within GitLab's Duo Agent Platform (DAP). The first is DAP onboarding , helping customers quickly set up a project so they can start using the Duo Agent Platform with minimal friction. The second is repository and background flows : autonomous and scheduled agents that work on a customer's codebase on their behalf. This work is central to how GitLab delivers AI-powered automation on repositories. As a Staff Software Engineer, you will set technical direction for the team, drive complex initiatives across team boundaries, and raise the engineering bar for how we
$171K – $240K/yr
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. AI at Brex AI Engineering at Brex is redefining how businesses run their finances by building intelligent, autonomous systems directly into the Brex platform. Our teams develop AI agents that don’t just surface insights—they take action, optimizing spend, managing workflows, and making real-time decisions on behalf of our customers. By deeply integrating proprietary financial data with product and platform infrastructure, we’re turning complex financial operations into simple, automated experiences and setting a new standard for how modern finance works. What you’ll do You'll be a product engineer building Brex's Audit Agent — an agentic system that reviews customer spend at scale and replaces the manual work traditionally done by BPO teams. The agent itself reasons; the surrounding product harness is what makes that reasoning useful, trustworthy, and operable for real customers. That product harness is where you'll live. You'll desig
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is about empowering enterprises to achieve their full potential — and people too. With a culture that’s all in on impact, innovation, and collaboration, Snowflake is the sweet spot for building big, moving fast, and taking technology — and careers — to the next level. ABOUT THE ROLE: You will work on critical business initiatives in the core database engine, bring an AI-forward approach to software development and accelerate roadmap cycles for the benefit of our customers. Your work will directly impact how developers and businesses build with data. You'll own the full AI engineering lifecycle: design, prompt/tool engineering, evals, deployment, measurement, and optimization. You'll work with a small, high-powered engineering team. What you will do in this role: Own features end-to-end for Snowflake Database Engineering products. Build agentic workflows, coding harnesses, evaluation pipelines. Build enterprise-grade context engineering: function calling, tool schemas, guardrails, agent teams, and verification/repair. Design evals and hillclimb : create golden sets, create rubrics and metrics, analyze errors, run experiments to hill climb on the metrics. Partner with product and infra: translate customer problems into products and experiments. Collaborate with infr
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Cortex CoWork team is defining the future of AI for enterprise data. Our mission is to transform how the world’s largest enterprises interact with their data through flagship products like Snowflake (CoWork) Intelligence . As a Principal AI Engineer , you will be a technical North Star for our AI initiatives. You won't just execute on a roadmap; you will help define it. You will tackle the most complex, "frontier" problems in agentic reasoning, NL-to-SQL, and enterprise-scale RAG, ensuring our AI products are not only innovative but fundamentally reliable and scalable for the Fortune 500. What you will do in this role: Technical Strategy & Architecture: Define the long-term technical vision for Snowflake Intelligence. Lead the architectural design of multi-agent systems, complex tool-use frameworks, and self-correcting NL-to-SQL engines. Drive Industry-Leading Reliability: Move beyond simple evals to build world-class, automated "hill-climbing" infrastructure. You will establish the methodology for how Snowflake measures and guarantees LLM performance across diverse customer schemas. Cross-Functional Influence: Partner with Product and Engineering leadership to align AI capabilities with business goals. You will bridge the gap between Research (modeling) and Product
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As the Head of AI Platform Engineering at Postman, you will lead the alignment of AI development with our growing API platform. You will drive the AI roadmap with a focus on expanding AI-driven API collaboration and agentic capabilities across the platform. This role requires a leader who can identify market opportunities, coordinate cross-functional AI initiatives, and foster strong partnerships to amplify the Postman AI platform's impact What You’ll Do Lead the development and execution of Postman’s AI platform strategy, focused on API ecosystem growth and platform innovation. Drive the AI roadmap, concentrating on API integration, platform expansion, and AI-driven agent functionality. Identify and capitalize on market opportunities for AI-enhanced API collaboration and intelligent agent features. Collaborate closely with business units, product teams, engineering, and external partners to ensure alignment and successful AI initiatives deployment. Oversee implementation with core AI platforms (OpenAI, Anthropic, AWS, etc)), ensuring technical and strategic alignment with AI features and API lifecycle improvements.
Other cities to consider
More places hiring for this role
Get new agent ai engineer jobs in United States by email
Daily job updates · Unsubscribe anytime