About the Team The Agent Infrastructure team at OpenAI is responsible for building systems that enable training and deployment of highly useful AI agents, both internally and for the world. We work hand-in-hand with researchers to design and scale the environment in which agentic models are trained – providing a workspace for AI models to execute code, debug issues, and develop software just as human SWEs do. Our training environment for agentic models operates at an extremely high scale and has the flexibility to emulate any environment in which an agent might work. At the same time, our team builds and maintains OpenAI’s core platform for the deployment and execution of agents in production. Our systems power products such as Codex, Operator, tool use in ChatGPT, and future agentic products. Some of the most challenging technical problems in scaling the capabilities and utility of agents and agentic models lie in the infrastructure layer – and our team is focused on building the research and production systems that enable OpenAI to train the most capable models in the world, and maximize the utility of our agentic products for users around the world. About the Role As a Software Engineer on the Agent Infrastructure team, you will have the opportunity to work closely with both research and product at OpenAI - building and scaling systems to train highly capable agentic models, and building the platform and integrations to launch new agents to hundreds of millions of users worldwide. Your work will consist of both building new capabilities - standing up the infrastructure and integrations needed to train more complex agentic models - and rapidly scaling these new capabilities to some of the largest compute clusters in the world. At the same time, you’ll be instrumental to the launch of agentic products at OpenAI - building, maintaining, and scaling the production platform on which all agents run. We’re looking for people with deep experience building AI infrastructu
Jobs in United States
Ai Agent Engineer in United States
5,246 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai agent engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The team’s mission is to accelerate the secure evolution of agentic AI systems at OpenAI. To achieve this, the team designs, implements, and continuously refines security policies, frameworks, and controls that defend OpenAI’s most critical assets—including the user and customer data embedded within them—against the unique risks introduced by agentic AI. About the Role As a Security Engineer on the Agent Security Team , you will be at the forefront of securing OpenAI’s cutting-edge agentic AI systems. Your role will involve designing and implementing robust security frameworks, policies, and controls to safeguard OpenAI’s critical assets and ensure the safe deployment of agentic systems. You will develop comprehensive threat models, partner tightly with our Agent Infrastructure group to fortify the platforms that power OpenAI’s most advanced agentic systems, and lead efforts to enhance safety monitoring pipelines at scale. We are looking for a versatile engineer who thrives in ambiguity and can make meaningful contributions from day one. You should be prepared to ship solutions quickly while maintaining a high standard of quality and security. We’re looking for people who can drive innovative solutions that will set the industry standard for agent security. You will need to bring your expertise in securing complex systems and designing robust isolation strategies for emerging AI technologies, all while being mindful of usability. You will communicate effectively across various teams and functions, ensuring your solutions are scalable and robust while working collaboratively in an innovative environment. In this fast-paced setting, you will have the opportunity to solve complex security challenges, influence OpenAI’s security strategy, and play a pivotal role in advancing the safe and responsible deployment of agentic AI systems. You’ll be responsible for: Architecting security controls for agentic AI – design, implement, and iterate on identity, netwo
About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role We’re seeking an exceptional Principal-level Offensive Security Engineer focused on deep, hands-on penetration testing of OpenAI’s agent-powered products, infrastructure, and model-integrated application surfaces. You’ll assess complex systems end to end, identify realistic vulnerabilities, validate exploitability and impact, and partner closely with engineering teams to drive durable fixes. This role will be primarily focused on continuously testing our agent-powered products like Codex and Operator. These systems are uniquely valuable targets because they’re rapidly evolving, can perform sensitive actions on behalf of users, and have large, diverse attack surfaces. You will play a crucial role in securing our agents by finding vulnerabilities that emerge from the interactions between the applications, infrastructure, tools, and models that power them. You’ll have the chance to not only find vulnerabilities, but actively drive their resolution, build reusable testing approaches, automate offensive security workflows with cutting-edge technologies, and use your attacker perspective to improve the security of OpenAI’s products. In this role you will: Conduct deep penetration tests of OpenAI’s agent-powered products, including web applications, APIs, cloud services, identity and authorization flows, CI/CD systems, and model-integrated product surfaces. Continuously hunt for exploitable vulnerabilities in the interactions between the appli
AI Systems Engineer - Codex Core Agents About The Team The Codex Core Agents team builds the agent harness that turns model capability into real-world action. We own the systems around the model: prompting and interpreting model outputs, executing actions safely in real environments, and feeding production experience back into better models and better agent behavior. This team sits close to research and works across the stack: harness, model interaction, inference, sandboxed execution, orchestration, evals, production reliability, and the performance envelope around tokens, latency, cost, capacity, and quality. The harness is open source and increasingly part of how models are trained and evaluated, making this one of the highest-leverage layers in Codex. About The Role We’re looking for engineers to build the AI systems that make Codex agents dependable in production. The ideal candidate is an agent-systems builder: hands-on across low-level systems and ML workflows, able to debug Codex behavior end to end across the harness, model behavior, inference/runtime stack, GPU fleet, and product surface. You’ll work with research, infrastructure, and product to design agent harness capabilities, run experiments and ablations across the model + system prompt + harness stack, build frameworks for assessing production agent performance, and turn messy failures into durable improvements. What You’ll Do Design and build the core agent harness and execution loop that lets Codex agents interpret model outputs, use tools, execute code, and complete long-horizon tasks safely. Build sandboxing, isolation, orchestration, state, and workflow infrastructure for agents operating in real development environments. Develop evaluation, experimentation, and debugging systems that distinguish harness issues, model behavior, inference/runtime issues, and product failures. Run ablations across prompts, model-facing interfaces, context construction, tool-use strategies, and harness behavior to
We are investing in agentic AI and need a Senior AI Engineer to lead the design and delivery of these systems. This is a foundational hire: you will own both the agent-facing workstreams — pipelines, orchestration, conversational interfaces — and the underlying context layer that makes them reliable, including memory management, knowledge graph integration, and retrieval infrastructure. You will work closely with data engineers, project leads, and client stakeholders, and play a key role in shaping how Lynx builds and ships AI solutions at scale. What This Involves: Lead the architecture and delivery of agentic AI systems end-to-end: agents, orchestration, tool use, and multi-step reasoning workflows. Own the context layer: design and implement memory architectures (episodic, semantic, working memory) and integrate GraphRAG and knowledge graph retrieval into agentic pipelines. Build robust RAG systems — including vector retrieval, graph traversal, and hybrid search — and ensure retrieval quality through evaluation frameworks. Translate client requirements into technical designs, presenting approaches and trade-offs to both technical and non-technical stakeholders. Define standards and reusable patterns for agentic AI development that other engineers at Lynx can build on. Set up observability, evaluation, and monitoring pipelines to ensure AI systems perform correctly in production. Requirements: 5–8 years of software or ML engineering experience, with at least 2–3 years building LLM-based or agentic AI systems in production. Deep hands-on experience with agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI, or similar) and LLM APIs (OpenAI, Anthropic, etc.). Strong understanding of agent design patterns: ReAct, planning loops, tool use, multi-agent coordination, and memory architectures. Practical experience with GraphRAG or knowledge graph-based retrieval (e.g., Neo4j, Microsoft GraphRAG) and vector databases (Pinecone, Weaviate, Qdrant, etc.). Proficiency in
We’re looking for a Staff Software Engineer to help shape AI governance for developer tooling at Coder. This role sits on our AI Governance team, which builds and maintains two enterprise-grade components of Coder's AI governance stack. AI Gateway is a centralized LLM gateway that sits between coding agents and providers such as OpenAI or Anthropic, providing organizations with audit trails, token tracking, cost control, and centralized authentication. Agent Firewall wraps those agents with default-deny network policies, controlling which domains and methods they can reach inside workspaces. This team works across the full stack - from Go backend and React frontend to integrating with LLM provider APIs. Day to day, you'll be shipping features, hardening security boundaries, collaborating with enterprise customers on real-world policy needs, and contributing to Coder's open-source ecosystem. What you'll do here Design and build product features that push the standard for remote development in self-hosted environments Create and improve upon popular open source projects that integrate with VS Code, JetBrains, and other developer tools Champion best practices to both internal team members and external contributors Collaborate with Product and Design teams at Coder, as well as with partners like JetBrains, to execute key product integrations Document the design, implementation, and operations of systems for knowledge sharing within the team Work alongside Customer Success teams to support Coder’s enterprise user base Work with cutting-edge AI technologies to create seamless, painless developer experiences Rapidly iterate from prototype to implementation in a highly adaptive, reactive team environment What we're looking for 8+ years of full-stack experience writing code in a professional setting, with 1+ year(s) writing Go (ideally in current or most recent position) Proficiency in building distributed systems in Go Excellent verbal and written communication skills Excep
About the Team OpenAI’s AI Success Engineer team partners with the world’s most ambitious government & partner organizations to translate cutting edge AI into real business and mission impact for governments of all levels from Local, State, Federal, and International. We guide customers and users journey from the first time they try ChatGPT Enterprise, automate a workflow, develop and execute a new skill, and create their first agent to scaled enterprise adoption of ChatGPT, Codex, our API and other novel capabilities. Our work spans technical integration and enablement, workflow transformation, inspiring and upskilling AI literacy and confidence across the workforce, sustained program, product and new capability delivery. Most importantly, we help each member of our customer's workforce, their teams, programs and missions meet their total potential. Our government customers have vital missions, and we must meet them with game-changing technology. Every engagement is an opportunity to shape how AI changes work, productivity, and innovation. This role sits at the center of that mission. About the Role Governments work at a scale that is truly exponential on missions that are of critical importance to people, communities and nations. The AI Success Engineer role is the primary post-sales relationship for OpenAI’s most important customers. You are responsible for the end-to-end account management of critical Government and Partner customers. You will be helping Government Leaders/Partners appropriately and effectively use AI for their mission, while simultaneously investing in ensuring their people are AI-enabled and ready to advance positive outcomes that their constituents depend on them for. You will drive: the impact of our tools on their mission, account health and adoption, ensuring technical readiness, creating and executing on the deployment strategy, enabling, educating and training their workforce, identifying new use cases and upsell opportunities, and d
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”
Become a part of our caring community The Lead Software Engineer codes software applications based on business requirements. The Lead Software Engineer works on problems of diverse scope and complexity ranging from moderate to substantial. The Lead Software Engineer standardizes the quality assurance procedure for software. Oversees testing and debugging and develops fixes. Researches complaints and makes necessary adjustments and/or recommendations to resolve complex software related issues. Advises executives to develop functional strategies (often segment specific) on matters of significance. Exercises independent judgment and decision making on complex issues regarding job duties and related tasks, and works under minimal supervision, Uses independent judgment requiring analysis of variable factors and determining the best course of action. Key Responsibilities Technical Architecture and Ownership:** Design and own the end-to-end architecture of Centerwell's AI systems, including LLM-powered clinical tools, RAG pipelines, harnesses, agent-based workflows, and intelligent automation. Make and communicate foundational technical decisions in close collaboration with the broader engineering team. Model Development and Fine-Tuning:** Evaluate, select, and where appropriate guide the fine-tuning of foundation models. Establish model evaluation frameworks that prioritize safety, accuracy, and clinical relevance. Clinical and Product Partnership:** Collaborate closely with product managers, designers, clinicians, and data stakeholders to understand care delivery workflows and translate them into well-scoped, high-impact AI features. HIPAA Compliance and Responsible AI:** Ensure all AI systems are designed, deployed, and monitored in compliance with HIPAA and Humana's Responsible AI standards, including participation i
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a rock star leader for our AI DevEx team. This team plays a crucial part in transforming how the Snowflake product is developed, evaluated and supported. The job is no less than building the ecosystem that powers our AI-pilled engineers and our agentic organizations and reinvents the SDLC for Snowflake. As the manager for this growing team, you will lead the group of engineers working to reinvent how code is written, reviewed and runs in production to usher scale and velocity for the AI era. With context as the new source code, you will make it possible for every engineer at Snowflake to deliver through intent and direction. In this role, you will: Lead, coach, and grow the ES AI DevEx team while creating a high-energy, cohesive environment with strong planning, ownership, and career development. Deliver the vision and roadmap for transforming the SDLC for all Snowflake engineers. Drive measurable improvements in code review latency, ensure Snowflake developers can use the latest harnesses and models, and create the golden paths that agentic codebases are built upon. Act as an agent of clarity in a space that is moving fast, making high quality decisions quickly by keeping up with the latest developments in the industry. Own operational excellence for the area
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Are you the person on your team who builds the agent everyone else ends up using? We're looking for an AI Engineer to join our Training Product team and do that at Baseten. You'll build AI-driven product features for the customers training and post-training frontier models on our platform, and you'll raise the ceiling on how Baseten itself uses AI internally, turning manual workflows into agentic ones that make every other team faster. You'll work directly with our research engineers to scope and build products, taking ideas from a research loop that already works internally to something customers can run themselves. This is a hands-on role with real autonomy. You'll pick the problems worth solving, build the harnesses, execution flows, and guardrails that make AI systems reliable, and own the results. If you've been shipping agents and want that to be the job, let's talk. EXAMPLE INITIATIVES: Take a look at these blog posts written by members of our team: Baseten Training: an autoresearch substrate Introducing Baseten Loops Harnesses are everything. Here's how to optimize yours. Building with NVIDIA Nemotron 3 Ultra and LangChain Deep Agents Code on Baseten RESPONSIBILITIES: Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to GA. Design the harnesses, execution flows, and guardrails that make AI systems reliable in production. Build internal autom
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. At Ema, we build AI Employees that operate inside the enterprise. Healthcare is where the bar is highest: the output has to be accurate, auditable, and clinically sound. We're opening a part-time role focused on agent measurement and improvement. You'll instrument how our healthcare AI Employees perform in production, identify where quality breaks down, and design the experiments that close the gap — partnering with clinical experts to ensure improvements translate into better patient outcomes, not just better metrics. We're looking for strong analytical judgment (Python, agent lead development), comfort operating with incomplete information, and genuine interest in the healthcare domain. Clinical experience is valued but not required. This engagement is structured as a paid internship or contract engagement, with weekly syncs in our Bay Area office. Compensation offered will be determined by factors such as location, level, job-related knowledge, skills, and experience. Certain roles may be eligible for variable compensation, equity, and benefits. Ema Unlimited is an equal opportunity employer and is committed to providing equal employment opportunities to all employees and applic
From $152.8K/yr
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role The DAP Repository Flows team owns two areas within GitLab's Duo Agent Platform (DAP). The first is DAP onboarding , helping customers quickly set up a project so they can start using the Duo Agent Platform with minimal friction. The second is repository and background flows : autonomous and scheduled agents that work on a customer's codebase on their behalf. This work is central to how GitLab delivers AI-powered automation on repositories. As a Staff Software Engineer, you will set technical direction for the team, drive complex initiatives across team boundaries, and raise the engineering bar for how we
Other cities to consider
More places hiring for this role
Get new ai agent engineer jobs in United States by email
Daily job updates · Unsubscribe anytime