Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity (Summer 2026 AI Internship - Applications Open Now) We're seeking an AI Engineer Intern to work alongside our AI team on large-scale AI and Agentic systems from data pipeline to production deployment. This role is scoped for someone with foundational experience who wants to deepen it: you'll own discrete pieces of real systems under the mentorship of senior engineers, not shadow work or isolated coursework-style projects. What You'll Do You’ll work directly with the AI team, taking responsibility for well-scoped pieces of real systems, with mentorship from senior engineers. Benchmarks & Evaluation Contribute to APIFlow-Bench , our open-source benchmark for real API-development work: design and review benchmark tasks and their mock API environments, extend the evaluation harness and task-generation pipeline in Python, and help maintain the public multi-model leaderboard with statistical confidence intervals. Help build a new action-level AI safety benchmark: instead of grading what a model says, it scores what an agent actually does inside a simulated enterprise API environment. You’ll work on scenari
Jobs in United States
Ai Deployment Engineer in United States
5,082 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai deployment engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. About Our Team Our team builds and enables scalable, cloud-native software platforms that power critical manufacturing and business operations. We use modern full stack engineering practices and AI-driven technologies to create innovative solutions that improve automation, decision making, and operational efficiency. Position Overview We are seeking an AI-Centric Full Stack Software Engineer to design, build, and support modern software solutions with a strong focus on AI-enabled applications and intelligent automation. This role goes beyond simply using AI coding assistants. You will have strong understanding of Prompt Engineering, Vibe Coding, Rework Rate Reduction and leverage Custom Agents, and integrations built around the Model Context Protocol (MCP). You will apply strong software engineering fundamentals to assemble, integrate, and operationalize AI capabilities into real production systems being accountable for Full-Stack solutions. The goal of this role is to significantly shorten development and feedback cycles by automating routine engineering work, augmenting human decision-making, and embedding intelligence directly into our development workflows and the applications we deliver! Responsibilities Design, develop, test, deploy, and maintain scalable full stack applications and platform services. Own software solutions end-to-end, including system design, architecture, implementation, testing, observability, deployment, and lifecycle support. Develop an
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. About Our Team Our team builds and enables scalable, cloud-native software platforms that power critical manufacturing and business operations. We use modern full stack engineering practices and AI-driven technologies to create innovative solutions that improve automation, decision making, and operational efficiency. Position Overview We are seeking an AI-Centric Full Stack Software Engineer to design, build, and support modern software solutions with a strong focus on AI-enabled applications and intelligent automation. This role goes beyond simply using AI coding assistants. You will have strong understanding of Prompt Engineering, Vibe Coding, Rework Rate Reduction and leverage Custom Agents, and integrations built around the Model Context Protocol (MCP). You will apply strong software engineering fundamentals to assemble, integrate, and operationalize AI capabilities into real production systems being accountable for Full-Stack solutions. The goal of this role is to significantly shorten development and feedback cycles by automating routine engineering work, augmenting human decision-making, and embedding intelligence directly into our development workflows and the applications we deliver! Responsibilities Design, develop, test, deploy, and maintain scalable full stack applications and platform services. Own software solutions end-to-end, including system design, architecture, implementation, testing, observability, deployment, and lifecycle support. Deve
Become a part of our caring community The Senior Full Stack Engineer Performs software engineering activities in all layers of the stack, from setting up the database to programming in the back-end and the appearance at the front-end. The Senior Full Stack Engineer work assignments involve moderately complex to complex issues where the analysis of situations or data requires an in-depth evaluation of variable factors. As Centerwell builds its AI engineering function from the ground up, we need a platform foundation strong enough to support everything that comes next. As Lead Full-Stack Engineer focused on platform and API engineering, you will design and build the service layer that connects AI capabilities, data systems, and product frontends—setting the standards for how services are built, secured, and operated across the team. You will work with meaningful architectural scope, making decisions that span API design, security patterns, and deployment practices. The platform you build will serve care teams and patients across hundreds of Centerwell clinics. If you want to build platforms that others build on—and do it in service of better primary care—this is the role for you. Key Responsibilities Platform and API Architecture: ** Design and lead development of core backend services, REST and GraphQL APIs, and service-to-service integrations that connect all layers of Centerwell's AI product stack. Security and Compliance by Design: ** Establish patterns for authentication, authorization, rate limiting, PHI access control, and audit logging. Ensure HIPAA compliance is embedded in platform design from day one—not bolted on after the fact. AI and LLM Integration Patterns: ** Define and implement reusable patterns for integrating AI capabilities into product
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. The AI Engineer, People Technology is responsible for designing, developing, and deploying AI-powered solutions that transform work across the People Organization. This role combines software engineering, AI development, and HR domain expertise to build intelligent applications, automations, and agents that improve employee experiences, increase operational efficiency, and accelerate workforce transformation. The ideal candidate is a hands-on builder with demonstrated experience using AI Assisted Software Development Platforms to develop enterprise AI solutions. You must have successfully designed and deployed AI agents that collaborate across multiple platforms, systems, and business functions while operating within enterprise governance, security, and compliance standards. This role requires deep knowledge of HR technologies, including Workday, ServiceNow, and the Microsoft Copilot ecosystem, along with a passion for applying AI to solve complex business challenges. Responsibilities: AI Product & Solution Development: Experience with the end-to-end product lifecycle turning a vision into a roadmap while driving adoption and value delivery. Design, develop, test, and deploy AI-powered products, applications, and intelligent workflow solutions. Apply AI Assisted Development Tools to accelerate software development, solution building, and deployment activities. Support authorities in translating business needs into developed solutions and prototypes. Develop reusable frameworks, ser
Become a part of our caring community Most AI engineering jobs are a thin wrapper around a model API. This role is different. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to the source document, and route ambiguous cases to human experts for review. Our users make decisions that impact real healthcare outcomes, so “good enough” is not good enough. Building AI systems that are accurate, reliable, auditable, and scalable is at the core of this role. As a Senior AI Applied Engineer, you will design, build, deploy, and operate production AI systems used at scale within one of the largest health insurers in the United States. You will own solutions end-to-end, from user experience and APIs to model orchestration, evaluation frameworks, infrastructure, and production operations. Why Join Us Build production AI systems where LLMs are in the critical path, not just demos or proofs of concept. Work on extraction, retrieval, agentic workflows, and human-review systems that process real healthcare data at scale. Own projects end-to-end across frontend, backend, AI orchestration, infrastructure, deployment, and operations. Solve challenging problems around accuracy, explainability, traceability, and reliability in regulated environments. Ship quickly in a small, high-impact team that embraces AI-assisted development and rigorous quality standards. Build systems that continuously improve through expert feedback, evaluations, and human-in-the-loop workflows. Key Responsibilities Design, develop, and deploy full-stack AI-powered application
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is about empowering enterprises to achieve their full potential — and people too. With a culture that’s all in on impact, innovation, and collaboration, Snowflake is the sweet spot for building big, moving fast, and taking technology — and careers — to the next level. ABOUT THE ROLE: You will work on critical business initiatives in the core database engine, bring an AI-forward approach to software development and accelerate roadmap cycles for the benefit of our customers. Your work will directly impact how developers and businesses build with data. You'll own the full AI engineering lifecycle: design, prompt/tool engineering, evals, deployment, measurement, and optimization. You'll work with a small, high-powered engineering team. What you will do in this role: Own features end-to-end for Snowflake Database Engineering products. Build agentic workflows, coding harnesses, evaluation pipelines. Build enterprise-grade context engineering: function calling, tool schemas, guardrails, agent teams, and verification/repair. Design evals and hillclimb : create golden sets, create rubrics and metrics, analyze errors, run experiments to hill climb on the metrics. Partner with product and infra: translate customer problems into products and experiments. Collaborate with infr
About the Team The Recursive Self-Improvement (RSI) team works across research, engineering, product, and infrastructure to build AI systems that accelerate and ultimately conduct high-quality research at OpenAI. We work to automate real research workflows and improve research productivity by building systems and feedback loops, designing evaluations, and training models to develop missing capabilities. Our work spans the full lifecycle of model training, evaluation, and deployment to help researchers move faster and tackle increasingly ambitious problems. About the Role We’re hiring research scientists , research engineers , and AI systems engineers to work on automating research at OpenAI. This role is based in San Francisco, CA. In this role, you will: Design evaluations for research judgment, hypothesis generation and testing, and long-horizon experiment execution. Turn real research workflows and model failures into data and evaluation flywheels. Improve model research capabilities through agent harnesses, synthetic data, RL environments, and model training. Build and maintain safe, reliable integrations between our models and OpenAI’s research infrastructure. Develop research agents, experiment-orchestration systems, and sandboxed runtimes that support real research workflows. Create metrics and economic models to understand RSI’s current and future effects on research productivity, model capabilities, and the safety of internal deployments. This is a high-ownership role for researchers and engineers who thrive in ambiguity, move fluidly between research and implementation, and turn emerging opportunities into rigorous, reliable, scalable results. You might thrive in this role if you: Have research or engineering experience across LLM training, model evaluations, agent systems, synthetic data, research infrastructure, or large-scale distributed systems. Are a strong generalist who can move between open-ended research and practical implementation, turning ambig
OpenAI’s charter calls on us to ensure the benefits of AI are distributed broadly and safely. Our Health AI team focuses on expanding access to high-quality medical expertise and aims to set a high standard for deploying AI responsibly in high-stakes domains. Improving health will be one of the defining impacts of AGI. Today, millions of people lack access to reliable medical information, and clinicians around the world face increasing time and resource constraints. We are building AI systems that support patients, clinicians, and health workers, while meeting the highest standards for safety, reliability, and privacy. We are seeking full stack software engineers to help build and scale products used by consumers and care providers globally. You will work closely with product, design, and research teams to ship real systems in a fast-moving, high-impact environment. In this role, you will: Design and build scalable fullstack systems for consumer and enterprise health. Own end-to-end feature development—from early design and implementation through deployment, monitoring, and iteration. Build and maintain data pipelines and services that meet strict privacy, security, and compliance requirements (e.g., HIPAA). Collaborate closely with researchers and safety teams to integrate reliability, evaluation, and guardrails into production systems. Debug, optimize, and harden systems to support high availability, performance, and global scale. Take ownership of ambiguous problems and drive them to practical, high-quality solutions. You might thrive in this role if you: Are deeply motivated by improving health outcomes and expanding access to medical expertise. Are a strong engineer who enjoys building durable, well-designed systems. Have 5+ years of experience writing maintainable, production-quality code. Can operate with high agency—owning problems end-to-end with minimal supervision. Enjoy working in fast-moving, cross-functional teams with engineers, product managers, desi
About the Team The Applied AI Engineering team partners closely with customers to help them move from experimentation to production with OpenAI’s technologies. We act as trusted technical advisors, working across customer strategy, architecture, deployment, and adoption to help organizations realize meaningful impact from frontier AI. The Startups segment serves fast-moving, high-growth companies that are often building new products, workflows, and businesses directly on top of AI. These customers move quickly, operate with high ambiguity, and expect practical, creative, and technically rigorous partnership. About the Role We are looking for an Applied AI Engineering Manager, Startups to lead and scale the Startups Applied AI Engineering motion. This team helps high-growth startups move quickly from experimentation to production, unlock meaningful usage, and build durable technical partnerships with OpenAI. This leader will operate in a high-velocity customer segment where founders, CTOs, and technical teams expect speed, judgment, and hands-on problem-solving. They will balance team leadership, technical depth, customer prioritization, and cross-functional influence across Sales, Product, Engineering, Research, and broader go-to-market teams. In this role, you will define how OpenAI supports startup customers at scale: identifying where deep technical engagement can unlock outsized impact, building repeatable deployment mechanisms, and ensuring the team can serve a broad and dynamic customer base without losing quality or strategic focus. In this role, you will: Craft and continuously refine the strategic vision and operating model for the Startups Applied AI Engineering team, aligning it with OpenAI’s broader company objectives and the evolving needs of high-growth startup customers. Lead, mentor, and grow a team of high-performing technical ICs supporting startup customers across AI-native, developer-led, and product-led companies. Help startups move from early e
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's engineers want to work in an AI-first way. What's missing isn't enthusiasm — it's the platform underneath it. Today everyone assembles their own agent config, context files, and MCP servers, so the good patterns stay trapped in individual setups instead of becoming defaults everyone inherits. You'll build that platform: the agent configurations tuned to our monorepo, the context and tooling layer that makes agents competent in our codebase, the evals that tell us which approaches actually work, and the rollout mechanics that get a new engineer productive with agents in week one. You are not here to mandate how engineers use AI — you're here to make the good path the easy path. Success looks like teams adopting what you build because it beats what they'd cobble together themselves, not because a policy requires it. Platform engineer, not AI evangelist. Ship infrastructure, measure it, kill what doesn't work, let adoption be the referee. The playbook for AI-first SDLC doesn't exist at any company yet. You'll write ours. WHAT YOU'LL BUILD Agent substrate — Repo-level context infrastructure that makes agents competent in our codebase ( CLAUDE.md/AGENTS.md conventions, architecture and domain context, and the tooling to keep it accurate as code moves). Internal MCP servers giving agents scoped access to CI, observability, incident tooling, deployment state, and docs. Shared skills, subagents, and hooks th
About the Team OpenAI’s Infrastructure Operations team is responsible for the availability, reliability, and operational excellence of one of the world’s largest AI infrastructure networks. The team owns day-to-day operations of production AI networks across Industrial Compute's data centers, working with colocation providers, deployment teams, and hardware vendors to deliver highly available GPU infrastructure for AI training and inference workloads. About the Role We are seeking an Infrastructure Operations Engineer to operate and improve the large-scale Ethernet fabrics that support GPU clusters, storage systems, and management infrastructure. This role combines hands-on production operations with automation, observability, and incident response across a global AI network. The ideal candidate has experience operating high-availability data center, cloud, AI, or HPC networks and can move comfortably from physical-layer troubleshooting to routing and fabric behavior, change execution, and root-cause analysis. You will partner closely with network architecture, systems engineering, GPU engineering, storage engineering, security, deployment, site operations, service providers, colocation partners, and hardware vendors to raise reliability and reduce operational toil. Key Responsibilities Own the operational health, availability, and reliability of production AI network infrastructure across Industrial Compute's data centers. Monitor, troubleshoot, and resolve network incidents while meeting service-level objectives (SLOs), reducing Mean Time to Detect (MTTD), and minimizing Mean Time to Recovery (MTTR). Operate and maintain large-scale Ethernet fabrics supporting GPU compute, storage, and management networks. Execute production network changes, maintenance windows, and capacity expansions with minimal customer impact. Manage the hardware lifecycle, including switch and optics replacements, RMA coordination, software upgrades, and preventive maintenance. Support new A
About the Team The Safety Systems team is dedicated to ensuring the safety, robustness, and reliability of AI models and their deployment in the real world. Learn more about OpenAI’s approach to safety. Building on the many years of our practical alignment work and applied safety efforts, Safety Systems addresses emerging safety issues and develops new fundamental solutions to enable the safe deployment of our most advanced models and future AGI, to make AI that is beneficial and trustworthy. About the Role At OpenAI, we're dedicated to advancing artificial intelligence, and we know that creating a secure and reliable platform is vital to our mission. That's why we're seeking a software engineer to help us build out our trust and safety capabilities. In this role, you'll work with our entire engineering team to design and implement systems that detect and prevent abuse, promote user safety, and reduce risk across our platform. You'll be at the forefront of our efforts to ensure that the immense potential of AI is harnessed in a responsible and sustainable manner. Your Responsibilities: Architect, build, and maintain anti-abuse and content moderation infrastructure designed to protect us and end users from unwanted behavior. Work closely with our other engineers and researchers to utilize both industry standard and novel AI techniques to measure, monitor and improve AI models’ alignment to human values. . Diagnose and remediate active incidents on the platform and build new tooling and infrastructure that address the root causes of system failure. You might thrive in this role if: You have built and run production services in a high growth, rapidly scaling environment. You can debug live issues and restore systems quickly. You have worked on content safety, fraud, or abuse, or are motivated and excited to work on present-day (“now-term”) AI safety. You have experience with Python or with modern languages such as C++, Rust, or Go, and are able to quickly ramp up on Py
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Data Clean Rooms team is Leading the market shift from traditional 2-party data sharing to multi-party collaboration hubs . Our vision is to provide a seamless, "safe-room" environment where enterprises can collaborate on shared datasets while maintaining absolute governance. We ensure that no party can exfiltrate another's underlying content, even while running complex joint workloads and getting high-value results. You will join a fast-paced, collaborative team of engineers on a journey to provide customers with an integrated set of innovative, AI-enabled capabilities to analyze data in a privacy-preserving way. You will have a real opportunity to impact and shape the future of secure data collaboration at Snowflake. AS A SOFTWARE ENGINEER IN DATA CLEAN ROOMS, YOU WILL: Architect and build highly scalable infrastructure that enables secure, multi-party collaboration. Design and implement core clean room features and services, intelligent agents, and robust developer APIs to expand platform capabilities and support custom AI/ML workflows. Partner closely with Product Management and cross-functional teams to drive complex projects from ideation and system design through to production deployment. Mentor peers and foster a warm, supportive culture of innovation, cross-tea
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Imagine stepping into a role where your code directly empowers the world's largest enterprises to safely adopt superintelligence. As a software engineer, generative ai at WRITER, you'll be at the forefront of expanding human capacity by building the secure, scalable foundation that allows our generative AI solutions to thrive in complex corporate environments. The impact of this work is massive – for example, in the consumer packaged goods industry alone, our AI adoption is driving 69% revenue increases and 72% cost reductions. This role is designed for a well-rounded engineering generalist who leans heavily into generative AI while bringing a whole-systems mindset to architectural design. If you thrive on proactivity without red tape and love owning projects from proposal to deployment, you'll shape the future of AI and contribute to a product that’s changing how the world works. This is a hybrid role based out of our San Francisco, New York City, or London hu
Other cities to consider
More places hiring for this role
Get new ai deployment engineer jobs in United States by email
Daily job updates · Unsubscribe anytime