Jobs in United States

Ai Deployment Engineer in United States

5,082 active opportunities · Updated October 2026

Explore current ai deployment engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

25/100

cooling · 8 related jobs

Hiring trend

-40%

Job postings compared with the previous 30 days

Remote options

37.5%

Share of matching jobs listed as remote

SF
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $200K/yr

Quick readStrong listing-quality and freshness signals

About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role At Stitch Fix, we are at the forefront of innovation, creating cutting-edge solutions that blend fashion, technology, and data science. Our data science team combines machine learning with expert human judgment to generate innovative recommendations and insights that transform the way our clients discover what they love. We believe in a curiosity-driven data science culture where members are empowered to deliver impact through end-to-end model development. The diversity of the problems that we work on and the data-rich environment of our business make it possible, even essential, to bring the tools of multiple disciplines to bear on our hardest problems. We are looking for an experienced Styling Algorithms Team Manager to lead a group of talented machine learning engineers and data scientists. In this role, you will shape the future of fashion technology by driving the development and deployment of our styling algorithms, which empower our human stylists to delight clients by nailing their fit and style. This includes ML-, AI-, and product-driven feature curation and testing for our proprietary styling platform, as well as client-facing AI personalization experiences, such as Stitch Fix Vision, our virtual try-on. Responsibilities: Champion bold AI and ML interventions to improve our styling experiences, enabling our stylists to have a multiplicative impact on their client connection points. Likewise, actively shape the product roadmap for direct client-facing styling experiences, expand

PythonRestMachine LearningAI
O
📍 United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. The Industrial Compute team is central to this mission, setting the core infra strategy and implementing this vision. From site selection to the buildout process, this team sits at the intersection of commercial, technical, strategy, and operations, interacting with teams and executives inside and outside of OpenAI. About the Role Responsible for validating that proposed sites are buildable, compliant, and cost-effective. You will lead diligence across civil, geotechnical, environmental, and entitlement dimensions, identifying risks and driving mitigation strategies. Key Responsibilities Lead all technical diligence: geotech, soils, title/ALTA surveys, mineral rights, and access. Oversee permitting/entitlement path and schedule governance with agencies. Evaluate generator air permits, wetlands, floodplain, and stormwater constraints. Manage consultants performing feasibility studies and environmental assessments. Deliver go/no-go recommendations with risk and mitigation options. Build diligence templates and playbooks to scale future site reviews. Qualifications 8+ years in land development, civil/environmental engineering, or data center diligence. Knowledge of permitting, entitlements, and AHJ engagement. Strong project management and technical review skills. Experience managing consultants and interpreting complex studies. Regularly communicate site readiness updates, risks, and milestones to executive stakeholders Establish and track key performance indicators to assess the effectiveness of the site selection program and the contributions of external vendors and partners. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely depl

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Synthetic RL team develops reinforcement learning methods that leverage synthetic data, environments, and feedback to train and evaluate frontier AI models. The team explores approaches such as self-play, simulators, and other synthetic evaluations to push model capability, generalization, and alignment beyond what is possible with the current prevailing methodology. About the Role As a Research Scientist on the Synthetic RL team, you will develop novel reinforcement learning techniques that use synthetic environments and feedback to improve large-scale models. You’ll work closely with other researchers to design experiments, analyze learning dynamics, and translate research insights into training approaches used in production systems. We’re looking for researchers who enjoy working on open-ended problems, value fast iteration, and want their work to directly shape how frontier models are trained. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Research and develop reinforcement learning algorithms Design and run experiments to study training dynamics and model behavior at scale Collaborate with engineers and researchers to integrate successful approaches into model training pipelines You might thrive in this role if you: Have a strong background in reinforcement learning, machine learning research, or related fields Have strong engineering and statistical analysis skills Enjoy exploring new problem spaces where data, objectives, and evaluation are imperfect or evolving Are motivated by seeing research ideas influence real-world AI systems About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an ex

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Forward Deployed Engineering (FDE) team turns research breakthroughs into production-grade systems. We embed deeply with customers to solve high-leverage problems and act as the delivery engine for our most complex large-scale engagements. We move quickly from prototype to production and surface reusable patterns that shape our platform. We operate at the intersection of deployment and development – working closely with OpenAI Research, Product and Partnerships. About the Role As a Technical Deployment Lead (TDL), you will define how OpenAI delivers complex systems to customers. You will own how they are built, shipped, and adopted. You’ll translate business outcomes into a technical plan, run day-to-day execution across FDEs, Researchers, and Customer Engineers, and partner with customer teams to ensure delivery supports their goals. You will own delivery end-to-end: embedding with customers to map workflows and success criteria, ensuring components ship on time, and leading readiness and change management for adoption. You’ll track progress, manage dependencies, make sequencing decisions, and drive 0→1 prototypes through MVP and scale. You will also share field insights with Product and Research to guide roadmap and priorities. Success will be measured first and foremost by impact - deployments that deliver measurable value against customer goals, drive adoption, and become critical to their workflows. Additional measures of success include delivery reliability (milestones hit, low reopen/churn), operating leverage (patterns reused across deployments), judgment under pressure, and product impact (field signal that shifts roadmaps/architectures). This is a high-trust, high-autonomy role. Success requires deep technical project management expertise, extreme ownership of outcomes, and an ability to immerse in customer workflows and partner with customer teams to solve complex engineering problems at pace. This role is based in San Francisco. W

AWSRestAIGo
O
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Forward Deployed Engineering (FDE) team turns research breakthroughs into production-grade systems. We embed deeply with customers to solve high-leverage problems and act as the delivery engine for our most complex large-scale engagements. We move quickly from prototype to production and surface reusable patterns that shape our platform. We operate at the intersection of deployment and development – working closely with OpenAI Research, Product and Partnerships. About the Role As a Technical Deployment Lead (TDL), you will define how OpenAI delivers complex systems to customers. You will own how they are built, shipped, and adopted. You’ll translate business outcomes into a technical plan, run day-to-day execution across FDEs, Researchers, and Customer Engineers, and partner with customer teams to ensure delivery supports their goals. You will own delivery end-to-end: embedding with customers to map workflows and success criteria, ensuring components ship on time, and leading readiness and change management for adoption. You’ll track progress, manage dependencies, make sequencing decisions, and drive 0→1 prototypes through MVP and scale. You will also share field insights with Product and Research to guide roadmap and priorities. Success will be measured first and foremost by impact - deployments that deliver measurable value against customer goals, drive adoption, and become critical to their workflows. Additional measures of success include delivery reliability (milestones hit, low reopen/churn), operating leverage (patterns reused across deployments), judgment under pressure, and product impact (field signal that shifts roadmaps/architectures). This is a high-trust, high-autonomy role. Success requires deep technical project management expertise, extreme ownership of outcomes, and an ability to immerse in customer workflows and partner with customer teams to solve complex engineering problems at pace. This role is based in NYC. We use a hy

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team The Applied AI Engineering team is responsible for ensuring the safe and effective deployment of Generative AI applications for developers and startups. We act as a trusted advisor and thought partner for our customers, working to build an effective backlog of GenAI use cases for their industry and drive them to production through strong technical guidance. OpenAI's customers represent a range of diverse backgrounds and maturity, from early-stage startups to late-stage startups. About the Role We are seeking a technically proficient, business-minded Applied AI Engineer to help push the frontier of advanced AI with our strategic startup customers. You'll work with some of the most exciting AI startups in the world, guiding them through ideation, development, delivery, and scaling to accelerate and maximize the value of what they build on our platform. You will have the opportunity to work on the most novel and creative use cases being built on our API, serving as a critical partner in collecting and delivering high-fidelity product and model feedback internally. You will collaborate closely with Sales, Solutions Engineering, Applied Research, and Product teams, and you will report to the Startups Applied AI Lead. This role is based in our San Francisco or New York offices. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Partner closely with strategic startup customers as their technical thought partner to build novel applications on our API, helping them rapidly move from ideation to scale. Provide proactive guidance to maximize business impact and accelerate application development. Experiment and prototype alongside customers, demonstrating practical use cases. Contribute to open-source resources and scale the function by sharing knowledge, codifying best practices, and publishing useful resources. Synthesize and deliver valuable feedback to the Product and Research

JavaScriptPythonJavaAWS
O
📍 Washington, District of Columbia, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s AI Success Engineer team partners with the world’s most ambitious government & partner organizations to translate cutting edge AI into real business and mission impact for governments of all levels from Local, State, Federal, and International. We guide customers and users journey from the first time they try ChatGPT Enterprise, automate a workflow, develop and execute a new skill, and create their first agent to scaled enterprise adoption of ChatGPT, Codex, our API and other novel capabilities. Our work spans technical integration and enablement, workflow transformation, inspiring and upskilling AI literacy and confidence across the workforce, sustained program, product and new capability delivery. Most importantly, we help each member of our customer's workforce, their teams, programs and missions meet their total potential. Our government customers have vital missions, and we must meet them with game-changing technology. Every engagement is an opportunity to shape how AI changes work, productivity, and innovation. This role sits at the center of that mission. About the Role Governments work at a scale that is truly exponential on missions that are of critical importance to people, communities and nations. The AI Success Engineer role is the primary post-sales relationship for OpenAI’s most important customers. You are responsible for the end-to-end account management of critical Government and Partner customers. You will be helping Government Leaders/Partners appropriately and effectively use AI for their mission, while simultaneously investing in ensuring their people are AI-enabled and ready to advance positive outcomes that their constituents depend on them for. You will drive: the impact of our tools on their mission, account health and adoption, ensuring technical readiness, creating and executing on the deployment strategy, enabling, educating and training their workforce, identifying new use cases and upsell opportunities, and d

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s AI Success Engineer team partners with the world’s most ambitious organizations to translate cutting edge AI into real business value. We guide customers from first deployment through scaled enterprise adoption. Our work spans technical integration and enablement, workflow transformation, and sustained program and product delivery. Our customers range from fast growing digital natives to the largest global enterprises, government agencies and educational institutions. Every engagement is an opportunity to shape how AI changes work, productivity, and innovation. This role sits at the center of that mission. About the Role The Success Engineer role is the primary post-sales point of contact for a portfolio of education institutions. You are responsible for driving account health and adoption, ensuring technical readiness, identifying high-impact academic and administrative use cases, and delivering measurable value to our education customers using OpenAI’s platform. This role blends technical leadership, program management, customer advisory, and product influence. You will partner deeply with customer teams, map workflows, lead configuration and enablement, oversee deployment plans, and guide institutions toward high impact use cases that showcase the full value of our platform. You will work closely with Sales, Solutions Architecture, Product, and Research to ensure the customer experience is connected and successful across every touchpoint. Success in this role means accelerating adoption, increasing customer activation depth, guiding strategic use cases that get to production, and helping customers demonstrate tangible business impact. This role is based in SF or NYC; we offer relocation benefits to new hires. In this role, you will: Lead the technical relationship for post-sale customers and act as their trusted advisor on deployment, adoption, and value realization Own account health, adoption velocity, and ongoing technical deployment an

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team The Applied AI Engineer - Digital Natives team is responsible for ensuring the safe and effective deployment of Generative AI applications for developers and enterprises. We act as a trusted advisor and thought partner for our customers, working to build an effective backlog of frontier AI use cases for their industry and drive them to production through strong technical guidance. As an Applied AI Engineer in the Digital Native segment, you’ll help large and highly sophisticated companies transform their business through custom AI solutions applications such as customer service, automated content generation, contextual search, personalization, and other novel use cases leveraging OpenAI’s newest, most exciting models and latest capabilities. About the role We are looking for a driven solutions leader with a product mindset to partner with our customers and ensure they achieve tangible business value with frontier AI. You will pair with senior customer leaders to establish AI strategic roadmaps and identify the highest value applications. You’ll then partner with their engineering and product teams to move from prototype through production. You’ll take a holistic view of their needs and design an enterprise architecture using OpenAI APIs and other services to maximize customer value. You will collaborate closely with Sales, Solutions Engineering, Applied Research, and Product. This role is based in our SF or Seattle office. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Deeply embedded with our most sophisticated and technical platform customers, serving as their technical thought partner in ideating and building novel applications on our APIs. Proactively provide guidance to our customers on how to maximize business impact from their applications, accelerating their time to value. Experiment and prototype solutions with and for your customers. Forge and manage relat

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team pAGI Infra team builds and operates the systems that make large-scale model training and evaluation reliable, efficient, and easy to run. Our work spans distributed training infrastructure, inference and grading platforms, compute scheduling, and research tooling. We partner closely with researchers and engineering teams to turn new research needs into dependable infrastructure, improve GPU efficiency, and shorten the path from an experiment to a validated model. About the Role We’re looking for an AI Systems Engineer to help scale the infrastructure behind our training and evaluation workflows. You’ll own projects from identifying bottlenecks and designing solutions through deployment and operation. The work combines distributed systems engineering, performance optimization, and close collaboration with researchers. You might build a shared grading service, improve resource allocation across workloads, or bring a new training stack into production — directly improving how quickly and reliably research moves forward. In this role, you will: Build and operate infrastructure for large-scale training and evaluation, improving reliability, throughput, and resource efficiency. Develop shared inference and grading platforms with automated capacity management, health monitoring, and visibility into performance. Improve compute scheduling and resource allocation to reduce idle GPU time and help workloads recover quickly from failures. Diagnose bottlenecks across training, inference, and orchestration, and work across teams to improve end-to-end performance. Build self-service tools, automated validation, and observability that help researchers launch experiments, diagnose issues, and compare results with less manual intervention. You might thrive in this role if you: Are excited about the potential of personal AGI and want to build the infrastructure that enables it. Have strong software engineering fundamentals and experience building or operating large-scal

AWSRestAIRust
O
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team The Applied AI Engineer, Digital Natives team is responsible for ensuring the safe and effective deployment of Generative AI applications for developers and enterprises. We act as a trusted advisor and thought partner for our customers, working to build an effective backlog of frontier AI use cases for their industry and drive them to production through strong technical guidance. As an Applied AI Engineer in the Digital Native segment, you’ll help large and highly sophisticated companies transform their business through custom AI solutions applications such as customer service, automated content generation, contextual search, personalization, and other novel use cases leveraging OpenAI’s newest, most exciting models and latest capabilities. About the role We are looking for a driven solutions leader with a product mindset to partner with our customers and ensure they achieve tangible business value with frontier AI. You will pair with senior customer leaders to establish AI strategic roadmaps and identify the highest value applications. You’ll then partner with their engineering and product teams to move from prototype through production. You’ll take a holistic view of their needs and design an enterprise architecture using OpenAI APIs and other services to maximize customer value. You will collaborate closely with Sales, Solutions Engineering, Applied Research, and Product. This role is based in our NYC office. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Deeply embedded with our most sophisticated and technical platform customers, serving as their technical thought partner in ideating and building novel applications on our APIs. Proactively provide guidance to our customers on how to maximize business impact from their applications, accelerating their time to value. Experiment and prototype solutions with and for your customers. Forge and manage relationships wi

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products—including next-generation ads experiences—that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, and Research to bring research breakthroughs into real-world systems at global scale. About the Role As a Research Engineer in OpenAI's Monetization Group, you will have the opportunity to work with some of the brightest minds in AI. You'll contribute to deploying state-of-the-art models in production environments, helping turn research breakthroughs into tangible solutions. If you're excited about making AI technology accessible and impactful, this role is your chance to make a significant mark. In this role, you will: Innovate and Deploy: Design and deploy advanced machine learning models that solve real-world problems. Bring OpenAI's research from concept to implementation, creating AI-driven applications with a direct impact. Collaborate with the Best: Work closely with researchers, software engineers, and product managers to understand complex business challenges and deliver AI-powered solutions. Be part of a dynamic team where ideas flow freely and creativity thrives. Optimize and Scale: Implement scalable data pipelines, optimize mod

Machine LearningArtificial IntelligenceAI
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin

PythonDockerMachine LearningAI
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview: We are seeking a skilled and experienced Senior AI Engineer – AI Platform to join our ClickUp Engineering team. In this role, you will play a critical part in both building the core AI platform and directly applying large language models (LLMs) to deliver intelligent features across ClickUp. You will focus on backend systems that enable scalable, reliable, and secure AI-powered capabilities, while also working hands-on with LLMs to solve real user problems and drive product innovation. Key Responsibilities: Architect, design, and implement scalable AI platform services that support the deployment, orchestration, and lifecycle management of LLMs and other AI models. Apply LLMs and other AI technologies directly to build and enhance ClickUp’s intelligent features, working closely with product and engineering teams to deliver impactful solutions. Build and maintain robust APIs and backend systems that enable seamless integration of AI-powered features into ClickUp’s core platform. Develop infrastructure for model serving, monitoring, logging, and automated evaluation to ensure high reliability and performance of AI services in production. Integrate with multiple LLM providers (e.g., OpenAI, Anthropic, Google) and manage model selection, routing, and fallback strategies for optimal performance and cost. Drive the adoption of best practices in AI privacy, security, and compliance, including data anonymization, secure data handling, and regulatory adherence. Optimize platform performance, scalability, and cost-efficiency, leveraging cloud-native technologies and distributed systems. Stay curre

TypeScriptPythonAWSAzure
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview: We are seeking a highly skilled Staff AI Engineer – AI Platform to join our ClickUp Engineering team. In this role, you will play a critical part in both building the core AI platform and directly applying large language models (LLMs) to deliver intelligent features across ClickUp. You will focus on backend systems that enable scalable, reliable, and secure AI-powered capabilities, while also working hands-on with LLMs to solve real user problems and drive product innovation. Key Responsibilities: Architect, design, and implement scalable AI platform services that support the deployment, orchestration, and lifecycle management of LLMs and other AI models. Apply LLMs and other AI technologies directly to build and enhance ClickUp’s intelligent features, working closely with product and engineering teams to deliver impactful solutions. Build and maintain robust APIs and backend systems that enable seamless integration of AI-powered features into ClickUp’s core platform. Develop infrastructure for model serving, monitoring, logging, and automated evaluation to ensure high reliability and performance of AI services in production. Integrate with multiple LLM providers (e.g., OpenAI, Anthropic, Google) and manage model selection, routing, and fallback strategies for optimal performance and cost. Drive the adoption of best practices in AI privacy, security, and compliance, including data anonymization, secure data handling, and regulatory adherence. Optimize platform performance, scalability, and cost-efficiency, leveraging cloud-native technologies and distributed systems. Stay current with ad

TypeScriptPythonAWSAzure
🔔

Get new ai deployment engineer jobs in United States by email

Daily job updates · Unsubscribe anytime