About the Team The Future of Computing Research team is an applied research team in the Consumer Devices group focused on developing new methods and models to support our vision as we advance forward in our mission of building AGI that benefits all of humanity. About the Role As a Technical Lead on the Future of Computing Research team, you will work together with both the best ML researchers in the world and the greatest design talent of our generation to push the frontier of model capabilities. This role is based in San Francisco, CA. We follow a hybrid model with 3 days a week in the office and offer relocation assistance to new employees. In this role, you will: Evaluate and select silicon platforms (GPUs, NPUs, and specialized accelerators) for on-device and edge deployment of OpenAI models. Work closely with research teams to co-design model architectures that meet real-world deployment constraints such as latency, memory, power, and bandwidth. Analyze and model system performance, identifying tradeoffs between model design, memory hierarchy, compute throughput, and hardware capabilities. Partner with hardware vendors and internal infrastructure teams to bring up new accelerators and ensure efficient execution of transformer workloads. Build and lead a team of engineers responsible for implementing the low-level inference stack, including kernel development and runtime systems. Run through the necessary walls to take nascent research capabilities and turn them into capabilities we can build on top of. You might thrive in this role if you: Have experience evaluating or deploying workloads on GPUs, NPUs, or other specialized accelerators. Understand the performance characteristics of transformer models, including attention, KV-cache behavior, and memory bandwidth requirements. Have designed or optimized high-performance compute systems, such as inference engines, distributed runtimes, or hardware-aware ML pipelines. Have experience building or leading teams work
Jobs in United States
Ai Deployment Engineer in United States
5,082 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai deployment engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.
About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We translate real-world user and operational signals into timely decisions, practical interventions, and improvements to our products and systems. This role will take on new, ambiguous, or underdeveloped operational risks and help mature them into scalable capabilities. We work across USRO and partner closely with Product, Engineering, Data Science, Product Policy, Legal, Safety, Support, and external vendors or partnership stakeholders. About the Role We are seeking a Senior Operations Analyst to take on complex, ambiguous safety and risk problems and turn them into practical operational solutions that can scale. This is a senior individual-contributor role for a versatile operator who is comfortable moving between queues, investigation, analysis, workflow design, hands-on execution, and cross-functional leadership. Depending on team needs, the role may focus on emerging-risk incubation, cloud deployment partnerships, or other new operational areas. You will be expected to move quickly, work hands-on, and create structure without waiting for perfect requirements or a large support team. The work starts with the problem, not a prescribed process. You may investigate unstructured user signals, stand up a lightweight workflow, build an AI-assisted tool, improve an existing operation, or help a new launch become operationally ready. The goal is to produce durable systems that other people can run, not simply complete a series of individual tasks. The portfolio will change with company priorities and may span established harm areas, emerging-risk incubation, cloud deployments and partnerships, device safety, or new product launches. Some hires may focus primarily on cloud deployment operations, including launch readiness, partner coordination, safety workflows, and operational monitoring. You will ty
About the Role As a Sales Manager, Energy, you will build and lead a team of Account Directors focused on strategic growth across utilities, oil and gas, renewables, power generation, and energy services. The team will partner with complex organizations modernizing operations, improving reliability, accelerating the energy transition, and adopting enterprise AI responsibly at scale. You’ll help the team navigate regulated enterprise sales cycles, deepen relationships with business, technology, operations, engineering, security, and risk leaders, and drive adoption of OpenAI’s platform across safety-conscious, asset-intensive organizations. Key Responsibilities Recruit, develop, and lead a high-performing team of Energy Account Directors. Create a strong coaching culture through deal reviews, account strategy sessions, ride-alongs, and structured 1:1s. Define the Energy GTM strategy, including subsector segmentation, account prioritization, partner strategy, executive engagement, and territory planning. Drive disciplined pipeline generation, forecast accuracy, and operational rigor. Guide multi-stakeholder opportunities involving operations, engineering, digital, data, security, legal, risk, procurement, and executive leadership. Help customers translate AI and API capabilities into measurable outcomes across asset and field operations, grid and generation planning, engineering knowledge, customer service, commercial workflows, and enterprise productivity. Partner with Product, Solutions Architecture, Technical Success, Legal, Security, Finance, and policy experts to support responsible deployment. Provide structured feedback on customer requirements, integration blockers, reliability and governance needs, and emerging industry trends. What We’re Looking For 15+ years of enterprise sales, GTM, or sales leadership experience. Proven experience building and scaling enterprise sales teams responsible for complex strategic accounts and large revenue targets. Deep underst
About the Team Like every team at OpenAI, the Marketing team contributes to our broader mission of ensuring responsible and widespread adoption of artificial intelligence. With that aim in mind, we are responsible for developing and executing strategies that drive awareness, engagement, and usage for OpenAI’s products and platform amongst our core audiences. We take a data-driven approach to understand our customers' needs and challenges, ensuring that their voices are reflected in product development and messaging. We then partner closely with Product, Engineering, Research, Comms, and Design teams to create a cohesive customer experience across all our channels. Our focus extends beyond just promoting product features; we aim to provide valuable insights and resources that help our users make the most out of AI technologies. About the Role We are seeking a Government Marketing Manager to lead marketing campaigns and programs for OpenAI’s government business, OpenAI for Government. This role will define how we communicate the impact of our work with governments and public institutions, turning technical products and partnership outcomes into clear stories that show how AI can help public servants better serve people and communities. The ideal candidate has experience working with government entities and leading in complex, multi-stakeholder environments. They bring strong strategic and creative judgment, translate technology and policy priorities for diverse audiences, and build trust and engagement across paid, owned, and earned channels. In this role, you will: Build the OpenAI for Government Narrative and Brand: Build a distinct government narrative aligned with the OpenAI brand, centered on mission outcomes, public servants, security, and responsible deployment. Build a Government Storytelling Engine: Create a repeatable process to source, prioritize, approve, and refresh customer stories, public use cases, metrics, and quotes for web, video, social, sales, eve
About the Team OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. The Industrial Compute team is central to this mission, setting the core infra strategy and implementing this vision. From site selection to the buildout process, this team sits at the intersection of commercial, technical, strategy, and operations, interacting with teams and executives inside and outside of OpenAI. About the Role The Clean Energy and New Technology Lead will own infrastructure clean energy and emerging energy technology strategy and execution, identifying and deploying scalable solutions that enable resilient, low-carbon compute and data center growth. The role will work closely with regulatory and policy teams to align infrastructure expansion with OpenAI’s long-term environmental and operational objectives. This is an individual contributor lead role and does not have direct reports initially. The role will evaluate where emerging energy technologies can materially improve reliability, cost, carbon, speed, or resilience; translate those options into practical deployment pathways; and help ensure OpenAI’s infrastructure growth remains aligned with sustainability considerations. Key Responsibilities Evaluate emerging energy solutions such as clean firm power, advanced storage, grid flexibility, low-carbon backup power, heat reuse, water-related energy efficiency, and other scalable technologies where relevant. Identify pilot opportunities and deployment pathways that can move promising energy technologies from concept to commercially and operationally credible execution. Translate technical options into clear reliability, cost, schedule, carbon, regulatory, and operational implications for infrastructure decision-making. Partner with energy regulatory, policy, procurement, engineering, deployment, finance, legal, and site-readiness teams to align technology and sustainability choices with
About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency. Within Safety Systems, the Model Policy team aligns model behavior with desired human values and norms. We co-design policy with models and for models by driving rapid policy taxonomy iteration based on data and defining evaluation criteria for foundational models’ ability to reason about safety. About the Role Frontier AI systems are rapidly expanding what is possible in cybersecurity and software engineering. These capabilities create major defensive opportunities, but they also raise serious dual-use and misuse risks across areas such as malware development, exploit discovery, vulnerability chaining, credential abuse, cyber intrusion, and autonomous offensive operations. In this role, you will help define how OpenAI’s models should behave in high-risk cybersecurity contexts. You will develop policy frameworks, threat models, taxonomies, evaluations, and behavioral specifications that guide model behavior across training, deployment, and monitoring systems. This role sits at the intersection of cybersecurity, AI safety, threat modeling, evaluation science, and policy implementation. You will work closely with research, engineering, safety training, preparedness, and product teams to build policies that are technically grounded, measurable, enforceable, and responsive to real-world cyber risk. Your Responsibilities: Design and maintain model policies for cybersecurity and frontier-risk domains, especially dual-use and high-risk cyber capabilities. Translate cybersecurity threat models into clear behavioral specifications, evaluation criteria, grading guidance, and system-level mitigations. Define practical boundaries between legitimate security research, defensive workflows, and assistance that could materially enable harmful activity. Build policy artifacts that support i
About the Team OpenAI’s Industrial Compute team is responsible for building and scaling large-scale compute capacity across first-party data centers, strategic partners, and industrial infrastructure environments. We focus on converting power, land, hardware, and operational execution into reliable compute capacity that can support frontier AI training and inference workloads. This team operates at the intersection of infrastructure delivery, hardware systems, utilities, supply chain, and capacity strategy—ensuring OpenAI can scale compute faster than traditional models allow. About the Role We are seeking a Tokens-as-a-Service (TaaS) Lead to drive the end-to-end conversion of industrial-scale infrastructure investments into usable token capacity for OpenAI workloads. In this role, you will own execution across complex compute programs where raw infrastructure capacity must be transformed into operational GPU throughput. You will coordinate across data center delivery, power, networking, hardware deployment, workload enablement, finance, and external partners to ensure capacity becomes productive tokens as quickly and efficiently as possible. This role is ideal for someone who can bridge physical infrastructure delivery with compute utilization outcomes. Success requires strong systems thinking, elite program leadership, and the ability to drive accountability across internal teams and strategic partners. In this role, you will Lead Tokens-as-a-Service programs across industrial compute environments, including first-party and partner-owned capacity. Convert delivered power, space, and hardware capacity into production-ready token throughput. Build integrated execution plans spanning construction, power energization, rack deployment, networking, cluster readiness, and workload onboarding. Partner with infrastructure engineering, hardware, networking, finance, supply chain, and operations teams. Drive external providers, EPCs, OEMs, utilities, and strategic partners t
The Applications Development Technology Senior Lead Analyst is a senior level position responsible for establishing and implementing new or revised application systems and programs in coordination with the Technology Team. The overall objective of this role is to lead applications systems analysis and programming activities. Responsibilities: Lead the architecture, design, development, and delivery of enterprise UI applications. Strong hands-on experience with Angular and Ext JS for developing scalable and responsive user interfaces. Strong experience with Java and Spring Boot for designing and developing backend services and APIs. Experience deploying, supporting, and troubleshooting applications in ECS/containerized environments . Define application architecture, technical standards, reusable components, and development best practices. Provide technical leadership to development teams, including design reviews, code reviews, performance optimization, and troubleshooting . Strong understanding of CI/CD pipelines , including automated build, testing, deployment, and release processes. Hands-on experience with GitHub , including source-code management, branching strategies, pull requests, code reviews, and integration with CI/CD pipelines. Strong knowledge of open-source technologies and frameworks , with hands-on implementation experience. Ability to evaluate and select appropriate open-source libraries, frameworks, and tools , considering security, licensing, maintainability, vulnerabilities, and enterprise standards. Drive application modernization, technical improvements, and adoption of engineering best practices. Work closely with architects, developers, infrastructure teams, product owners, and business stakeholders to deliver solutions successfully. Prov
At Freddie Mac, our mission of Making Home Possible is what motivates us, and it’s at the core of everything we do. Since our charter in 1970, we have made home possible for more than 90 million families across the country. Join an organization where your work contributes to a greater purpose. Position Overview: We need a highly innovative Technical Lead! How confident are you that you can build sophisticated analytic systems? If you believe you could contribute to the development of innovative principles and ideas in a matrixed environment, please keep reading as we are seeking an individual contributor who has experience with Java and Python and can lead and nurture an inspiring environment in our Virginia office. Our Impact: The Investments and Capital Markets (I&CM) division is looking for a capable technology lead for its trading and analytics development team. This could be you! To thrive in this division, you must have a comprehensive understanding of system implementation and design, experience working in capital markets, and be enthusiastic about leading development of new paradigms in software system architecture. Your Impact: As a Trading Analytics Development Tech Lead, you will develop and maintain software using Java and Python tech stack that adheres to software engineering best practices. You will influence technical decisions, mentor developers, resolve engineering blockers, and partner with engineering managers to help teams deliver secure, reliable, and maintainable solutions. You will provide hands-on directions for full-stack applications, APIs, microservices, and integration services while reinforcing engineering discipline across design, development, testing, deployment, observability, and production readiness. Partner closely with Product Owners, engineering managers, architecture, business stakeholders, and cross-functional tec
We are seeking a highly motivated and experienced Lead Java Developer to join our Credit Risk Technology team. This is a hands-on leadership role where you will be instrumental in designing, developing, and deploying robust, high-performance, and large-scale software solutions for Counterparty Credit Risk Management. You will be at the forefront of firm-wide initiatives, leveraging best-in-class software stacks and cutting-edge AI tools to solve complex challenges. You will lead a team of talented engineers, providing technical guidance, fostering innovation, and ensuring the delivery of high-quality, scalable, and resilient software. The ideal candidate is passionate about modern software development practices, proficient with contemporary tech stacks, and eager to utilize artificial intelligence to manage Citi’s exposure to financial institutions, governments, and corporates. This role offers the opportunity to make a significant impact by contributing to critical systems that provide an integrated view of trades, collateral, and market data from numerous sources. Key Responsibilities: Technical Leadership & Hands-On Development: Lead by example, actively contributing to the design, architecture, and hands-on development of critical software components for Counterparty Credit Risk Management leveraging AI tools such as Devin and Co-Pilot Drive technical excellence, ensuring best practices in coding, testing, and deployment are followed for high-performance and large-scale applications. Conduct code reviews, provide constructive feedback, and mentor team members in advanced development techniques. Team Leadership & Management: Supervise work of a team of software engineers Foster a collaborative, innovative, and inclusive team environment focused on tackling challenging firm-wide initiatives. Allocate resources effectively to meet project deadlines an
About the Team OpenAI's Go-To-Market team helps organizations understand, adopt, and scale the use of advanced AI across their businesses. Our team brings together Sales, Solutions, Customer Success, Marketing, Partnerships, and other cross-functional leaders to support customers in deploying AI responsibly and delivering meaningful business outcomes. About the Role As an Account Director focused on Media, you will lead relationships with a select portfolio of the largest and most strategically important media, entertainment, publishing, streaming, and digital content companies. You will help these organizations apply OpenAI's technology to transform content creation, audience engagement, advertising, personalization, customer experience, employee productivity, and new AI-powered products and services. This role combines deep enterprise sales experience, executive relationship management, media industry knowledge, and technical fluency. You will work across complex, highly matrixed organizations to identify strategic opportunities, build alignment with senior decision-makers, and lead enterprise-wide adoption of OpenAI's products. You will own the full customer lifecycle, from strategic account planning and opportunity development through negotiation, deployment, consumption growth, and long-term expansion. In close partnership with Solutions, Product, Engineering, Research, and other internal teams, you will help shape OpenAI's media go-to-market strategy and establish durable relationships with some of the industry's most influential companies. In this role, you will: Manage a small portfolio of large, strategic media accounts and develop comprehensive, multiyear account plans. Build and maintain executive relationships with C-level and senior business leaders across technology, digital, editorial, content, advertising, audience development, product, data, and AI. Lead complex, enterprise-wide sales cycles involving multiple stakeholders, business units, procureme
As the Salesforce Administrator at Coder, you’ll help keep our Salesforce environment reliable, efficient, and ready to support the business. You will help keep our Salesforce environment running cleanly and efficiently, and you'll work closely with Finance and our Sales Ops Salesforce Admin. The work spans day-to-day administration, Flow builds, and DealHub CPQ support — with occasional AI-assisted development tasks mixed in. You'll be working alongside someone who knows the environment well, so the expectation is that you can plug in quickly and operate with minimal hand-holding. What you’ll do here Own day-to-day Salesforce administration including user management, configuration, and data integrity Design, implement, and troubleshoot Automations using Salesforce Flow Configure and support DealHub CPQ including pricing rules, quote templates, and related requests Build and maintain reports and dashboards, with a focus on pipeline and revenue visibility for Finance Partner cross functionally on projects, sandbox management, and incoming support work Assist with Apex AI-assisted development tasks as needed Support data quality efforts including deduplication and cleanup as needed Manage end to end implementation of Salesforce including Okta provisioning, sales tools, and CPQ Own data governance across Salesforce by enforcing data quality standards, managing field-level security and access, and building validation rules that catch bad data at entry. You'll audit integrity regularly and partner with RevOps and Engineering to keep CRM data trustworthy Lead technical delivery of system improvements: scoping, building, testing, and deploying changes end-to-end Manage Salesforce releases, sandbox environments, and deployment processes Develop and maintain Salesforce customizations using Apex, Visualforce, and Lightning Web Components Own integrations between Salesforce and connected systems including architecture, build, testing, and maintenance Perform unit testing, trou
From $164.7K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . As a Senior Product Manager for Signal Lifecycle within Trust & Safety, you'll own the product strategy for the ML platform that powers how Pinterest trains, evaluates, deploys, and measures content safety models at scale. You'll lead the development of ML Signal Management — making ML signals first-class entities with unified metadata and identity across systems. Partnering deeply with ML engineering, data science, content safety, and enforcement systems, you'll drive a platform whose scope is expanding from T&S into content quality, ads safety, and beyond. What you'll do: Own and drive the Signal Lifecycle product roadmap, including ML Flywheel infrastructure, auto-deployment, model onboarding, golden dataset management, and signal performance measurement Define and ship ML Signal Management — a unified backbone that elevates ML
About the Team The Agent Post-Training team creates the frontier agents OpenAI ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that can operate computers, collaborate with people and other agents, and expand what people and organizations can imagine, attempt, and achieve. We define what the next generation of agents should be able to do, build the training signal that teaches those abilities, and run the experiments that make them real. Our work spans coding, tool use, computer use, multi-agent coordination, long-horizon execution, factuality, instruction following, calibrated reasoning, and taste. Our team is where new model capabilities get made. We build the data, environments, graders, training methods, and feedback loops that shape what OpenAI's next agents can do, then carry those capabilities through major training runs and into the products people use. About the Role We believe that the final enabler for AGI is spending compute on context. As a Context Researcher on Agent Post-Training, you will scale compute spent on context. You will get to work in our frontier training stack on enabling the next paradigm of model training with a clear product interface for iterative deployment (Codex Chronicle). You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measure whether it worked, and ship improvements into products used by real people. This is a high-agency role for people who want their work to land directly in frontier models. In this role, you might Design and run experiments that improve scaling of compute on context. Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, diagnostics, and model-behavior analysis. Build evals and environments that expose the next set of model failures,
Other cities to consider
More places hiring for this role
Get new ai deployment engineer jobs in United States by email
Daily job updates · Unsubscribe anytime