Jobs in United States

Production Operator in United States

1,337 active opportunities · Updated October 2026

Explore current production operator jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

M
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -100%

What you’ll do Design and implement secure cloud pipelines that ingest very large scan datasets (multi-terabyte), reliably and resumably. Build orchestration for GPU-accelerated reconstruction and analysis with strong retry semantics, idempotency, and cost controls. Define end-to-end data lifecycle for medical imaging: raw vs intermediate vs derived artifacts, retention policies, and reproducibility. Implement security + compliance primitives appropriate for HIPAA/PHI: encryption in transit/at rest, key management, least privilege, audit logs, and access reviews. Build operational tooling: monitoring, alerting, runbooks, and incident-driven improvements for a growing device fleet. What we’re looking for Strong experience with cloud batch/queueing/orchestration, storage systems, and data pipeline reliability. Experience shipping production systems that handle large data volumes and failure-prone networks. Practical security mindset (least privilege, secrets, audit logging) and comfort operating in compliance-constrained environments. Useful experience Building reliable data pipelines at scale (queues/orchestration, resumable uploads, GPU batch execution) with strong observability. Security + privacy by default: encryption, least-privilege access, auditing, and practical HIPAA/PHI guardrails. Owning the “boring” backend details that keep a lean team moving: schemas/migrations, cost controls, retries, and runbooks. Understanding compute tradeoffs across hardware options, and specifying appropriate cloud resources.

E(
📍 San Francisco Bay Area, California, United States· Full-time
✓ Quality checkedCompany trend -100%

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. At Ema, we build AI Employees that operate inside the enterprise. Healthcare is where the bar is highest: the output has to be accurate, auditable, and clinically sound. We're opening a part-time role focused on agent measurement and improvement. You'll instrument how our healthcare AI Employees perform in production, identify where quality breaks down, and design the experiments that close the gap — partnering with clinical experts to ensure improvements translate into better patient outcomes, not just better metrics. We're looking for strong analytical judgment (Python, agent lead development), comfort operating with incomplete information, and genuine interest in the healthcare domain. Clinical experience is valued but not required. This engagement is structured as a paid internship or contract engagement, with weekly syncs in our Bay Area office. Compensation offered will be determined by factors such as location, level, job-related knowledge, skills, and experience. Certain roles may be eligible for variable compensation, equity, and benefits. Ema Unlimited is an equal opportunity employer and is committed to providing equal employment opportunities to all employees and applic

PythonRestAIGo
G
📍 United States· Full-time
✓ High-confidence listingCompany trend -100%

From $154K/yr

Quick readStrong listing-quality and freshness signals

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s

PythonKubernetesAISwift
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $192K/yr

Quick readStrong listing-quality and freshness signals

You will lead a small, hands-on engineering team building the secure, scalable Core Analytics Data Access Platform that accelerates Datadog’s Applied AI and analytics capabilities. The team owns the Data Access Platform — a unified interface that lets AI and analytics teams discover and self-serve production-ready datasets while abstracting underlying systems and embedding required legal and compliance guardrails. In this role you’ll own technical direction, contribute to design and code, and partner closely with Applied AI, Product Analytics, and internal platform teams to provide reliable datasets and APIs for model training and analysis. This role balances day-to-day engineering leadership with long-term platform planning. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead a Hands-On Engineering Team: Manage, mentor, and grow a small team of 2–4 data engineers (mix of senior and junior) across Paris and NYC, fostering technical excellence and career development. Own Technical Direction and Delivery: Define architecture, engineering priorities, and the team roadmap for the Data Access Platform, driving implementation of scalable, secure data pipelines and platform services. Contribute to Design and Code: Spend substantial time coding, reviewing, and shipping critical platform components to ensure performance, reliability, and operational excellence. Partner with Internal Stakeholders: Work closely with Applied AI, Internal Product Analytics, product managers, and platform teams to define data contracts, APIs, SLAs, observability, and curated analytical datasets. Ensure Data Security, Governance, and Reliability: Implement access controls, lineage, monitoring, and compliance guardrails to support safe model training and repeatable analytics workflows.

AWSAIGoRust
D
📍 Massachusetts, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $192K/yr

Quick readStrong listing-quality and freshness signals

About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. You will: Solve a scaling bottleneck in a critical service Deploy a new feature to production, progressively rolling it out with feature flags Investigate and fix a production issue from a service your team owns Design a way to scale up a service for more traffic With your team, plan the most important projects to work on next Who You Are: You have significant experience in one or more languages You value code simplicity and performance You can design architecture to solve problems at high scale You have a BS/MS/PhD in a scientific field or equivalent experience You want to work in a fast, high-growth startup environment that respects its engineers and customers You have demonstrated ability to use AI coding tools in day-to-day workflows and build, validate, and refine AI-generated output in products You can design AI Backend systems, with awareness of quality, cost, and latency tradeoffs 6+ years of experience Bonus points: You've worked at high scale with systems like Redis, Cassandra, Kafka You wrote your own data pipelines once or twice before You have a strong background in statistics You have significant experience with Go, C, or Python You’re excited about leveraging AI tools to enhance how you code, solve problems, and build – or eager to learn how You’re motivated to push the boundaries of how AI can improve software engineering best practices and contribute to building AI-enabled products Datadog values people from all walks of life. We understand not everyone will meet all the above qualificat

PythonRedisAIGo
D
📍 Massachusetts, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $156K/yr

Quick readStrong listing-quality and freshness signals

As organizations rapidly adopt AI applications and agentic systems, security teams need visibility and control over how these technologies are being used. Datadog's AI & Data Security product helps customers discover, secure, and govern AI usage across their environments ensuring sensitive data is properly managed from model training through production. As a Product Manager II for AI & Data Security, you will own capabilities for AI discovery, posture management, and data security; giving customers complete visibility into their AI applications, agents, and enabling ecosystem, with prioritized actions to operate AI systems securely. You'll partner with engineering, design, security research, and GTM teams to define and ship platform capabilities that help organizations adopt AI at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the roadmap for AI & Data Security capabilities, including AI and data discovery & posture management. Define how security teams can assess and manage the security posture of AI-enabled systems, including configuration risks, sensitive data exposure, and policy violations. Work closely with engineers and designers to deliver new product capabilities end-to-end, from early concept through launch and iteration. Partner with Datadog security researchers to identify emerging risks in AI systems and translate them into actionable product features. Engage with customers to understand how they are adopting AI and validate solutions that help them operate these systems securely. Collaborate with go-to-market teams to enable adoption and communicate the value of AI security capabilities to customers. Who You Are: You have 3+ years of product management experience building technical products, ideally in security,

RestMachine LearningAIGo
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $192K/yr

Quick readStrong listing-quality and freshness signals

As the Senior Product Manager for the Actions & Automations team, you will own the ecosystem that enables customers, partners, and Datadog teams to build, deploy, and operate AI agents on Datadog. You will drive the strategy and execution for the platform capabilities, developer experience, integrations, and extensibility model that make Datadog the best place to build agents that understand and act on production systems. Modern engineering organizations are entering a new era where software is not only monitored and operated by humans, but increasingly by AI-powered agents. As agentic workflows reshape how teams build, operate, secure, and troubleshoot systems, customers need a platform for creating specialized agents, connecting them to business and engineering systems, governing their behavior, and extending them to solve unique organizational problems. You will define and build the ecosystem that makes this possible. Agent Builder sits at the intersection of Datadog's products, AI capabilities, and ecosystem strategy. You will have the opportunity to work across the breadth of the Datadog platform, partner with teams throughout the company, and help establish Datadog as the foundation for operational AI. At Datadog, we place value in our office culture, the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Define the vision, strategy, and roadmap for Datadog's Agent Builder platform and ecosystem. Own the core platform capabilities that enable customers and partners to create, customize, deploy, and manage AI agents. Drive the extensibility model for agents, including integrations, tools, actions, context sources, APIs, SDKs, and developer workflows. Shape how agents perform actions across Datadog products and third-party systems. Partner closely with AI, platform, infrastructure, and product tea

AIGoRustSpring
M
📍 United States· Full-time
✓ High-confidence listingCompany trend -93.7%

From $104K/yr

Quick readStrong listing-quality and freshness signals

MongoDB is hiring a Staff Product Marketing Manager to build and own our go-to-market narrative for the Public Sector vertical, with a focus on Federal Government and the broader public sector market. This is a foundational hire for MongoDB’s Industry Verticals product marketing function: you will define how MongoDB’s unified data platform, spanning cloud, on-premises, and hybrid database deployments with integrated, production-ready AI capabilities, shows up for government buyers. You’ll turn a major compliance milestone into a durable competitive differentiator: developing the positioning, messaging, and sales-ready content that helps government agencies, systems integrators, and cloud/public-sector resellers understand why MongoDB is the right data platform for mission-critical, regulated workloads. You do not need prior government or public-sector work experience to succeed in this role — you need to be an excellent product marketer who can get fluent in a new domain quickly and partner closely with the compliance, product, and sales experts who already are. This role can be based in one of our MongoDB hub offices in the U.S. or remotely in the U.S. What you’ll do Own positioning and messaging for MongoDB’s Public Sector go-to-market, leading the federal GTM and launch related activities Translate MongoDB’s data platform capabilities — document database, search, vector search, stream processing, and integrated AI — into mission-relevant outcomes and value propositions for government buyers and the systems integrators who serve them Partner with Compliance, Security, Industry Solutions and Product teams to accurately represent related certification requirements in external-facing content, staying current as MongoDB pursues additional authorizations (e.g., DoD Impact Levels) Build the public sector sales enablement toolkit: battlecards, pitch decks, discovery guides, ROI/value models, and competitive intelligence tailored to federal buying processes and procuremen

MongoDBAWSAzureAI
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $192K/yr

Quick readStrong listing-quality and freshness signals

As Engineering Manager for Threat Detection, you will lead a high-performing team that powers Datadog's detection program. Threat Detection is the organization responsible for keeping Datadog ahead of an evolving threat environment: closing coverage gaps faster, raising the bar on signal quality, and shipping detections that hold up under the scale and complexity of cloud-native infrastructure. Your team will combine direct detection expertise, platform engineering, and applied AI to ship detections at a pace and scale traditional rule-writing alone cannot match. Examples of what your team will work on include detection-authoring agents, the detection platform that powers every rule in production, coverage analysis, alert triage and response automation, and the evaluation infrastructure that holds these systems to a high bar of fidelity. Detection authorship is a shared responsibility across the organization, and your team will contribute both by building the systems that scale our authoring capacity and by writing detections directly when their domain expertise is the right tool. You will partner closely with our Security Incident & Response Team (SIRT), Cyber Threat Intelligence (CTI), AI Engineering teams, and Datadog's broader Security organization. This is a high-impact leadership role: you will grow a team of security and software engineers responsible for building and executing our detection and AI strategy. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the strategy, roadmap, and execution of Datadog Security's shift to AI-accelerated detection and response. Drive development of high-fidelity detections as a shared responsibility across the organization, ensuring your team's systems and direct contributions raise the bar on coverage and

PythonCI/CDRestAI
M
📍 United States· Full-time
✓ High-confidence listingCompany trend -93.7%

From $90K/yr

Quick readStrong listing-quality and freshness signals

MongoDB Atlas is the premier multi-cloud database-as-a-service built and operated by the makers of MongoDB. The Cloud Operations Engineering team at MongoDB is a worldwide team responsible for the consistent operational success of every MongoDB Atlas customer. As a Cloud Operations Engineer, you will help ensure the success of our Atlas customers, whether they are early startups or large multinational companies, cloud-native or just getting started with a digital transformation to the cloud. You are excited about the core mission of MongoDB, and the opportunity to join the team responsible for operating Atlas, the fastest-growing multi-cloud database-as-a-service in the world. You are prepared to be one of the founding members of a 24/7/365 global cloud operations team. Cloud Operations Engineers will be responsible for day-to-day duties such as creating and monitoring systems alert dashboards, reviewing critical event and system logs, accessing customer instances that underpin their production databases and performing server administration duties including performance troubleshooting. Applicants must be critical thinkers who are quick to detect, resolve, or escalate issues that are sometimes broad in scope and difficult to trace. FedRamp engineers are specifically tasked with supporting our government customers in our FedRamp Atlas environment. This includes SLED (State and Local Government and Education), various federal agencies, and other customers that leverage FedRamp. At MongoDB you will grow your career and skills, wear multiple hats, and be part of an operations team that works at the frontier of Cloud services and database systems. This role will be based remotely in Colorado. Responsibilities Successfully coordinate with a global team of Cloud Operations Engineers who are tasked with ensuring our uptime guarantees to our Atlas customer base Help scale the worldwide Cloud Operations Engineering team with the strategic implementation of new processes and to

JavaScriptPythonJavaMongoDB
S
📍 United States· Full-time
✓ Quality checkedCompany trend -94.3%

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe Capital provides access to fast, flexible financing to small-and-medium businesses on Stripe to accelerate their growth, and we lent over $1B in 2024. Businesses use the funds for marketing, team growth, geographic expansion, working capital, new equipment purchases, and much more. Machine learning is core to Stripe Capital’s business—we use information about businesses from their activity within and outside of Stripe and our models to automatically underwrite uniquely tailored financing offers to their needs, which banks are often unable to do. We are doing so through models with an established performance history, data infrastructure that is Stripe scale, and a strong feedback loop that includes explainability, anomaly detection and a risk portfolio management layer. We're an end-to-end team going from ideas to models to shipping in production. What you’ll do As a machine learning engineer for Stripe Capital, you'll be responsible for designing, building, training, evaluating, deploying, and owning ML models in production with the goals of providing financing opportunities to as many users as possible while satisfying financial performance goals. You'll work closely with software engineers, data scientists, product managers, and risk managers to operate Stripe’s ML powered systems, features, and products. You'll also contribute to and influence ML architecture at Stripe and be a part of a larger ML community. Responsibilities Design

Machine LearningAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As an AI Accelerator Systems Software Technical Program manager at OpenAI, you will help bring our chips/system hardware roadmap to life, navigating an array of technical and partnership challenges. We’re looking for people excited to push the frontiers of computing by navigating technical explorations and are passionate about building. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Manage the end-to-end software development from design to implementation for our AI acceleration systems, working across technical, cross-functional and external stakeholders Lead planning and scheduling of AI system software designs with our strategic partners and vendors Coordinate and lead internal resources and communication for efficient interaction with partners and vendors. You might thrive in this role if you: Have experience as a software technical program manager for data center system products (server, GPU, TPU, networking, storage and so on) taking products from concept to volume in a data center environment ensuring the systems scale with high quality Know end-to-end software development program management techniques from concept, design, production, deployment into the data center Want to help design some of the world’s largest supercomputing systems, working at the edge of complex hardware challenges Enjoy working with and enabling world-clas

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI for Financial Services is part of OpenAI's Verticals organization, which focuses on accelerating the economy and knowledge work. We build AI products for financial institutions and the professionals who power them, from investment bankers and research analysts to investors and other financial services teams. We combine OpenAI's models, financial data, and enterprise knowledge to help professionals research companies, analyze markets, and produce high-quality work in the tools they use every day. We're a small, entrepreneurial team working closely with customers and partners across research, product, design, engineering, and go-to-market to bring new capabilities from idea to production. About the Role We're looking for full-stack engineers to build new, AI-native products on top of ChatGPT Work and Codex. This is zero-to-one work: you'll help define how financial professionals research, analyze, and make decisions alongside AI. You'll own the experience across the stack, work directly with customers to understand their workflows, and collaborate across OpenAI to turn new model capabilities into products that professionals can trust. Your work will shape how some of the world's largest financial institutions adopt AI and how financial knowledge work gets done. In this role, you will: Build a new financial services app within ChatGPT Work and Codex, creating intuitive AI-native experiences for company research, financial analysis, document review, and professional work products. Develop new product experiences around enterprise memory that learn from an organization's knowledge, workflows, and context, and adapt to how its teams work. Develop the APIs, services, and integrations required to connect user experiences with financial data providers, enterprise systems, and OpenAI's models. Work closely with product and design to turn ambiguous customer problems into polished, useful, and reliable products. Work directly with financial institutions to

TypeScriptPythonReactAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI for Financial Services is part of OpenAI's Verticals organization, which focuses on accelerating the economy and knowledge work. We build AI products for financial institutions and the professionals who power them, from investment bankers and research analysts to investors and other financial services teams. We combine OpenAI's models, financial data, and enterprise knowledge to help professionals research companies, analyze markets, and produce high-quality work in the tools they use every day. We're a small, entrepreneurial team working closely with customers and partners across research, product, design, engineering, and go-to-market to bring new capabilities from idea to production. About the Role We're looking for backend engineers to build the systems that make advanced AI useful, reliable, and trustworthy in financial services. You'll build the data systems, agentic workflows, and enterprise integrations behind our products. You'll also help bring them into production at some of the world's largest financial institutions. This is a product-minded engineering role with significant ownership and zero-to-one building. You'll shape new products from the ground up, work directly with customers to understand their workflows, and collaborate across OpenAI to turn new model capabilities into products that professionals can trust with high-stakes work. In this role, you will: Design and build backend systems that power AI-native financial workflows across ChatGPT Work and Codex. Build infrastructure to ingest, index, retrieve, and serve financial data, company filings, market information, and firm-specific knowledge at scale. Develop integrations with financial data providers, enterprise knowledge systems, and customer environments, including the authentication, authorization, and entitlements required to use them securely. Build the systems that let models and agents use the right tools and data, preserve source provenance, and produce accurate,

TypeScriptPythonAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We’re looking for a software engineer to help build the design methodology, software abstractions, and infrastructure that enable a small silicon team to develop complex chips rapidly and with high confidence. You will turn evolving architecture and design needs into reusable tools and workflows that improve iteration speed, quality, and then apply those tools to help construct world-class silicon. You’ll work closely across architecture, design, verification, performance modeling, and systems software. This role is well suited for an engineer who enjoys building high quality software and is motivated by the challenge of improving velocity and quality of the silicon development process. In this role, you will: Develop and scale design methodologies for rapid first-party chip development and apply them to construct complex custom chips Create abstractions that allow hardware structures, configurations, experiments, and results to be represented consistently across tools. Automate high-value engineering workflows and improve their reproducibility, observability, testability, and ease of use. Partner with architects, RTL designers, verification engineers, compiler engineers, and systems software engineers to gather requirements and then implement solutions. Use methodology and tooling to identify design risks early, accelerate iteration, and improve confidence in performance and implementation tradeoffs. Contribute across multiple aspects of software and hardware

PythonAWSGitRest
🔔

Get new production operator jobs in United States by email

Daily job updates · Unsubscribe anytime