Jobs in United States

Senior Machine Learning Operations Engineer in United States

1,941 active opportunities · Updated October 2026

Explore current senior machine learning operations engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $280.5K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Every day, Roblox users play over 7M experiences, make over 270M avatar updates, send 6+ billion chat messages and talk for 1+ million hours using voice. Safety is Roblox’s top priority. As part of the Safety Group, the Content and Communications Safety team builds AI features and tools to automate review and moderation of all of these activities on the platform. Areas that the Content Safety Team works on include: Experience reviews The components that make up those experiences (textures, 3D meshes, models, scripts) Avatars, Avatar items Audio, Video Text chat and voice conversations and other content types As a Senior Product Manager on the team, you will work closely with leadership, product, engineering, data science, and operations teams across Roblox. In addition, you will drive critical company-level KPIs that affect how Roblox keeps the platform safe and civil. As we bring new immersive features and content types to Roblox, you will be responsible for leading the product in one of the most challenging and vital roles at Roblox. You will Lead safety features focused on automation of content reviews by leveraging AI Stay one step ahead of malicious actors and their adversarial tactics

AWSGitMachine LearningAI
C
📍 Woonsocket, United States
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. At CVS Health, Site Reliability Engineering (SRE) is fundamental to delivering the reliable, secure, and scalable technology experiences that support millions of patients, customers, pharmacists, and healthcare professionals every day. Our SRE organization drives operational excellence across critical healthcare and retail platforms through innovation, automation, observability, and engineering best practices. The Executive Director, Site Reliability Engineering serves as the strategic leader responsible for the reliability, resilience, and performance of CVS Health's retail and pharmacy technology ecosystem. This executive will define and execute a comprehensive reliability strategy, oversee large global engineering teams, and establish a long-term vision for observability, automation, and operational excellence across thousands of store locations. Working closely with senior business and technology leaders, the Executive Director will champion modern SRE practices, accelerate incident response capabilities, and deliver real-time operational visibility that enables proactive issue prevention and exceptional customer and patient experiences. Key Responsibilities Strategic Leadership & Vision Define and lead the enterprise-wide Site Reliability Engineering strategy supporting CVS Health's retail and pharmacy operations. Align reliability and operational

AWSAzureGCPKubernetes
G
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -97.9%

From $86.4K/yr

Quick readStrong listing-quality and freshness signals

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As a Senior Internal Auditor reporting to the Senior Manager, Technology Internal Audit, you’ll help GitLab assess risk and strengthen controls across a technology landscape that includes multi-cloud infrastructure, artificial intelligence and machine learning systems, and modern development practices. This USA-based role supports our Sarbanes-Oxley Act (SOX) program while partnering with Engineering, IT Operations, Security, and business teams to build controls that work in practice, not just on paper. You’ll execute technology audits, turn findings into practical improvements, and use data analytics, au

GitRestAgileMachine Learning
O
📍 Atlanta, Georgia, United States· Full-time
✓ High-confidence listing

From $116.5K/yr

Quick readStrong listing-quality and freshness signals

Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Senior Software Engineer that will report to the Development Manager / R&D Head. In this role, you will part of the R&D Team that works on mission-critical applications. Your Mission Engage and partner with various Engineering, Operations, and Product teams to design, deliver, and maintain a highly available and performant application platform. Build and implement application observability and platform monitoring tools to continuously improve the customer experience Eliminate toil by automating processes, tuning alerts, and improving code where it is most needed Frequently evaluate new ideas and trends to identify potentially useful tools and techniques Collaborate with different functional groups to identify gaps, prioritize, and resolve issues Defining, implementing, and maintaining SLIs and SLOs aligned with customer experience. Design and instrument SLIs such as latency, error rates, and availability across critical services Manage and enforce error budgets to balance system reliability with product feature v

PythonJavaSQLAWS
O
📍 Atlanta, Georgia, United States· Full-time
✓ High-confidence listing

From $116.5K/yr

Quick readStrong listing-quality and freshness signals

Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Senior Software Engineer that will report to the Development Manager / R&D Head. In this role, you will part of the R&D Team that works on mission-critical applications. Your Mission Engage and partner with various Engineering, Operations, and Product teams to design, deliver, and maintain a highly available and performant application platform. Build and implement application observability and platform monitoring tools to continuously improve the customer experience Eliminate toil by automating processes, tuning alerts, and improving code where it is most needed Frequently evaluate new ideas and trends to identify potentially useful tools and techniques Collaborate with different functional groups to identify gaps, prioritize, and resolve issues Defining, implementing, and maintaining SLIs and SLOs aligned with customer experience. Design and instrument SLIs such as latency, error rates, and availability across critical services Manage and enforce error budgets to balance system reliability with product feature v

PythonJavaSQLAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. Safely delivering increasingly capable AI systems requires scalable technical safeguards, clear ownership of emerging risks, rigorous deployment readiness, and close coordination across research, engineering, product, operations, legal, policy, and external partners. Our Technical Program Managers lead complex, high-stakes initiatives that turn safety commitments into deployed systems and measurable outcomes. We work across model development, infrastructure, product, and operational response to help ensure our technology is deployed responsibly and cannot be used to cause serious real-world harm. About the Role We’re seeking Technical Program Managers to drive complex product, platform, and safety initiatives across ChatGPT, API, enterprise, and related deployment environments. These roles operate at the intersection of technical strategy and execution: you will turn safety and product priorities into actionable plans, influence architectural and operational decisions, and deliver durable capabilities across model, infrastructure, application, and platform layers. Depending on the role, you may enable sensitive or high-impact model deployments, integrate safeguards into cloud and API platforms, prevent violent misuse and other serious harms, improve detection and enforcement systems, create platform solutions for safety or establish new programs as risks evolve. You will partner deeply with engineers, researchers, product managers, and operational teams while communicating technical tradeoffs and program decisions to senior leadership. You bring technical fluency, product judgment, and a strong execution record. You’re comfortable navigating ambiguity, advocating for users and developers, balancing safety with model usefulness, and leading cross-functional work with urgency, rigor, and empathy. Specific focus areas and scope will vary by opening and level. Thi

AWSRestMachine LearningAI
LA
📍 New York, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

SUMMARY STATEMENT We are looking for a Solution Architect to design the technical solutions behind our client engagements and give delivery teams a clear, workable path from concept to production. You will work across enterprise data, software applications and GenAI - translating complex business problems into practical architectures that delivery teams can build and scale. This could include architecting an agentic workflow for clinical operations, a conversational analytics product grounded in enterprise data, or an AI-enabled decision platform for commercial teams. You will work directly with clients, define the architecture, test the most important technical decisions yourself and establish the foundations for successful delivery. This is an architecture-first role with meaningful hands-on engineering: you will stay close enough to implementation to prove the architecture works and support it through production delivery, without becoming the primary engineer for every component. You will also help shape the reusable patterns, technical standards and accelerators behind Lynx’s growing AI-native life sciences practice. KEY RESPONSIBILITIES Solution Architecture Own the end-to-end solution architecture for client engagements, including data models, system design, integration patterns and technology choices. Translate business requirements into clear technical designs and implementation paths that delivery teams can build from. Design solutions spanning enterprise data, APIs, applications, cloud platforms and GenAI capabilities. Lead technical discovery with clients: understand requirements, assess existing systems and identify dependencies, constraints and delivery risks. Present architectural options and trade-offs clearly to technical teams, business stakeholders and senior leaders. Make pragmatic decisions across build speed, cost, scalability, security and maintainability. Review key implementation decisions and remain

TypeScriptPythonAWSAzure
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Strategic Finance team at OpenAI plays a critical role in shaping the company’s long-term trajectory. We partner closely with Product, Engineering, and Go-To-Market teams to inform high-stakes decisions through rigorous data science and economic modeling. As part of our expanding Data Science function, we’re building a best-in-class Forecasting capability to drive real-time, data-driven decision-making across user growth, revenue, compute infrastructure, and more. We are developing scalable forecasting infrastructure to help us understand and anticipate business dynamics in an increasingly complex, usage-based world. Our models are foundational to planning, pricing, operational efficiency, and growth strategy - supporting key investment decisions and unlocking OpenAI’s full potential. About the Role We’re looking for a senior Machine Learning Data Scientist to lead our forecasting initiatives. You’ll be one of the founding members of the Forecasting pillar within Strategic Finance Data Science, responsible for building and scaling robust, interpretable, and production-ready forecasting systems. Your models will power critical business decisions by predicting core metrics such as DAU/WAU, revenue, LTV, compute consumption, and profitability. This is a highly cross-functional role, requiring technical excellence, strong product intuition, and business acumen. You’ll collaborate with product managers, researchers, engineers, and finance leaders to operationalize forecasting insights, influence company-wide strategy, and build foundational forecasting capabilities at OpenAI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Build statistical and machine learning models to solve forecasting needs across product, finance, infrastructure, and GTM domains. Own the end-to-end modeling lifecycle , including scoping, feature engineerin

PythonSQLAWSRest
S
📍 Bellevue, Washington, United States· Full-time
✓ High-confidence listingCompany trend -92.9%
Quick readStrong listing-quality and freshness signals

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Where Data Does More. Join the Snowflake team. Join our ML Feature Store team where we're building cutting-edge product capabilities that power complex feature transformations and low latency feature serving. We're revolutionizing machine learning feature management and serving capabilities as part of the Snowflake ML suite of products. In the era of GenAI and agents, our team delivers high-quality, fresh feature solutions that make a real difference for our customers. IN THIS ROLE AT SNOWFLAKE, YOU WILL: Help define and own the roadmap for Snowflake Feature Store, working collaboratively with senior architects and ML team leadership Build and execute a vision for incorporating new advances in machine learning Ensure operational excellence of services and meet reliability, availability, and performance commitments Collaborate across ML partner teams to improve development velocity and capabilities Support team members in delivering high technical quality WE WOULD LOVE TO HEAR FROM YOU IF YOU HAVE: 10+ years of experience in designing and building data serving infrastructure and/or machine learning platforms. Strong track record working with machine learning systems and platforms. Strong understanding of computer science fundamentals. B.Sc . in Computer Science Fluency in Ja

PythonJavaMachine LearningAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $345K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Senior Engineering Manager, Safety Platform You will lead engineering pods within the Safety Platform organization, driving the technical vision and execution for the foundational systems that power every safety workflow at Roblox. You will own the core safety platform end-to-end: from the shared infrastructure and APIs that enable trust & safety capabilities across the company, to the tooling ecosystem that empowers internal operators to investigate, intervene, and resolve issues at scale. The scope also includes the Safety agentic platform — AI-powered systems that automate and augment safety workflows across detection, enforcement, and review pipelines. As the platform layer beneath every safety surface, your work will define how quickly and reliably Roblox can respond to emerging threats. This role requires a leader with a strong platform mindset who can balance technical rigor with broad organizational impact — working across User Safety, Trust & Safety Policy, Data Science, and Machine Learning to deliver scalable, extensible, and high-availability systems for one of the world's largest platforms. You Will: Lead and Develop: Recruit, hire, mentor, and inspire a diverse team of

AWSGitMachine LearningAI
M
📍 O Fallon, Missouri, United States
✓ Quality checkedCompany trend +212.5%

Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior AI Platform Engineer (DevOps) Who is Mastercard? Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payment choices, making transactions secure, simple, smart, and accessible. Our technology and innovation, partnerships, and networks combined to deliver a unique set of products and services that help people, businesses, and governments realize their greatest potential. Our decency quotient (DQ) drives our culture and everything we do inside and outside our company. We cultivate an environment where individuals can thrive, collaborate, and contribute to innovations that power the global economy. Overview: The AI Platform Engineering team is responsible for building, operating, and evolving Mastercard's enterprise AI platforms and capabilities. Our mission is to provide scalable, secure, and reliable AI infrastructure that enables teams across Mastercard to accelerate the development and deployment of AI-powered solutions. As a Senior AI Engineer, you will help design, implement, and operate the foundational platforms that support AI and machine learning workloads across the enterprise. You will work at the intersect

PythonDockerKubernetesMachine Learning
C
📍 Work From Hom, United States
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a Senior Full Stack Engineer, Agentic AI Platform to join and lead engineering efforts for Lumina, CVS Health's enterprise knowledge and agentic AI platform. The Senior Full Stack Engineer, Agentic AI Platform will provide both technical and people leadership while driving the evolution of a platform that delivers trusted, secure, and governed AI experiences across the enterprise. As a Senior Full Stack Engineer, Agentic AI Platform, you will own the technical direction, architecture, scalability, and operational excellence of Lumina while leading a team of engineers and data scientists responsible for building and supporting the platform. This is a hands-on leadership role requiring active contribution to production code, architectural decision-making, code reviews, and engineering best practices, while simultaneously coaching and developing team members. The Senior Full Stack Engineer, Agentic AI Platform will play a critical role in advancing agentic AI capabilities, retrieval-augmented generation (RAG) systems, Model Context Protocol (MCP) integrations, enterprise search, knowledge ingestion, and AI governance. You will drive platform enhancements that improve retrieval quality, strengthen security and access controls, expand automation capabilities, and ensure the platform remains reliable, observable, and scalable as adoption accelerates across CVS

JavaScriptPythonReactSQL
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend +66.7%

From $224K/yr

Quick readStrong listing-quality and freshness signals

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Senior Solutions Engineer (Enterprise Pre-Sales) Secure Every Identity, from Human to AI Agent About the Role At Okta, we believe that identity is the foundation of security and digital transformation. As a Senior Solutions Engineer, you will be the trusted technical advisor to our largest, most complex enterprise customers and a critical driver of our sales organization. We are looking for a highly strategic Pre-Sales professional whose consultative expertise, emotional intelligence, and enterprise sales acumen are their defining strengths. While technical agility is required, your primary focus will be owning the technical sales cycle by anchoring technical features to positive business outcomes. You will partner closely with Enterprise Account Executives to uncover top-of-mind business challenges—specifically around mitigating risk, reducing costs, and driving operational efficiency. By establishing value-based conversations, you will prove how Okta’s independent, neutral, and end-to-end identity platform can transform their architecture and secure the "Tech Win" on large-scale deals. What You'll Be Doing Master the Discovery Process: Leverage exceptional active listening to dig deep into customer pain points. You will uncover the 'why' behind the initiative, focusing on how Okta's vast pre-built integrations and scalable architecture can solve their most complex, diverse environmental challenges. Navigate Complex Organizations: Translate highly technica

AWSGitRestMachine Learning
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

We at NVIDIA seek an Senior Developer Relations, Automated Synthetic Chemistry Science Lead who can coordinate science streams. This role exists at the crossroads of experimental science, optimization, analytical characterization, AI modeling, laboratory automation, and data infrastructure. It demands a practitioner with a background in research who can lead interdisciplinary science teams and accelerate experimentation without compromising scientific quality. What you'll be doing: Coordinate the operating plan across research priorities, technical execution, and program achievements. Convert research objectives into experimental priorities, parameter-space development, campaign planning, and success criteria. Lead research initiatives encompassing synthetic chemistry, catalysis, process chemistry, machine learning, automated systems, and data processing. Build standardized experimental traces capturing successful and unsuccessful outcomes, metadata, quality-control signals, analytical summaries. Guide AI systems for feasibility assessment, outcome prediction, optimization, scope exploration, and campaign orchestration. Integrate automated experimentation, analytical data streams, data curation, model retraining, and campaign decisions into a closed-loop operating model. Establish science-stream governance: build reviews, decision logs, risk and dependency tracking, quality thresholds, and paths for addressing blocking issues. Prepare recurring workstream readouts for program leadership, including scientific progress, critical decisions, cross-team dependencies, resource needs, and unresolved risks. What we need to see: PhD (or equivalent experience) with 10+ years of hands-on experience in experimental science, chemical engineering, material science, robotics, a

Machine LearningAI
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re hiring a People Business Partner to support our team during a critical phase of growth. This is a highly strategic, high-impact role for someone who has partnered closely with leadership teams and helped organizations scale with intention. You will be deeply embedded with managers, bringing clarity and rigor to how teams are structured, how leaders operate, and how talent is developed across the organization. You’ll help shape team effectiveness, identify critical talent gaps, drive talent and performance strategies that enable high-performing teams, and build the people practices and change management approaches that allow us to scale with both speed and discipline. This role requires strong business judgment and the ability to operate with deep context. You’ll partner closely with leaders to navigate complex organizational decisions, anticipate challenges before they surface, and bring a clear point of view on what great looks like at every level of the organization. RESPONSIBILITIES Strategic partnership to leadership Serve as the trusted people partner to leadership, maintaining deep business context and translating it into people priorities by proactively surfacing systemic issues, risks, and opportunities before they become urgent. Bring data-driven insights to advise management on org design, succession planning, performance, retention, and engagement. Build management capacity across the org, equ

Machine LearningAIGoRust
🔔

Get new senior machine learning operations engineer jobs in United States by email

Daily job updates · Unsubscribe anytime