Jobs in United States

Senior Machine Learning Operations Engineer in United States

1,941 active opportunities · Updated October 2026

Explore current senior machine learning operations engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $260.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Why Safety? At Roblox, we strive to connect a billion people with optimism and civility, and the Safety organization’s mission is to become the leader in civil immersive online communities. We systematically and proactively work to detect, remove, and prevent problematic content and behavior. We seek to influence and shape the product roadmap and prioritization, build safety products, and measure our impact on the community of

AWSGitMachine LearningAI
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? At Cohere, we believe in the power of multimodal AI to revolutionise the way we interact with technology. Our engineering teams push the boundaries of what's possible, and we're looking for talented individuals to join us on this exciting journey. With an exceptional ratio of compute resources to engineers, we provide an ideal environment for you to explore, innovate and shape the future of AI. July 31st 2025 - Cohere's Multimodal team Introduced Command A Vision: Multimodal AI Built for Business. At release our new flagship vision-language model: ● Consistently outperforms major models like Llama 4 Maverick, Mistral Medium/Pixtral Large, and GPT4.1 ● 83.1% average benchmark (73.5% MathVista, 90.9% ChartQA...) ● Built for the real world - 112B parameters running on just 2 GPUs ● Open weights live on HuggingFace With a focused team, breakthrough performance doesn't require breakthrough compute. Focus on the things that matter, and join the team. As a Member of Technical Staff with a focus on Multimodal AI, you will: Design and develop cutting-edge multimodal AI systems, integrating various modalities such as text,

PythonGitMachine LearningAI
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 ClickUp is seeking a talented and experienced Senior Search Engineer to join our team and help revolutionize our search capabilities. As a key member of our engineering team, you'll be responsible for optimizing and enhancing our search functionality, which is a critical feature of our platform and underpins our AI efforts. Your impact at ClickUp As a Senior Engineer on the Search squad, you'll be responsible for designing, optimizing, and scaling our search infrastructure. Our platform is built as backend services on top of Postgres and OpenSearch, focused on real-time search ingestion and serving. You'll work hands-on with these systems to ensure users can instantly find exactly what they need across their workspace. Your expertise will directly contribute to making ClickUp the most intuitive and responsive productivity platform available. Core Responsibilities Design and implement robust search solutions that scale with our rapidly growing user base Improve search relevance, accuracy, and speed to deliver the most relevant results to users at blazing fast speeds Improve our real-time indexing pipelines to ensure search results remain up-to-date Create measurement frameworks to evaluate and improve search quality Build and enhance vector search capabilities to power next-generation search experiences Collaborate with AI, backend, and product teams to integrate search into new features Troubleshoot complex search-related issues at scale Design and implement robust search solutions that scale with our rapidly growing user base Improve search relevance, accuracy, and speed to deliver the most relevant r

TypeScriptAWSMachine LearningAI
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 ABOUT THE ROLE We are hiring the next generation of product managers. AI has collapsed the work that used to sit between an idea and a working version of it, and for the first time, PMs can be true builders. The PMs who win in this era do not manage the work. They build the product, every day, with their own hands. A Senior Product Manager owns a meaningful surface of ClickUp and is accountable for whether it gets better. You decide what to build and why, then go build a working version of it. You prototype in Cursor and Claude when an idea is faster shown than written. You live in the data and the customer signal, and the time between spotting a problem and having something people can react to is measured in days, not weeks. The PMs who compound the most impact at this level are the ones who keep finding leverage that did not exist last quarter. KEY RESPONSIBILITIES Own the vision, roadmap, and delivery for a defined product surface, balancing near-term wins with longer-horizon investments. Run continuous discovery. Building and using agents to monitor and synthesize that part of your job. Write sharp PRDs and specs that give engineering and design what they need to move fast without ambiguity. Define success metrics before you build, track them after you ship, and iterate when outcomes do not match expectations. Prototype your own ideas using Cursor, Claude Code, and our Prototype Playground. Show the concept before filing the ticket. Collaborate cross-functionally with engineering, design, analytics, and GTM without needing to be managed through the process. Present your roadmap and results clearly

ReactAWSMachine LearningAI
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.6%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Senior Software Engineer on Sentry’s AI team, you’ll be directly responsible for developing the platform used by our debugging agents. This role is crucial; you will be at the forefront of integrating AI and machine learning into our core products, from issue triage and resolution to predictive analytics for application performance monitoring. Your work will help companies around the globe gain actionable insights into their software, enabling them to build better products, faster. In this role you will Build state-of-the-art agentic AI platforms to triage, debug, and solve real production issues Leverage Sentry’s novel (and massive) dataset of errors, spans, and profiles Own the development of major initiatives in the AI/ML space You'll love this job if you Are driven by impact and enjoy working on high-stakes, high-visibility projects Enjoy building things. You will have the opportunity to join the AI/ML team as one of its foundational members Thrive in cross-functional teams and enjoy building features alongside developers and product teams Qualifications Minimum 5+ years of professional experience with Bachelor’s degree in computer science, machine learning, or a related field Demonstrated expertise building production-grade agentic systems and tools You are comfortable writing production quality code (we use Python and Typescript) Familiarity with deep learning frameworks (we use PyTorch) Familiarity in deploying machine learning models at scale in production environments The base salary range (or hourly wage range, if applicable) that Sentry reasonably expects to pay for this position is $155,000 to

TypeScriptPythonMachine LearningAI
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.6%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Senior Software Engineer on Sentry’s AI/ML team, you’ll be responsible for building the evaluation infrastructure that measures the accuracy, reliability, and real-world performance of our AI systems. This role is critical to ensuring that our debugging agents and AI-powered features behave correctly, safely, and predictably as they scale. You’ll design datasets, benchmarks, and test harnesses that turn ambiguous AI behavior into measurable signals, helping the team ship AI with confidence. In this role you will Design and build robust evaluation frameworks to measure accuracy, reliability, regressions, and edge cases in AI systems Create and curate high-quality datasets, golden test cases, and benchmarks grounded in real production data Build automated test harnesses and metrics pipelines to continuously evaluate models, prompts, and agentic workflows Partner closely with applied AI engineers and product leaders to define what “good” looks like and translate it into measurable criteria Own the evaluation lifecycle for major AI initiatives, from early experimentation through production monitoring You’ll love this job if you Care deeply about correctness, rigor, and measurement in AI systems Enjoy turning fuzzy product goals and model behavior into concrete tests and metrics Like building foundational infrastructure that unlocks faster iteration and higher confidence for the entire AI team Thrive in cross-functional environments and enjoy influencing model design through better evaluation Qualifications Minimum 5+ years of professional experience with a Bachelor’s degree in computer science, machine learni

TypeScriptPythonMachine LearningAI
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Financial Management team develops market leading API products that power the most influential digital finance experiences that help millions of consumers and small businesses every day - think budgeting apps and financial management tools offered by banks and wealth platforms. Our mission is to unlock financial freedom for everyone by making it easier for customers and businesses to meet their financial goals. We do this through a state-of-the art, data aggregation and machine learning engine that enables our customers to seamlessly build delightful experiences on top of consumer-permissioned financial data. Every day, Plaid helps thousands of developers build a better financial future. You'll be responsible for defining the products which enable Plaid and our customers to shape the future of financial services. We are looking for the right product manager to own and grow Plaid’s financial management products, including Transactions, Investments, and Liabilities, even more valuable to our customers and partners. Plaid is on a journey to build insights on top of open banking data that drive value for customers, consumers and the fintech ecosystem in general. Additionally AI is changing how consu

AWSGitMachine LearningAI
SF
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $200K/yr

Quick readStrong listing-quality and freshness signals

About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role At Stitch Fix, we are at the forefront of innovation, creating cutting-edge solutions that blend fashion, technology, and data science. Our data science team combines machine learning with expert human judgment to generate innovative recommendations and insights that transform the way our clients discover what they love. We believe in a curiosity-driven data science culture where members are empowered to deliver impact through end-to-end model development. The diversity of the problems that we work on and the data-rich environment of our business make it possible, even essential, to bring the tools of multiple disciplines to bear on our hardest problems. We are looking for an experienced Styling Algorithms Team Manager to lead a group of talented machine learning engineers and data scientists. In this role, you will shape the future of fashion technology by driving the development and deployment of our styling algorithms, which empower our human stylists to delight clients by nailing their fit and style. This includes ML-, AI-, and product-driven feature curation and testing for our proprietary styling platform, as well as client-facing AI personalization experiences, such as Stitch Fix Vision, our virtual try-on. Responsibilities: Champion bold AI and ML interventions to improve our styling experiences, enabling our stylists to have a multiplicative impact on their client connection points. Likewise, actively shape the product roadmap for direct client-facing styling experiences, expand

PythonRestMachine LearningAI
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the Team The Finance Data Science team owns the forecasting systems that power Snowflake’s revenue planning and long-term financial strategy. Our work supports corporate planning, executive decision-making, and investor reporting, and we partner closely with Product and Sales to understand customer behavior and product impact. We operate at the intersection of machine learning, statistical research, and corporate finance, building production-grade forecasting infrastructure that is foundational to how the company plans and operates. The Role As a Senior Data Scientist, you will independently lead high-impact modeling initiatives and build production-ready forecasting systems for core financial metrics. You will work on complex, open-ended problems at the intersection of machine learning and business strategy, translating real-world financial questions into rigorous, scalable models. What You’ll Do Design and implement advanced time-series and probabilistic models (e.g., hierarchical models, state-space models, Bayesian approaches, multivariate forecasting). Contribute to internal tooling and shared infrastructure that enables scalable forecasting and analytics. Establish best practices for model evaluation, backtesting, uncertainty quantification, and scenario simulat

PythonSQLMachine LearningAI
S
📍 Chicago, Illinois, United States· Full-time
✓ High-confidence listingCompany trend -92.9%
Quick readStrong listing-quality and freshness signals

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our Solution Engineering organization is seeking an AI Specialist who can provide hands-on expertise and support while working with technical decision makers and data scientists to design and architect AI solutions built on the Snowflake AI Data Cloud. This is a strategic role that works closely with cross-functional teams, including product, engineering, and the broader field organization to ensure successful execution and customer adoption of Snowflake’s AI & ML solutions. IN THIS ROLE YOU WILL GET TO: Be the technical expert in the room that positions Snowflake’s AI and ML features and value to technical stakeholders at Snowflake’s customers across the Americas. Partner with Snowflake account team teams and customer champions to scope and drive POCs to success and technical wins that prove the value of Snowflake’s capabilities, including executive readouts and business value cases. Collaborate with Snowflake’s product and engineering teams to influence Snowflake’s AI and ML roadmaps based on customer feedback. Publish content that helps the team and company scale beyond your individual efforts, like blog posts, presentations at conferences, or technical collateral like notebooks and demos. Influence, tailor and maintain Sales Engineering AI and ML selling assets, inc

PythonAWSAzureGCP
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our Solution Engineering organization is seeking an AI Specialist who can provide hands-on expertise and support while working with technical decision makers and data scientists to design and architect AI solutions built on the Snowflake AI Data Cloud. This is a strategic role that works closely with cross-functional teams, including product, engineering, and the broader field organization to ensure successful execution and customer adoption of Snowflake’s AI & ML solutions. IN THIS ROLE YOU WILL GET TO: Be the technical expert in the room that positions Snowflake’s AI and ML features and value to technical stakeholders at Snowflake’s customers across the Americas. Partner with Snowflake account team teams and customer champions to scope and drive POCs to success and technical wins that prove the value of Snowflake’s capabilities, including executive readouts and business value cases. Collaborate with Snowflake’s product and engineering teams to influence Snowflake’s AI and ML roadmaps based on customer feedback. Publish content that helps the team and company scale beyond your individual efforts, like blog posts, presentations at conferences, or technical collateral like notebooks and demos. Influence, tailor and maintain Sales Engineering AI and ML selling assets, inc

PythonAWSAzureGCP
O
📍 San Francisco, California, United States
✓ High-confidence listingCompany trend +66.7%
Quick readStrong listing-quality and freshness signals

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta is the identity standard. The Okta Platform is an independent and neutral platform that securely connects the right people to the right technologies at the right time. We help organizations do two things - secure and manage their extended enterprise, and transform their users’ experiences. Okta's Core Engineering team is responsible for building and evolving shared infrastructure and services that lay the foundation for what other engineering teams build on. We're in charge of common shared services like search, cache, configuration management, frameworks for async job management, and email pipeline, to name a few. We're cloud native, where redundancy, multi-tenancy, scale, resource optimization and resiliency are first class citizens. With Okta's mantra of 'Always On!' there's never a dull moment. Our biggest asset is our team of passionate engineers and technically minded managers. We're looking for a staff level backend engineer to join a team of highly skilled and talented team players who're proud of what they own and deliver. Our elite team is fast, creative and flexible; with a weekly release cycle and individual ownership we expect great things from our engineers and reward them with stimulating new projects, new technologies and the chance to have significant equity in a company that is changing the cloud computing landscape forever. You will: Work with engineering teams to design, develop and deliver cloud based infrastructu

JavaRedisAWSDocker
C
📍 Work At Home Massachusetts, United States
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary Health100 is America's trusted front door to health and care. The Health100 platform integrates any participating health plan, PBM, pharmacy (retail and specialty), provider, digital health point solution provider, and employer, and addresses the top health care challenges for the consumer. The Senior Manager, Software Engineering will lead engineering teams building a best-in-class consumer health experience on the Health100 platform, focused on identifying, prioritizing, shaping, and executing complex platform initiatives. As a key member of our engineering organization, you will drive innovation, manage cross-functional teams, and deliver scalable cloud-native and AI-enabled solutions that improve consumer experiences and business outcomes. This role works within a top-notch organization of software developers who identify, design, and deliver technology solutions using Java, distributed systems, cloud platforms, and AI-powered capabilities to achieve defined business value. You will guide the integration and validation of complex technology solutions, drive engineering excellence, and leverage emerging technologies including Generative AI and intelligent automation to accelerate innovation and efficiency *Remote eligible within the United States. Preference for candidates in close proximity to our Woonsocket, RI headquarters. Resp

JavaGCPDockerKubernetes
M
📍 O Fallon, Missouri, United States
✓ High-confidence listingCompany trend +212.5%
Quick readStrong listing-quality and freshness signals

Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Software Engineer Who is Mastercard? We work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion for all employees that respect their individual strengths, views, and experiences. We believe that our differences enable us to be a better team – one that makes better decisions, drives innovation, and delivers better business results About the Role Senior Software Engineers at Mastercard design and code artificial intelligence, cloud, and machine learning platforms that provide mission-critical insights to many of the world’s leading organizations and governments. As a Senior Software Engineer, you will deliver these products and solutions with speed and agility as part of a small team. This will involve developing high-performing, highly scalable software solutions and products for some of the world’s top brands. Specific tasks vary depending on the project and the busi

JavaSQLPostgreSQLAWS
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

NVIDIA is seeking a Senior System Architect: Heterogeneous EDA Systems to solve a complex challenge in accelerated computing: Failure Attribution at Scale. As EDA or equivalent experience workloads scale across thousands of heterogeneous nodes, a single failure can cause massive resource waste. We need an engineer to develop and build an automated framework. This framework will ingest telemetry from CPU and GPU clusters to identify the root cause of job failures in real-time. It will distinguish between hardware faults, infrastructure instability, and software defects. What you'll be doing: Architect Failure Attribution Frameworks: Build a scalable "flight recorder" for EDA jobs that captures high-fidelity state across the CPU, GPU, and Fabric at the moment of failure. Build automated diagnostics that correlate GPU XID errors, PCIe bus failures, and CUDA memory exceptions. Connect these errors with system-level events such as OOM kills or NUMA-related hangs. Distributed Logging & Tracing: Implement low-overhead tracing mechanisms (using tracing tools or custom agents) that provide access to job execution across multi-node Slurm or Kubernetes clusters. Root Cause Automation: Develop heuristics and models based on machine learning to classify failures as "Hardware Fault," "Software Bug," or "Environment Issue." This reduces the Mean Time to Identify (MTTI) for R&D teams. Resiliency Engineering: Work closely with hardware and infrastructure teams to define "signals of impending failure," enabling proactive job migration or check-pointing before a crash occurs. What we need to see: Distributed Systems Mastery: BS, MS, or PhD in Computer Science or Electrical Engineering (or equivalent experience) with 6+ years in systems programming. Experience building automated

PythonKubernetesLinuxMachine Learning
🔔

Get new senior machine learning operations engineer jobs in United States by email

Daily job updates · Unsubscribe anytime