PagerDuty (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on PagerDuty to succeed with Digital Transformation, Cloud Migration, and DevOps Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, DoorDash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization. PagerDuty is looking for a Machine Learning Engineer who is passionate about collaborating with data scientists, product managers and engineers alike. As part of our team, you will help us accelerate the development and extension of products powered by Gen AI and many other shapes of Machine Learning. You’ll be contributing hands-on to the development of the services and pipelines that enable multiple ML/AI features in our product. You will have the opportunity to collaborate with multiple organizations, taking input and guidance from your senior stakeholders and helping bring our initiatives to reality. You’ll succeed by showcasing excellent capacity to manage time, demonstrating emotional intelligence as you navigate stakeholder relationships, and by continuously improving your technical skill set. Key Responsibilities Build and improve the capabilities that enable and accelerate the production of machine learning (ML) and generative AI (genAI) based solutions Partner with data scientists, effectively sharing engineering context and collaborating to support larger initiatives Incorporate the best ava
Jobiba hiring network
Senior Machine Learning Operations Engineer Jobs
7,292 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current senior machine learning operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
PagerDuty (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on PagerDuty to succeed with Digital Transformation, Cloud Migration, and DevOps Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, DoorDash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization. PagerDuty is looking for a Machine Learning Engineer who is passionate about collaborating with data scientists, product managers and engineers alike. As part of our team, you will help us accelerate the development and extension of products powered by Gen AI and many other shapes of Machine Learning. You’ll be contributing hands-on to the development of the services and pipelines that enable multiple ML/AI features in our product. You will have the opportunity to collaborate with multiple organizations, taking input and guidance from your senior stakeholders and helping bring our initiatives to reality. You’ll succeed by showcasing excellent capacity to manage time, demonstrating emotional intelligence as you navigate stakeholder relationships, and by continuously improving your technical skill set. Key Responsibilities Build and improve the capabilities that enable and accelerate the production of machine learning (ML) and generative AI (genAI) based solutions Partner with data scientists, effectively sharing engineering context and collaborating to support larger initiatives Incorporate the best ava
Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role As a Senior/Staff Software Engineer working on driving behavior verification, you are responsible for implementing metrics that evaluate the end-to-end behavior of the Nuro Driver. These metrics will be used to quantify the safety of the driving behavior in our target ODD. This requires prior experience with the development or verification of behavior planning/prediction systems for robots, and a collaborative nature to work closely with a variety of teams across Nuro: Systems, Onboard Software, Simulation, Product, and Operations. About the Work Develop and implement in Python generalizable metrics to verify the driving behavior of an autonomous vehicle. Leverage a combination of machine learning (ML) models and safety metrics from literature to evaluate the end-to-end driving behavior. Evaluate these metrics on a variety of tests: synthetic and log simulation, on-road logs, closed-course testing data, and third-party acc
GTM Operations & Strategy at MongoDB is a global team of builders and innovators focused on unleashing MongoDB’s sales greatness by pairing world‑class analytics with scalable operations. Within GTM Operations, the GTM Intelligence – Applied Science team turns complex GTM data into tools, models, and insights that help our sales organization make better, faster decisions across our people, segmentation, territory design, forecasting, and account prioritization. As a Senior Analyst on the Applied Science team, you will own high‑impact analytical workstreams end‑to‑end: from problem framing with senior GTM stakeholders, to data engineering and model design, through to productionalized workflows, dashboards, and executive‑ready narratives that drive concrete changes in the field. This role is based in Dublin, Ireland and supports a global stakeholder set across regions and GTM functions. We are looking to speak to candidates who are based in Dublin or Cork for our hybrid working model. What You’ll Do Translate GTM questions into analytical projects Partner with GTM Ops, Sales Strategy & Planning, Sales Leadership, and Central Analytics to scope problems, define success criteria, and prioritize work across areas like segmentation, territory design, account prioritization, and pipeline/forecast health. Structure ambiguous questions into hypotheses, analytical plans, and clear recommendations for senior stakeholders (SVPs, RVPs, functional leaders). Design and build scalable analytics & models Develop and maintain statistical and machine learning models (e.g., NARR prediction, deal qualification, account momentum, workload identification) that inform forecast expectations, territory assignments, and deal prioritization. Engineer robust data pipelines and features (SQL/Python) on top of our GTM data stack (Salesforce, product usage, call transcripts, marketing signals, etc.) in partnership with data and platform teams. Own core GTM analytics assets Contribute t
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview: We are seeking a skilled and experienced Senior AI Engineer - Multi-Agent Frameworks to join our AI Platform team. In this role, you will play a pivotal part in building a cutting-edge platform that empowers our users to create and deploy sophisticated intelligent agents, with a key focus on enabling collaborative and multi-agentic behaviors . This is a backend-focused role that requires deep expertise in AI, large language models (LLMs), and orchestration software. Key Responsibilities: Design, develop, and maintain a robust platform to enable users to create and manage AI agents and their interactions. Integrate and work with multiple LLMs, ensuring seamless orchestration and scalability for both individual and coordinated agent operations. Leverage orchestration frameworks like LangGraph and others to build complex workflows and pipelines that support diverse agent functionalities, including frameworks for multi-agent coordination . Develop and implement evaluation frameworks for testing AI agents in challenging and complex scenarios, focusing on individual performance and system-level dynamics. Stay at the forefront of AI advancements, incorporating the latest research and technologies into our platform to enhance agent capabilities and collaboration. Collaborate with cross-functional teams, including product managers, designers, and frontend engineers, to deliver a seamless user experience for building and deploying intelligent systems. Address challenging AI privacy scenarios, ensuring compliance with data protection regulations and best practices within agent-based applications. C
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 As Senior Manager, CX Operations at ClickUp, you will own strategy, execution, and operational leadership for Technical Account Management (TAM) Operations , spanning Professional Services and Customer Success. Partnering closely with TAM leadership, you will serve as the day-to-day operational leader for the systems, automation, planning, quality, and insights that power our Services & Success organization. You will set the operating rhythm and performance standards for TAM Operations while owning the AI agent harness and automation layer within your domain. You will shape the cadence of business for TAM, unlock data insights, coach organizational performance, transform internal capabilities, and influence strategic decisions across Services, Success, and cross-functional leadership. You will architect, ship, operate, and iterate on AI-driven workflows that enhance and automate customer health inspection, engagement management, services delivery, quality measurement, and decision support across TAM and CX. About the Role Strategy & Operations Own the vision and strategy for delivering world-class customer experience through our Services & Success operating model Drive cross-functional alignment across Sales, Product & Engineering, Finance, and Support by synthesizing customer health, retention, and services delivery opportunities into operational priorities Lead and execute strategic initiatives to optimize and transform customer engagement, services delivery, and internal collaboration processes Extract key business insights from qualitative and quantitative data, identify risks and o
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The role At Okta, we believe AI will fundamentally transform how design, research, and product development happen. We are looking for a Senior Design Operations Program Manager, Agentic Design to lead the operationalization and enablement of internal AI agents, plug-ins, and automated workflows across our Design team — reimaging how designers create in the AI era through internal AI agents. In this role, you will bridge the gap between AI tooling capabilities and practical design ops execution. You will partner closely with design, product management and design systems to govern, evaluate, deploy, and scale AI-driven design capabilities — making our 60+ person design org dramatically faster, more consistent, and highly innovative. This is a tactical, execution-focused builder role for an operations leader who is deeply curious about AI tools, understands modern UX workflows, and excels at change management, team enablement, and operationalizing new frameworks. What you'll do AI & Agentic Design Enablement Serve as the operational connective tissue driving AI adoption, evaluation, and workflow integration across Product Design, UX Research, and Content Design. Manage the intake, evaluation, and approval process for design plug-ins, Model Context Protocol (MCP) integrations, and generative AI tools (e.g., custom content design assistants, design-to-code agents). Partner with internal engineering teams to operationalize, ship, and drive adopt
At Lyft, our mission is to improve people's lives with the world's best transportation. To accomplish this, we start with our community by creating an open, inclusive, and diverse organization. About the Team The Risk Tech engineering organization is committed to tangibly reducing accident frequency, saving lives, and managing costs to enhance the safety and affordability of rides. Claim Management is a core financial function for Lyft. Each claim touches complex workflows, multiple stakeholders, sensitive data, financial reserves, regulatory processes, and significant financial liabilities. This role offers the opportunity to define a leading claims management system for the industry. Our vision is to establish a single, Unified Risk Platform where comprehensive claims workflows across all business lines are efficiently administered, communications are consolidated, and data is structured to facilitate data-driven insights and decisions.This enables cost-efficient claims operations, mitigates risks and expenses as Lyft scales, and ensures people receive assistance proactively and accurately. About the Role We are seeking a Senior Software Engineer to contribute to the technical direction, drive architectural decisions, and lead the development of a highly reliable, scalable, and intelligent Risk Management Information System that powers our insurance platform. You will collaborate with passionate colleagues from Engineering, Data Science, Product and Claim Operations to deliver end-to-end solutions. Responsibilities: Define and drive the long-term technical roadmap for claims management systems, aligning priori
At Lyft, our mission is to improve people's lives with the world's best transportation. To accomplish this, we start with our community by creating an open, inclusive, and diverse organization. About the Team The Risk Tech engineering organization is committed to tangibly reducing accident frequency, saving lives, and managing costs to enhance the safety and affordability of rides. Claim Management is a core financial function for Lyft. Each claim touches complex workflows, multiple stakeholders, sensitive data, financial reserves, regulatory processes, and significant financial liabilities. This role offers the opportunity to define a leading claims management system for the industry. Our vision is to establish a single, Unified Risk Platform where comprehensive claims workflows across all business lines are efficiently administered, communications are consolidated, and data is structured to facilitate data-driven insights and decisions.This enables cost-efficient claims operations, mitigates risks and expenses as Lyft scales, and ensures people receive assistance proactively and accurately. About the Role We are seeking a Senior Software Engineer to contribute to the technical direction, drive architectural decisions, and lead the development of a highly reliable, scalable, and intelligent Risk Management Information System that powers our insurance platform. You will collaborate with passionate colleagues from Engineering, Data Science, Product and Claim Operations to deliver end-to-end solutions. Responsibilities: Define and drive the long-term technical roadmap for claims management systems, aligning priori
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 About This Role ClickUp is looking for a Senior Database Reliability Engineer to join our Database Operations team. You’ll be responsible for the performance, integrity, security, and availability of our PostgreSQL databases running on Linux in AWS. This role focuses on day-to-day database administration, operational excellence, and ensuring data is consistently reliable and well-managed across environments. Key Responsibilities Administer, monitor, and maintain PostgreSQL databases (250GB+) in production environments (AWS RDS, Aurora, EC2) Ensure database availability, performance, and data integrity through proactive monitoring and maintenance Perform routine database administration tasks including patching, upgrades, vacuuming, reindexing, and statistics management Execute and manage backup, restore, and disaster recovery procedures, including regular testing of recovery plans Handle user access management, roles, and database security to ensure compliance with best practices Perform capacity planning, storage management, and growth forecasting Troubleshoot database issues, including performance bottlenecks, locking/contention, and failed jobs Support application teams with query tuning, schema changes, and release deployments Manage database migrations, upgrades, and change requests with minimal downtime Maintain documentation for database configurations, standards, and operational procedures Participate in on-call rotations and provide support for production incidents Qualifications 7+ years of experience in a senior database administrator or Database Engineering role Strong hands-on PostgreSQL ad
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Okta Privileged Access Management (PAM) is an identity-centric approach to a common and critical privileged access use case. Our elegant Zero Trust architecture is purpose-built for the modern cloud and helps customers solve challenging security and operations pain points at scale. We are looking for a software engineer to join our fast-growing team with a focus on scalability, reliability, and enhancing the core building blocks of the product. In this role you will: Be deeply involved in evolving the core architecture of PAM. Work in our product development teams to build scalable, composable components of our platform. Be responsible for designing and implementing scalable architecture patterns. Delight our customers by providing world class UX using our React-based design system Design and build APIs that customers rely on for access to production infrastructure. Work on backend components written in Go and frontend components written in React. You might be a good fit if you: Have 3-5 years of software development experience with a background in Golang or similar programming languages. Proficient in React or similar front-end UI stacks. Experienced working with relational databases like PostgreSQL or similar RDBMS technologies. have the ability to complete a feature end to end from designing database models to backend APIs and frontend UI components. Experienced working with any cloud provider such as AWS, GCP or Azure. Thrive in a collaborativ
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Okta Privileged Access Management (PAM) is an identity-centric approach to a common and critical privileged access use case. Our elegant Zero Trust architecture is purpose-built for the modern cloud and helps customers solve challenging security and operations pain points at scale. We're looking for a Senior level Platform Engineer to join a team of highly skilled and talented team players who are proud of what they own and deliver. Our elite team is fast, creative, and flexible; with a weekly release cycle and individual ownership, we expect great things from our engineers and reward them with stimulating new projects, new technologies, and the chance to have significant equity in a company that is changing the cloud computing landscape forever. What you’ll do Leverage cutting-edge AI pair-programmers and LLMs (such as Copilot and Claude) to accelerate the development of secure, enterprise-grade Privileged Access Management (PAM) products. Work with engineering teams to design, develop and deliver cloud-based infrastructure projects on a modern tech stack (Kubernetes/EKS, RDS, DynamoDB, Kinesis, MKS, Redis, OpenSearch, Docker, Terraform on AWS) Drive evaluation, development, and rollout of microservices Operate, support, and upgrade shared services and frameworks. Scale these as their usage invariably grows along with Okta's business. Evaluate and scale existing systems to meet specialized requirements and support Okta’s future business ne
About Graphcore Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Join our dynamic Software Infrastructure team and take a pivotal role in scaling and managing our infrastructure. You will develop essential tools and services that empower our broader software team. Your contributions will enhance the build, test, deployment, and productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed systems. The Team An exciting opportunity to join a new team within the Software Operations group. The Build Engineering team is a new function within Software Infrastructure, which focuses on the overall process of building and integration of the Machine Learn ing S oftware S tack. You will work closely with the QA and development teams to get an understanding of how our ML SW stack is built, helping to ensure good build practices, and proving that the stack works together and is reproducible in secure, sandboxed environments. Responsibilities and Duties Developing our internal t
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The BizTech team at Airbnb is crucial to the company's operations, handling critical data related to compliance with Tax, Payments, and Legal regulations. We also manage application data for tools such as CRM, Jira, and Workday, which are essential for Airbnb’s business. Joining this team means working with cross-functional stakeholders, designing scalable solutions, and contributing to a world-class data engineering environment with an emphasis on quality, scalability, and robust engineering practices. The Difference You Will Make: We are looking for a hands-on expert to provide technical leadership in addressing BizTech’s diverse data engineering needs and driving long-term strategies and best practices. This key leadership role requires strong collaboration and influence across teams. You'll play a crucial role in understanding business needs, identifying the right data sources, designing efficient data models, and building reliable, scalable data pipelines. As technology continues to evolve, you'll help shape and maintain significant parts of BizTech’s critical data ecosystem. Your contributions will not only address complex business challenges but also help refine and advance Airbnb’s Data Engineering Paved Path, benefiting the entire data community at Airbnb. We believe in solving problems and contributing back to our data community to continuously improve. A Typical Day /Responsibilities: Lead the design, implementation, and testing of data systems, from architecture to production. Build batch and real-time data systems that support business needs and critical products. Ensur
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The BizTech team at Airbnb is crucial to the company's operations, handling critical data related to compliance with Tax, Payments, and Legal regulations. We also manage application data for tools such as CRM, Jira, and Workday, which are essential for Airbnb’s business. Joining this team means working with cross-functional stakeholders, designing scalable solutions, and contributing to a world-class data engineering environment with an emphasis on quality, scalability, and robust engineering practices. The Difference You Will Make: We are looking for a hands-on expert to provide technical leadership in addressing BizTech’s diverse data engineering needs and driving long-term strategies and best practices. This key leadership role requires strong collaboration and influence across teams. You'll play a crucial role in understanding business needs, identifying the right data sources, designing efficient data models, and building reliable, scalable data pipelines. As technology continues to evolve, you'll help shape and maintain significant parts of BizTech’s critical data ecosystem. Your contributions will not only address complex business challenges but also help refine and advance Airbnb’s Data Engineering Paved Path, benefiting the entire data community at Airbnb. We believe in solving problems and contributing back to our data community to continuously improve. A Typical Day /Responsibilities: Lead the design, implementation, and testing of data systems, from architecture to production. Build batch and real-time data systems that support business needs and critical products. Ensur
Get new senior machine learning operations engineer jobs by email
Daily job updates · Unsubscribe anytime