Jobiba hiring network

Ml Platform Engineer Jobs

832 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current ml platform engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

P
1mo ago

PagerDuty (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on PagerDuty to succeed with Digital Transformation, Cloud Migration, and DevOps Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, DoorDash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization. PagerDuty is looking for a Machine Learning Engineer who is passionate about collaborating with data scientists, product managers and engineers alike. As part of our team, you will help us accelerate the development and extension of products powered by Gen AI and many other shapes of Machine Learning. You’ll be contributing hands-on to the development of the services and pipelines that enable multiple ML/AI features in our product. You will have the opportunity to collaborate with multiple organizations, taking input and guidance from your senior stakeholders and helping bring our initiatives to reality. You’ll succeed by showcasing excellent capacity to manage time, demonstrating emotional intelligence as you navigate stakeholder relationships, and by continuously improving your technical skill set. Key Responsibilities Build and improve the capabilities that enable and accelerate the production of machine learning (ML) and generative AI (genAI) based solutions Partner with data scientists, effectively sharing engineering context and collaborating to support larger initiatives Incorporate the best ava

pythonawskubernetes
View job →
P
1mo ago

PagerDuty (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on PagerDuty to succeed with Digital Transformation, Cloud Migration, and DevOps Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, DoorDash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization. PagerDuty is looking for a Machine Learning Engineer who is passionate about collaborating with data scientists, product managers and engineers alike. As part of our team, you will help us accelerate the development and extension of products powered by Gen AI and many other shapes of Machine Learning. You’ll be contributing hands-on to the development of the services and pipelines that enable multiple ML/AI features in our product. You will have the opportunity to collaborate with multiple organizations, taking input and guidance from your senior stakeholders and helping bring our initiatives to reality. You’ll succeed by showcasing excellent capacity to manage time, demonstrating emotional intelligence as you navigate stakeholder relationships, and by continuously improving your technical skill set. Key Responsibilities Build and improve the capabilities that enable and accelerate the production of machine learning (ML) and generative AI (genAI) based solutions Partner with data scientists, effectively sharing engineering context and collaborating to support larger initiatives Incorporate the best ava

pythonawskubernetes
View job →

About Anyscale At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray , a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI , Uber , Spotify , Instacart , Cruise , and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world. With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert. Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250+ million raised to date. About the role The Customer Engineer will play a crucial role in the customers’ post-sale journey - helping them to onboard, adopt and grow on Anyscale, troubleshooting and resolving open customer tickets and driving consumption. Anyscale is an ever evolving platform and hence will require close co-ordination with our engineering teams to debug complex issues. This is an exciting role for those who are technically curious and passionate about ML/AI, LLM, vLLM and the role of AI in next generation applications. It’s an opportunity to make a significant impact in a collaborative, fast-paced environment while building a new segment in this space. In this role, you’ll be able to Resolve customer issues and help in their successful adoption of Anyscale platform Be a technical advisor, and internal champion for our key customers Own customer issues end-to-end, from troubleshooting, triaging, escalations and eventual resolution Participate in our follow-the-sun customer support model to ensure continuity in resolving high priority tickets Keep track of open customer bugs and feature requests to influence prioritization and provide timely customer updates upon resolution Contribute towards improvement of internal tools an

awsazuregcp
View job →
C
Clickup
📍 United States• Full-time
1mo ago

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 We're looking for a Staff Data Engineer to own the architecture and technical vision of our data platform. This is a high-leverage, high-autonomy role where you'll set the technical bar for the team, drive cross-functional alignment on data infrastructure strategy, and solve our hardest engineering problems. You'll operate across AWS serverless technologies, Snowflake, dbt, and Terraform, but your impact goes well beyond any single tool: you'll shape how we think about reliability, scalability, cost, and developer experience at the platform level. This role is for someone who doesn't just build great systems, but makes the engineers around them better. The Role: Own the technical architecture of ClickUp's data platform, making design decisions that balance scalability, cost, reliability, and velocity. Define and drive the technical roadmap for data infrastructure in partnership with leadership. Design systems at scale : build frameworks, abstractions, and patterns that other engineers use daily. Lead complex, cross-team technical initiatives spanning data engineering, analytics engineering, data science, and data analytics. Drive cost optimization across cloud infrastructure and compute, turning efficiency into a competitive advantage. Build and evolve our data pipelines using AWS serverless (Lambda, Fargate, Step Functions, Kinesis, S3, DynamoDB, Aurora), Snowflake, and dbt. Establish and champion engineering standards : observability, testing, CI/CD, code review, and documentation practices. Design and maintain infrastructure for AI/ML workloads , including LLM frameworks, feature pipelines, training

pythonsqlaws
View job →
R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers. We’re looking for an EM to lead a team of exceptional ML infrastructure engineers, build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You Will: Lead strategic planning and roadmap execution of scalable production-ready ML systems including model training, data pipelines, feature engineering and model inference. Own the architecture, establish engineering best practices of scalability, reliability, and cost-effectiveness of ML infrastructure (e.g., training, serving, feature). Work closely with data scientists, ML engineers, platform teams, and product stakeholders to design, implement, and operate robust ML platforms that accelerate model development and deployment. Recruit, mentor, and grow a high-performing team of ML infrastructure engineers. You Have: 5+ years of experienc

awsgitmachine learning
View job →
R
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads & Discovery business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver more value to our users and our advertisers. As a Machine Learning Infrastructure Engineer, you’ll build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You will: You will co-design models and systems, working at the intersection of model architecture and ML infrastructure, partnering closely with core modelers, data and AI infrastructure engineers, and product teams to push the boundaries of large-scale training and serving. Your work will span recommendation, search, and agentic applications, including large transformer architectures, LLMs, generative rankers, and efficient offline and online content-understanding systems. You will investigate model, data, and systems tradeoffs end to end—from data pipelines and distributed training to low-latency inference and production serving. This includes designing efficient KV-cache strategies, applying p

awsgitmachine learning
View job →
F
Figma
📍 Ca New York• Full-time• From $153K/yr
1mo ago

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! The Data Platform team at Figma builds and operates the foundational systems that power analytics, AI/ML, and data-driven decision-making across the company. We serve a diverse set of stakeholders, including AI researchers, machine learning engineers, data scientists, product engineers, and business teams that rely on data for insights and strategy. Our team owns and scales critical platforms such as the Snowflake data warehouse, ML Datalake, orchestration and pipeline infrastructure, and large-scale data ingestion and processing systems, managing all data flowing into and out of these platforms. Despite being a small team, we take on high-scale, high-impact challenges. In the coming years, we're focused on building the data infrastructure layer for Figma's AI-powered products, driving cost and performance optimizations across our data stack, scaling our ingestion and reverse ETL capabilities for new product use cases, and strengthening data quality, reliability, and compliance at every layer. If you're passionate about building scalable, high-performance data platforms that empower teams across Figma, we'd love to hear from you! This is a full-time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Design and build large-scale distributed data systems that power analytics, AI/ML, and business intelligence across Figma. Develop batch and streaming solutions to ensure data is reliable, efficient, and scalable across the company. Manage and evo

pythonsqlaws
View job →
I
Instacart
📍 United States - Remote• Full-time• Remote• From $265K/yr
1mo ago

We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Overview Instacarts Data Infrastructure organization builds and operates the systems that power our company’s data ecosystem, including a modern open data lakehouse on Apache Iceberg, a multi-engine compute platform for stream and analytical workloads, and self-serve tooling that helps Product, Data Science, ML, Ads, Finance, and engineering teams move fast with data. We’re looking for a Staff Software Engineer, Data Infrastructure to join our Data Governance and Foundations Team. In this role, you’ll serve as a senior technical leader owning the architecture and delivery of our open lakehouse foundation, governance and access patterns, and multi-engine compute strategy—balancing today’s reliability with the next three to five years of scale, maturity, and cost efficiency. You’ll collaborate closely with engineering leadership and stakeholders across Data Science,

REMOTEpythonsqlaws
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Integrity team at OpenAI is dedicated to ensuring that our cutting-edge technology is not only revolutionary, but also secure from a myriad of adversarial threats. We strive to maintain the integrity of our platforms as they scale. The Integrity team is at the front lines of defending against misuse in all its forms: content abuse, scaled attacks, and other actions that could undermine the user experience or harm our operational stability. About the Role As a Machine Learning Engineer in OpenAI's Integrity team, you will have the opportunity to work with some of the brightest minds in AI. You’ll work on state-of-the-art models and classifiers, experiment with new architecture and approaches, and push forward our abilities in content and user understanding. You’ll help turn research breakthroughs into tangible solutions that improve the trust and safety of our platform. If you're excited about training LLMs and building ML models, this role is your chance to make a significant mark. In this role, you will: Innovate and Deploy: Design and deploy advanced machine learning models that solve real-world problems. Bring OpenAI's research from concept to implementation, creating AI-driven applications with a direct impact. Collaborate with the Best: Work closely with researchers, software engineers, and product managers to understand complex business challenges and deliver AI-powered solutions. Be part of a dynamic team where ideas flow freely and creativity thrives. Optimize and Scale: Implement scalable data pipelines, optimize models for performance and accuracy, and ensure they are production-ready. Contribute to projects that require cutting-edge technology and innovative approaches. Learn and Lead: Stay ahead of the curve by engaging with the latest developments in machine learning and AI. Take part in code reviews, share knowledge, and lead by example to maintain high-quality engineering practices. Make a Difference: Monitor and maintain deployed m

awsrestmachine learning
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role OpenAI is developing custom silicon to power the next generation of frontier AI models. We’re looking for experienced Design Verification (DV) Engineers to ensure functional correctness and robust design for our cutting-edge ML accelerators. You will play a key role in verifying complex hardware systems—ranging from individual IP blocks to subsystems and full SoC—working closely with architecture, RTL, software, and systems teams to deliver reliable silicon at scale. In this role you will: Own the verification of one or more of: custom IP blocks, subsystems (compute, interconnect, memory, etc.), or full-chip SoC-level functionality. Define verification plans based on architecture and microarchitecture specs. Develop constrained-random, directed, and system-level testbenches using SystemVerilog/UVM or equivalent methodologies. Build and maintain stimulus generators, checkers, monitors, and scoreboards to ensure high coverage and correctness. Drive bug triage, root cause analysis, and work closely with design teams on resolution. Contribute to regression infrastructure, coverage analysis, and closure for both block- and top-level environments. You might thrive in this role if you have: BS/MS in EE/CE/CS or equivalent with 3+ years of experience in hardware verification. Proven success verifying complex IP or SoC designs in industry-standard flows Proficient in SystemVerilog, UVM, and common simulation and debug tools (e.g., VCS, Questa, Verdi). Strong knowledge

awsrestai
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Customer Experience Engineering Team builds the internal and external technologies that scale Snowflake’s global support and sales organizations. We empower our technical experts by providing the advanced tools they need to resolve complex issues and drive customer success. Our team specializes in software engineering, data-driven decisions, ML, and LLM-based solutions . We build production-grade systems to automate manual processes and augment the capabilities of our technical staff. Our current focus includes: LLMs : Developing and deploying LLM and agent-based architectures for streamlining troubleshooting Scalable Evaluations : Implementing large-scale evaluations to ensure the quality and reliability of our internal and external tools Process Automation : Designing intelligent workflows that eliminate bottlenecks and allow our experts to focus on the most technical aspects of the Snowflake platform Incident discovery: using embeddings, LLMs, clustering, and agents to detect potential widespread issues more quickly Now, the team is growing, and we are looking for a Software Engineer to join us. In this role, you will work closely with the state of the art LLM models, fine-tune them, develop agents, apply various clusterings, summarizations, embeddings, and so on. Ev

pythonjavakubernetes
View job →
C
12 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary: We are seeking a highly skilled backend-focused Senior Software Engineer to join our modernization team focused on transforming the legacy CVS pharmacy systems into a cutting-edge, cloud-native platform. As part of the team, you will play a crucial role in designing and developing microservices that drive the modernization of critical patient and drug domains. You will work on integrating cloud-native solutions, enhancing performance, and ensuring seamless operation within a highly scalable and secure environment. The ideal candidate will have extensive experience in backend development, system design, and a strong understanding of cloud-native software engineering principles. Responsibilities: · Design, build, and maintain scalable data pipelines to support analytics, ML, and operational reporting · Develop robust data ingestion, transformation, and integration workflows using Python, SQL, and modern data engineering frameworks · Build and maintain batch and streaming data pipelines leveraging technologies such as Kafka (or similar pub/sub tools) · Work with Google Cloud Platform (GCP) services, including Cloud Storage, Dataflow, Pub/Sub, BigQuery, Cloud Spanner and Cloud Functions · Develop and manage data APIs and interfaces (REST and GraphQL) to enable high-performance data access across microservices · Implement CI/CD automation fo

pythonjavasql
View job →
D
DevRev
📍 Austin• Full-time
16 days ago

About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the team The Applied AI Engineering team ensures our customers get the optimal experience from DevRev. As customers go through their DevRev journey, they may identify needs for integration with existing enterprise systems and services, workflow and process automation, or customization of the DevRev platform to achieve their business objectives. Our team works with customers to understand requirements and design, develop, and implement solutions to meet customer goals. Your mission is to systematically help customers find value with DevRev by developing a thorough understanding of their needs, owning coordination between internal and external stakeholders, and engineering the solution to get the job done. You are a product expert and will use your application development and AI/ML skills to ensure our customers get the most out of the DevRev platform. About the role As a Forward Deployed Engineer at DevRev, you will work closely with our post-sales Customer Experience team to help customers unlock value through AI. You will design and deploy intelligent agents, automated workflows, and integrations that transform how customers operate — connecting DevRev to their broader ecosystem and putting AI to work across

javascripttypescriptpython
View job →
T
Tenstorrent
📍 Austin• Full-time• $100K – $500K/yr
16 days ago

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking talented Physical Design Engineers to implement high-performance blocks for our industry-leading CPU and AI/ML architectures. You'll own the complete implementation flow from synthesis to tapeout, working alongside world-class engineers to push the boundaries of performance, power, and area. If you're passionate about crafting silicon that powers the future of AI computing and thrive on solving complex design challenges, we want you on our team. This role is hybrid , based out of Austin, TX, Santa Clara, CA or Fort Collins, CO. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A hands-on engineer with deep expertise in SOC/ASIC physical design and a track record of successful tapeouts. Passionate about optimizing PPA through innovative implementation techniques and close RTL collaboration. Strong problem solver who excels at debugging complex issues across design hierarchies. Collaborative team player who thrives in fast-paced, technically challenging environments. What We Need BS/MS/PhD in EE/ECE/CE/CS with proven experience in synthesis, PnR, and timing closure on taped-out designs. Expertise with industry-

pythonawsai
View job →
T
Tenstorrent
📍 Boston• Full-time• $100K – $500K/yr
16 days ago

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. We’re looking for a Field Application Engineer who’s wired for AI/ML, fluent in real-world problem-solving, and excited to build with the people actually using what we make. You will collaborate closely with the sales team and enterprise customers, leveraging your deep technical knowledge in AI to drive the adoption of our products and solutions. This is a customer-facing role that requires both technical expertise and excellent communication skills to convey complex technical concepts to non-technical stakeholders. This role is remote based out of North America with preference near one of our main hubs Santa Clara, CA; Boston, MA; or Toronto,ON. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are You’ve lived in the AI/ML trenches, whether as a field engineer, a solutions architect, or the one tapped in when things needed to “just work.” You speak both machine and human. Whether it’s a researcher or a skeptical executive, you know how to break things down and bring them to life. You’re fired up about generative models, LLMs, and the edge of what’s possible when software meets purpose-built silicon. Work directly with customers in mee

awsmachine learningai
View job →
🔔

Get new ml platform engineer jobs by email

Daily job updates · Unsubscribe anytime