Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. We’re looking for a Principal Engineer to join the ML Platform team at Synthesia. Our team builds and operates the systems that allow researchers and product teams to train, serve, and deploy generative models reliably and efficiently . This includes research infrastructure, production serving systems, internal tooling, and the platform interfaces that connect them. A growing part of our mission is making these systems more automation-friendly and agent-oriented , so that workflows can increasingly be operated through reliable tooling rather than manual effort. We’re looking for a strong generalist with a systems mindset: someone who is comfortable working across infrastructure, backend systems, and tooling, and who has seen ML systems in practice. this is not a pure ML Engineer role. We’re especially interested in people who think deeply about reliability, scalability, performance, and resource efficiency in complex production environments. This is a hands-on IC role with significant ownership. You’ll help shape how our ML platform evolves as we scale the number of models, workloads, tools and teams relying on it. What you’ll do Design and improve the platform systems that support model training, evaluation, an
Jobiba hiring network
Principal Ml Platform Engineer Jobs
1,487 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current principal ml platform engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. ML Platform @ Roblox today supports hundreds of ML use cases and billions of inferences per day across Discovery, Safety, Engine, and much more. As an Infrastructure Engineer on the ML Platform team, you will design, scale, and maintain the foundational infrastructure powering our entire machine learning ecosystem. We are looking for accomplished engineers to spearhead the development of our next-generation ML tooling and platform capabilities. You will: Bootstrap and maintain Kubernetes and Cloud infrastructure for ML Platform components--Serving Layer, Metadata Store, Model Registry, and Pipeline Orchestrator. Set technical strategy and oversee development of high scale and reliable infrastructure systems. Propose and implement new platform tooling to improve time to production for MLEs and Data Scientists across the full ML lifecycle. Work on infrastructure projects such as GPU fleet management, hybrid-cloud orchestration, and writing custom Kubernetes controllers and resources. Stay abreast of industry trends in machine learning and infrastructure to ensure the adoption of leading-edge technologies and practices. Partner across organizations to build tooling, interfaces, and visualizati
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Role Overview & Key Responsibilities This is a high-leverage leadership role that spans architecture, execution, and org-building, and will shape the direction of our AI / ML initiatives at Ema. We are seeking an AI / ML technical leader who can take a vision and build it. As a Principal ML Engineer at Ema, you will be a senior technical leader responsible for shaping the machine learning roadmap, architecting large-scale ML systems, driving innovation, and ensuring our mixture of expert models (LLM + SLM + Custom Model) is accurate and performant at scale. You will collaborate across teams (research, product, infra, data, etc.), mentor senior engineers, and influence strategy and execution at company-wide levels. Responsibilities Lead the technical direction of GenAI and agentic ML systems that power enterprise-grade AI agents — spanning reasoning, retrieval, tool use, and integrations across various SaaS products. Architect, design, and implement scalable production pipelines for model training, fine-tuning, retrieval (RAG), agent orchestration, and evaluation — ensuring robustness, latency efficiency, and continuous learning. Define and own the multi-year ML roadmap for GenA
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. ML Platform @ Roblox today supports hundreds of ML use cases and billions of inferences per day across Discovery, Safety, Engine, and much more. As a Model Optimization engineer on ML Platform, you will be responsible for digging deep into model internals to optimize performance, for both training and inference. We are looking for accomplished engineers to help us maximize performance of our platform. You Will: Optimize machine learning models for performance on GPU architectures, focusing on both training and inference workflows. Conduct low-level performance profiling analysis to identify bottlenecks in existing machine learning pipelines and propose actionable improvements. Contribute to the development of best practices and tooling for model optimization and deployment. Collaborate with cross-functional teams, including data scientists and software engineers, to integrate and deploy optimized models into production environments. Partner across organizations to build tooling, interfaces, and visualizations that make the ML@Roblox a delight to use. You Have: 6+ years of professional experience and a tool chest of system design experience upon which to draw to build performant system
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Who We Are: Shape the future of Roblox’s virtual economy. The Economy ML team is building the machine learning backbone that powers Roblox’s Marketplace, Developer Monetization, and Payments ecosystems. From intelligent pricing and personalized storefronts to dynamic layout optimization and avatar understanding, we’re reimagining how the Roblox economy drives user engagement, monetization, and creator success at scale. As a Principal Software Engineer (Data Systems) , you will architect, build and deploy high-scale, reliable real-time and batch data systems for personalization, search and recommendation across various product surfaces in Marketplace, Developer Monetization and Payments. You will be involved in key data projects from architecting event taxonomies and logging interfaces to real-time feature serving across multiple search and recommendation surfaces. What You’ll Do Act as data engineering lead for Economy ML, setting standards for batch vs streaming feature pipelines, table design, observability, and documentation used across the Economy group. Work as a hands-on contributor on our data systems to power content recommendation, search and personalization across Economy product
What you’ll do Act as the technical lead for large parts of the scanner platform: system architecture, codebase structure, and long-term maintainability. Own core runtime foundations: distributed control, state management, fault handling, and reliability. Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. Build robust observability: logs, metrics, traces, and replayable diagnostics (with privacy constraints). Collaborate with hardware and recon/ML teams to define interfaces, data contracts, timing/synchronization, and failure modes. Lead complex refactors (e.g., message passing / RPC boundaries, modularization, concurrency model) without halting forward progress. What we’re looking for Deep software architecture experience for real-world systems: robotics, instrumentation, medical devices, or other complex distributed products. Strong Python and concurrency background (asyncio, multiprocessing, profiling, performance engineering). Track record of shipping systems that are observable, debuggable, and resilient. Strong technical leadership: clarity, pragmatic trade-offs, and mentoring. Useful experience Building but rock-solid systems: clear interfaces (gRPC/protobuf or equivalent), strong state modeling, and failure handling. High-leverage engineering habits on a lean team: good tests, CI, reproducible dev environments, and fast code review. Practical performance + concurrency work in Python (asyncio, profiling, multiprocessing) and comfort debugging distributed behavior. Security-minded device software: safe defaults, encrypted data paths, and disciplined handling of PII/PHI. Operational thinking: remote updates/management, excellent logging, and diagnostics that make real hardware debuggable.
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About the role: We’re looking for a Principal Engineer to set the technical vision and lead foundational work across Homefeed, Search, and our emerging AI Assistant experiences at Pinterest. This role sits at the intersection of relevance, ranking, retrieval, and generative AI, and will shape how hundreds of millions of Pinners discover and act on inspiration every day. You will operate as a cross-cutting technical leader: defining long-term architecture, raising the bar on ML and systems excellence, and partnering closely with product and executive leadership to drive step‑function improvements in engagement, quality, and creator value. What you’ll do: Define the long-term technical strategy and architecture for Homefeed, Search, and AI Assistant, including retrieval, ranking, personalization, and generative experiences. Drive a multi‑year road
Role Purpose: As a Principal iOS Engineer, you will make sure that our SDK stays cutting edge and therefore elegantly converts strangers into trusted users. You take pride in combining mobile native technologies with ML and CV. You feel comfortable switching hats between software development, software architecture, business requirements gathering, and working with our machine learning experts. You enjoy learning new technologies and sharing your code with others. Our team has already proven that our mobile SDK solution was perfect for far more than 100 Million users, join our team to make it even better. Example Responsibilities You’ll be responsible for the further development and architecture of our core products (mainly our native iOS SDK and sometimes on our end user-facing customer applications, cross-platform deliveries, …) Take ownership in delivering new functionality and fixes from idea creation, through the design phase until the public release. You write clean, readable and well documented code and you understand that your changes impact hundreds of well known applications in the app store and influence the experience of an impressive number of daily unique end-users using our SDK. Participate in resolving issues together with our customers. Collaborate with the whole R&D team / entire company to build features that matter Manage your own work by organizing your time, prioritizing tasks and taking ownership of topics. Experience and Qualifications 8+ years of experience in developing iOS applications in Swift Strong understanding of software design patterns as well as iOS specifics Capable of working on large scale projects Obsessive in delivering quality software Business driven and result-oriented individual Effective communication & collaboration skills as you will be working a high performing team who currently works in a distributed setup Experience with CI/CD Test Automation - writing Unit and Integration tests Android bas
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads & Discovery business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver more value to our users and our advertisers. As a Machine Learning Infrastructure Engineer, you’ll build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You will: You will co-design models and systems, working at the intersection of model architecture and ML infrastructure, partnering closely with core modelers, data and AI infrastructure engineers, and product teams to push the boundaries of large-scale training and serving. Your work will span recommendation, search, and agentic applications, including large transformer architectures, LLMs, generative rankers, and efficient offline and online content-understanding systems. You will investigate model, data, and systems tradeoffs end to end—from data pipelines and distributed training to low-latency inference and production serving. This includes designing efficient KV-cache strategies, applying p
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Why Content Safety? As a Principal Machine Learning Engineer for Content Safety, you will define the future of proactive moderation, driving immense social impact through cutting-edge, innovative ML solutions, focused on critical and ambiguous safety challenges. You will set the 3-5 year technical strategy and architectural blueprint for how Roblox uses machine learning for content moderation. You will own the architectural and execution roadmap of massive-scale ML systems that mitigate violative UGC content before it impacts our community. You will feel a deep sense of responsibility in proactively protecting our community thoughtfully and fairly, while balancing user freedom with platform civility. Your efforts will ensure Roblox remains one of the safest places on the internet for our broad community of over 100 million daily active users. You will: Define and Own the Technical Vision: Define and lead the multi-year technical vision, architectural strategy, and execution for machine learning solutions in Content Safety, ensuring these systems proactively and effectively detect and mitigate violative content at massive scale. Strategic Stakeholder Partnership: Collaborate with execu
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Machine Learning Engineer within the Creator Services Machine Intelligence team, you will focus on the research and development of Embodied AI and Behavioral Agents that revolutionize how games are created and played on Roblox. You will bridge the gap between cutting-edge research and massive-scale product application, building agents capable of complex 3D gameplay and unblocking many use cases across Roblox, from automated playtesting to ensure quality, to "ML Players" with human-like movement and strategic reasoning, playing with real players in games. You will work on feature extraction, model training, building validation / RL platform as well as inference set up leveraging methods from imitation learning to reinforcement learning. And you will create generalizable agents that can perceive 3D environments, understand game rules, plan long-term strategies, and execute complex physics-based actions in real-time. You Will: Design and implement foundation models end to end through the feature extraction to inference for embodied agents. Define the long-term roadmap for Game AI and Embodied Intelligence, acting as a technical bar-raiser for code quality and architectural desig
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox's data infrastructure processes petabytes of data daily, powering analytics, ML, and product decisions for a platform serving 200M+ daily active users. As a Principal Software Engineer in our Data Infra org, you will be the primary technical leader driving the strategic vision, long-term architecture, and massive scalability of our distributed data platforms that power Roblox. You will own and drive the next-generation architecture of our core platforms, which span Kafka, Flink, Spark, Trino, Druid, Airflow and Data Catalog. This role operates under high ambiguity, demanding unparalleled ownership to redefine the limits of infrastructure handling exabyte-scale workloads, and providing a unique opportunity to lead the future evolution of our global data ecosystem. You Will: Define Multi-Year Technical Strategy: Own and drive the end-to-end architectural vision for Roblox's core data platforms spanning Kafka, Flink, Spark, Trino, Druid, Airflow, and Data Catalog systems. Turn multi-year company strategies into concrete, production-grade infrastructure blueprints. Lead Cross-Functional Alignment: Partner closely with executive leadership, platform governance, data science, and product e
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is about empowering enterprises to achieve their full potential — and people too. With a culture that’s all in on impact, innovation, and collaboration, Snowflake is the sweet spot for building big, moving fast, and taking technology — and careers — to the next level. We are hiring a Principal Software Engineer for our AI team. Our team delivers an AI Functions product that is the key Cortex Platform feature used by Snowflake customers. We're delivering scalable, governed, managed, powerful and flexible transformation primitives that allow customers to build AI ETL pipelines on all data. We focus on solving the hard research and engineering problems required to make high quality multi-cloud service work. If you enjoy designing and building the AI services that run reliably at scale, this is the team for you. AS PRINCIPAL SOFTWARE ENGINEER IN AI & ML YOU WILL: Build customer facing AI Functions portfolio of products Design and implement highly scalable distributed platforms within the global Snowflake platform. Participate in decision-making processes on technical or business issues. Collaborate with engineers across teams to help deliver cross-functional initiatives. Ensure operational readiness of the services and meet the commitments to our customers regardi
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Product Security team ensures that Snowflake products are built and shipped with the highest level of security. Our team drives the security posture of Snowflake products and is responsible for embedding security into every stage of the product lifecycle, from design through deployment and beyond. We design and build frameworks, systems and services that keep Snowflake secure. As a Principal Software Engineer II on the Product Security team, you will be the senior technical authority for Product Security and play a critical leadership role in shaping and advancing Snowflake’s security. This is a unique opportunity to define and influence our long-term security strategy and have a direct impact on the security of the Snowflake platform and the trust of our customers. You will operate across organizational boundaries, guiding major security initiatives, influencing architectural decisions at the highest levels, setting the technical direction for the organization, and ensuring consistent security excellence across all product teams while working closely with business leaders to advance Snowflake’s business. The role requires deep expertise in security, software engineering, distributed systems, software infrastructure, AI/ML, applied cryptography, threat modeling and clou
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Snowhouse Foundation team builds our globally distributed data warehouse. We manage a vast array of petabyte scale data sets that are continuously ingested, processed and replicated from across all Snowflake environments and external data sources. Snowhouse powers all of Snowflake’s core business, engineering and data science needs and provides customers with full visibility into their account activities, usage, and resource consumption from all their global environments. The team is investing in multiple critical areas, including a pipeline authoring platform, high performance/high efficiency data export, ingestion and data layout. Our team is also responsible for a fundamental product for Snowflake’s customers: the Snowflake system database/application that provides customers with all usage insights they need to reason about their global Snowflake footprint as well as 1st party business logic such as ML powered functions and Budgeting applications. AS A PRINCIPAL SOFTWARE ENGINEER IN SNOWHOUSE FOUNDATION, YOU WILL: Design and implement innovative highly available distributed platforms and pipelines and enhance the overall Snowflake data infrastructure Lead and drive projects from idea formulation to design, implementation and successful productionization. Collaborate
Get new principal ml platform engineer jobs by email
Daily job updates · Unsubscribe anytime