Applying for Distinguished Machine Learning Engineer - Safety at Roblox? Compare this live job with your resume first.

🎯 Tailor my resume
←Jobiba.Me
R
Active1mo ago

Distinguished Machine Learning Engineer - Safety

Roblox·📍 San Mateo, CA, United States

Employment

FULL TIME

Work mode

On-site

Experience

SENIOR LEVEL

Salary

From $399.4K/yr

Role market pulse

How Machine Learning Engineer demand looks in United States

42/100 · watch

Live jobs

56

Posted 30d

10

30d movement

-78.3%

Remote share

30.4%

Salary listed

73.2%

Salary trend 1Y

Not enough history

Role overview

Job description

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Distinguished Machine Learning Engineer/Technical Director in the Safety organization at Roblox, you will drive th

…

What they are looking for

Skills & requirements

Department · Software Engineering

R

Hiring company

Roblox

Technology

Be part of an amazing team that’s building the world’s largest social platform for play. Every month, over 100 million players come to Roblox to immerse themselves in any game or experience imaginable. Players can create the ultimate theme park, compete as a professional race car driver, star in a fashion show, become a superhero, or simply build a dream home and hang out with friends. Our Mission: to become the global leader in user-generated gaming and creation for all ages. Our Products: The Roblox Imagination Platform™ is a vibrant ecosystem of creators, builders, game players, and adventurers imagining with their friends. Our Roblox Studio IDE allows powerful and intuitive creation, published with a single click to the cloud, and instantly playable with your friends across PC, Mac, iOS, Android, Xbox and Oculus Rift. Integrated matchmaking, persistence, and billing support provide developers rich functionality for building great games and experiences. The Roblox.com website ties this ecosystem together with search and discovery, social networking, and a massive catalog of user made content. Check out our latest news at blog.roblox.com!

Keep exploring

Similar active roles

Fresh roles matched to this title and market.

View all →
R
📍 San Mateo, CA, United States· Full-time

From $397.5K/yr

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. At Roblox , we’re building the tools and platform that empower a global community of creators and developers to build immersive experiences and a dynamic virtual economy. Our Economy ML team sits at the heart of this mission, delivering scalable machine learning systems that power personalization, pricing, search, and content understanding across all Economy surfaces: Marketplace, Developer Monetization, Payments, and Avatar. We’re looking for a Distinguished Engineer/Technical Director to lead the strategy and technical direction for ML systems , with a focus on large-scale recommendations, infrastructure, and emerging Generative AI applications. You’ll help build the systems that support retrieval, ranking, generative modeling, and LLM-powered personalization, all at massive scale. This role requires deep systems thinking, hands-on ML expertise, and a vision for how traditional ML and GenAI come together to power the future of the Roblox economy. Why Roblox for ML Systems AI/ML is a top company priority , with long-term investment. Real-world scale : Power millions of daily economic interactions across ranking, pricing, fraud, and search. Full-system ownership : Build and optimize end-to-

AWSGitMachine LearningAI
N
📍 Santa Clara, United States

NVIDIA is seeking a Senior System Architect: Heterogeneous EDA Systems to solve a complex challenge in accelerated computing: Failure Attribution at Scale. As EDA or equivalent experience workloads scale across thousands of heterogeneous nodes, a single failure can cause massive resource waste. We need an engineer to develop and build an automated framework. This framework will ingest telemetry from CPU and GPU clusters to identify the root cause of job failures in real-time. It will distinguish between hardware faults, infrastructure instability, and software defects. What you'll be doing: Architect Failure Attribution Frameworks: Build a scalable "flight recorder" for EDA jobs that captures high-fidelity state across the CPU, GPU, and Fabric at the moment of failure. Build automated diagnostics that correlate GPU XID errors, PCIe bus failures, and CUDA memory exceptions. Connect these errors with system-level events such as OOM kills or NUMA-related hangs. Distributed Logging & Tracing: Implement low-overhead tracing mechanisms (using tracing tools or custom agents) that provide access to job execution across multi-node Slurm or Kubernetes clusters. Root Cause Automation: Develop heuristics and models based on machine learning to classify failures as "Hardware Fault," "Software Bug," or "Environment Issue." This reduces the Mean Time to Identify (MTTI) for R&D teams. Resiliency Engineering: Work closely with hardware and infrastructure teams to define "signals of impending failure," enabling proactive job migration or check-pointing before a crash occurs. What we need to see: Distributed Systems Mastery: BS, MS, or PhD in Computer Science or Electrical Engineering (or equivalent experience) with 6+ years in systems programming. Experience building automated

PythonKubernetesLinuxMachine Learning
D
📍 New York, California, United States· Full-time

From $330K/yr

Datadog is expanding the Technical Solutions (TS) organization by seeking a customer-focused, deeply technical Distinguished Architect to join our Product Solutions Architecture (PSA) team. In this role, you will act as a technical multiplier for the world's leading AI labs and AI-native companies. You will bridge the gap between their bleeding-edge infrastructure aspirations and Datadog’s technology roadmap, ensuring our platform natively solves the unique observability challenges of training and deploying foundational models at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Leadership : Demonstrate thought leadership in the AI/LLM space. Influence key decision makers and stakeholders by connecting technical capabilities to organizational and business impact. Advisory : Strategically partner with highly technical Founders, Heads of Infrastructure, and Research Lead peers. Guide them on best practices and emerging industry trends in the AI/LLM space. Lead high-level technical and architectural conversations around AI adoption. Presentations : Lead deep-dive architecture reviews and design engagements with customer teams and their leaders to share industry trends, best practices, and demonstrate how Datadog can support high-throughput hyper scale AI workloads. GTM : Identify emerging AI-native technology shifts and feed them directly back to Datadog Product Management. Co-create custom observability integrations and solutions alongside Product SAs to keep Datadog at the absolute forefront of the AI stack. Collaboration : Collaborate with Product Solutions Architecture (PSA), Sales, Sales Engineering and Marketing in providing high-quality technical resources to a broad audience of practitioners and economic buyers. Hiring : Assis

G
📍 Austin, Texas, United States· Full-time

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore fosters continuous learning and innovation. Job Summary Reporting into the Systems Engineering organisation, the Distinguished Engineer, End-to-End Security Architect will define and lead the security architecture for Graphcore’s inference service platform. This role is responsible for establishing a comprehensive security strategy spanning platform, infrastructure, networking, service operations, customer assurance, and compliance readiness. Working across multiple engineering and operational functions, the successful candidate will provide technical leadership, drive security requirements, and ensure the platform delivers robust protection, resilience, and trust for customers. The Team You will work closely with teams across security architecture, infrastructure engineering, networking, site reliability engineering, platform software, firmware, data centre operations, compliance, legal, customer engineering, and customer security. The team collaborates across the business to deliver secure, reliable, and scalable AI infrastructure and services while supporting customer assurance, regulatory requirements, and operational excellence. Responsibilities and Duties Own the end-to-end security a

AIRustExcelRecruitment
H
📍 Texas, United States of America, United States

Distinguished Technologist - Principal Architect, Platform Security and Firmware Systems Description - HP is seeking a senior technical leader to serve as a lead architect for hardware rooted Platform Root of Trust for HP's commercial PC platform. In this role, you will be responsible for envisioning, defining, and driving the architecture of hardware, firmware, software, and cloud-integrated security capabilities that protect HP commercial PCs from current and emerging threats. You will also work across the hardware and security architecture teams, collaborate with HP's Security Research Labs, Firmware (BIOS) team, our security business that creates and deploys industry leading enterprise security and manageability solutions, product management, manufacturing, supply chain, customer enablement, and field organizations to create cohesive system solutions that deliver unique customer value. This position requires deep expertise in embedded systems, platform firmware, hardware, security, device manageability, and PC system architecture. The role also requires the ability to translate complex technical capabilities into customer-relevant value propositions and to influence internal and external stakeholders without relying solely on formal authority. Responsibilities Serve as principal architect for HP Endpoint Security Controller roadmap, associated hardware and firmware based capabilities, cryptographic identity, attestation, secure communication, event logging, and future controller generations. Define and evolve HP Sure Start architecture and related firmware resilience capabilities, including hardware-enforced firmware authenticity, recovery, update, and protection mechanisms. Drive secure BIOS and firmware configuration management strategies, including HP Sure Admin ecosystem capabilities, key management, zero-touch provision

PythonJavaC#Supply Chain
N
📍 Remote, United States· Remote

NVIDIA DGX Cloud is an AI Factory designed to power the next generation of AI and industrial-scale breakthroughs. As the Distinguished Engineer for Security Architecture, within our Security Engineering organization, you will set the security design bar for an AI factory of hundreds of thousands of GPUs, and then build against it alongside the teams. This is the founding architecture seat in a new organization. Security Engineering is a new organization at DGX Cloud, accountable for the security outcome of the platform, and this is the architecture function inside it. You will define the security design standard for DGX Cloud, a bar that sits above the company floor, and hold it from inside the teams doing the building. Security here is fleet horizontal and stack vertical, so your scope runs from the hardware root of trust and the hardened baseline, through tenancy and GPU workload isolation, to the services and APIs built on top, across every DGX Cloud engineering organization. A small team of Principal Engineers will report to you and hold the bar at domain depth. This is still a hands-on seat, and you stay in the design with them. You will also serve as DGX Cloud's technical interface into NVIDIA's central security organization. There is no architecture review board here and no approval queue; the bar holds because the strongest security engineers in the room helped set it and helped ship it. What You Will Be Doing: Set the DGX Cloud Security Bar: Own the security design standard across DGX Cloud (tenancy, GPU workloads, identity, supply chain, and isolation) and make it concrete. Reference architectures, golden paths, and requirements engineers can actually build against, not a policy library. Hold the Bar by Building: Embed with engineering teams on real work: join the design, learn the code, help ship the thing rather than grade it afterward.

KubernetesLinuxArtificial IntelligenceAI

🔔 Get job alerts

New Distinguished Machine Learning Engineer - Safety jobs in San Mateo, CA, United States, straight to your inbox.

No spam · Unsubscribe anytime