Jobs in United States

Principal Model Optimization Engineer in United States

718 active opportunities · Updated October 2026

Explore current principal model optimization engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingExact matchCompany trend -100%

From $295.3K/yr

Quick readExact title match for your search

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. ML Platform @ Roblox today supports hundreds of ML use cases and billions of inferences per day across Discovery, Safety, Engine, and much more. As a Model Optimization engineer on ML Platform, you will be responsible for digging deep into model internals to optimize performance, for both training and inference. We are looking for accomplished engineers to help us maximize performance of our platform. You Will: Optimize machine learning models for performance on GPU architectures, focusing on both training and inference workflows. Conduct low-level performance profiling analysis to identify bottlenecks in existing machine learning pipelines and propose actionable improvements. Contribute to the development of best practices and tooling for model optimization and deployment. Collaborate with cross-functional teams, including data scientists and software engineers, to integrate and deploy optimized models into production environments. Partner across organizations to build tooling, interfaces, and visualizations that make the ML@Roblox a delight to use. You Have: 6+ years of professional experience and a tool chest of system design experience upon which to draw to build performant system

AWSGitMachine LearningAI
A
📍 United States· Full-time
✓ High-confidence listingCompany trend -98.9%

From $292K/yr

Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: Machine Learning and Artificial Intelligence are at the heart of the Airbnb product. From Trust to Payments, and from Customer Service to Marketing we rely on ML to ensure that guests and hosts have the best possible experience with Airbnb. The CS AI product team is responsible for driving CSxAI (Customer Support x Artificial Intelligence) initiatives by adopting the Generative AI technologies to enable an intelligent, scalable and exceptional service experience. The team develops and enhances various AI models, ML services and tools including LLM fine-tuning, alignment and optimization, RAG/Search, LLM evaluation and testing automation, feedback-based learning and guardrail for a wide range of applications in Airbnb. What you will do: As a principal machine learning engineer, you will be responsible for fine-tuning state-of-the-art LLMs for diverse use cases while optimizing models for high-performance deployment on Airbnb’s ML Infrastructure. You will partner with product managers, software engineers, data scientists and operation teams to brainstorm, design and develop AI products such as AI Assistant, Autonomous agent, recommendation, travel planning, and many more products that make meaningful

PythonAgileMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83.9%

About the Team Our Cyber team builds AI systems and products that help trusted defenders understand and respond to cyber threats while improving the safety and reliability of frontier models in security-sensitive settings. The team works across product engineering, model training, evaluations, safeguards, and deployment to make advanced cyber capabilities useful to defenders and responsibly managed. We collaborate closely with Safety/Preparedness, Research, Security, Legal, Communications, GTM, and external partners across OpenAI’s broader cyber work. About the Role We’re looking for research and software engineers to join Codex Cyber. You’ll help define and ship security products, work with trusted defenders and customers, shape model training and access patterns, and build research and evaluation systems for assessing cyber capabilities, validating safeguards, and improving training data. This role is hands-on and cross-functional, connecting product launches, model development, safety work, and real-world security use cases. In this role, you will: Help define and execute the technical roadmap for Codex Cyber’s security products, including evaluations, safeguards, trusted-defender workflows, and deployment decisions. Work with trusted defenders, customers, and partner teams to understand cyber use cases, evaluate risk, and turn feedback into product and research priorities. Shape cyber-specific model training and access patterns, including data, evaluations, validation, and deployment criteria. Build and validate systems for measuring cyber capabilities, monitoring misuse risk, and proving safeguards work in practice. Collaborate with Safety/Preparedness, Research, Security, Legal, Communications, Go-to-Market, and external partners on company-wide cyber priorities. Translate frontier cyber research into launch-ready tools, operational playbooks, and durable infrastructure for Codex and security products. You might thrive in this role if you: Enjoy 0 -> 1 envi

JavaScriptTypeScriptPythonJava
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83.9%

About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu

PythonAWSRestAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -12.7%

We are hiring senior engineers to work on the CUDA driver, a core component of our platform for accelerating general purpose computation on the GPU. Our team delivers features and improvements to better realize the potential of NVIDIA hardware for a growing range of computational workloads, ranging from deep learning, scientific computation, and self-driving cars to video games and virtual reality! CUDA defines a unified programming model across a range of system configurations and hardware capabilities. To accomplish this, the CUDA driver interacts with GPU hardware, kernel mode drivers, switches and the operating system. What you'll be doing: As a member of our team, you will use your design abilities, coding expertise, and creativity to deliver the best Compute platform in the world. You will craft elegant solutions to exciting problems and craft the future direction of CUDA as you collaborate with your peers across NVIDIA. You will evangelize, architect, and implement new CUDA features You'll oversee and drive development efforts across multiple teams Collaborate with members of hardware architecture teams Help define forward-looking improvements to the CUDA APIs and programming model Design and maintain performance and precision modeling Write effective, maintainable, and well-tested code Develop code for multiple operating systems What we need to see: Bachelor of Science or Master of Science degree in Computer Science, Electrical Engineering, or related field (or equivalent experience) 15&#43; years of relevant systems software development experience Strong C programming skills </

Artificial IntelligenceAI
V
📍 United States· Full-time· Remote
✓ Quality checkedCompany trend -92.7%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's Governance and Compliance platform is the operating layer enterprises trust to run their security programs. We are hiring an engineer who will own the architectural foundation that makes it work for their most complex organizational structures. Program Structure and Trust is a newly chartered group at Vanta with a focused mission: build the enterprise org model that lets customers bring their compliance and security structure — product lines, business units, isolated data environments, cross-cutting audits, scoped approvals, and data residency requirements — natively into the platform. The problems this team solves determine whether Vanta can serve the enterprise customers it's increasingly winning. This is the defining technical role of the group. The Principal Engineer owns the design, phased delivery, and long-term technical direction of Vanta's enterprise org model — a multi-quarter initiative that cuts across the platform and establishes the foundation for how enterprise customers structure, segment, and operate inside Vanta. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as a Principal Engineer at Vanta: Own the design and multi-quarter delivery of Vanta's enterprise org model, including hierarchical product lines and business units, isolated data access and ownership, cross-cutting audit workflows, scoped approvals, and EU and GovCloud data residency support Define and evolve the core abstractions that let the platform absorb structurally diverse, often conflicting enterprise requirements — solving for the general case rather than one-off customer accommodations R

RestAIGoRust
H
📍 New York, NY, United States
✓ Quality checkedCompany trend +310%

Become a part of our caring community Every large organization is making critical decisions today about how it will leverage AI over the next decade. Few have leaders who can both define that vision and demonstrate its viability through hands-on engineering. This role requires both. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations to source documents, and route difficult cases to human experts. These capabilities support decisions that impact real healthcare outcomes for members. As a Principal AI Applied Engineer, you will define the technical strategy, architectural standards, and long-term vision for AI-enabled products across the organization. You will influence enterprise-wide decisions regarding AI platforms, model strategies, engineering standards, and technology investments while remaining deeply hands-on in prototyping, experimentation, architecture, and software development. This is the highest-level individual contributor role within the AI Applied Engineering organization. Success requires exceptional technical depth, organizational influence, strategic thinking, and the ability to translate emerging AI capabilities into scalable, reliable, and responsible production systems. Why Join Us Shape the long-term AI architecture and engineering direction for a large enterprise healthcare organization. Influence how AI-enabled products are designed, built, evaluated, deployed, and governed across multiple teams. Drive strategic decisions involving models, vendors, platforms, infrastructure, and shared capabilities. Prototype and validate emerging technologies before the organization invests at scale.</

JavaScriptTypeScriptPythonReact
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 We’re hiring a highly influential product analytics leader who can turn ambiguous questions into sharp insight, scalable measurement, and recommendations that directly shape what we build. This person will partner closely with Product, Engineering, Design, and Growth to raise the bar on decision quality and establish a more AI-native analytics operating model. Mission Drive the product insight agenda by helping ClickUp make faster, smarter product decisions through rigorous analysis, strong product judgment, and AI-enabled analytics workflows. What You'll Do Own the product analytics agenda across product usage, activation, feature adoption, retention, and expansion, and translate open-ended business questions into structured analyses and clear recommendations Partner with Product, Engineering, and Design to define success metrics early, improve instrumentation quality, and ensure important product surfaces are measurable from launch Build reusable analysis frameworks, semantic layers, metric definitions, and self-serve resources that help product teams answer routine questions faster and more consistently Apply AI-first methods across the analytics workflow, using large language models, coding agents, and automation for tasks like query drafting, QA, validation, documentation, and first-pass synthesis while keeping human judgment at the center of final recommendations Design and interpret experiments, observational analyses, and trend investigations, including situations where data is incomplete or traditional experimentation is not feasible Surface meaningful patterns in behavioral, subscription, and

PythonSQLAWSMachine Learning
M
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -100%

What you’ll do Act as the technical lead for large parts of the scanner platform: system architecture, codebase structure, and long-term maintainability. Own core runtime foundations: distributed control, state management, fault handling, and reliability. Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. Build robust observability: logs, metrics, traces, and replayable diagnostics (with privacy constraints). Collaborate with hardware and recon/ML teams to define interfaces, data contracts, timing/synchronization, and failure modes. Lead complex refactors (e.g., message passing / RPC boundaries, modularization, concurrency model) without halting forward progress. What we’re looking for Deep software architecture experience for real-world systems: robotics, instrumentation, medical devices, or other complex distributed products. Strong Python and concurrency background (asyncio, multiprocessing, profiling, performance engineering). Track record of shipping systems that are observable, debuggable, and resilient. Strong technical leadership: clarity, pragmatic trade-offs, and mentoring. Useful experience Building but rock-solid systems: clear interfaces (gRPC/protobuf or equivalent), strong state modeling, and failure handling. High-leverage engineering habits on a lean team: good tests, CI, reproducible dev environments, and fast code review. Practical performance + concurrency work in Python (asyncio, profiling, multiprocessing) and comfort debugging distributed behavior. Security-minded device software: safe defaults, encrypted data paths, and disciplined handling of PII/PHI. Operational thinking: remote updates/management, excellent logging, and diagnostics that make real hardware debuggable.

PythonAIGo
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -12.7%

NVIDIA’s Silicon Co-Design Group sits at the crossroads of architecture, silicon, systems, and manufacturing, where first-principles thinking and engineering judgment at the highest level translate directly into product outcomes at scale. We are looking for a Principal Performance and Manufacturing Architect who has built the models, defined the specs, and seen them validated through silicon. You have owned the connection between design intent and manufacturing reality, not as a reviewer or a contributor, but as the person who set the methodology and proved it worked. You turn ambiguous physical phenomena into quantified, defensible margin terms. You do not wait for data to confirm your hypothesis; you design the experiment that gets it. You improve how the organization ships products after every program. The exceptional hire also uses AI deliberately — with proven workflow impact and the judgment to know where it compresses real work and where it introduces risk. What you'll be doing: Own the physics, from mechanism to margin. Build first-principles models connecting AVF, defect mechanisms, and DVFS transients to field FIT, system-level yield, and DPPM vs. coverage — calibrated per node and population shift — so every margin term in the V/F curve and P-state table is named, sourced, and defensible. Set the screen that resolves escapes. Specify ATE and SLT voltage, frequency, and timing conditions that capture worst-case transient VF windows — making it unambiguous whether a marginal defect or timing violation is detected or escapes at every manufacturing stage. Make the POR the authoritative source. Author the methodology document for each program and drive alignment across build, product definition, reliability, and test engineering — so every team is making decisions from the same model. Prove the model before produc

H
📍 Texas, United States of America, United States
✓ Quality checkedCompany trend +19.6%

Principal Data Privacy Architect Description - Job Summary - Role Purpose • Lead and oversee complex, cross-functional privacy and data protection programs from strategy through implementation, ensuring alignment across business, technical, legal, and compliance stakeholders. • This role will design and implement scalable, AI-ready data privacy architecture across enterprise data environments, applications, and AI-enabled workflows. • The Principal Data Privacy Architect will serve as a hands-on subject matter expert responsible for embedding privacy-by-design, consent enforcement, data sovereignty, data loss prevention, and compliance controls into large, complex global data environments. • The architect will partner closely with Data Engineering, Cybersecurity, Legal, Privacy, AI Governance, Product, and Enterprise Architecture teams to ensure customer, employee, partner, and sensitive enterprise data is accessed, processed, shared, retained, and protected in a compliant, secure, and trustworthy manner. - Why This Role Matters • Architect for Trust & Scale: Build reusable privacy architecture patterns that enable secure, compliant, and scalable data usage across platforms, products, and regions. • Enable Responsible AI: Design privacy guardrails for AI agents, generative AI, RAG pipelines, model inputs and outputs, embeddings, vector stores, and automated data workflows. • Reduce Risk While Enabling Innovation: Translate privacy, consent, regulatory, and data sovereignty obligations into practical engineering controls that accelerate business outcomes. Responsibilities - Think Customer First • Embed customer trust, transparency, and privacy-by-design principles into enterprise data platforms and customer-facing applications. • Design consent-aware data access and usage p

PythonJavaSQLAWS
A
📍 United States
✓ Quality checkedCompany trend +50.7%

Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Position Overview The Principal Data Engineer (IC) is a senior individual contributor and the accountable technical leader for assigned cross-domain initiatives and enterprise data engineering capabilities. The role owns integrated technical direction and technical outcomes for work spanning multiple data domains, defines and stewards enterprise engineering standards and reference architectures, and drives convergence where duplicated or inconsistent solutions create enterprise cost, risk, or operational burden. The role advises on scope, sequencing, capacity, dependencies, and technical debt, but does not independently commit domain resources or business delivery dates. This position has no people-management responsibility. Enterprise Data operates a domain-aligned model built on Databricks and Unity Catalog. Working with Domain Leaders, Staff Engineers, Platform Engineering, and partner organizations, the role converts ambiguous enterprise needs into executable architecture and carries the most complex or highest-risk work through validation and production. The role remains hands-on through prototyping, reference implementations, critical-path development, design and code review, and production problem solving. This role is based in Madison, WI. Essential Duties Include, but are not limited to, the following: Cross-domain technical leadership and delivery Own the technical outcome of assigned cross-domain initiatives from initial ambigu

PythonSQLAWSAzure
P
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, CA, United States· Remote
✓ High-confidence listingCompany trend -86.4%
Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Millions of people across the world come to Pinterest to find new ideas every day. It’s where they get inspiration, dream about new possibilities and plan for what matters most. Our mission is to help those people find their inspiration and create a life they love. As a Pinterest employee, you’ll be challenged to take on work that upholds this mission and pushes Pinterest forward. As a Principal Engineer on the AI Platform team, you'll help architect the infrastructure that powers both Generative AI and Recommender Systems across Pinterest's entire product suite. Our team builds the end-to-end engines for petabyte-scale data orchestration, model training and fine-tuning, and high-performance inference, ensuring our models scale seamlessly to hundreds of millions of inferences per second in service of over 600 million monthly active users.

M
📍 O Fallon, Missouri, United States
✓ Quality checkedCompany trend +212.5%

Our Purpose Mastercard powers economies and empowers people in 200&#43; countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Principal Software Engineer Who is Mastercard? Mastercard is a global technology company in the payments industry. Our mission is to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. With connections across more than 210 countries and territories, we are building a sustainable world that unlocks priceless possibilities for all. Overview We are seeking a hands-on Principal Forward Deployed Engineer to accelerate delivery across Mastercard Services. Modeled on the forward deployed engineering role that leading technology companies embed alongside their most important customers, this position turns that model inward. You will deploy into Services programs as an internal partner who writes production code, unblocks delivery, and puts AI to work in how we build. Reporting to the SVP of Developer Enablement within the Services Enablement & Transformation (SET) organization, you will move to wherever the need is greatest, adding s

AzureDockerKubernetesAI
N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -12.7%

NVIDIA has transformed computer graphics, PC gaming, and accelerated computing for more than 25 years through exceptional technology and the people who build it. In semiconductor manufacturing, our role is to enable the ecosystem, not compete within it. We partner with fabs, equipment manufacturers, and software providers to make inspection, metrology, and manufacturing intelligence dramatically faster on the NVIDIA platform. Our team builds the software that makes this possible: models, adaptation and evaluation workflows, and deployable inference capabilities that partners integrate into their own tools. We work in environments where labeled data is limited and proprietary, distributions shift across tools and fabs, production budgets are tight, and software must operate inside air-gapped facilities. We’re seeking a Principal Systems Software Engineer for Semiconductor Inspection in Santa Clara. This is a hands-on architect role: you will define the approach, build it, evaluate it, and demonstrate the results. You will work across computer vision, time-series modeling, multimodal AI, anomaly detection, model adaptation, evaluation, and production inference. Success means technology that a fab or equipment vendor can integrate, operate, and trust—not only a successful internal demonstration. What you’ll be doing: Define and prototype AI system architectures spanning optical and e-beam inspection, wafer and mask inspection, metrology, defect review, equipment signals, and process data. Advance world foundation model capabilities for semiconductor manufacturing, including vision, time-series and multimodal representation learning, model adaptation, domain transfer, and data-scarce defect understanding. Develop workflows for defect detection, classification, localization, segmentation, nuisance filtering, ADC, AD

PythonMachine LearningAI
🔔

Get new principal model optimization engineer jobs in United States by email

Daily job updates · Unsubscribe anytime