Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including frontier model training, enterprise adoption, defense applications, and autonomous vehicles. Our mission is to develop reliable AI systems for the world’s most important decisions. We’re looking for an AI Product Manager to own the Finance vertical within our Agents Data & Reinforcement Learning Environments team. In this role, you’ll own both the development of RL environments (the realistic, high-fidelity simulations of financial software and workflows that labs use to train and evaluate agents) and the “data as a product” strategy that powers them. You’ll understand where AI is being used in the Finance industry, decide what financial tasks are worth modeling, how to source and structure the underlying data, and how to turn deep domain knowledge into a defensible product. The ideal candidate has lived inside the Finance industry, and is able to pair that domain understanding with a sense for AI research and current agent capabilities in Finance workflows. You’ll translate that expertise into environments and datasets that teach AI agents to perform real financial work, and you’ll be the domain expert Scale’s most important customers and their leading researchers turn to. A strong entrepreneurial & go-to-market mindset will be necessary. What You’ll Do Own the Finance AI roadmap & data strategy: Set product direction for the Finance agents training stack and the data strategy behind it. Establish a vision for where AI is continuing to transform the Finance industry (including investment banking, private equity, public markets, corporate finance, FP&A, etc), driving execution across engineering, operations, and go-to-market teams. Build partnerships with research teams at frontier labs: Work directly with researchers at leading AI labs to understand where their Finance agentic capabilities fall short and shape new product lines and competit
Jobs in Canada
Senior Ai Research Engineer in San Francisco
128 active opportunities · Updated October 2026
Showing
15 jobs
Explore current senior ai research engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
From $264.8K/yr
Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with leading enterprises and government organizations to accelerate their AI initiatives through our data annotation platform, generative AI solutions, and enterprise AI capabilities. About the General Agents Team The General Agents team, part of Scale’s Enterprise organization, builds robust general agents for customer use cases and applications. The team sits at the intersection of frontier agent development and real-world deployment, translating state-of-the-art reasoning and agentic capabilities into reliable, production-grade systems that drive real economic value. Our agents are scalable systems built around recurring enterprise problem domains, with a strong emphasis on generalization, extensibility, and deployment across many customers. About the Role As a Senior/Staff Machine Learning Engineer (MLE) on the General Agents team, you’ll play a critical role in designing, building, and deploying production-ready AI agents that solve high-impact enterprise problems. You will work across the full agent lifecycle—from model and system design to evaluation, deployment, and iteration—bridging cutting-edge agentic techniques with the constraints and requirements of real customer environments. You will: Design and implement end-to-end agent systems that combine LLM reasoning, tool use, memory, and control logic to solve recurring enterprise use cases. Build scalable, reliable agent architectures that can be deployed across many customers with varying data, tools, and constraints. Develop evaluation frameworks, datasets, environments, and metrics to measure agent performance, reliability, and business impact in production settings. Collaborate closely with product managers, customers, data annotators, and other engineering teams to translate enterprise requirements into robust agent designs. Productionize frontier agent techniques (e.g.,
From $250K/yr
About Scale Scale’s mission is to develop reliable AI systems for the world’s most important decisions. As the leading AI data foundry, we provide the high-quality data and full-stack technologies that power the world’s most advanced models — fueling breakthroughs in generative AI, defense, and autonomous vehicles. We partner with leading enterprises and governments to bring AI into production that performs when it matters most, combining rigorous evaluation with full-stack deployment so our customers can build AI they can trust. About the Team Applied Intelligence Systems (AIS) is part of the Scale Generative AI Platform (SGP), focused on pushing the frontier of what agentic applications can do across diverse enterprise and government use cases. We build the infrastructure and tooling that power agentic AI in production, paired with applied ML research, design, and evaluation to ensure these systems perform reliably at the scale our customers demand. AIS spans multiple workstreams — agent evaluation and oversight, orchestration and tool-use infrastructure, model and systems optimization, and applied research on new agent capabilities — and this role is not scoped to any single one of them. We’re growing fast, with increasing traction across both commercial and public sector customers, and we’re just getting started — this team will define what dependable, production-grade agentic AI looks like. About the Role As a Staff Machine Learning Research Engineer, you will operate across the full breadth of AIS’s technical needs — wherever the hardest ML problem in agentic AI happens to be that quarter. This could mean training and fine-tuning models, designing evaluation and observability systems, building improvement loops from production data, prototyping novel agent architectures, or designing internal systems and tooling that boost productivity across teams. You’re not tied to one team’s roadmap; you’re expected to move to where the technical leverage is highest, and t
From $216K/yr
About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with the world's leading enterprises and government organizations to accelerate their AI transformation through frontier AI systems that solve real business problems. Every day, we work with organizations across finance, healthcare, manufacturing, media and telecommunications to build production AI agents that automate complex workflows, help humans, reason over enterprise knowledge, and operate safely at scale. The Opportunity Applied AI is moving faster than ever. New foundation models, reasoning techniques, agent architectures, and research papers emerge every week. Yet building AI systems that reliably solve real-world problems remains one of the hardest engineering challenges. As a Senior Frontier Agent Engineer (Applied AI) , you'll bridge the gap between cutting-edge AI research and production deployment. You'll work directly with enterprise customers to design, evaluate, and deploy intelligent systems that combine frontier models with structured knowledge, retrieval, traditional machine learning, and enterprise software. Unlike traditional ML roles that focus on a single model or product, you'll work across a diverse portfolio of AI challenges spanning multiple industries and use cases. You may build a multi-agent research system and then participate in designing a customer intelligence platform, a healthcare copilot, or an autonomous workflow for a Fortune 100 company. If you enjoy reading new AI papers, experimenting with the latest models, and shipping production systems that create measurable business impact, you'll fit right in. What You'll Build Frontier AI Systems Design and deploy production AI agents that leverage the latest advances in large language models, reasoning, retrieval, memory, and tool use. Architect intelligent systems that combine LLMs, traditional machine learning, structured knowledge, enterpri
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! In September 2026 we raised a $350 million Series E at a $3.5 billion valuation , and we are scaling our engineering and research teams to meet demand. The role Frontier AI data is expensive to make and hard to measure. Every task we deliver is tested against the strongest models, often through many long-running agent rollouts. Your job is to make that process faster, cheaper, and more rigorous with ML and AI You will be one of the early members of ML & Research Engineering at Snorkel. You will study how frontier-grade data is generated and evaluated, form hypotheses, validate them against real production data, and ship the winners at scale. You will shape the discipline's direction, its standards, and the team that grows around it. What you'll work on Efficient agentic evals. Cut the cost of long-horizon agent evaluation with adaptive sampling, statistically grounded early stopping, model cascades, caching, and cheap-first gating. AI model routing. Route every eval and judge call to the cheapest model that clears the quality bar, with fallback, monitoring, and cost attribution. Fine-tuned small models. Fine-tune and serve open-weight models (LoRA and other
From $302.4K/yr
About Scale At Scale AI, our mission is to accelerate the development of AI applications. For 8 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including: generative AI, defense applications, and autonomous vehicles. With our recent Series F round, we’re accelerating the abundance of frontier data to pave the road to Artificial General Intelligence (AGI), and building upon our prior model evaluation work with enterprise customers and governments, to deepen our capabilities and offerings for both public and private evaluations. About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer-facing Researchers and Applied AI Engineers. Our core mission includes research on agent environments and RL reward signals, benchmarking autonomous agent performance across real-world scenarios and environments, creating robust data programs to improve Large Language Models (LLMs) agentic capabilities and building foundational tools and frameworks for evaluating models as agents. ACE focuses on autonomous agents that dynamically interact with diverse external environments, including code repositories, GUI interfaces, browsers, and more. About This Role This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate t
From $102K/yr
About the Team The DoorDash Research Fellowship is a 3-month program (extendable to 6 months) looking for Summer and Fall 2026 cohorts, for researchers and engineers who want to work on the hardest applied ML and AI problems in local commerce. Fellows are given the resources, autonomy, and access to real-world operational data needed to pursue ambitious research directions — with the goal of producing work that influences both the field and how DoorDash operates at scale. This program is modeled on the best external research fellowships: fellows are treated as independent researchers, not as junior employees on a product team. You pick the problem (within a set of priority areas), you own the direction, and you publish or ship the outcome. You’re excited about this opportunity because you will receive… Dedicated compute allocation sized to the research agenda — GPU clusters for training and inference budgets for experimentation Full access to DoorDash's research infrastructure — our internal RL stack, training and evaluation pipelines, RL environments built on real operational systems, agent evaluation harnesses, and the tooling our own research teams use day-to-day. Fellows are first-class users, not sandboxed visitors. Access to DoorDash operational data — real-world datasets spanning logistics, merchant operations, consumer behavior, and marketplace dynamics, under appropriate data governance Research mentorship from senior researchers and engineering leaders at DoorDash, plus a named research sponsor for each fellow who meets with you weekly and is accountable for unblocking your work Speaker series featuring leading researchers and practitioners from academia and industry — faculty from top ML programs, research leads from frontier AI labs, and senior operators from across tech. Fellows get dedicated 1:1 time with speakers when possible. A cohort of fellows working alongside you — a small, tight-knit group of researchers tackling different problems but sharing
From $302.4K/yr
Director of Engineering, Physical AI Role Overview The Director of Engineering will report to the General Manager of Physical AI, and will be responsible for leading a multi-disciplinary engineering organization. In this senior leadership role, you will own the execution of the Physical AI Data Engine — the platform powering the next generation of Physical AI/Embodied AI. You will collaborate closely with Operations and GTM to guide product direction and help solve the data bottleneck that stands between today's robotics research and real-world deployment. This role requires significant ownership in a fast-paced environment and you will motivate internal teams to set the pace for business growth. Travel will come into play. Key Responsibilities: Set and drive the technical vision across data collection infrastructure, teleoperation systems, ML training pipelines, model evaluation frameworks, annotation tooling, and research Lead a multidisciplinary engineering organization—spanning engineering managers, software engineers, ML engineers, and ML research scientists—while designing the organizational structure, talent strategy, and culture required to scale rapidly without compromising on quality or strategic alignment Maintain exceptional technical and operational excellence by deeply understanding team deliverables, asking incisive questions, identifying slipping standards early, and knowing precisely when to step in Drive cross-functional alignment across Engineering, Operations, and GTM on platform architecture, release processes, and shared priorities Collaborate with researchers and clients to architect and deliver scalable, production-grade data infrastructure tailored for complex robotics workloads Required Qualifications: Bachelor's degree in Engineering, Robotics, Computer Science, or a related technical field 8+ years of engineering experience in fast-paced environments, including 4+ years direct people management demonstrated history of recruiting, mentorin
From $130K/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About the Team Marketing at Snorkel is growing rapidly and anchored by high-ownership operators who lead independently, align cross-functionally, and consistently deliver outsized results. We partner across Sales, Product, Research, and the executive team to translate complex AI and data value into differentiated positioning, integrated programs, and field-ready enablement that accelerates growth. The culture is high standards, high autonomy, and high collaboration. About the Role Reporting to the Sr. Director of Product Marketing, the Product Marketing Manager, Frontier Labs will own the GTM execution for our frontier lab business. You will partner closely with Research, FDE, and frontier-facing sales teams to translate technical work into research-credible positioning, repeatable sales plays, and high-quality GTM programs that resonate with research and procurement leaders inside frontier labs. You will keep the frontier asset library (battlecards, technical one-pagers, decks, benchmark and eval narratives) sharp against a fast-moving market, and bring competitive and customer insight into every motion. This is a hands-on role for a product marketer who wants
$160K – $200K/yr
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About the Role HP IQ is looking for a highly organized Senior Product Design Producer to support and scale Product Design work across key verticals. This role is a liaison between Product Design, Engineering, and Partnerships teams, promoting cross-functional communication and collaboration. This position also manages product pilots and qualitative research schedules. Our ideal candidate is able to effortlessly manage multiple product work streams, competing priorities, and able to adapt to changing circumstances in a fast-paced environment while keeping all design deliverables on track. What You Might Do Partner with a product lead to define scope, develop roadmaps, and prioritize product design deliverables Collaborate with Engineering Project Managers to develop design project schedules, coordinating and tracking designer activity to ensure timely delivery Support a hybrid creative project management methodology in tandem with an Agile software development approach, moving with ease between organizational systems Manage product qualitative research schedules, product pilots, schedules and i
From $182.4K/yr
Join the team shaping the future of AI at Scale. Scale builds RL environments: sandboxed replicas of the digital spheres where real knowledge work occurs and built from the operating data of the companies that actually hold it. As the Data Acquisition Lead , you will own the commercial motion that gets us that data end to end. You will figure out which companies sit on the data for the next domain worth owning and then go and get it. This is a zero-to-one function with no playbook. You should be prepared to wear many hats, from thesis-driven dealmaker to hands-on operator to technical translator between commercial and Research teams. You will: Map the supply side. Work backwards from where labs are pushing to the specific organizations holding the underlying data. Build a thesis on which domains are worth owning and in what order. Invent the deal structures. You'll work with Scale’s legal team to define the first version of how these transactions get priced. Close. Own it from cold outreach to signature. Close the loop with the technical side. You need to hold a real conversation about what makes a dataset trainable and become an expert in what makes this underlying data valuable. Build the machine. Do the work by hand first, then turn what you learn into a repeatable pipeline. Ideally, you’d have: 5+ years across some mix of business development, corp dev, commercial strategy, or early-stage GTM. The label matters less than a track record of building a commercial motion that didn't exist before you got there A strong track record of managing important external relationships Strong business judgment and the ability to evaluate partnership value quickly Clear communication skills and comfort working with senior stakeholders Ability to operate independently while staying closely connected to cross-functional teams A practical, hands-on approach to building new functions from the ground up Comfort working in fast-moving, ambiguous environments Experience in
$149K – $240K/yr
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The AI team is building cutting-edge solutions that bring the power of AI directly to edge devices while seamlessly integrating with cloud infrastructure. We are looking for a Senior Software Engineer to design and develop high-performance, scalable services to support AI workloads across edge and cloud environments. What You Might Do Design, build, and maintain services that power AI-driven applications, ensuring scalability and performance. Develop APIs and microservices that facilitate seamless integration between cloud-based AI models and edge devices. Optimize data pipelines and storage solutions for real-time AI inference and processing. Implement security and privacy best practices for distributed AI systems. Work closely with AI researchers, infrastructure engineers, and frontend developers to deliver end-to-end AI-driven solutions. Build and optimize an agent orchestration runtime that enables tool use, memory management, and multi-step reasoning across LLMs, APIs, and edge-connected systems. Develop robust logging, monitoring, and alerting systems to ensure system reliabilit
$240K – $270K/yr
About the Role At Sigma, we’re not just adding AI—we’re building the future of how people work with data. Our platform already lets users explore billions of rows of data in seconds with a spreadsheet-like interface, analyze and present their data in workbooks, and build data apps and workflows. Now we’re pushing further, applying AI to reshape how people build in Sigma, discover insights, and make smarter decisions—fast. That’s where you come in. As an AI/ML Engineer, you’ll join a growing team focused on building the AI foundation that will power Sigma for the future. Your work will become an integral part of the workflow for the thousands of enterprises that run on Sigma. What You’ll Do Partner with product, design, and engineering teams to identify high-impact AI/ML opportunities Prototype and productionize AI systems that feel intuitive but do a lot under the hood—recommendations, natural language interfaces, agentic workflows, and more Develop and scale AI/ML infrastructure that powers both internal tooling and customer-facing features Tackle novel UX problems at the intersection of AI, BI, and apps What You Bring Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related field (required) 10+ years of experience building and deploying production-grade AI/ML systems Deep knowledge of machine learning, deep learning, and applied AI Experience across the full ML lifecycle: data curation, training, deployment, monitoring A track record of building things that ship—whether it’s recommendations, search, machine translation, or something equally complex Experience adapting or training foundation models (language or multimodal) for novel domains Bonus Points (or skills you’ll build here) You've built agents that can plan, reason, and use tools You know your way around cloud infrastructure (AWS, GCP, Azure) You’ve worked in a fast-moving startup or high-growth environment Additional Job details The base salary range for this posit
From $200K/yr
AI only answers correctly when it can trust the data underneath it. This role owns two connected parts of how Sigma shows up in that world: where Sigma's experience lives outside its own product (MCP, a CLI, the Claude and ChatGPT marketplaces, integrations like Slack, Teams, and Glean), and the semantic layer that makes every one of those surfaces trustworthy. The first mandate is Sigma's AI ecosystem: defining Sigma's approach to MCP, giving external agents structured, governed access to Sigma's data model; owning the CLI, giving developers a fast way to work with Sigma outside the UI; and building Sigma's presence in the Claude and ChatGPT marketplaces, plus integrations for Slack, Teams, Glean, and similar surfaces, so people can reach Sigma's data wherever they already work. This is some of the most visible, fastest-growing surface area in the product. The second mandate is the semantic layer underneath it all. Every agent, chat answer, and integration is only as reliable as the data model behind it — get a metric definition wrong here, and every surface built on top inherits the mistake. This includes setting the roadmap for how semantic views connect across data platforms and how the model evolves as new AI capabilities emerge. What you'll do Define Sigma's approach to MCP, giving external agents and tools structured, governed access to Sigma's data model. Own the CLI roadmap, giving developers a fast, scriptable way to work with Sigma outside the UI. Build Sigma's presence in the Claude and ChatGPT marketplaces, along with integrations for Slack, Teams, Glean, and similar surfaces, so people can reach Sigma's data wherever they're already working. Set the roadmap for Sigma's data modeling strategy, including how semantic views connect across data platforms and how the semantic layer evolves as new AI capabilities emerge. Partner with engineering and design to ship integration and semantic modeling capabilities that hold up at enterprise scale.
$210K – $250K/yr
Sigma is transforming how businesses allow customers to build apps, agents and dashboards on top of governed enterprise data. Hence, we are growing the design team and looking for designers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of designers with a shared mission to make data easily accessible for all users. We're looking for a Senior Product Designer / Design Engineer who sits at the intersection of interaction design and AI engineering: someone who uses AI to ship faster, builds the skills and evals that make AI more effective, and invents new interaction paradigms for how people work alongside intelligent systems. This isn't a traditional design role. Yes you'll be using Figma, but also writing code with AI, training it, evaluating it, and questioning every assumption about what a "UI" can be when the interface itself reasons. Please note this is a 4 day on-site role in our San Francisco office. What You'll Do Start with AI, stay with AI. Use LLMs to clarify scope, draft specs, surface edge cases, and align your team before committing to a direction, use AI coding tools to build and iterate on the solution itself, and merge code to prod when fits. Prototype in code. Build working interfaces with Cursor and Claude Code, guiding structure, behavior, interaction, motion and UX quality while AI handles implementation. Partner directly with engineering to decide what moves into the product and what stays as a validated spike. Bring it to production. Fix small interaction and refinement issues directly on prod code. Design new AI interaction paradigms for conversational interfaces. Invent and validate novel patterns for how users converse with, direct, and trust AI systems - especially in data contexts where precision and confidence matter. Write evals, skills, and help on tools. Build the scaffolding that makes AI reliabl
Other cities to consider
More places hiring for this role
Get new senior ai research engineer jobs in San Francisco, Canada by email
Daily job updates · Unsubscribe anytime