Jobs in Canada

Senior Research Scientist in San Francisco

128 active opportunities · Updated October 2026

Explore current senior research scientist jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $302.4K/yr

Quick readStrong listing-quality and freshness signals

About Scale At Scale AI, our mission is to accelerate the development of AI applications. For 8 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including: generative AI, defense applications, and autonomous vehicles. With our recent Series F round, we’re accelerating the abundance of frontier data to pave the road to Artificial General Intelligence (AGI), and building upon our prior model evaluation work with enterprise customers and governments, to deepen our capabilities and offerings for both public and private evaluations. About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer-facing Researchers and Applied AI Engineers. Our core mission includes research on agent environments and RL reward signals, benchmarking autonomous agent performance across real-world scenarios and environments, creating robust data programs to improve Large Language Models (LLMs) agentic capabilities and building foundational tools and frameworks for evaluating models as agents. ACE focuses on autonomous agents that dynamically interact with diverse external environments, including code repositories, GUI interfaces, browsers, and more. About This Role This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate t

SQLAWSGCPRest
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $302.4K/yr

Quick readStrong listing-quality and freshness signals

Director of Engineering, Physical AI Role Overview The Director of Engineering will report to the General Manager of Physical AI, and will be responsible for leading a multi-disciplinary engineering organization. In this senior leadership role, you will own the execution of the Physical AI Data Engine — the platform powering the next generation of Physical AI/Embodied AI. You will collaborate closely with Operations and GTM to guide product direction and help solve the data bottleneck that stands between today's robotics research and real-world deployment. This role requires significant ownership in a fast-paced environment and you will motivate internal teams to set the pace for business growth. Travel will come into play. Key Responsibilities: Set and drive the technical vision across data collection infrastructure, teleoperation systems, ML training pipelines, model evaluation frameworks, annotation tooling, and research Lead a multidisciplinary engineering organization—spanning engineering managers, software engineers, ML engineers, and ML research scientists—while designing the organizational structure, talent strategy, and culture required to scale rapidly without compromising on quality or strategic alignment Maintain exceptional technical and operational excellence by deeply understanding team deliverables, asking incisive questions, identifying slipping standards early, and knowing precisely when to step in Drive cross-functional alignment across Engineering, Operations, and GTM on platform architecture, release processes, and shared priorities Collaborate with researchers and clients to architect and deliver scalable, production-grade data infrastructure tailored for complex robotics workloads Required Qualifications: Bachelor's degree in Engineering, Robotics, Computer Science, or a related technical field 8+ years of engineering experience in fast-paced environments, including 4+ years direct people management demonstrated history of recruiting, mentorin

TypeScriptPythonAWSKubernetes
SA
📍 San Francisco, Canada· Hybrid
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! In September 2026 we raised a $350 million Series E at a $3.5 billion valuation , and we are scaling our engineering and research teams to meet demand. The role Frontier AI data is expensive to make and hard to measure. Every task we deliver is tested against the strongest models, often through many long-running agent rollouts. Your job is to make that process faster, cheaper, and more rigorous with ML and AI You will be one of the early members of ML & Research Engineering at Snorkel. You will study how frontier-grade data is generated and evaluated, form hypotheses, validate them against real production data, and ship the winners at scale. You will shape the discipline's direction, its standards, and the team that grows around it. What you'll work on Efficient agentic evals. Cut the cost of long-horizon agent evaluation with adaptive sampling, statistically grounded early stopping, model cascades, caching, and cheap-first gating. AI model routing. Route every eval and judge call to the cheapest model that clears the quality bar, with fallback, monitoring, and cost attribution. Fine-tuned small models. Fine-tune and serve open-weight models (LoRA and other

PythonMachine LearningAI
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $130K/yr

Quick readStrong listing-quality and freshness signals

About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About the Team Marketing at Snorkel is growing rapidly and anchored by high-ownership operators who lead independently, align cross-functionally, and consistently deliver outsized results. We partner across Sales, Product, Research, and the executive team to translate complex AI and data value into differentiated positioning, integrated programs, and field-ready enablement that accelerates growth. The culture is high standards, high autonomy, and high collaboration. About the Role Reporting to the Sr. Director of Product Marketing, the Product Marketing Manager, Frontier Labs will own the GTM execution for our frontier lab business. You will partner closely with Research, FDE, and frontier-facing sales teams to translate technical work into research-credible positioning, repeatable sales plays, and high-quality GTM programs that resonate with research and procurement leaders inside frontier labs. You will keep the frontier asset library (battlecards, technical one-pagers, decks, benchmark and eval narratives) sharp against a fast-moving market, and bring competitive and customer insight into every motion. This is a hands-on role for a product marketer who wants

AIGoExcelMarketing
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $250K/yr

Quick readStrong listing-quality and freshness signals

About Scale Scale’s mission is to develop reliable AI systems for the world’s most important decisions. As the leading AI data foundry, we provide the high-quality data and full-stack technologies that power the world’s most advanced models — fueling breakthroughs in generative AI, defense, and autonomous vehicles. We partner with leading enterprises and governments to bring AI into production that performs when it matters most, combining rigorous evaluation with full-stack deployment so our customers can build AI they can trust. About the Team Applied Intelligence Systems (AIS) is part of the Scale Generative AI Platform (SGP), focused on pushing the frontier of what agentic applications can do across diverse enterprise and government use cases. We build the infrastructure and tooling that power agentic AI in production, paired with applied ML research, design, and evaluation to ensure these systems perform reliably at the scale our customers demand. AIS spans multiple workstreams — agent evaluation and oversight, orchestration and tool-use infrastructure, model and systems optimization, and applied research on new agent capabilities — and this role is not scoped to any single one of them. We’re growing fast, with increasing traction across both commercial and public sector customers, and we’re just getting started — this team will define what dependable, production-grade agentic AI looks like. About the Role As a Staff Machine Learning Research Engineer, you will operate across the full breadth of AIS’s technical needs — wherever the hardest ML problem in agentic AI happens to be that quarter. This could mean training and fine-tuning models, designing evaluation and observability systems, building improvement loops from production data, prototyping novel agent architectures, or designing internal systems and tooling that boost productivity across teams. You’re not tied to one team’s roadmap; you’re expected to move to where the technical leverage is highest, and t

AWSRestMachine LearningAI
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $264.8K/yr

Quick readStrong listing-quality and freshness signals

Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with leading enterprises and government organizations to accelerate their AI initiatives through our data annotation platform, generative AI solutions, and enterprise AI capabilities. About the General Agents Team The General Agents team, part of Scale’s Enterprise organization, builds robust general agents for customer use cases and applications. The team sits at the intersection of frontier agent development and real-world deployment, translating state-of-the-art reasoning and agentic capabilities into reliable, production-grade systems that drive real economic value. Our agents are scalable systems built around recurring enterprise problem domains, with a strong emphasis on generalization, extensibility, and deployment across many customers. About the Role As a Senior/Staff Machine Learning Engineer (MLE) on the General Agents team, you’ll play a critical role in designing, building, and deploying production-ready AI agents that solve high-impact enterprise problems. You will work across the full agent lifecycle—from model and system design to evaluation, deployment, and iteration—bridging cutting-edge agentic techniques with the constraints and requirements of real customer environments. You will: Design and implement end-to-end agent systems that combine LLM reasoning, tool use, memory, and control logic to solve recurring enterprise use cases. Build scalable, reliable agent architectures that can be deployed across many customers with varying data, tools, and constraints. Develop evaluation frameworks, datasets, environments, and metrics to measure agent performance, reliability, and business impact in production settings. Collaborate closely with product managers, customers, data annotators, and other engineering teams to translate enterprise requirements into robust agent designs. Productionize frontier agent techniques (e.g.,

PythonAWSRestMachine Learning
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $216K/yr

Quick readStrong listing-quality and freshness signals

About Scale AI Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with the world's leading enterprises and government organizations to accelerate their AI transformation through frontier AI systems that solve real business problems. Every day, we work with organizations across finance, healthcare, manufacturing, media and telecommunications to build production AI agents that automate complex workflows, help humans, reason over enterprise knowledge, and operate safely at scale. The Opportunity Applied AI is moving faster than ever. New foundation models, reasoning techniques, agent architectures, and research papers emerge every week. Yet building AI systems that reliably solve real-world problems remains one of the hardest engineering challenges. As a Senior Frontier Agent Engineer (Applied AI) , you'll bridge the gap between cutting-edge AI research and production deployment. You'll work directly with enterprise customers to design, evaluate, and deploy intelligent systems that combine frontier models with structured knowledge, retrieval, traditional machine learning, and enterprise software. Unlike traditional ML roles that focus on a single model or product, you'll work across a diverse portfolio of AI challenges spanning multiple industries and use cases. You may build a multi-agent research system and then participate in designing a customer intelligence platform, a healthcare copilot, or an autonomous workflow for a Fortune 100 company. If you enjoy reading new AI papers, experimenting with the latest models, and shipping production systems that create measurable business impact, you'll fit right in. What You'll Build Frontier AI Systems Design and deploy production AI agents that leverage the latest advances in large language models, reasoning, retrieval, memory, and tool use. Architect intelligent systems that combine LLMs, traditional machine learning, structured knowledge, enterpri

PythonAWSAzureGCP
HI
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$160K – $200K/yr

Quick readStrong listing-quality and freshness signals

Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About the Role HP IQ is looking for a highly organized Senior Product Design Producer to support and scale Product Design work across key verticals. This role is a liaison between Product Design, Engineering, and Partnerships teams, promoting cross-functional communication and collaboration. This position also manages product pilots and qualitative research schedules. Our ideal candidate is able to effortlessly manage multiple product work streams, competing priorities, and able to adapt to changing circumstances in a fast-paced environment while keeping all design deliverables on track. What You Might Do Partner with a product lead to define scope, develop roadmaps, and prioritize product design deliverables Collaborate with Engineering Project Managers to develop design project schedules, coordinating and tracking designer activity to ensure timely delivery Support a hybrid creative project management methodology in tandem with an Agile software development approach, moving with ease between organizational systems Manage product qualitative research schedules, product pilots, schedules and i

RedisGitAgileAI
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including frontier model training, enterprise adoption, defense applications, and autonomous vehicles. Our mission is to develop reliable AI systems for the world’s most important decisions. We’re looking for an AI Product Manager to own the Finance vertical within our Agents Data & Reinforcement Learning Environments team. In this role, you’ll own both the development of RL environments (the realistic, high-fidelity simulations of financial software and workflows that labs use to train and evaluate agents) and the “data as a product” strategy that powers them. You’ll understand where AI is being used in the Finance industry, decide what financial tasks are worth modeling, how to source and structure the underlying data, and how to turn deep domain knowledge into a defensible product. The ideal candidate has lived inside the Finance industry, and is able to pair that domain understanding with a sense for AI research and current agent capabilities in Finance workflows. You’ll translate that expertise into environments and datasets that teach AI agents to perform real financial work, and you’ll be the domain expert Scale’s most important customers and their leading researchers turn to. A strong entrepreneurial & go-to-market mindset will be necessary. What You’ll Do Own the Finance AI roadmap & data strategy: Set product direction for the Finance agents training stack and the data strategy behind it. Establish a vision for where AI is continuing to transform the Finance industry (including investment banking, private equity, public markets, corporate finance, FP&A, etc), driving execution across engineering, operations, and go-to-market teams. Build partnerships with research teams at frontier labs: Work directly with researchers at leading AI labs to understand where their Finance agentic capabilities fall short and shape new product lines and competit

AWSRestAIGo
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $102K/yr

Quick readStrong listing-quality and freshness signals

About the Team The DoorDash Research Fellowship is a 3-month program (extendable to 6 months) looking for Summer and Fall 2026 cohorts, for researchers and engineers who want to work on the hardest applied ML and AI problems in local commerce. Fellows are given the resources, autonomy, and access to real-world operational data needed to pursue ambitious research directions — with the goal of producing work that influences both the field and how DoorDash operates at scale. This program is modeled on the best external research fellowships: fellows are treated as independent researchers, not as junior employees on a product team. You pick the problem (within a set of priority areas), you own the direction, and you publish or ship the outcome. You’re excited about this opportunity because you will receive… Dedicated compute allocation sized to the research agenda — GPU clusters for training and inference budgets for experimentation Full access to DoorDash's research infrastructure — our internal RL stack, training and evaluation pipelines, RL environments built on real operational systems, agent evaluation harnesses, and the tooling our own research teams use day-to-day. Fellows are first-class users, not sandboxed visitors. Access to DoorDash operational data — real-world datasets spanning logistics, merchant operations, consumer behavior, and marketplace dynamics, under appropriate data governance Research mentorship from senior researchers and engineering leaders at DoorDash, plus a named research sponsor for each fellow who meets with you weekly and is accountable for unblocking your work Speaker series featuring leading researchers and practitioners from academia and industry — faculty from top ML programs, research leads from frontier AI labs, and senior operators from across tech. Fellows get dedicated 1:1 time with speakers when possible. A cohort of fellows working alongside you — a small, tight-knit group of researchers tackling different problems but sharing

GitRestAIGo
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $182.4K/yr

Quick readStrong listing-quality and freshness signals

Join the team shaping the future of AI at Scale. Scale builds RL environments: sandboxed replicas of the digital spheres where real knowledge work occurs and built from the operating data of the companies that actually hold it. As the Data Acquisition Lead , you will own the commercial motion that gets us that data end to end. You will figure out which companies sit on the data for the next domain worth owning and then go and get it. This is a zero-to-one function with no playbook. You should be prepared to wear many hats, from thesis-driven dealmaker to hands-on operator to technical translator between commercial and Research teams. You will: Map the supply side. Work backwards from where labs are pushing to the specific organizations holding the underlying data. Build a thesis on which domains are worth owning and in what order. Invent the deal structures. You'll work with Scale’s legal team to define the first version of how these transactions get priced. Close. Own it from cold outreach to signature. Close the loop with the technical side. You need to hold a real conversation about what makes a dataset trainable and become an expert in what makes this underlying data valuable. Build the machine. Do the work by hand first, then turn what you learn into a repeatable pipeline. Ideally, you’d have: 5+ years across some mix of business development, corp dev, commercial strategy, or early-stage GTM. The label matters less than a track record of building a commercial motion that didn't exist before you got there A strong track record of managing important external relationships Strong business judgment and the ability to evaluate partnership value quickly Clear communication skills and comfort working with senior stakeholders Ability to operate independently while staying closely connected to cross-functional teams A practical, hands-on approach to building new functions from the ground up Comfort working in fast-moving, ambiguous environments Experience in

AWSGitRestAI
T
📍 San Francisco, Canada
✓ High-confidence listingCompany trend -94.4%
Quick readStrong listing-quality and freshness signals

About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Role Twitch Security Platform builds and operates the technology at the intersection of security, privacy, software engineering, and data engineering. As a Senior Engineering Manager, Security Platform, you will guide a diverse group of software engineers to build critical tools and automation solutions used across Twitch. You will report to the Director of Security, Identity, and Privacy Platforms. You will manage engineers responsible for operating our security data platform handling billions of weekly events, a multi-petabyte data lake, and numerous analysis tools that provide critical data across the organization to support important security programs. You will also lead engineers responsible for writing software to accomplish security at scale through automation, libraries, and services. With this team, you will lead the full software development lifecycle from product discovery, roadmap prioritization, and execution. Programs you support will include Fraud, Service-to-Service Auth, Detection and Incident Response, Vulnerability Management, Engineering Intelligence, and Privacy. You are experienced in software and data engineering, you bring awareness of the industry's cutting edge, and you're passionate about security. If this sounds accurate, come join us! You can work from San Francisc

A
📍 San Francisco, Canada
✓ High-confidence listingCompany trend -86.2%
Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: Everyone at Airbnb thinks about trust, but our team obsesses over it daily. At the core of trust is safety, and thus we spend a significant amount of our time and energy keeping the community safe. The Trust team is responsible for developing the technology that helps protect our community and platform from fraud while also ensuring our hosts, guests, homes, and experiences meet our high standards. We constantly work to fight against online fraud (such as monetary loss, compromised accounts, spam and scam in messages, fake inventory, etc.) as well as offline fraud (theft, property damage, personal safety, etc.). We also work on onboarding and screening of users, and think about complex topics like identity and reputation to ensure that every interaction with Airbnb helps build trust in us and our community. You'll work side-by-side with talented product managers, data scientists, software engineers, fraud intelligence, and operations teams. Together, you'll design and build ML solutions that have direct, meaningful impact on user trust, business success, and the global Airbnb community. The Difference You Will Make: As a Senior Machine Learning Engineer on the Trust team, you will actively contribute code and ideas that shape the ML systems protecting millions of Airbnb users. You'll own and deliver ML projects end-to-end — from designing and training models to productionizing and operating them at scale, while collaborating closely with cross-functional partners. You'll tackle real-world challenges such as account takeover, fake accounts, payment fraud, and bot detection. Your work

PythonJavaMachine LearningRecruitment
L
📍 San Francisco, Canada
✓ High-confidence listingCompany trend -72.4%
Quick readStrong listing-quality and freshness signals

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Lyft Ads’ vision is to build the largest transportation media network. We leverage media to elevate the transportation experience for our users while building an exciting and sustainable advertising business. Our assets include Lyft in-app sponsorship opportunities, the largest bikeshare network in the country, off network audience programmatic solutions, as well as experiential activations. Our work is fun, challenging, and rewarding and we are looking for a customer-focused team member passionate about solving problems, delighting our agency partners, and winning new business. As a Senior Digital Account Executive for Lyft Media, you will be the primary point of contact for major advertising agencies and marketing clients. You will have discretion to shape the future of our growing media business while managing the day-to-day sales processes and relationships with advertising agencies. Responsibilities: Lead C Suite level conversations to increase revenue and grow strategic partnerships Demonstrate in-depth insights of the sales process and product combination Lead account planning process that aligns brand and Lyft’s resources to maximize opportunities Grow revenue, educate and lead strategic conversations with clients and navigate complex relationships Demonstrate expertise in all matters relevant to your book of business, including escalation and troubleshooting to resolve client issues Spearhead client education on products and give product updates to advise on the best approach to drive business outcomes for clients and agencies Lead collaboration with external clients and internal stakeholders Create compelling media packages and custom offerings to close new media business Build and manage sales pipelines while achieving quarterly revenue goals Work autonomously and independently to seek new

HI
📍 San Francisco, Canada
✓ High-confidence listing

$149.9K – $270K/yr

Quick readStrong listing-quality and freshness signals

Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role As a Senior Platform Engineer at HP IQ, you will help build and evolve the infrastructure, tooling, and shared platform capabilities that enable our engineering teams to develop and operate reliable, secure, and scalable services across cloud and edge environments . You will work closely with application, services, AI/ML, and security teams to improve developer velocity, production readiness, reliability, and operational efficiency across a heterogeneous infrastructure footprint. What You Might Do Design, build, and maintain shared infrastructure and platform capabilities across cloud and edge environments. Build automation and self-service tooling that improves engineering velocity and operational consistency. Develop and maintain Infrastructure-as-Code, deployment workflows, and environment provisioning. Partner with engineering teams on production readiness, including reliability, security, observability, scalability, and recovery. Improve monitoring, alerting, incident response, and operational tooling across distributed environments. Automate repetitive operational t

PythonKubernetesAI
🔔

Get new senior research scientist jobs in San Francisco, Canada by email

Daily job updates · Unsubscribe anytime