Jobs in Canada

Test in San Francisco

77 active opportunities · Updated October 2026

Explore current test jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $216K/yr

Quick readStrong listing-quality and freshness signals

At Scale, our mission is to develop reliable AI systems for the world's most important decisions. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. Scale Frontier Data is the organization behind the training and evaluation data that frontier labs depend on. We build the systems, tooling, and expert workflows that turn hard human expertise into signals that models can learn from, across reasoning, coding, agentic tool use, and domain expertise. About our Customer Platform team: Our Customer Platform Team plays a pivotal role in integrating our platform with external systems and ensuring seamless, reliable connectivity for both internal users and customers. As the leader of this team, you’ll drive the strategy, architecture, and development of our connectivity solutions, focusing on API integration, distributed systems, and a robust data platform. Your role will be crucial in maintaining and enhancing our platform’s ability to meet the needs of both our internal and external stakeholders. Responsibilities: Own large areas within our product Comfortable working cross functionally, whether that be internal or external customers Build features end-to-end: front-end, back-end, system design, debugging and testing Deliver experiments at a high velocity and level of quality to engage our customers Work across the entire product lifecycle from conceptualization through production Influence the culture, values, and processes of a growing engineering team Inspire and mentor less experienced engineers Collaborating with cross-functional teams to define, design, and ship new product features and experiences. Requirements: At least 7-10 years of relevant experience is preferred Track record of shipping high-quality products and features at scale Desire to work in a very fast-paced environment Abil

AWSRestAIGo
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

C$60 – C$80/hr

Quick readStrong listing-quality and freshness signals

As a member of our Frontier Tech Consultant team, you will play a critical role in advancing cutting-edge AI innovations by conducting high-impact experiments and ensuring seamless execution at the highest quality standards. Your work will directly contribute to Scale AI’s growth, shaping the future of artificial intelligence. In this role, you will be working on various types of projects, including but not limited to: research experiments, dataset generation, data quality improvements, and in-depth technical analysis. You will tackle complex, technical and operational challenges while collaborating closely with Scale’s ML research scientists and SPM team. The ideal candidate is analytical, detail-oriented, and results-driven, with strong problem-solving abilities and excellent communication skills. We are looking for someone who thrives in a fast-paced environment, is proactive in overcoming challenges, and is committed to delivering exceptional outcomes. If you are eager to contribute to the forefront of AI innovation, we encourage you to apply. You will be responsible for: Design and execute research experiments Build and evaluate frontier LLM datasets Develop training and testing material for frontier pipelines Improve quality of existing and new products Ideally you’d have: Strong machine learning knowledge, either by being in the final years of a ML PhD career or having already graduated Strong writing and verbal communication skills An action-oriented mindset that balances creative problem solving with the scrappiness to ultimately deliver results Analytical, planning, and process improvement capability Experience working in a fast-paced, entrepreneurial environment Technical skills including familiarity with Python, GPU, AWS, API, LLM, ML, and SQL Pay: $60-80/hr Commitment: This is a fully remote, US-based part-time (10-20 hours per week), on-going contract position staffed via HireArt. HireArt values diversity and is an Equal Opportunity E

PythonSQLAWSRest
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $216K/yr

Quick readStrong listing-quality and freshness signals

Software is eating the world, but AI is eating software. We live in unprecedented times – AI has the potential to exponentially augment human intelligence. Every person will have a personal tutor, coach, assistant, personal shopper, travel guide, and therapist throughout life. As the world adjusts to this new reality, leading platform companies are scrambling to build LLMs at billion scale, while large enterprises figure out how to add it to their products. To make them safe, aligned and actually useful, these models need human eval and reinforcement learning through human feedback (RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation that’s enabled ChatGPT to get such a large headstart among competition. At Scale, our products include the Generative AI Data Engine, SGP, Donovan, and others that power the most advanced LLMs and generative models in the world through world-class RLHF, human data generation, model evaluation, safety, and alignment. The data we are producing is some of the most important work for how humanity will interact with AI. At the foundation of these products is the Platform Engineering team. In this role, you will support the design and development of shared platforms used across Scale. This includes designing our foundational data platforms and lifecycle, architecting Scale’s core cloud infrastructure and orchestration stack, and redefining how engineers develop, build, test, and deploy software at Scale. You’ll also get widespread exposure to the forefront of the AI race as Scale sees it in enterprises, startups, governments, and large tech companies. You will: Drive the design, and implementation of our foundational platforms and systems, working closely with stakeholders and internal customers to understand and refine requirements. Collaborating with cross-functional teams to define, design, and deliver new features. Proactively identifying opportunities for, and driving improvements to, current p

SQLMongoDBAWSDocker
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

At Scale, our mission is to develop reliable AI systems for the world's most important decisions. For 10 years, Scale has provided the high-quality data and full-stack technologies that power the world's leading models, and has helped enterprises and governments build, deploy, and oversee AI applications that deliver real impact. We work closely with industry leaders like Meta, Ernst & Young, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force. Scale's internship is not a side project. Interns own real, shipped work on the same roadmaps as full-time engineers, with mentorship from world-class talent and a culture that values ownership, speed, and truth-seeking. Many of our interns return as full-time Scaliens. Example Projects Build reinforcement learning and post-training data pipelines that power frontier model development Develop evaluation infrastructure that measures model reliability for enterprise and public sector customers Ship agentic AI applications and the tooling that makes them observable, testable, and safe to deploy Ship tools that accelerate the growth of new qualified contributors on Scale's platform Build fraud-detection systems that remove bad actors and keep Scale's contributor base safe and trusted Use models to estimate the quality of tasks and contributors, and guarantee quality on requests at large scale Devise advanced matching algorithms that pair contributors to customers for optimal turnaround and accuracy Create optimized and efficient UI/UX tooling, in combination with ML algorithms, for 100k+ contributors completing billions of complex tasks Develop new AI infrastructure products to visualize, query, and explore Scale data Requirements A graduation date in Fall 2027 or Spring 2028 with a Bachelor's degree (or equivalent) in a relevant field (Computer Science, EECS, Computer Engineering, Statistics) Available for a Summer 2027 internship (May/June start dates) in San Franci

TypeScriptPythonReactMongoDB
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $216K/yr

Quick readStrong listing-quality and freshness signals

Software is eating the world, but AI is eating software. We live in unprecedented times – AI has the potential to exponentially augment human intelligence. Every person will have a personal tutor, coach, assistant, personal shopper, travel guide, and therapist throughout life. As the world adjusts to this new reality, leading platform companies are scrambling to build LLMs at billion scale, while large enterprises figure out how to add it to their products. To make them safe, aligned and actually useful, these models need human eval and reinforcement learning through human feedback (RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation that’s enabled ChatGPT to get such a large headstart among competition. At Scale, our products include the Generative AI Data Engine, SGP, Donovan, and others that power the most advanced LLMs and generative models in the world through world-class RLHF, human data generation, model evaluation, safety, and alignment. The data we are producing is some of the most important work for how humanity will interact with AI. At the foundation of these products is the Platform Engineering team. In this role, you will support the design and development of shared platforms used across Scale. This includes designing our foundational data platforms and lifecycle, architecting Scale’s core cloud infrastructure and orchestration stack, and redefining how engineers develop, build, test, and deploy software at Scale. You’ll also get widespread exposure to the forefront of the AI race as Scale sees it in enterprises, startups, governments, and large tech companies. You will: Drive the design, and implementation of our foundational platforms and systems, working closely with stakeholders and internal customers to understand and refine requirements. Collaborating with cross-functional teams to define, design, and deliver new features. Proactively identifying opportunities for, and driving improvements to, current p

SQLMongoDBAWSDocker
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $216K/yr

Quick readStrong listing-quality and freshness signals

Scale Labs, Research Scientist — Frontier Risk Evaluations As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs has launched a new team focused on policy research, to bridge the gap between AI research and global policymakers to make informed, scientific decisions about AI risks and capabilities. Our research tackles the hardest problems in agent robustness, AI control protocols, and AI risk evaluations to help governments, industry, and the public understand and mitigate AI risk while maximizing AI adoption. This team collaborates broadly across industry, the public sector, and academia and regularly publishes our findings. We are actively seeking talented researchers to join us in shaping this vision. As a Research Scientist focused on Frontier Risk Evaluations, you will design and create evaluation measures, harnesses and datasets for measuring the risks posed by frontier AI systems. For example, you might do any or all of the following: Design and build harnesses to test AI models and systems (including agents) for dangerous capabilities such as security vulnerability exploitation, CBRN uplift, and other high-risk activities; Work with government agencies or other labs to collectively scope and design evaluations to measure and mitigate risks posed by advanced AI systems; Publish evaluation methodologies and write technical reports for policymakers. Ideally you’d have: Commitment to our mission of promoting safe, secure, and trustworthy AI deployments in the industry as frontier AI capabilities continue to advance. Practical experience conducting technical research collaboratively. You should be comfortable building and instrumenting ML pipelines, writing evaluation harnesses, and quickly turning new ideas from the research literature into working prototypes. A track record of published research in m

AWSRestMachine LearningAI
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $1.5M/yr

Quick readStrong listing-quality and freshness signals

About the Team At DoorDash, we’re reimagining how people connect with the things they need — whether it’s a meal, a grocery run, and anything in between. Our audiences — Consumers, Dashers, and Merchants — are at the heart of everything the Design org builds. Our content design team plays a big part in this, with each content designer shaping the experience and bringing our vision to life through clear, thoughtful language that makes our three-sided marketplace easier, faster, and more human. About the Role We’re looking for a full stack UX Design Engineer to be the technical backbone of our global Content Design org and leader who builds the platforms, tooling, and ML systems that make world-class, localized product content effortless across DoorDash, Wolt, and Deliveroo. You’ll work at the intersection of design and engineering to build content tooling and embed LLM-driven workflows into product experiences. You’ll accelerate content design AI tooling so that CD’s can focus on the highest leverage strategic work, while your tools enable product designers and cross functional partners to ship high-quality content faster across DoorDash, Wolt, and Deliveroo. You're excited about this opportunity because... Design and build internal content tools that: Help PMs and designers generate, manage, localize, and deploy product content at scale. Plug custom GPTs and other LLMs into everyday product workflows for content iteration and improvements. Own full-stack development of these tools: Build and maintain frontend experiences using React and modern JavaScript/TypeScript. Design and implement backend services and APIs (e.g., Kotlin/Java) to support content tooling, experimentation, and automation. Integrate tooling with experimentation and infra: Embed content tools into A/B testing platforms so teams can test and ship variants with minimal engineering dependency. Build pipelines from variant generation → experiment → automated deployment of winners. Operationalize

JavaScriptTypeScriptJavaReact
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by millions of people, then we want to talk to you! About the Role We are hiring a Software Integration Engineer for our Platform Integration team. This is a critical role with impact across the robot lifecycle, from manufacturing to daily operation. The Platform Integration team owns making sure the robot works as one cohesive system. The focus is on the interfaces between subsystems; this role in particular is focused on the software side handling interaction between: OS & software stack; firmware; networking; timing; and calibration. In this role, you will work cross functionally with our electrical, hardware, firmware, and autonomy engineers to support new functionality both in both hardware and software. This includes creating provisioning tools, functional tests, and supporting integration into the autonomy software stack. You will report to the Autonomy Platform Lead on our Autonomy Platform Team at DoorDash Labs. We expect this role to be hybrid with some time in-office and some time remote. You’re excited about this opportunity because you will… Play an integral role on a small and focused team. Lead system-level debug when an issue crosses subsystem boundaries or no single team can isolate it. Support early integration of new sensor and software component designs by identifying interface requirements, risks, dependencies and required checks. Design and maintain integration tests, test setups, and procedures to ensure subsystems, once combined, satisfy requirements and design intent. Build and maintain the mission-readiness checks used before manufacturing signoff, validation, field testing, or mission use for different robot platforms. Create the tools, checks, and debug guidance that Manufacturing Integration, Validation,

PythonAWSGitRest
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team Data is at the foundation of DoorDash success. The Data Engineering team builds database solutions for various use cases including reporting, product analytics, marketing optimization and financial reporting. By implementing pipelines, data structures, and data warehouse architectures; this team serves as the foundation for decision-making at DoorDash. About the Role DoorDash is looking for a Softare Engineer II to be a technical powerhouse to help us scale our data infrastructure, automation and tools to meet growing business needs. You're excited about this opportunity because you will… Work with business partners and stakeholders to understand data requirements Work with engineering, product teams and 3rd parties to collect required data Design, develop and implement large scale, high volume, high performance data models and pipelines for Data Lake and Data Warehouse Develop and implement data quality checks, conduct QA and implement monitoring routines Improve the reliability and scalability of our ETL processes Manage a portfolio of data products that deliver high-quality, trustworthy data Help onboard and support other engineers as they join the team We're excited about you because… 3+ years of professional experience working in data engineering, business intelligence, or a similar role You have proficiency in using AI coding tools (e.g., Claude Code, Codex, Cursor) in the full software development lifecycle, including designing, generating code, testing, monitoring and releasing software Proficiency in programming languages such as Python/Java 3+ years of experience in ETL orchestration and workflow management tools like Airflow, Flink, Oozie and Azkaban using AWS/GCP Expert in Database fundamentals, SQL and distributed computing 3+ years of experience with the Distributed data/similar ecosystem (Spark, Hive, Druid, Presto) and streaming technologies such as Kafka/Flink. Experience working with Snowflake, Redshift, PostgreSQL and/or other DBMS

PythonJavaSQLPostgreSQL
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team Come help us build and develop tools serving hundreds of engineers internally! We’re looking for a Fullstack Software Engineer to join our Developer Insights team. About the Role Our mission is to improve the developer experience of engineers at DoorDash by building various internal products, including our internal developer portal, Developer Insights. Our success as a platform team depends on the success of the product teams we serve. Because of this, we invest in building a strong community that encourages participation and promotes best practices. You’re excited about this opportunity because you will… Introduce cutting edge technologies to our engineering organization, including tools built on LLMs Build new features for Developer Insights (using Backstage.io) Improve the developer experience for all of our engineers Work and collaborate across team boundaries. Contribute features and bug fixes to upstream open-source projects. Mentor and educate your peers. Lead the team in a technical fashion and assist in roadmap planning and measurement of existing features. Represent the team at large in OKR and engineering all-hands presentations. Context switch from frontend to backend to data depending on the need that arises. We’re excited about you because… You have at least 2 years of experience in web technologies using Typescript with React on the frontend with Java, Kotlin, Python or Go backend experience. You have a product mindset and apply that to how you would build out platform services. You love systems and software, and you're proficient in both. You’re curious and dive deep into different system architectures. You are an organized and excellent written and verbal communicator. You have proficiency in using AI coding tools (e.g., Claude Code, Codex, Cursor) in the full software development lifecycle, including designing, generating code, testing, monitoring and releasing software Compensation The successful candidate’s starti

TypeScriptPythonJavaReact
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team Our mission is to provide a world-class development experience that makes DoorDash's web engineers among the most productive in the industry. We achieve this by creating the tools that enable all teams at the company to ship features quickly and reliably. Because their success is our success, we are deeply invested in building a strong, collaborative web community that champions best practices and welcomes participation. About the Role As a Software Engineer on the Developer Experience team, you will build the foundational pieces for all DoorDash, Wolt and Deliveroo Web applications. These include monorepos, build & CI systems and agent-first development tooling. You will work closely with engineers and other internal stakeholders to deliver large and impactful initiatives. Additionally, you will be a culture carrier for our Web engineers through mentorship, education, and engagement of your peers. You will report into the Engineering Manager of our Web Developer Experience team in our Developer Platform organization. You must be located in either San Francisco, CA, Sunnyvale, CA, Los Angeles, CA, Seattle, WA, or New York, NY. You’re excited about this opportunity because you will… Shape the future of Web Development. You will have a direct and meaningful impact on the daily workflows of every web engineer at the company, enhancing their productivity and overall developer experience. Build from the ground up. You will architect and implement foundational libraries, cutting-edge build systems, and innovative development tools that serve as the bedrock for all of our web applications. Solve complex, high-impact challenges. You will tackle some of the most significant technical hurdles in web engineering, and the solutions you deliver will be leveraged by hundreds of employees across numerous product teams. Act as a force multiplier. Your work will directly empower product teams to build, test, and release new features to our customers fa

TypeScriptNode.jsAWSGit
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team Data is at the foundation of DoorDash success. The Data Engineering team builds database solutions for various use cases including reporting, product analytics, marketing optimization and fi nancial reporting. Team serves as the foundation for decision-making at DoorDash. About the Role DoorDash is looking for a Sta ff Software Engineer,Data to be a technical lead and help architect and scale our data reliability, data infrastructure, automation and tools to meet growing business needs. You’re excited about this opportunity because you will... Own critical data systems that support multiple products/teams Develop, implement and enforce best practices for data infrastructure and automation Design, develop and implement large scale, high volume, high performance data models and pipelines for Data Lake and Data Warehouse Improve the reliability and scalability of our Ingestion, data processing, ETLs, Reporting tools and data ecosystem services Manage a portfolio of data products that deliver high-quality, trustworthy data Help onboard and support other engineers as they join the team We’re excited about you because... 8+ years of professional experience as a hands-on engineer and technical leader leading multiple projects 6+ years experience working in data platform and data engineering or a similar role You have proficiency in using AI coding tools (e.g., Claude Code, Codex, Cursor) in the full software development lifecycle, including designing, generating code, testing, monitoring and releasing software Pro fi ciency in programming languages such as Python/Kotlin/Scala 4+ years of experience in ETL orchestration and work fl ow management tools like Air fl ow Expert in database fundamentals, SQL, data reliability practices and distributed computing 4+ years of experience with the Distributed data/similar ecosystem (Spark, Presto) and streaming technologies such as Kaa/Flink/Spark Streaming Excellent communication skills and experience working

PythonSQLAWSGit
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $1.3M/yr

Quick readStrong listing-quality and freshness signals

About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last-mile logistics in the long term. If you want to work on commercializing autonomy and robotics in a service used by millions of people — and on bringing the merchant partners who power that service along with us — then we want to talk to you! About the Role Come help us redefine last-mile logistics through robotics, automation, and other advanced technologies. Autonomy only works when merchants — restaurants, retailers, and other partners — can reliably interact with our robots: handing off orders, troubleshooting edge cases, and trusting the experience enough to keep using it. This role owns that side of the equation. We're looking for a Merchant Success & Growth lead to build the strategy and operational mechanisms that get merchants onboard, keep them performing, and turn their day-to-day reality into a tight feedback loop for product and engineering. You're excited about this opportunity because you will… Own merchant adoption and performance KPIs for autonomy end-to-end — defining what "successful merchant interaction with a robot" means, instrumenting it, and driving improvement against it. Build the playbooks and operational mechanisms to onboard merchants to autonomy — from first conversation through training, go-live, and steady-state ops — and scale them across markets. Partner with sales, account management, and field ops to recruit and ramp the right merchant cohorts for each stage of the product, and design experiments that test new merchant-facing features and handoff models. Define merchant performance benchmarks (handoff success rate, dwell time, dasher/robot interaction quality, merchant CSAT) and run the cadence that holds partners and internal teams accountable to them. Stand up the feedback loop from the field back to product and engineering — turning merchant complaints, edge cases, and frontline observations into prioriti

AWSGitRestAI
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $1.5M/yr

Quick readStrong listing-quality and freshness signals

About the Team The Analytics team is looking for experienced Data Scientists and Senior Data Scientists to guide measurement, strategy, and tactical decision-making across the company across a variety of teams and levels. Data Scientists at DoorDash work to uncover insights and turn them into relevant recommendations, driving decisions for the entire organization. Analytics is integral to all operational areas at DoorDash. Please apply here for all non-managerial levels within the following analytics teams: Consumer & Growth Business Operations Dasher & Logistics Customer Experience & Integrity Merchant, Ads & Sales New Verticals International Data Science About the Role As a Data Scientist at DoorDash, you'll use your quantitative background to mentor other scientists and dive into large datasets to guide decision-making. We solve a multitude of exciting challenges including customer acquisition, fraud and support, marketing, balancing supply and demand, new city launches, marketplace efficiency, and more. If you enjoy finding patterns amidst chaos, and have experience using analytics to affect revenue, growth, operations or beyond, we're looking for someone like you! You're excited about this opportunity because you will… Use quantitative analysis and the presentation of data to see beyond the numbers and understand what drives our business Build full-cycle analytics experiments, reports, and dashboards using SQL, R, Python, or other scripting and statistical tools Work with and mentor junior analysts on how to use more advanced methods and solve challenges Produce recommendations and use statistical techniques and hypothesis testing to validate your findings Provide insights to help business and product leaders understand marketplace dynamics, user behaviors, and long-term trends Identify and measure levers to help move essential metrics and make recommendations Work backwards from understanding and sizing problems to ideating solutions Report aga

PythonSQLAWSGit
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $1M/yr

Quick readStrong listing-quality and freshness signals

About the Team DoorDash is looking for an analytical and entrepreneurial operator to join our US Marketplace Growth team that is responsible for accelerating new growth bets, most notably with College students and Gen Z. You will dive into data to explain performance at the lowest level of detail, use AI tools to drive reporting efficiency, work with cross-functional partners to report out and explain performance, and help manage a team of College Ambassadors and leads who support our on-the-ground presence at colleges around the country. About the Role The Senior Associate, Marketplace Audience Strategy & Operations role is for truth seekers who enjoy getting scrappy while also examining data, creatively problem-solving on a daily basis, driving insights on business performance trends, using AI to drive efficiency, and working across a cross-functional organization to bring strategies to life. You will directly support new Audience penetration for DoorDash’s largest business unit. On a typical day, you will dive into data to explain performance at the lowest level of detail, broker and manage external ecosystem partnerships, think through how to solve blockers to performance at individual College campuses, work with cross functional teams to build reporting and performance updates for senior leadership, think up new ideas and translate them into an actionable test or new reporting mechanism, and drive the creating of strategic plans that drive execution throughout the business. Our best Strategy & Operations Sr. Associates are motivated, analytical, creative, and have exceptional interpersonal and relationship-building skills. We are currently operating in a hybrid model, some time in-office and some time remote but it is quite flexible! You're excited about this opportunity because you will… Collaborate - Work cross-functionally with teams across DoorDash to understand, report, and drive business performance Strategize - Support the

AWSGitRestAI
🔔

Get new test jobs in San Francisco, Canada by email

Daily job updates · Unsubscribe anytime