Jobs in Canada

Software Reliability Engineer in San Francisco

129 active opportunities · Updated October 2026

Explore current software reliability engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

A
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About the Team DevX owns engineering velocity at Amplitude: build systems, CI/CD, developer environments, and internal tooling. We're building a software factory: automated workflows that remove manual bottlenecks from how engineers ship code. The Role We're looking for a Staff Software Engineer – DevX (Hybrid – San Francisco) who bridges infrastructure and application thinking and can accelerate how the whole team develops, tests, and ships in the cloud. You'll set architecture for our developer platform, lead our software factory work, and push our development model toward cloud-first workflows. This is a high-leverage, low-oversight role. You'll own initiatives end to end, from an ambiguous problem to production, and set technical direction for a foundational team. What You'll Do Cloud development platform: Unify and scale our existing loca

CI/CDGitAIGo
V
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$190K – $220K/yr

Quick readStrong listing-quality and freshness signals

About VSCO For years, we've helped photographers create their work. Now we're building what comes next. VSCO exists for photographers. Not as a side feature, not as an afterthought, but as the whole point. We build the connected system photographers rely on, and we've spent over a decade earning the trust of a global creative community that takes the craft seriously. Photography is at an inflection point. AI is reshaping what's possible for creative work, and that's where our mission shines. VSCO is building the full photographer's workflow: from creating and editing your work, to delivering it to clients, to running your business. All of it built thoughtfully, with photographers leading the way. We believe the future of photography tools creates more space for creativity, handling the busy work so photographers can focus on the craft. If you care about craft, community, and what technology can unlock for creative people, this is the work. We're a mission-driven and focused company where your work ships quickly, is meaningful, and reaches tens of millions of people worldwide. You'll have a real say in what we build and how we build it. We hire people who don't wait to be asked, naturally connect the dots, care about the quality of what they ship, and believe the best outcomes come from building together. About The Role VSCO is hiring a Senior Software Engineer to be a primary contributor on Reflex, our production GPU image-processing engine, and significantly guide the path it takes into Studio Pro on iOS and macOS. You will deepen a stack that is already in customers’ hands: new capabilities (RAW, large-image export, desktop), production quality (color correctness, memory, performance), and the contracts that let product engineers ship on top of the engine without becoming graphics programmers. In This Role, You Will Be a primary contributor on the Reflex imaging engine: DAG execution, WGSL operators, GPU resource/memory budgets, and color management (linear workin

ReactAIGoRust
V
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$230K – $260K/yr

Quick readStrong listing-quality and freshness signals

About VSCO For years, we've helped photographers create their work. Now we're building what comes next. VSCO exists for photographers. Not as a side feature, not as an afterthought, but as the whole point. We build the connected system photographers rely on, and we've spent over a decade earning the trust of a global creative community that takes the craft seriously. Photography is at an inflection point. AI is reshaping what's possible for creative work, and that's where our mission shines. VSCO is building the full photographer's workflow: from creating and editing your work, to delivering it to clients, to running your business. All of it built thoughtfully, with photographers leading the way. We believe the future of photography tools creates more space for creativity, handling the busy work so photographers can focus on the craft. If you care about craft, community, and what technology can unlock for creative people, this is the work. We're a mission-driven and focused company where your work ships quickly, is meaningful, and reaches tens of millions of people worldwide. You'll have a real say in what we build and how we build it. We hire people who don't wait to be asked, naturally connect the dots, care about the quality of what they ship, and believe the best outcomes come from building together. About The Role If you are passionate about iOS development and excited to join a dynamic team of creators, we encourage you to apply for this position and help us empower creators worldwide. At VSCO, you'll have the opportunity to collaborate with talented individuals who share your passion for innovation and creativity. We're building an ecosystem of end-to-end tools for photographers, enhancing our existing product while creating entirely new tools and services. This is a crucial role for someone who will help define and drive that vision forward, shaping the way our users create, connect and earn. Join us in shaping the future of mobile technology and making a me

ReactRestAISwift
V
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$150K – $240K/yr

Quick readStrong listing-quality and freshness signals

Location: San Francisco, CA (Remote/Hybrid Available) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role As a Software Engineer focused on Fleet Telemetry & Control at Verse, you will be working closely with our energy solutions partners to design, implement, and test distributed energy resource controls and telemetry software on customer hardware at sites around the world. You will be part of a dynamic, high-performance team building applications directly on bare-metal or on hardware-level virtualization platforms. As an advanced technical leader in network programming and state management development, engineering teams will look to you for best standards and practices for interfacing with on-premises grid assets using solutions you will build and maintain. Key Responsibilities Foster a culture and mindset of well-designed systems, test-driven software, and transparent communication with a high caliber of mutual respect and consideration for stakeholders Mentor and support career and junior level engineers in their fleet telemetry and control software career development

PythonSQLAIC++
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$240K – $270K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on the modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing Solve challenging problems that arise in providing an interactive experience on data warehouses for data exploration and analysis Build with modern tools and languages like Rust, Go, GraphQL, Node, and Kubernetes Build backend distributed services, new algorithms and modern API to support a cloud application Triage product or system issues and debug/track/resolve by analyzing the sources of issues Design and implement new software features to support our fast growing user base Collaborate with peers and stakeholders through design and code reviews to ensure best practices amongst available technologies Qualifications We Need 5+ years industry experience building and maintaining high-quality software Experience building and deploying robust and secure web applications in a continuous deployment environment Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong Computer Science fundamentals Qualifications We Want (also, skills you’ll learn!) Data driven aptitude and its application to solve distributed system problems Data model design, and API development experience SQL query optimization and database internals Administered cloud service infrastructure (GCP, AWS, Azure) Prior experience working at high growth company solving technical problems to enable continued success Additional Job details The base salary range for this position is $240k - $270

PythonSQLAWSAzure
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$240K – $270K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on the modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing Solve challenging problems that arise in providing high performance interactive experience to enable analytics and workflows use cases on top of modern warehouses Build software using the latest developer tools and using programming languages like Rust, Go, GraphQL, Typescript Develop new algorithms and techniques for improving the performance and interactivity for enabling analytics and workflows for the world largest companies Triage product or system issues and debug/track/resolve by analyzing the sources of issues Design and implement new software features to support our fast growing user growth Collaborate with cross-functional groups - infrastructure, design, product, customer support, sales and marketing to build an innovative product capabilities Qualifications We Need 10+ years industry experience building and maintaining high-quality software Experience building and deploying robust and secure web applications in a continuous deployment environment Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong Computer Science fundamentals Qualifications We Want (also, skills you’ll learn!) Experience building software capabilities for analyzing large scale data web applications Data driven aptitude and its application to solve distributed system problems Data model design, and API development experience to enable customer facing cap

TypeScriptPythonSQLGraphql
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$240K – $270K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing You will be responsible for developing elegant and responsive user experience using the latest front-end technologies. You'll own substantial pieces of the product, from design to launch Working with our product, UX design, and backend development teams, you will develop new features and technologies that make our product experience awesome and radically simplify the user experience for non-technical users You will leverage your technical expertise in front-end application development in the creation of novel visualizations for structured and unstructured data and develop new techniques for improving the performance and interactivity of the application Use modern frontend frameworks like React, GraphQL, TypeScript and Node.js Qualifications We Need 10+ years industry experience building and maintaining high-quality software An eye for great design and a passion for building products that provide a great user experience The ability to make the right trade-offs between functionality and delivery speed that supports delivering value to customers, all the while iterating based on feedback and roadmap priorities Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong computer science fundamentals Qualifications We Want (also, skills you’ll learn!) Experience building software capabilities for analyzing large scale data web applications Prior ex

TypeScriptPythonReactNode.js
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$150K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on the modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing Solve challenging problems that arise in providing high performance interactive experience to enable analytics and workflows use cases on top of modern warehouses Build software using the latest developer tools and using programming languages like Rust, Go, GraphQL, Typescript Develop new algorithms and techniques for improving the performance and interactivity for enabling analytics and workflows for the world largest companies Triage product or system issues and debug/track/resolve by analyzing the sources of issues Design and implement new software features to support our fast growing user growth Collaborate with cross-functional groups - infrastructure, design, product, customer support, sales and marketing to build an innovative product capabilities Qualifications We Need 5+ years industry experience building and maintaining high-quality software Experience building and deploying robust and secure web applications in a continuous deployment environment Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong Computer Science fundamentals Qualifications We Want (also, skills you’ll learn!) Experience building software capabilities for analyzing large scale data web applications Data driven aptitude and its application to solve distributed system problems Data model design, and API development experience to enable customer facin

TypeScriptPythonSQLGraphql
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$170K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing You will be responsible for developing elegant and responsive user experience using the latest front-end technologies. You'll own substantial pieces of the product, from design to launch Working with our product, UX design, and backend development teams, you will develop new features and technologies that make our product experience awesome and radically simplify the user experience for non-technical users You will leverage your technical expertise in front-end application development in the creation of novel visualizations for structured and unstructured data and develop new techniques for improving the performance and interactivity of the application Use modern frontend frameworks like React, GraphQL, TypeScript and Node.js Qualifications We Need 5+ years industry experience building and maintaining high-quality software An eye for great design and a passion for building products that provide a great user experience The ability to make the right trade-offs between functionality and delivery speed that supports delivering value to customers, all the while iterating based on feedback and roadmap priorities Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong computer science fundamentals Qualifications We Want (also, skills you’ll learn!) Experience building software capabilities for analyzing large scale data web applications Prior exp

TypeScriptPythonReactNode.js
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$170K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on the modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing Solve challenging problems that arise in providing an interactive experience on data warehouses for data exploration and analysis Build with modern tools and languages like Rust, Go, GraphQL, Node, and Kubernetes Build backend distributed services, new algorithms and modern API to support a cloud application Triage product or system issues and debug/track/resolve by analyzing the sources of issues Design and implement new software features to support our fast growing user base Collaborate with peers and stakeholders through design and code reviews to ensure best practices amongst available technologies Qualifications We Need 5+ years industry experience building and maintaining high-quality software Experience building and deploying robust and secure web applications in a continuous deployment environment Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong Computer Science fundamentals Qualifications We Want (also, skills you’ll learn!) Data driven aptitude and its application to solve distributed system problems Data model design, and API development experience SQL query optimization and database internals Administered cloud service infrastructure (GCP, AWS, Azure) Prior experience working at high growth company solving technical problems to enable continued success Additional Job details The base salary range for this position is $170k - $240

PythonSQLAWSAzure
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $252K/yr

Quick readStrong listing-quality and freshness signals

About Scale AI At Scale, our mission is to develop reliable AI systems for the world's most important decisions. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. Scale Frontier Data is the organization behind the training and evaluation data that frontier labs depend on. We build the systems, tooling, and expert workflows that turn hard human expertise into signals that models can learn from, across reasoning, coding, agentic tool use, and domain expertise. Reinforcement learning environments are now the center of gravity for that work: the difference between a model that demos well and a model that reliably completes long-horizon work is almost always the quality of the environments and reward signals it was trained against. Responsibilities As a Staff Software Engineer, RL Environments, you'll own the technical foundation for how Scale builds, runs, verifies, and delivers RL environments at scale. An RL environment is a real piece of software: a containerized world with real dependencies, real state, real tools, and a grader that has to be correct even when the agent is creative about breaking it. Building one is a full-stack engineering problem. Building thousands of them reproducibly, cheaply, with trustworthy reward signals and throughput measured in millions of rollouts is a systems problem that very few people have solved. You'll work on both. You'll design the platform: sandboxed execution, environment packaging and versioning, rollout orchestration, trajectory capture, verifier frameworks, and the authoring surfaces that let engineers and domain experts produce environments without reinventing infrastructure each time. And you'll go deep on the environments themselves by instrumenting real applications, designing task suites that expose specific capability gaps, and building graders that

TypeScriptPythonReactAWS
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $179.4K/yr

Quick readStrong listing-quality and freshness signals

Scale GP (Scale Generative AI Platform) is an enterprise-grade AI platform that provides APIs for knowledge retrieval, inference, evaluation, and more. We are looking for a strong engineer to join our team and help us build and scale our core infrastructure in a fast-paced environment. The ideal candidate will have a strong understanding of software engineering principles and practices, as well as experience with large-scale distributed systems. You will implement solutions across multiple cloud providers (GCP, Azure, AWS) for customers in diverse, highly-regulated industries like healthcare, telecom, finance, and retail. What You’ll Do: Architect multi-cloud systems and abstractions to allow the SGP platform to run on top of existing Cloud providers Implement custom integrations between Scale AI's platform and customer data environments (cloud platforms, data warehouses, internal APIs) Collaborate with platform, product teams and our customers directly to develop and implement innovative infrastructure that scales to meet evolving needs. Deliver experiments at a high velocity and level of quality to engage our customers Work across the entire product lifecycle from conceptualization through production Be able, and willing, to multi-task and learn new technologies quickly What We’re Looking For: 4+ years of full-time engineering experience, post-graduation Experience scaling products at hyper growth startups Experience tinkering with or productizing LLMs, vector databases, and the other latest AI technologies Proficient in Python or Javascript/Typescript, and SQL Experience with Kubernetes Experience with major cloud providers (AWS, Azure, GCP) Excellent communication skills with the ability to explain technical concepts to both technical and non-technical audiences Compensation packages at Scale for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries fo

JavaScriptTypeScriptPythonJava
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $216K/yr

Quick readStrong listing-quality and freshness signals

The Public Sector software engineers (SWEs) create the core product building blocks forward-deployed teams use to develop agentic capabilities that function across multiple domains. SWEs responsibilities include building the systems required to ingest and process federal datasets to support real-time decision-making in contested environments. We develop novel agentic enabling capabilities that includes: Create multi-layered guardrails around agents Optimize data retrieval for agents Orchestrate fleets of asynchronous agents Automatically alerts users to deviations in data Illustrating how an agent reached a decision As a Senior Software Engineer, you will lead the development of a vertical feature or a horizontal capability to include defining requirements with stakeholders and implementation until it is accepted by the stakeholders. You will: Lead the design and implementation of scalable backend systems and distributed architectures for Federal customers. Manage the full lifecycle of feature development from requirement definition to deployment on classified networks. Direct the orchestration of asynchronous agent fleets to meet mission requirements. Lead customer engagements to translate mission needs into technical requirements. Own the communication with stakeholders to ensure implementation meets defined acceptance criteria. Conduct technical reviews and identify risks within machine learning infrastructure and model serving. Drive the platform roadmap by providing technical specifications for Federal product offerings. Ideally you will have: Full Stack Development: Proficiency in front-end, back-end development and infrastructure, including experience with modern web development frameworks, programming languages, and databases Cloud-Native Technologies: Familiarity with cloud platforms (e.g., AWS, Azure, GCP) and experience in developing and deploying applications in a cloud-native environment. Understanding of containerization (e.g., Docker) and contai

AWSAzureGCPDocker
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $180K/yr

Quick readStrong listing-quality and freshness signals

About Scale At Scale AI, our mission is to accelerate the development of AI applications. For 8 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including: generative AI, defense applications, and autonomous vehicles. With our recent Series F round, we’re accelerating the abundance of frontier data to pave the road to Artificial General Intelligence (AGI), and building upon our prior model evaluation work with enterprise customers and governments, to deepen our capabilities and offerings for both public and private evaluations. About Data Engine Our Generative AI Data Engine powers the world’s most advanced LLMs and generative models through world-class RLHF (Reinforcement Learning with Human Feedback), human data generation, model evaluation, safety, and alignment. The data we are producing is some of the most important work for how humanity will interact with AI. Our Approach As part of the interview process, you’ll be considered for opportunities across several teams within the GenAI Engineering organization, based on your interests, expertise, and business needs. Potential team placements include Allocation, Growth, Frontier Data, Trust & Safety, Pay, Operator, or Tasking Experience. Together, these teams power Scale’s AI data operations - from building high-impact datasets that push the boundaries of LLM capabilities, to optimizing contributor onboarding and incentives, to safeguarding data integrity through advanced trust, safety, and security measures. They work at the intersection of ML, operations, and analytics to ensure we deliver the highest-quality data at scale. Responsibilities: Design, build, and maintain robust, scalable systems across the full stack, including front-end, back-end, and infrastructure layers Implement high-impact features using modern technologies such as TypeScript, React, Node.js, MongoDB, Elasticsearch, and Temporal Collaborate closely with internal operators (your use

TypeScriptPythonReactNode.js
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team DoorDash’s GenAI Platform team sits within Machine Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and Deliveroo teams safely bring GenAI-powered products, agents, automation, and personalization to production. Our mission is to increase the velocity of business impact from GenAI. A central pillar of that work is running frontier open-weight LLMs and VLMs (such as GLM, Qwen, Kimi, and DeepSeek) ourselves — real-time GPU serving, high-throughput batch inference, and fine-tuning on autoscaling GPUs — delivering large cost and latency wins (for example, a billion embeddings produced roughly 20× cheaper and visual models served roughly 72% cheaper). We also own core platform surfaces including the LLM Gateway, Agent Gateway, evals infrastructure, guardrails, and cost attribution. About the Role You will join a small, high-leverage team building production infrastructure for Generative AI at DoorDash, leading the design and architecture of our open-weights model platform spanning inference and fine-tuning: real-time GPU serving, high-throughput batch inference, and model fine-tuning. You’ll set technical direction across model serving and inference engines, fine-tuning and training pipelines, GPU autoscaling and utilization, batch pipelines, backend services, and observability, and mentor engineers as you go. This role is ideal for a senior engineer who enjoys owning ambiguous, high-impact systems and pushing the cost/performance frontier of GPU inference and fine-tuning in a fast-moving technical area where product needs, model capabilities, vendor ecosystems, and cost/performance tradeoffs are evolving quickly. You’re excited about this opportunity because you will… Lead the design of infrastructure that helps DoorDash teams move GenAI ideas from prototype to production, increasing the velocity of business impact from AI across the company. Own and evolve our open-weights serving stack — real-time GPU endpoints, high-thr

PythonAWSGCPKubernetes
🔔

Get new software reliability engineer jobs in San Francisco, Canada by email

Daily job updates · Unsubscribe anytime