Jobs in Canada

Software Reliability Engineer in Toronto

155 active opportunities · Updated October 2026

Explore current software reliability engineer jobs in Toronto. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 Toronto, Ontario, Canada· Full-time
✓ High-confidence listingCompany trend -63.6%

From C$160K/yr

Quick readStrong listing-quality and freshness signals

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Streaming Foundations team builds services and operates data pipeline infrastructure to support event streaming, messaging, and analytics use cases. We are looking for a Software Engineer who is passionate about distributed systems, platform engineering, and solving data-intensive problems at scale. In this high-impact role, you will get to work with engineers throughout the organization to build foundational infrastructure that allows Auth0 to scale for years to come. What you’ll be doing Help set the technical direction for the team and influence the engineering roadmap for the Platform’s streaming capabilities Design and lead the implementation of our most complex and critical systems for data-intensive use cases. Research and champion new technologies and architectural patterns to solve strategic challenges and scale the platform. Lead and influence cross-functional initiatives, ensuring technical alignment and successful execution across multiple teams. Improve the operational posture of our systems by designing for observability, reliability, and scalability, and by mentoring others in operational best practices. Coach and mentor senior engineers and act as a technical leader across the engineering organization. Collaborate with different stakeholders like product teams whenever needed. What you’ll bring to our teams 7+ years of software development experience in a fast-paced, agile environment Experience working with Golang or Java is preferred H

TypeScriptJavaReactAWS
L
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -72.4%

From C$108K/yr

Quick readStrong listing-quality and freshness signals

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are building and maintaining a highly scalable asynchronous platform that empowers our organization to handle critical business cases. As a software engineering team, our mission is to create robust and innovative solutions that drive the success of our business and deliver unparalleled value to our customers. We adopt Infrastructure as Code practice to automate the provisioning and configuration of our resources, which helps reduce manual configuration and improve consistency. Our team culture is built on collaboration, open communication, and a supportive environment where each member's ideas are valued and contributions are recognized. We believe in the importance of fostering a positive workplace culture that inspires innovation and creativity. Responsibilities: Maintain and analyze metrics from; operating systems; control planes; and applications to assist in fault detection and performance enhancement Design, develop and deploy tooling and systems that continually improve the reliability, scalability and efficiency of our platform Balance feature development speed and reliability with service-level objectives Operate and improve our Infrastructure using industry best practices and tools Participate in design and production readiness reviews, platform management and capacity planning ceremonies with cross-functional teams Document Infrastructure operations process and insights, identify repeatable actions and ruthlessly automate repetitive tasks Participate in our teams on-call rotations, respond to incidents and support other teams mitigate customer impacting events Experience: 5+ years experience working on teams responsible for software development, automation and systems engineering Experience building large-scale infrastructure, distributed systems or networks. Knowledge with SQS,

PythonAWSAzureGCP
L
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -72.4%

From C$108K/yr

Quick readStrong listing-quality and freshness signals

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Our Infrastructure team is passionate about building software to solve problems at massive scale. We do this often, and when we believe our solution is worth sharing with the community, such as Envoy Proxy , we open source our ideas for the benefit of others. As an Observability team member, you are responsible for the operation and maintenance of our logging and metrics infrastructure. You ensure all teams at Lyft are aware of the operational health of their products by monitoring system availability and take a holistic view of our platform performance. You build software and platforms to automate infrastructure platform operations and management. By measuring and monitoring our operations you find opportunities to improve our systems in order to push our platform forward. You provide our partners with the support they need to help them build robust large scale distributed systems. We count on the reliability of our infrastructure to empower Lyft teams to provide our customers rich experiences that are highly available with rock solid performance to ensure our transportation platform continues to connect people and places. As we grow our team, we are seeking experienced Infrastructure Engineer to ensure that as our Infrastructure continues to scale, our platform continues to provide an essential and dependable service that transports millions of people every day. Specifically we are searching for someone who brings fresh perspectives, enjoys collaborating with cross-functional teams in order to continually improve our products and services for our customers. Responsibilities: Maintain, improve, and develop tooling and systems that enhance the reliability, scalability, and efficiency of our platform. Assist engineering teams in defining service-level objectives (SLOs) and provide the necessary toolin

PythonAWSKubernetesAI
S
📍 Toronto, Ontario, Canada· Full-time
✓ Quality checkedCompany trend -85.7%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. AS A SENIOR SOFTWARE ENGINEER YOU WILL: Drive high-impact initiatives that span our product areas and full tech stack, including golang and Python on the backend and TypeScript/React on the frontend. Own and deliver features across the notebook service, container runtimes, and UI — designing and shipping medium-to-large projects independently, from ambiguous problem statements through production and post-launch. Advance core platform initiatives such as runtime management and patching, environment reproducibility and replication, security and compliance, and observability for notebooks. Extend the product to operate reliably in regulated and air-gapped environments, where security, compliance, and operational rigor are paramount. Promote strong collaboration within a cross-functional team and partner closely with embedded product managers and designers, as well as platform organizations across Snowflake Be a strong contributor to the product vision and drive team planning. Build for scale, reliability, and high performance, and participate in the on-call rotation to keep a Tier-1 production service healthy. Mentor, coach, and empower more junior team members, and raise the engineering bar through high-quality design and code review. OUR IDEAL CANDIDATE WILL HAVE: 7+ years o

TypeScriptPythonReactSQL
V
📍 Toronto, Ontario, Canada· Full-time
✓ Quality checkedCompany trend -100%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's Core Platform team provides the foundational infrastructure that powers all engineering at Vanta. We're expanding upmarket to support enterprise customers, which requires strategic investment in platform systems that ensure security, reliability, and developer productivity at scale. As we expand upmarket to support enterprise and regulated customers, we’re investing heavily in platform capabilities that scale securely while reducing cognitive load for product teams. As the Engineering Manager, Core Platform at Vanta, you'll own the foundational infrastructure that every engineer builds on, ensuring it scales with company growth while remaining fast, simple, and reliable. This team’s ownership spans shared services infrastructure, observability and monitoring, datastore management, and async work systems. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow high-performing platform engineering teams that deliver reliable, scalable infrastructure and operational excellence for Vanta’s products and customers Set technical direction and drive multi-quarter platform initiatives spanning infrastructure reliability, security, scalability, and developer experience across shared systems and services Partner closely with product engineerin

MongoDBAWSRestAI
R
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -76.2%
Quick readStrong listing-quality and freshness signals

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The International engineering team's mission is to expand Robinhood's products globally and scale our platform to support millions of new customers in diverse markets. We build and maintain experiences across onboarding, funding, account setup, localized product experiences, growth incentives, and platform capabilities to operate in new jurisdictions. Our engineering team partners closely with product, design, compliance, marketing teams, along with an array of platform engineering teams to launch new products and optimize user acquisition globally. We prioritize technical rigor, system reliability, and quick execution to deliver reliable products to our growing customer base. Our culture is centered on clear communication, strong team partnership, and a commitment to helping people manage their financial lives! As an Engineering Manager , you will lead a team of software developers to drive some of Robinhood’s most important international launch initiatives. You will be responsible for driving the architecture and execution required to launch crypto and other financial products in new regions, while shaping the reusable systems that make future launches faster, safer, and m

AWSAIGoRust
L
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -72.4%

From C$118.8K/yr

Quick readStrong listing-quality and freshness signals

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Machine Learning is at the heart of Lyft’s products and decision-making. Machine Learning Engineers at Lyft operate in dynamic environments, moving quickly to build the world’s best transportation solutions. We tackle a wide range of challenges, from pricing and marketplace frameworks that ensure reliability and competitiveness, to agentic AI platforms that automate analytical workflows, to behavioral detection systems that protect the integrity of our network. We operate at the intersection of applied ML and real business impact, shipping models that directly influence revenue, rider experience, and partner trust. Lyft Business builds products that help organizations move the people who matter most—employees, customers, patients, and guests—easily and efficiently. Our offerings include Business Travel, Lyft Pass, and Concierge (for healthcare and non-healthcare rides), enabling companies to manage transportation at scale through APIs, integrations (e.g., Concur, Expensify), and dedicated tools. These platforms power high-impact B2B use cases across corporate travel, healthcare access, customer experience, and community programs. We're looking for a Machine Learning Engineer to design, build, and deploy ML systems across Lyft Business. This is a high-scope role: you won't be siloed into one problem area. Instead, you'll move across pricing algorithms, fraud and behavior detection, agentic AI systems, and emerging ML applications as the business evolves. You'll write production-quality code, own models end-to-end from prototyping through deployment, and collaborate closely with Data Scientists, Product Managers, and Software Engineers to translate complex business problems into scalable ML solutions. This role is ideal for someone who is technically versatile, energized by variety, and wants to see th

AWSMachine LearningAIGo
T
📍 Toronto, Ontario, Canada
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. As a PCB Power Design Engineer, you will design and optimize power distribution systems for Tenstorrent’s next-generation AI accelerator cards, balancing performance, area, cost, and reliability. You will contribute across the full power design lifecycle, from component selection and simulation through board bring-up, validation, debugging, and production readiness. Working closely with hardware, firmware/software, thermal, mechanical, and validation teams, you will help deliver robust power architectures for high-performance AI systems. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An electrical engineer with experience designing, analyzing, and debugging power distribution systems for high-performance or high-power products. A hands-on engineer who enjoys moving from schematics and simulations into lab bring-up, testing, troubleshooting, and design validation. A systems-oriented collaborator who can work effectively across hardware design, firmware/software, thermal, mechanical, and validation teams. A detail-oriented problem solver who balances electrical performance, power integr

T
📍 Toronto, Ontario, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Join Tenstorrent and help bring next-generation AI accelerator technology from silicon bring-up to production. You’ll work at the forefront of hardware innovation, diagnosing complex issues across chips, systems, firmware, and software while collaborating with some of the brightest engineers in the industry. This role offers the opportunity to solve challenging technical problems, build impactful debug solutions, and directly influence the reliability and performance of cutting-edge AI compute platforms. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A hands-on hardware debug engineer who thrives on solving complex, cross-functional problems at the intersection of silicon, firmware, and software. A curious and analytical problem solver who enjoys digging into failures, identifying root causes, and driving issues from initial discovery through resolution. An engineer with strong post-silicon validation and bring-up experience who is comfortable working in the lab and getting deep into system-level behavior. Someone who enjoys building tools, improving debug methodologies, and creating

PythonAWSAISEM
T
📍 Toronto, Ontario, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is building next-generation CPU and AI silicon. You’ll work at the forefront of hardware innovation, diagnosing complex issues across chips, systems, firmware, and software while collaborating with some of the brightest engineers in the industry. This role offers the opportunity to solve challenging technical problems, build impactful debug solutions, and directly influence the reliability and performance of cutting-edge AI compute platforms. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Experienced in hardware debug and post-silicon bring-up for CPU, SoC, or ASIC systems. Strong understanding of processor architecture and microarchitecture (RISC-V, x86, or ARM) with familiarity in debug and trace methodologies (e.g., iJTAG). Hands-on engineer who excels at diagnosing complex hardware, firmware, and software issues through root-cause analysis. Comfortable working in the lab with a passion for building debug tools, automation, and scalable methodologies. Collaborative team player with experience partnering across ASIC, firmware, software, and validation teams. What We Need

PythonAWSAIExcel
R
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -76.2%
Quick readStrong listing-quality and freshness signals

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team of bold thinkers and sharp problem-solvers who are wired to make an impact. The Ops Platform organization develops internal platforms that replace repetitive manual processes with AI-driven systems. These tools support key areas such as Fraud Operations, Account Operations, Financial Crimes Operations, and Retirement Services. The team works closely with product, data science, and operations partners to deliver reliable systems that improve decision-making and efficiency! As a Software Developer, you will design and build platforms that enable operational teams to investigate and resolve issues more quickly and accurately. You will work with large datasets and signals to create tooling that supports fraud investigation and other operational workflows. You will collaborate with data scientists and machine learning engineers to translate manual processes into automated systems. Your work will focus on improving system reliability, reducing operational effort, and increasing the speed at which new products and features can be supported across Robinhood’s offerings. This role is based in our Toronto, ON office, with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams. What you’ll do You will define technical direction and make architectural decisions for systems that support operational workflows across multiple product lines You will build tools that proces

AWSMachine LearningAI
R
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -76.2%
Quick readStrong listing-quality and freshness signals

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world's biggest financial problems. We're looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn't a place for complacency, it's where ambitious people do the best work of their careers. We're a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Software Platform team accelerates developer velocity and increases system reliability by building the foundational platforms and tools that power Robinhood engineering. Within this group, the Kubernetes Compute team focuses on building and operating a highly available, scalable Kubernetes-powered container platform. We ensure that our infrastructure seamlessly supports reliable application deployments, integrates core platform capabilities, and enables multi-region scalability. We are expanding our core container systems to support our next phase of technical growth! As a Software Developer, you will focus on building, maintaining, and scaling our container provisioning platforms. Working alongside senior engineers, you will write code to improve our infrastructure capabilities and actively participate in our technical transition to Amazon EKS. In this role, you will collaborate with teams across the organization to ensure robust platform integrations for everyday application needs like security and networking. Your efforts will directly improve system visibility, automation, and reliability across the platform. This role is based in our Toronto office(s), with in-perso

AWSKubernetesAI
R
📍 Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -76.2%
Quick readStrong listing-quality and freshness signals

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold builders and sharp problem-solvers who are wired to deliver great outcomes. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. The DevX team’s mission is to build and operate the core developer infrastructure at Robinhood. Our team owns and scales the systems that thousands of engineers rely on daily, partnering with software developers across the company to make development fast, reliable, and cost-efficient! As a Staff Software Developer, you will act as a technical leader for our build and developer infrastructure, driving the strategy and execution of the systems thousands engineers depend on every day. Your work will span our build systems, CI pipelines, and remote development environments, ensuring engineers can code, test, and build with speed, safety, and reliability at scale. In this role, you will collaborate with teams across Robinhood to eliminate developer friction and raise the bar for engineering productivity. This is a high-visibility leadership opportunity to shape our developer ecosystem and set new standards of engineering efficiency! This role is based in our Toronto, ON office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams. What you’ll do Architect the long-te

PythonAWSCI/CDAI
F
📍 Toronto, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! Senior Staff Backend Engineer About the Team We build enterprise software that helps organizations optimize sales performance, enabling go-to-market agility. Our engineering organization includes multiple product application teams responsible for delivering core customer-facing capabilities. We are seeking a Senior Staff Backend Engineer to join our application teams and help set technical direction across multiple domains within engineering. You'll work alongside staff, senior, and early-career engineers, and partner closely with engineering leadership to define, evolve, and scale the systems that power enterprise-grade product workflows. This is an opportunity to own complex, multi-domain technical problems and shape product direction beyond a single team. We are low on meetings, high on accountability. Most of the teams are in the EST time zone, but we have a few located in AST, PST, and Central as well. What you'll be doing You will play a pivotal role in shaping the technical direction of our application stack across multiple domains. You will lead development efforts for our most complex initiatives, the kind that span two or more teams or product areas, and serve as a technical benchmark for system design, code quality, and long-term maintainability. You'll operate at the intersection of data modelling, business logic, and enterprise-scale reliability, and your work will often set standards that neighboring teams adopt. This remains a hands-on

JavaScriptTypeScriptPythonJava
C
📍 Toronto, Ontario, Canada· Full-time· Remote
✓ High-confidence listingCompany trend -91.5%
Quick readStrong listing-quality and freshness signals

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? This role is for people who love building tools for their coworkers. The Internal Applications team creates tools that help us create better models. In this role you will collaborate with internal stakeholders, which include annotators, ML researchers, product managers and more. Join our team of builders who create tooling that will pave the way for the next generation of large language models! As a Full-Stack Software Engineer on the Internal Applications team, you will: Work with a small talented and enthusiastic team of software engineers Contribute to delightful experiences for our user-facing products, meticulously crafting code for browsers and servers Collaborate and grow with your engineering colleagues of all levels through direct pairing sessions, architectural designs, documentation and talks Identify and remove roadblocks to enable your team to increase its engineering velocity. Build resilient systems that are mission-critical Keep up with the cutting edge and adopt new technologies to improve performance and reliability You may be a good fit if: You have experience shipping products with a large numb

TypeScriptPythonReactSQL
🔔

Get new software reliability engineer jobs in Toronto, Canada by email

Daily job updates · Unsubscribe anytime