Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Data Engineering team builds and maintains the foundational datasets that power decision-making across Robinhood. We design reliable, scalable data systems that support product analytics, growth strategy, financial reporting, experimentation, and machine learning. The team partners closely with Product, Engineering, Data Science, and Finance to ensure accurate, well-modeled data is available to teams across the company. Our work directly influences how Robinhood measures performance, improves customer experience, and scales its products. As a Senior Data Engineer, you will design, build, and evolve core datasets that track product performance and company-wide metrics. You will develop scalable data pipelines that ingest application events and database snapshots into our data lake, ensuring high data quality and reliability. You’ll collaborate with application engineers to improve data generation patterns and with analytics teams to design intuitive, well-documented data models. This is an opportunity to shape the technical foundation that supports data-informed decisions across the organization! This role is based in our Menlo Park, CA office, with in-person attend
Jobs in Canada
Reliability Engineer Iii in Canada
138 active opportunities · Updated October 2026
Showing
15 jobs
Explore current reliability engineer iii jobs across Canada. Filter by work mode, employment type, experience, department, date posted and distance.
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. The Lyft Business Product Platform team builds the systems and experiences that power Lyft's B2B products — enabling companies, organizations, and their employees to seamlessly access Lyft's transportation network. We sit at the intersection of product and platform, owning both the customer-facing features and the underlying infrastructure that makes them reliable at scale. Our work directly impacts how businesses integrate with Lyft, how admins manage their programs, and how millions of riders get where they need to go. Responsibilities: Drive architecture and technical design for systems that are highly available, scalable, and built to last — not just for today's requirements but for where the product is heading Own features end-to-end: from shaping the technical spec and design through to production rollout and operational health Think critically about how AI capabilities can be incorporated into Lyft Business products to improve the experience for business admins and riders — and bring that perspective into roadmap and architecture conversations Make well-reasoned trade-off decisions and communicate them clearly to peers, leads, and cross-functional partners Write clean, well-tested, maintainable code and hold a high bar for the same in code reviews Partner across engineering, product, and design to align on direction and get buy-in on technical approaches Proactively engage in incident response, contributing both to resolution and to long-term reliability improvements Grow the team's technical culture through design reviews, tech talks, and mentorship Experience: 5+ years of software engineering experience, with a track record of designing and shipping production systems at scale Strong system design instincts — you can reason through distributed systems trade-offs, identify failure modes, and
$212K – $318K/yr
Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Design state-of-the-art ML models and large-scale ML systems for underwriting and portfolio management for Stripe Capital based on ML principles, domain knowledge, risk, regulatory and engineering constraints. Design systems to speed up the time from idea to deployment of new models. Experiment and iterate on ML models (using tools including PyTorch and TensorFlow) to achieve key business goals and drive efficiency. Develop pipelines and automated processes to train and evaluate models in offline and online environments. Integrate ML models into production systems and ensure their scalability and reliability. Collaborate with product and strategy partners to propose, prioritize, and implement new product features. Engage with the latest developments in ML/AI and take calculated risks in transforming innovative ML ideas into productionized solutions. Who you are Minimum requirements Must have a Bachelor's degree or foreign equivalent in Computer Science, Machine Learning, Mathematics, Physics, Statistics, or a related field, plus two (2) years of experience in Building and shipping ML systems in production. Must have two (2) years of experience in each of the following: ML algorithms and model architectures; Designing, training and evaluating machine learning models; Productionizing and deploying machine learning models at scale; Orchestrating data pipelines and leveraging large-s
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. AS A SENIOR SOFTWARE ENGINEER YOU WILL: Drive high-impact initiatives that span our product areas and full tech stack, including golang and Python on the backend and TypeScript/React on the frontend. Own and deliver features across the notebook service, container runtimes, and UI — designing and shipping medium-to-large projects independently, from ambiguous problem statements through production and post-launch. Advance core platform initiatives such as runtime management and patching, environment reproducibility and replication, security and compliance, and observability for notebooks. Extend the product to operate reliably in regulated and air-gapped environments, where security, compliance, and operational rigor are paramount. Promote strong collaboration within a cross-functional team and partner closely with embedded product managers and designers, as well as platform organizations across Snowflake Be a strong contributor to the product vision and drive team planning. Build for scale, reliability, and high performance, and participate in the on-call rotation to keep a Tier-1 production service healthy. Mentor, coach, and empower more junior team members, and raise the engineering bar through high-quality design and code review. OUR IDEAL CANDIDATE WILL HAVE: 7+ years o
From C$1.4M/yr
About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu
$190K – $270K/yr
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The AI team is building cutting-edge solutions that bring the power of AI directly to edge devices while seamlessly integrating with cloud infrastructure. We are looking for a Lead Software Engineer to design and develop high-performance, scalable services to support AI workloads across edge and cloud environments. What You Might Do Design, build, and maintain services that power AI-driven applications, ensuring scalability and performance. Develop APIs and microservices that facilitate seamless integration between cloud-based AI models and edge devices. Optimize data pipelines and storage solutions for real-time AI inference and processing. Implement security and privacy best practices for distributed AI systems. Work closely with AI researchers, infrastructure engineers, and frontend developers to deliver end-to-end AI-driven solutions. Build and optimize an agent orchestration runtime that enables tool use, memory management, and multi-step reasoning across LLMs, APIs, and edge-connected systems. Develop robust logging, monitoring, and alerting systems to ensure system reliability.
$140K – $225K/yr
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role HP IQ’s Connectivity team is seeking a Software Engineer with deep expertise in device software development. The ideal candidate will bring strong knowledge of connectivity stack and hands-on experience developing, integrating, and optimizing device software to deliver industry-leading user experiences. You will work at the intersection of Wi-Fi, Bluetooth, and emerging device-to-device transport technologies, tackling complex challenges in performance, reliability, and low-latency communication. This role offers the opportunity to contribute to cutting-edge innovations that are redefining how people and devices seamlessly connect across the modern enterprise. What You Might Do Design, develop, and integrate connectivity software features across Android, Windows, and embedded platforms, including SDKs, frameworks, and system services. Implement, optimize, and tune wireless networking protocols and sensing algorithms with a focus on enterprise-scale architectures and deployments. Design, implement, and troubleshoot peer-to-peer technologies to deliver secure, reliable, and low-latency
From C$190K/yr
About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! About the Team Engineers on this team build our rules-based calculation engine for processing sales commissions. This might sound simple if you have never been exposed to sales compensation plans, it is not. We are low on meetings and high on accountability. Most of the team is in the EST time zone, with a few located in PST and Central as well. We are still evolving many areas of the platform, which means there is meaningful room to improve the design, reliability, and scalability of the systems we build. What you’ll be doing Reporting to the Manager of Data Platform, you will play an important role in the evolution of our Spark-based data platform. You’ll design and build data-rich platform capabilities, contribute to system design discussions, and help ensure our data systems remain reliable, maintainable, and scalable as Forma grows. As a Senior Engineer, Data Platform, you are expected to operate with strong ownership and sound technical judgment. This includes identifying risks in the work you own, surfacing edge cases, asking thoughtful questions, and proposing improvements that strengthen the quality and reliability of the platform. You will: Design, build, and improve Spark-based data pipelines and platform services. Work with complex data models representing sales compensation plans, hierarchies, relationships, and enterprise datasets. Build reliable, deterministic data systems that customers and internal teams can trust. Improve testing, observability, data quality,
About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! Senior Staff Backend Engineer About the Team We build enterprise software that helps organizations optimize sales performance, enabling go-to-market agility. Our engineering organization includes multiple product application teams responsible for delivering core customer-facing capabilities. We are seeking a Senior Staff Backend Engineer to join our application teams and help set technical direction across multiple domains within engineering. You'll work alongside staff, senior, and early-career engineers, and partner closely with engineering leadership to define, evolve, and scale the systems that power enterprise-grade product workflows. This is an opportunity to own complex, multi-domain technical problems and shape product direction beyond a single team. We are low on meetings, high on accountability. Most of the teams are in the EST time zone, but we have a few located in AST, PST, and Central as well. What you'll be doing You will play a pivotal role in shaping the technical direction of our application stack across multiple domains. You will lead development efforts for our most complex initiatives, the kind that span two or more teams or product areas, and serve as a technical benchmark for system design, code quality, and long-term maintainability. You'll operate at the intersection of data modelling, business logic, and enterprise-scale reliability, and your work will often set standards that neighboring teams adopt. This remains a hands-on
We take play seriously. We’re looking for curious adventurers ready to find their party, fueled by imagination and drive to build what’s never been built before. At Hasbro and Wizards of the Coast, you’ll collaborate with passionate teams to reimagine our iconic brands and create experiences that spark joy, connection, and community through the magic of play. This is your chance to shape legendary play that lasts a lifetime. Step Into the Multiverse: Your Next Adventure Starts Here At Wizards of the Coast, we harness the power of imagination and connection to create unforgettable experiences. We create entertainment that inspires creativity, sparks passion, forges friendships, and fosters communities around the globe. In every pursuit our mission is to inspire a lifetime love of games. Whether it's through the strategic depth of Magic: The Gathering®, the rich storytelling of Dungeons & Dragons®, or our AAA digital game studios, we build worlds that bring people together, spark creativity, and fuel adventure. As we continue to grow and explore new realms, we're seeking passionate, curious, and innovative minds to join the adventure. Do you have the versatility to operate at the intersection of game development, cloud infrastructure, and platform reliability? We are seeking a Senior Cloud Platform Engineer to join our Central Technology Build Engineering team as we develop and curate an Unreal Engine ecosystem to share across our internal game studios. You will join a specialized team responsible for unblocking developers, ensuring build stability, and optimizing iteration workflows across multiple concurrent titles. This is a hands-on IC role where engineering excellence meets operational reliability. You’ll collaborate closely with game development teams to ensure a seamless experience for all Unreal-powered studios under our banner. In this role, you will design, build, and run scalable AWS architectures that serve as the backbone
From $264.8K/yr
Scale AI is the data foundation for AI, helping organizations build and deploy reliable production AI applications. We partner with leading enterprises and government organizations to accelerate their AI initiatives through our data annotation platform, generative AI solutions, and enterprise AI capabilities. About the General Agents Team The General Agents team, part of Scale’s Enterprise organization, builds robust general agents for customer use cases and applications. The team sits at the intersection of frontier agent development and real-world deployment, translating state-of-the-art reasoning and agentic capabilities into reliable, production-grade systems that drive real economic value. Our agents are scalable systems built around recurring enterprise problem domains, with a strong emphasis on generalization, extensibility, and deployment across many customers. About the Role As a Senior/Staff Machine Learning Engineer (MLE) on the General Agents team, you’ll play a critical role in designing, building, and deploying production-ready AI agents that solve high-impact enterprise problems. You will work across the full agent lifecycle—from model and system design to evaluation, deployment, and iteration—bridging cutting-edge agentic techniques with the constraints and requirements of real customer environments. You will: Design and implement end-to-end agent systems that combine LLM reasoning, tool use, memory, and control logic to solve recurring enterprise use cases. Build scalable, reliable agent architectures that can be deployed across many customers with varying data, tools, and constraints. Develop evaluation frameworks, datasets, environments, and metrics to measure agent performance, reliability, and business impact in production settings. Collaborate closely with product managers, customers, data annotators, and other engineering teams to translate enterprise requirements into robust agent designs. Productionize frontier agent techniques (e.g.,
About the Team DoorDash Labs is an independent innovation hub within DoorDash, focusing on developing automation and robotics solutions to enhance last-mile logistics. Our mission is to create technologies that support and augment human networks, aiming to improve efficiency for Dashers, merchants, and consumers alike. We are a highly senior team composed of former pioneers from a variety of different robotics industries. If you have a passion for applying AI and robotics solutions in a service used by millions of people, then we want to talk to you! Please Note: This role is based in San Francisco, CA, and requires being in-person 5 days a week. About the Role We are seeking a Senior RF Engineer to join our multidisciplinary team developing mission-critical RF/electrical systems for unmanned platforms.This role offers a rare opportunity to gain deep insight into RF design and validation, contributing across the full product lifecycle, from early concept and prototyping through production. In this role, you will apply deep expertise in electromagnetic theory and wireless system design to ensure autonomous platforms operate with exceptional reliability and precision. You will design, optimize and validate RF / mixed-signal PCBAs with a strong focus on interference mitigation, troubleshoot system-level desense, and leverage a thorough understanding of wireless technologies including LTE, 5G NR, Wi-Fi, GPS, and UWB. You’re excited about this opportunity because you will… Lead the development of innovative solutions that span across RF, Mixed signal and mechanical domain. Design RF hardware as well as wireless testbeds to support system validation, and lead the execution and documentation of system-level design and verification Provide expertise, mentorship to partner teams regarding EMI mitigation strategies, and cross-functional debug processes Characterize EMI caused by internal & external aggressors by tracing root causes using EMC probes
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? This role is for people who love building tools for their coworkers. The Internal Applications team creates tools that help us create better models. In this role you will collaborate with internal stakeholders, which include annotators, ML researchers, product managers and more. Join our team of builders who create tooling that will pave the way for the next generation of large language models! As a Full-Stack Software Engineer on the Internal Applications team, you will: Work with a small talented and enthusiastic team of software engineers Contribute to delightful experiences for our user-facing products, meticulously crafting code for browsers and servers Collaborate and grow with your engineering colleagues of all levels through direct pairing sessions, architectural designs, documentation and talks Identify and remove roadblocks to enable your team to increase its engineering velocity. Build resilient systems that are mission-critical Keep up with the cutting edge and adopt new technologies to improve performance and reliability You may be a good fit if: You have experience shipping products with a large numb
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. As a PCB Power Design Engineer, you will design and optimize power distribution systems for Tenstorrent’s next-generation AI accelerator cards, balancing performance, area, cost, and reliability. You will contribute across the full power design lifecycle, from component selection and simulation through board bring-up, validation, debugging, and production readiness. Working closely with hardware, firmware/software, thermal, mechanical, and validation teams, you will help deliver robust power architectures for high-performance AI systems. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An electrical engineer with experience designing, analyzing, and debugging power distribution systems for high-performance or high-power products. A hands-on engineer who enjoys moving from schematics and simulations into lab bring-up, testing, troubleshooting, and design validation. A systems-oriented collaborator who can work effectively across hardware design, firmware/software, thermal, mechanical, and validation teams. A detail-oriented problem solver who balances electrical performance, power integr
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Join Tenstorrent and help bring next-generation AI accelerator technology from silicon bring-up to production. You’ll work at the forefront of hardware innovation, diagnosing complex issues across chips, systems, firmware, and software while collaborating with some of the brightest engineers in the industry. This role offers the opportunity to solve challenging technical problems, build impactful debug solutions, and directly influence the reliability and performance of cutting-edge AI compute platforms. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A hands-on hardware debug engineer who thrives on solving complex, cross-functional problems at the intersection of silicon, firmware, and software. A curious and analytical problem solver who enjoys digging into failures, identifying root causes, and driving issues from initial discovery through resolution. An engineer with strong post-silicon validation and bring-up experience who is comfortable working in the lab and getting deep into system-level behavior. Someone who enjoys building tools, improving debug methodologies, and creating
Other cities to consider
More places hiring for this role
Get new reliability engineer iii jobs in Canada by email
Daily job updates · Unsubscribe anytime