Jobs in Canada

Aws And Tooling Platform Lead in Canada

545 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs across Canada. Filter by work mode, employment type, experience, department, date posted and distance.

SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $302.4K/yr

Quick readStrong listing-quality and freshness signals

Director of Engineering, Physical AI Role Overview The Director of Engineering will report to the General Manager of Physical AI, and will be responsible for leading a multi-disciplinary engineering organization. In this senior leadership role, you will own the execution of the Physical AI Data Engine — the platform powering the next generation of Physical AI/Embodied AI. You will collaborate closely with Operations and GTM to guide product direction and help solve the data bottleneck that stands between today's robotics research and real-world deployment. This role requires significant ownership in a fast-paced environment and you will motivate internal teams to set the pace for business growth. Travel will come into play. Key Responsibilities: Set and drive the technical vision across data collection infrastructure, teleoperation systems, ML training pipelines, model evaluation frameworks, annotation tooling, and research Lead a multidisciplinary engineering organization—spanning engineering managers, software engineers, ML engineers, and ML research scientists—while designing the organizational structure, talent strategy, and culture required to scale rapidly without compromising on quality or strategic alignment Maintain exceptional technical and operational excellence by deeply understanding team deliverables, asking incisive questions, identifying slipping standards early, and knowing precisely when to step in Drive cross-functional alignment across Engineering, Operations, and GTM on platform architecture, release processes, and shared priorities Collaborate with researchers and clients to architect and deliver scalable, production-grade data infrastructure tailored for complex robotics workloads Required Qualifications: Bachelor's degree in Engineering, Robotics, Computer Science, or a related technical field 8+ years of engineering experience in fast-paced environments, including 4+ years direct people management demonstrated history of recruiting, mentorin

TypeScriptPythonAWSKubernetes
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $252K/yr

Quick readStrong listing-quality and freshness signals

About Scale AI At Scale, our mission is to develop reliable AI systems for the world's most important decisions. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. Scale Frontier Data is the organization behind the training and evaluation data that frontier labs depend on. We build the systems, tooling, and expert workflows that turn hard human expertise into signals that models can learn from, across reasoning, coding, agentic tool use, and domain expertise. Reinforcement learning environments are now the center of gravity for that work: the difference between a model that demos well and a model that reliably completes long-horizon work is almost always the quality of the environments and reward signals it was trained against. Responsibilities As a Staff Software Engineer, RL Environments, you'll own the technical foundation for how Scale builds, runs, verifies, and delivers RL environments at scale. An RL environment is a real piece of software: a containerized world with real dependencies, real state, real tools, and a grader that has to be correct even when the agent is creative about breaking it. Building one is a full-stack engineering problem. Building thousands of them reproducibly, cheaply, with trustworthy reward signals and throughput measured in millions of rollouts is a systems problem that very few people have solved. You'll work on both. You'll design the platform: sandboxed execution, environment packaging and versioning, rollout orchestration, trajectory capture, verifier frameworks, and the authoring surfaces that let engineers and domain experts produce environments without reinventing infrastructure each time. And you'll go deep on the environments themselves by instrumenting real applications, designing task suites that expose specific capability gaps, and building graders that

TypeScriptPythonReactAWS
J
📍 New York; Toronto, Canada· Full-time
✓ High-confidence listingCompany trend -100%

$167.5K – $211.5K/yr

Quick readStrong listing-quality and freshness signals

Who We Are At Justworks, you’ll enjoy a welcoming and casual environment, great benefits, wellness program offerings, company retreats, and the ability to interact with and learn from leaders in the startup community. We work hard and care about our most prized asset - our people. We’re helping businesses get off the ground by enabling them to focus on running their business. We solve HR issues. We’re data-driven and never stop iterating. If you’d like to work in a supportive, entrepreneurial environment, are interested in building something meaningful and having fun while doing it, we’d love to hear from you. We're united by shared goals and shared motivations at Justworks. These are best summed up in our company values, which are reflected in our product and in our team. Our Values If this sounds like you, you’ll fit right in. Department Platform Engineering Who You Are You are a tooling-focused engineer who obsesses over developer productivity. You’ve built and maintained CI/CD pipelines, developer tooling, and automation that makes engineering teams faster. You understand that the best developer tools disappear into the background - they just work. You’re equally comfortable debugging a flaky GitHub Actions workflow and designing a new internal CLI. You care deeply about reducing friction and cognitive load for your fellow engineers. This is a foundational role. You’ll be one of the first dedicated engineers on a platform organization supporting 300+ engineers. You’ll shape not just the technical foundations but the culture and practices of the team. Your Success Profile What You Will Work On Own and evolve CI/CD infrastructure (GitHub Actions, deployment pipelines) to improve build times and deployment frequency Build and maintain our developer portal (Backstage) - service catalog, golden path templates, and TechDocs integration Create golden path templates for Go, Rails, and Vue.js services that bake in best practices automatically Build and maintain internal

TypeScriptPythonVueAWS
O
📍 Toronto, Ontario, Canada· Full-time
✓ High-confidence listingCompany trend -63.6%

From C$146K/yr

Quick readStrong listing-quality and freshness signals

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity Reporting to the Director of Quality & Performance, this role as Engineering Manager of Performance & Resilience will drive performance and resiliency improvements for the Auth0 product at Okta. Here you'll be working with some of the most advanced technology in the space, helping to streamline and secure billions of access requests a year. In this role, you will work closely with architects, platform team members, and product engineers to build the testing infrastructure that keeps Auth0 performant at scale — including the frameworks, tooling, and realistic datasets that make that testing meaningful. The ideal candidate is passionate about software quality and architecture, a self-starter, intellectually curious, and brings deep experience with performance testing, load testing frameworks, dataset generation, performance analysis, monitoring tooling, and chaos engineering. What you’ll be doing Collaborate with architects, tech lead, product owners, security and operations engineers to implement best practices related to performance and resiliency Communicate and organize cross-team projects with high business

JavaScriptTypeScriptJavaNode.js
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $1.5M/yr

Quick readStrong listing-quality and freshness signals

About the Team At DoorDash, we’re reimagining how people connect with the things they need — whether it’s a meal, a grocery run, and anything in between. Our audiences — Consumers, Dashers, and Merchants — are at the heart of everything the Design org builds. Our content design team plays a big part in this, with each content designer shaping the experience and bringing our vision to life through clear, thoughtful language that makes our three-sided marketplace easier, faster, and more human. About the Role We’re looking for a full stack UX Design Engineer to be the technical backbone of our global Content Design org and leader who builds the platforms, tooling, and ML systems that make world-class, localized product content effortless across DoorDash, Wolt, and Deliveroo. You’ll work at the intersection of design and engineering to build content tooling and embed LLM-driven workflows into product experiences. You’ll accelerate content design AI tooling so that CD’s can focus on the highest leverage strategic work, while your tools enable product designers and cross functional partners to ship high-quality content faster across DoorDash, Wolt, and Deliveroo. You're excited about this opportunity because... Design and build internal content tools that: Help PMs and designers generate, manage, localize, and deploy product content at scale. Plug custom GPTs and other LLMs into everyday product workflows for content iteration and improvements. Own full-stack development of these tools: Build and maintain frontend experiences using React and modern JavaScript/TypeScript. Design and implement backend services and APIs (e.g., Kotlin/Java) to support content tooling, experimentation, and automation. Integrate tooling with experimentation and infra: Embed content tools into A/B testing platforms so teams can test and ship variants with minimal engineering dependency. Build pipelines from variant generation → experiment → automated deployment of winners. Operationalize

JavaScriptTypeScriptJavaReact
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $165.6K/yr

Quick readStrong listing-quality and freshness signals

Overview Scale’s Finance Systems and Automation team is looking for a builder-oriented team member to help design and develop integrations, automations, and AI agents that streamline workflows across Finance, Accounting, People, and Recruiting. In this role, you will work closely with stakeholders across Finance, Accounting, People Operations, and Recruiting to understand their workflows and build integrations, automations, and agentic workflows that reduce manual effort and accelerate execution across these teams. You will leverage our internal data infrastructure, system integration tooling, and emerging AI platforms to architect scalable solutions — from traditional system integrations to intelligent agent-driven workflows — that serve as the foundation for long-term operational efficiency. We’re looking for someone who thrives on connecting systems, automating repetitive processes, and pushing toward more autonomous, AI-assisted operations. You should be comfortable navigating ambiguity, designing solutions that scale, and rigorously validating outcomes end to end. What You’ll Do Design and build agent-driven workflows and automation systems across People Operations, Recruiting, Finance, and Accounting Identify opportunities to replace manual or rules-based processes with agentic workflows Partner with stakeholders to translate business processes into scalable, automated solutions Lead the implementation of end-to-end workflows, from requirements through deployment and validation Automate candidate-to-employee transitions (e.g., Greenhouse → HRIS → provisioning systems) Build workflows to manage employee lifecycle events such as onboarding, transfers, and offboarding Automate approval flows and data synchronization across People, Finance, and Recruiting systems Support accounting and finance workflows through scalable integrations and automation Design and implement the underlying integrations and data flows that enable reliable automation and agent behavior Est

AWSRestAIGo
T
📍 San Francisco, CA· Full-time
✓ High-confidence listingCompany trend -94.4%

From $151.2K/yr

Quick readStrong listing-quality and freshness signals

About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Role Platform Core Engineering (PCE) builds the foundational systems, frameworks, and tooling that Twitch's engineers build on every day. As a Senior Product Manager on PCE, your customers are developers, and your success depends on earning their trust. You will own a platform product area and be measured on whether developers actually adopt, rely on, and advocate for what you ship. This is a deeply technical, senior individual-contributor role. You are expected to be hands-on with modern AI development tools, to understand the technology you are shaping at a real level (not just at the surface), and to hold credibility in a room full of engineers. You can be based out of one of the Twitch offices including San Francisco, Seattle, Los Angeles, Irvine, or New York City. You Will: Own the strategy and roadmap for a PCE platform area, balancing near-term developer needs with a 1 to 2 year technical vision. Treat developers as your customers: understand their workflows, earn their trust, and prioritize the capabilities that make them faster and more effective. Work directly with Senior Engineering Managers and Principal Engineers to define technical and organizational needs and build a cohesive platform ecosystem. Define and drive the platform metrics that matter. Drive alignment across dependent

AWSRestAIGo
T
📍 Irvine, CA· Full-time
✓ High-confidence listingCompany trend -94.4%

From $151.2K/yr

Quick readStrong listing-quality and freshness signals

About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Role Platform Core Engineering (PCE) builds the foundational systems, frameworks, and tooling that Twitch's engineers build on every day. As a Senior Product Manager on PCE, your customers are developers, and your success depends on earning their trust. You will own a platform product area and be measured on whether developers actually adopt, rely on, and advocate for what you ship. This is a deeply technical, senior individual-contributor role. You are expected to be hands-on with modern AI development tools, to understand the technology you are shaping at a real level (not just at the surface), and to hold credibility in a room full of engineers. You can be based out of one of the Twitch offices including San Francisco, Seattle, Los Angeles, Irvine, or New York City. You Will: Own the strategy and roadmap for a PCE platform area, balancing near-term developer needs with a 1 to 2 year technical vision. Treat developers as your customers: understand their workflows, earn their trust, and prioritize the capabilities that make them faster and more effective. Work directly with Senior Engineering Managers and Principal Engineers to define technical and organizational needs and build a cohesive platform ecosystem. Define and drive the platform metrics that matter. Drive alignment across dependent

AWSRestAIGo
T
📍 San Francisco, Canada
✓ Quality checkedCompany trend -94.4%

About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Team Twitch Security Platform builds and operates the foundational software, data, and automation that enable security at scale across Twitch. As a Software Development Engineer II (SDE II) on the Security Platform team, you will design, build, and operate critical services, pipelines, and tooling that power Twitch's security, privacy, and compliance programs. About the Role Twitch Security Platform builds and operates the foundational software, data, and automation that enable security at scale across Twitch. As a Software Development Engineer II (SDE II) on the Security Platform team, you will design, build, and operate critical services, pipelines, and tooling that power Twitch's security, privacy, and compliance programs. In this role, you will work at the intersection of software, security, privacy, and data engineering, contributing production systems that handle large-scale security telemetry, automate security workflows, and provide reliable data and services to internal teams. You will partner closely with engineers and product teams to solve real-world security problems through well-designed software. You will own projects end-to-end from design and implementation through deployment and operational support for the systems that are business-critical and highly visible. The problems yo

PythonJavaAWSAI
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team The Code Quality team sits within the Developer Platform organization and owns the systems that keep DoorDash's codebase healthy and secure as it scales: static analysis, quality gates, test frameworks, regression infrastructure, and tooling. Our job is to make sure the signals engineers rely on before shipping — test results, coverage, performance feedback etc — are fast and trustworthy. The decisions we make about tooling and standards directly shape how confidently and quickly engineering teams at DoorDash can ship to production. About the Role We're looking for Software Engineers to help build and maintain the systems that validate code quality across DoorDash's engineering org, treating our tooling as a critical product for the engineers who rely on it every day: static analysis and quality gates, test frameworks and regression infrastructure. You’ll design the tooling and automation that will help derive trustworthy quality signals, integrate them into the development lifecycle, and make it easy for engineers to execute reliable, repeatable workflows. You will collaborate across the engineering org, partnering directly with the teams who use what you build to understand the accuracy, reliability and performance of their functionality. You will report into the Engineering Manager on our Code Quality team in our Developer Platform organization. You must be located in either San Francisco, CA, Sunnyvale, CA, Los Angeles, CA, Seattle, WA, or New York, NY. You're excited about this opportunity because you will… Build and maintain quality tooling — static analysis, quality gates, coverage reporting, test frameworks, regression infrastructure — and integrate it directly into our developer workflows and CI/CD pipelines Define and derive quality signals - flakiness, pass rate, coverage, performance, scale readiness etc - Build tooling that improves everyday engineering workflows, including local development, CI/CD, debugging, and rollou

AWSCI/CDGitRest
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $897.6K/yr

Quick readStrong listing-quality and freshness signals

About the Team As one of DoorDash's core operations teams, Customer Experience ensures that when issues arise across the platform, there is a reliable and effective support system in place. Our team designs, manages, and continuously improves DoorDash's global support network, with the goal of delivering a high-quality and consistent customer experience. This role sits within the Safety Customer Experience team, focused on the most critical and high-risk incidents on the platform. The team owns some of the most sensitive customer interactions at DoorDash, where thoughtful operational and product decisions directly improve customer trust and platform safety. About the Role You'll operate at the intersection of Product, Operations, and Customer Experience to improve how DoorDash prevents, identifies, and responds to the most critical safety incidents on the platform. You'll partner closely with Product, Policy, Analytics, and Operations to design scalable solutions that deliver accurate, timely support during customers' highest-stakes moments. This role requires a detail-oriented operator who can navigate complex problem spaces, design scalable processes, and execute with precision in a fast-paced environment. You’ll be expected to take ownership of ambiguous, high-stakes problems and translate them into structured, actionable solutions. You’re excited about this opportunity because you will… Drive Safety Strategy – Partner with cross-functional teams to identify and execute initiatives that improve how DoorDash prevents, identifies, and responds to safety incidents. Own End-to-End Experience – Design and optimize the full lifecycle of safety incident handling, including reporting, workflows, support execution, and tooling. Drive Data-Driven Decisions – Leverage data and case-level insights to identify root causes, measure performance, and prioritize opportunities to improve the safety customer experience. Influence Cross-Functionally – Collaborate with Product, Engin

AWSGitRestAI
P
📍 Toronto, ON, CA· Full-time
✓ High-confidence listingCompany trend -100%

From C$139.4K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . As a Software Engineer at Pinterest, you will contribute to initiatives that empower developers to build modern, world-class web experiences for Pinners with high productivity, performance, and quality. We're looking for someone who is passionate about front-end, has a growth mindset, and isn't afraid to voice their opinion on how we can improve. You'll find creative solutions to thought-provoking problems, and because we value the kind of courageous thinking required for big bets and smart risks to pay off, you'll help drive new initiatives from technical design through implementation and release. What you’ll do: Build the web platform that powers Pinterest.com for hundreds of millions of users, along with the frameworks and tooling that empower hundreds of engineers to build it. Improve foundational web architecture, DevEx , builds, CI/CD , an

JavaScriptTypeScriptJavaReact
R
📍 Bellevue, WA; Menlo Park, CA· Full-time
✓ Quality checkedCompany trend -76.2%

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Proactive Security team builds and operates systems that help engineers identify and resolve security risks earlier in the software development lifecycle. The team creates practical safeguards, including shared libraries, frameworks, and automated checks that make secure development easier to adopt. You’ll work closely with product and engineering teams to reduce risk at scale while supporting efficient product development. The team values clear communication, sound judgment, and security practices that are useful for engineers building production systems. This is a meaningful opportunity to help strengthen application security across Robinhood! As a Security Engineer on the Application Security Team, you will help design and implement systems that embed security into development practices across Robinhood. You will review system designs, contribute to threat models, and guide secure implementation for product and platform teams. You will also build and improve security solutions and guardrails that give engineers clear patterns for developing secure services. You will utilize frontier AI models and tooling to help define the future of security at Robinhood. This r

PythonVueAWSAI
H
📍 Montreal, Quebec, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

We take play seriously. We’re looking for curious adventurers ready to find their party, fueled by imagination and drive to build what’s never been built before. At Hasbro and Wizards of the Coast, you’ll collaborate with passionate teams to reimagine our iconic brands and create experiences that spark joy, connection, and community through the magic of play. This is your chance to shape legendary play that lasts a lifetime. At Wizards of the Coast, we connect people around the world through play and imagination. From our genre-defining games like Magic: The Gathering® and Dungeons & Dragons® to our growing multiverse, we continue to innovate and build new ways to foster friendship and connection. That’s where you come in! The cloud platform underpins how services are built, deployed, secured, and operated across the organization. As a Principal Cloud Infrastructure Engineer, you will define and drive the governance model and automation strategy for cloud infrastructure. This includes setting policy-as-code standards, automating compliance and cost controls, and ensuring infrastructure is provisioned and operated through consistent, auditable, self-service pathways rather than tailored or manual processes. You will operate at both a strategic and hands-on level and will set direction for how cloud resources are governed and automated while also building the tooling and guardrails that make that direction real. Success in this role means engineering teams can self-service the build and operation of infrastructure with confidence that it is secure, cost-aware, and consistent by default, without slowing delivery down. What you'll do Cloud Governance Define and enforce policy-as-code standards (tagging, naming, encryption, network segmentation, access boundaries) across cloud accounts/subscriptions Establish guardrails using tools such as OPA, Sentinel, AWS Config/Service Control Policies to report on and prevent drift from

AWSAzureGCPCI/CD
T-
📍 Toronto, Canada· Full-time
✓ High-confidence listing

From C$1.4M/yr

Quick readStrong listing-quality and freshness signals

About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu

PythonJavaAWSKubernetes
🔔

Get new aws and tooling platform lead jobs in Canada by email

Daily job updates · Unsubscribe anytime