About the Team The Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company at significant scale and continue to expand the workloads, capabilities, and consumer base we serve. Orchestrating and operating thousands of Spark cluster deployments is a complex distributed system problem which the team invests heavily in runtime optimization, systems architecture, multi-tenant scheduling, and end-user tooling. About the Role As a Software Engineer on Spark Platform, you will execute across the surfaces of our in-house Spark deployment that serves the entire company. The work spans Spark runtime upgrades and performance, multi-tenant scheduling and executor bin-packing on Kubernetes, cluster lifecycle automation, and the observability and incident automation that keep the platform sustainable. You will move between layers as the work demands — picking up the next high-leverage problem regardless of where it sits — and partner closely with the rest of the team and with platform consumers across the company. You must be located in San Francisco, Sunnyvale, Seattle, or New York City for this hybrid position. You will report into the Engineering Manager on our Spark Platform team. You're excited about this opportunity because you will… Build and operate an in-house Spark platform that runs at company-wide scale, spanning runtime, scheduler, reliability, and user-facing tooling. Drive multi-tenant scheduling, executor bin-packing, and cost-aware placement that let a small team serve dozens of consumer teams. Own pieces of cluster lifecycle automation — provisioning, upgrades, capacity changes, and node-failure handling — at a scale where these stop being manual events. Build the observability and incident automation that make the platform debuggable end-to-end and keep on-call sus
Jobs in Canada
Deployment Strategist Lead in Canada
141 active opportunities · Updated October 2026
Showing
15 jobs
Explore current deployment strategist lead jobs across Canada. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The Merchant (Mx) AI/ML team is a cornerstone of DoorDash’s merchant organization. We empower our restaurant partners to thrive on DoorDash by building intelligent, scalable AI systems that simplify operations, elevate their digital presence, and enhance customer engagement. Our world-class machine learning engineers develop production-grade AI solutions that power the entire merchant lifecycle — from onboarding and store setup to menu management, growth, and real-time order operations. The team is deeply customer-obsessed and impact-driven, focused on turning cutting-edge research in LLMs, multimodal learning, generative AI, and agentic automation into products that make our merchants more successful every day. About the Role We’re looking for an Engineering Manager to lead the Mx AI/ML team, driving the design and deployment of next-generation AI product solutions that power our merchant experiences. This is a highly cross-functional leadership role — you’ll collaborate with product, design, operations, and data science teams to define the AI roadmap, guide technical direction, and deliver end-to-end AI/ML solutions at massive scale. You’ll lead a team of talented ML engineers who are building real AI products delivered to merchants’ fingertips, helping them become more successful on DoorDash. You’re excited about this opportunity because you will… Lead and grow a team of exceptional AI/ML engineers developing production-grade machine learning and generative AI solutions for the merchant ecosystem. Define a multi-year strategy and execute the AI roadmap across key Mx pillars — Onboarding, Menu Media Understanding & Generation, Menu Metadata Intelligence, and Agentic Task Automation. Partner with product and operations to translate merchant pain points into scalable AI-powered solutions. Own the end-to-end lifecycle of ML systems — from ideation and experimentation to productionization and continuous improvement. Drive innovation in multimodal AI
About the Team The Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company at significant scale and continue to expand the workloads, capabilities, and consumer base we serve. Orchestrating and operating thousands of Spark cluster deployments is a complex distributed system problem which the team invests heavily in runtime optimization, systems architecture, multi-tenant scheduling, and end-user tooling. About the Role As a Senior Software Engineer on Spark Platform, you will set the technical direction for our in-house Spark deployment and shape the architecture that will run DoorDash's data, analytics, and ML compute for the next five years and beyond. You will own the deep, cross-cutting problems that span the runtime, the shuffle service, the scheduler, and the overall service reliability — making the architectural calls that compound across the platform's lifetime. You will partner with the Engineering Manager on technical roadmap, hiring, and team shape, and act as the senior technical voice in cross-team partnerships with Data Engineering, ML Platform, and product engineering teams that depend on the platform. You must be located in San Francisco, Sunnyvale, Seattle, or New York City for this hybrid position. You will report into the Engineering Manager on our Spark Platform team. You're excited about this opportunity because you will… Set the multi-year technical direction for an in-house Spark-on-Kubernetes platform — runtime, shuffle, scheduler, reliability — and make the architectural calls that compound for years. Own the deepest distributed-systems problems on the team: shuffle architecture, multi-tenant scheduling, runtime performance, and the failure modes that only show up at scale. Partner with the Engineering Manager on technical roadmap, hiring, inte
C$34 – C$36/hr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Interns work side-by-side with top engineers in the industry while having autonomy from the get-go. They contribute to user-facing products and are able to see their work go live quickly. Lyft fosters a collaborative environment in the office, so there's always a sharp mind eager to hear about your next idea. So what's yours? Responsibilities: Own your project, while checking in with other team members throughout the day with questions and updates You leave the code in a better state than when you found it (progressive refactor) You value reliability, ensured by testing (unit, integration and load tests) Participate in code reviews to ensure code quality and distribute knowledge Continuous integration and deployment Go home knowing that your work today is meaningfully improving the lives of every Lyft driver and every Lyft passenger! Experience: Currently pursuing a Bachelor's or Master's degree in Computer Science from a university in Canada (required) , with a graduation date between December 2027 and Summer 2028 (required). For any candidates who are master's students who worked between their bachelor's and master's programs: candidates should also have less than 2 years of relevant full-time work experience Available during Summer 2027 for the internship in Montreal Strong knowledge of CS fundamentals Excellent communication skills Passion for community, sustainability, and/or transportation Ability to thrive in a startup environment Contributions to open source projects Experience working with databases Experience solving real-time technology problems Experience with mobile development Must be fluent in spoken and written English and have a working proficiency in French Benefits: Mental health benefits In addition to holidays, interns receive 2 days paid time off and 3 days sick time off Subsidi
From C$40/hr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Interns work side-by-side with top engineers in the industry while having autonomy from the get-go. They contribute to user-facing products and are able to see their work go live quickly. Lyft fosters a collaborative environment in the office, so there's always a sharp mind eager to hear about your next idea. So what's yours? Responsibilities: Own your project, while checking in with other team members throughout the day with questions and updates You leave the code in a better state than when you found it (progressive refactor) You value reliability, ensured by testing (unit, integration and load tests) Participate in code reviews to ensure code quality and distribute knowledge Continuous integration and deployment Go home knowing that your work today is meaningfully improving the lives of every Lyft driver and every Lyft passenger! Experience: Currently pursuing a Bachelor's or Master's degree in Computer Science from a university in Canada (required) , with a graduation date between December 2027 and Summer 2028 (required). For any candidates who are master's students who worked between their bachelor's and master's programs: candidates should also have less than 2 years of relevant full-time work experience Available during Summer 2027 for the internship in Toronto Strong knowledge of CS fundamentals Excellent communication skills Passion for community, sustainability, and/or transportation Ability to thrive in a startup environment Experience with real-time technology problems Contributions to open source projects Experience working with databases Experience solving real-time technology problems Experience with mobile development Benefits: Mental health benefits In addition to holidays, interns receive 2 days paid time off and 3 days sick time off Subsidized commuter benefits and Lyft ride credi
From C$40/hr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Interns work side-by-side with top engineers in the industry while having autonomy from the get-go. They contribute to user-facing products and are able to see their work go live quickly. Lyft fosters a collaborative environment in the office, so there's always a sharp mind eager to hear about your next idea. So what's yours? Responsibilities: Own your project, while checking in with other team members throughout the day with questions and updates You leave the code in a better state than when you found it (progressive refactor) You value reliability, ensured by testing Participate in code reviews to ensure code quality and distribute knowledge Continuous integration and deployment Go home knowing that your work today is meaningfully improving the lives of every Lyft driver and every Lyft passenger! Experience: Currently pursuing a Bachelor's or Master's degree in Computer Science, Data Science or related major from a university in Canada (required) , with a graduation date between December 2027 and Summer 2028 (required) . For any candidates who are master's students who worked between their bachelor's and master's programs: candidates should also have less than 2 years of relevant full-time work experience Available during Summer 2027 for an internship in Toronto Strong knowledge of CS fundamentals Knowledge of SQL and data modeling fundamentals Experience working with databases Excellent communication skills Interest in solving large scale data problems in a real world scenario Passion for community, sustainability, and/or transportation Benefits: Mental health benefits In addition to holidays, interns receive 2 days paid time off and 3 days sick time off Subsidized commuter benefits and Lyft ride credits Lyft is committed to creating an inclusive workforce that fosters belonging. Lyft belie
From C$40/hr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Interns work side-by-side with top engineers in the industry while having autonomy from the get-go. They contribute to user-facing products and are able to see their work go live quickly. Lyft fosters a collaborative environment in the office, so there's always a sharp mind eager to hear about your next idea. So what's yours? Responsibilities: Own your project, while checking in with other team members throughout the day with questions and updates You leave the code in a better state than when you found it (progressive refactor) You value reliability, ensured by testing (unit, integration and load tests) Participate in code reviews to ensure code quality and distribute knowledge Continuous integration and deployment Go home knowing that your work today is meaningfully improving the lives of every Lyft driver and every Lyft passenger! Experience: Currently pursuing a Bachelor's or Master's degree in Computer Science from a university in Canada (required) , with a graduation date between December 2027 and Summer 2028 (required). For any candidates who are master's students who worked between their bachelor's and master's programs: candidates should also have less than 2 years of relevant full-time work experience Available during Summer 2027 for an internship in Toronto Strong knowledge of CS fundamentals Knowledge of Python, JavaScript, CSS, and HTML Experience working with leading JavaScript frameworks, like React Experience with modern frontend testing tools, such as Webpack, Babel, Jest, Jasmine, Protractor, and WebDriver Understanding of how browsers and DOM work Experience with Git or other distributed version control systems Experience with browser developer tools Experience with the Unix command line interface Solid understanding of web performance Experience with TypeScript Experience with CSS
From C$142K/yr
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. COMPANY DESCRIPTION: Okta is the foundation for secure connections between people and technology. By harnessing the power of the cloud, Okta allows people to access applications on any device at any time, while allowing for strong security protection. The Okta service integrates directly with an organization’s existing directories and identity systems, as well as 6,500+ applications. Because Okta runs on an integrated platform, organizations can implement Okta quickly at large scale and low total cost. Thousands of customers, including Adobe, Allergan, Chiquita, LinkedIn, and Western Union, trust Okta to help their organizations work faster, boost revenue, and stay secure. To learn more about Okta, visit https://www.okta.com . JOB PURPOSE: In this role, you will be tasked with providing hands-on implementation and deployment services to our customers. We are looking for an experienced, enthusiastic and hands-on leader who can rapidly learn the Okta platform, our technology and the value proposition that we bring to customers of all sizes. You will lead multiple concurrent customer deployment projects with the number one goal to ensure long-term customer success. Additionally you will engage with the full Customer Success team in order to assure a smooth transition post-deployment to the support/maintenance phases. Finally, you will also be responsible for the continuous improvement of delivery processes and m
From C$176K/yr
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. COMPANY DESCRIPTION: Okta is the foundation for secure connections between people and technology. By harnessing the power of the cloud, Okta allows people to access applications on any device at any time, while allowing for strong security protection. The Okta service integrates directly with an organization’s existing directories and identity systems, as well as 6,500+ applications. Because Okta runs on an integrated platform, organizations can implement Okta quickly at large scale and low total cost. Thousands of customers, including Adobe, Allergan, Chiquita, LinkedIn, and Western Union, trust Okta to help their organizations work faster, boost revenue, and stay secure. To learn more about Okta, visit https://www.okta.com . JOB PURPOSE: In this role, you will be tasked with providing hands-on implementation and deployment services to our customers. We are looking for an experienced, enthusiastic and hands-on leader who can rapidly learn the Okta platform, our technology and the value proposition that we bring to customers of all sizes. You will lead multiple concurrent customer deployment projects with the number one goal to ensure long-term customer success. Additionally you will engage with the full Customer Success team in order to assure a smooth transition post-deployment to the support/maintenance phases. Finally, you will also be responsible for the continuous improvement of delivery processes and m
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role We are a software engineering team with expertise in enabling ML models in production. We deploy AI models to run in variety of environments: air-gapped government networks, forward-deployed defense environments, edge nodes, and enterprises with strict data sovereignty requirements. Our customers rely on us for frontier AI capabilities running on hardware they control, often with constrained GPU resources and limited direct access. Rising to that challenge and meeting those expectations is what Palantir's excels at. We treat models like any other software: continuously tested, continually delivered, packaged for reproducible deployment, and built for long-term maintainability. You will own services end-to-end, and work across the full stack, from inference engines, GPU scheduling to deployment pipelines, observability, and integration with Palantir's platform. The goal is to deliver new models and capabilities quickly and continuously. Join us if you want to solve problems at the intersection of infrastructure and machine learning that directly enable critical customers.
From C$46.1K/yr
Job Title: Co-Op, IT Operations - Summer 2026 Location : Vancouver, BC (on-site) Term : 4 months Start Date: September 14th, 2026 Openings : 1 Overview : We’re looking for a Co-op on the team to help us provide IT support to Hootsuite owls and solve both small as well as large technology problems. Under close supervision, you’ll support the Helpdesk ticket queue, deploy laptops plus hardware peripherals, AV challenges and work closely with the Global IT team to execute IT projects. This position is based in Vancouver, Canada. WHAT YOU’LL DO: Support the build and deployment of laptops for new users as well as end users as needed Support closing of our Macbook leases by configuring, wiping, recording laptop data as required Set up, install, modify, configure, maintain desk hardware in the office Receive offboarded hardware and update workflows as well as management system based on that Input and update assets in the hardware asset management system as governed by Helpdesk’s processes and procedures Work closely with Hootsuite’s global Helpdesk team to identify and resolve issues by following processes Document procedures and solutions in our ticketing system and follow up with clients to ensure full resolution of issues Set up, install, and troubleshoot Google Meet Hardware in the office Assist with setting up loaner laptops to AV gear for interviews or smaller company events Keep team documentation updated as processes and stakeholder need change Assist in departmental projects that require IT time and resources Creating tickets and moving tasks through agile workflow as work is completed WHAT YOU’LL NEED: Experience troubleshooting basic to advanced Mac problems Experience troubleshooting Windows OS problems and working knowledge of Dell laptops Experience working with meeting room hardware, AV gear to support meetings Experience working in agile environment a bonus Proven ability to translate technical issues into understandable language for end-
$212K – $318K/yr
Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Design state-of-the-art ML models and large-scale ML systems for underwriting and portfolio management for Stripe Capital based on ML principles, domain knowledge, risk, regulatory and engineering constraints. Design systems to speed up the time from idea to deployment of new models. Experiment and iterate on ML models (using tools including PyTorch and TensorFlow) to achieve key business goals and drive efficiency. Develop pipelines and automated processes to train and evaluate models in offline and online environments. Integrate ML models into production systems and ensure their scalability and reliability. Collaborate with product and strategy partners to propose, prioritize, and implement new product features. Engage with the latest developments in ML/AI and take calculated risks in transforming innovative ML ideas into productionized solutions. Who you are Minimum requirements Must have a Bachelor's degree or foreign equivalent in Computer Science, Machine Learning, Mathematics, Physics, Statistics, or a related field, plus two (2) years of experience in Building and shipping ML systems in production. Must have two (2) years of experience in each of the following: ML algorithms and model architectures; Designing, training and evaluating machine learning models; Productionizing and deploying machine learning models at scale; Orchestrating data pipelines and leveraging large-s
$149.9K – $270K/yr
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role As a Senior Platform Engineer at HP IQ, you will help build and evolve the infrastructure, tooling, and shared platform capabilities that enable our engineering teams to develop and operate reliable, secure, and scalable services across cloud and edge environments . You will work closely with application, services, AI/ML, and security teams to improve developer velocity, production readiness, reliability, and operational efficiency across a heterogeneous infrastructure footprint. What You Might Do Design, build, and maintain shared infrastructure and platform capabilities across cloud and edge environments. Build automation and self-service tooling that improves engineering velocity and operational consistency. Develop and maintain Infrastructure-as-Code, deployment workflows, and environment provisioning. Partner with engineering teams on production readiness, including reliability, security, observability, scalability, and recovery. Improve monitoring, alerting, incident response, and operational tooling across distributed environments. Automate repetitive operational t
About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by millions of people, then we want to talk to you! About the Role We are hiring an Autonomy Platform Engineer to build and evolve the foundational software that runs our autonomy stack across robot and compute platforms. The Autonomy Platform team works across embedded Linux, compute and sensor enablement, robotics middleware, process orchestration, data capture and replay, system observability, and performance. You will develop production software and tooling that enables autonomy engineers to bring up new hardware, deploy services reliably, diagnose failures, and validate system performance across robot generations. You will work closely with autonomy, firmware, electrical, hardware, manufacturing, and validation engineers and report to the Autonomy Platform Lead. We expect this role to be hybrid with some time in-office and some time remote. You’re excited about this opportunity because you will… Build and maintain core runtime, middleware, and platform services used by autonomy applications. Enable new compute, camera, lidar, and other sensor platforms. Improve process orchestration, messaging, configuration, startup and shutdown behavior, resource isolation, and fault recovery. Develop system observability, tracing, performance measurement, diagnostics, and regression-detection capabilities. Build reliable data capture, replay, and debugging workflows. Create provisioning, packaging, deployment, integration-test, and platform-readiness tooling. Lead complex debugging across application, middleware, OS, driver, networking, timing, and hardware boundaries. We’re excited about you because… Strong production C++ and Python experience. Experience with embedded Linux, robotics, autonomous vehicles, or complex mechatronic systems. Solid u
From C$1.4M/yr
About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu
Other cities to consider
More places hiring for this role
Get new deployment strategist lead jobs in Canada by email
Daily job updates · Unsubscribe anytime