About the Team: Tubi's Internal Tools team is at the forefront of AI integration, developing everything from developer resources to production-grade AI for business operations. We are the group responsible for turning AI from an experiment into an operating capability: training, infrastructure, developer agents, and AI-powered business systems. Engineers operate with high ownership and autonomy, collaborating on shared architectural decisions and AI infrastructure. What You'll Do: Own systems end to end — design them, build them, and support them in production. Lead the projects you own: sequence the work, decide what lands first, and set technical direction for the engineers working with you. Sit with the people who use what you build, and turn what you learn there into a system. Design the service boundaries, contracts and schema evolution that let our platforms grow without breaking the teams depending on them. Make our AI systems dependable in production: evaluation harnesses, human approval steps before an agent acts, retries that handle a model returning something unexpected, and cost tracking that tells you what a task costs before you run it. Build what other engineers build on — agent skills, tool and MCP integrations, shared libraries — and raise the bar through code review, design discussion and mentoring. Spot the platform work nobody has asked for yet, make the case for it, and build it. Your Background: 5+ years of professional experience building and operating production systems, from design through production ownership. A system you designed and can walk us through end to end — where its boundaries sit, what constrained it, and what you chose against. Strong programming proficiency in a statically typed language such as Rust, Go, C++, Java, Kotlin, C#, or TypeScript. Production Rust is a plus rather than a requirement. You have owned a service in production: you wrote the runbooks, you knew what it cost, and you were the one paged when it broke. Expe
Jobs in Canada
Operating Engineer in Toronto
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current operating engineer jobs in Toronto. Filter by work mode, employment type, experience, department, date posted and distance.
About the Role: Tubi is one of the largest free streaming platforms in the US, serving a large-scale streaming audience across Web, iOS, Android, Roku, Fire TV, Apple TV, and game consoles. Quality at this scale isn't a checkbox — it's a competitive advantage. We're looking for an Automation Engineering Manager to lead the team responsible for building and operating Tubi's multi-platform test automation infrastructure. You will own the strategy, tooling, and execution quality across our client surfaces — from video playback and ad delivery to content discovery and onboarding. This role is for a hands-on technical leader who can set direction, influence cross-functional roadmaps, and stay close enough to the code to guide architecture, review critical implementation decisions, and unblock complex technical issues. You will build a team that ships reliable automation at speed — and you will help the team move toward AI-native automation practices: fluent in AI tooling, proactive about applying it, and disciplined about using it responsibly. This is a hybrid role based out of either our San Francisco or Toronto office. You must be willing to travel to either location at least 2 days a week. What You'll Do: Test Strategy & Quality Planning Define and own Tubi's multi-platform automation strategy — covering Web, iOS, Android, CTV (Roku, Fire TV, Apple TV, Smart TVs, game consoles), and API layers. Establish testing standards, coverage targets, and quality gate policies across the CI/CD pipeline to protect release confidence and production reliability. Design specialized test strategies for business-critical scenarios: video playback (HLS/DASH), ad insertion, content recommendation surfaces, and user authentication flows. Use AI-assisted analysis (e.g., failure pattern clustering, test gap detection) to continuously improve test strategy based on real production signal and defect trends — not gut instinct. Automation Framework & Infrastructure Lead the
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's Governance and Compliance platform is the operating layer enterprises trust to run their security programs. We are hiring an engineer who will own the architectural foundation that makes it work for their most complex organizational structures. Program Structure and Trust is a newly chartered group at Vanta with a focused mission: build the enterprise org model that lets customers bring their compliance and security structure — product lines, business units, isolated data environments, cross-cutting audits, scoped approvals, and data residency requirements — natively into the platform. The problems this team solves determine whether Vanta can serve the enterprise customers it's increasingly winning. This is the defining technical role of the group. The Principal Engineer owns the design, phased delivery, and long-term technical direction of Vanta's enterprise org model — a multi-quarter initiative that cuts across the platform and establishes the foundation for how enterprise customers structure, segment, and operate inside Vanta. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as a Principal Engineer at Vanta: Own the design and multi-quarter delivery of Vanta's enterprise org model, including hierarchical product lines and business units, isolated data access and ownership, cross-cutting audit workflows, scoped approvals, and EU and GovCloud data residency support Define and evolve the core abstractions that let the platform absorb structurally diverse, often conflicting enterprise requirements — solving for the general case rather than one-off customer accommodations R
From C$108K/yr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are building and maintaining a highly scalable asynchronous platform that empowers our organization to handle critical business cases. As a software engineering team, our mission is to create robust and innovative solutions that drive the success of our business and deliver unparalleled value to our customers. We adopt Infrastructure as Code practice to automate the provisioning and configuration of our resources, which helps reduce manual configuration and improve consistency. Our team culture is built on collaboration, open communication, and a supportive environment where each member's ideas are valued and contributions are recognized. We believe in the importance of fostering a positive workplace culture that inspires innovation and creativity. Responsibilities: Maintain and analyze metrics from; operating systems; control planes; and applications to assist in fault detection and performance enhancement Design, develop and deploy tooling and systems that continually improve the reliability, scalability and efficiency of our platform Balance feature development speed and reliability with service-level objectives Operate and improve our Infrastructure using industry best practices and tools Participate in design and production readiness reviews, platform management and capacity planning ceremonies with cross-functional teams Document Infrastructure operations process and insights, identify repeatable actions and ruthlessly automate repetitive tasks Participate in our teams on-call rotations, respond to incidents and support other teams mitigate customer impacting events Experience: 5+ years experience working on teams responsible for software development, automation and systems engineering Experience building large-scale infrastructure, distributed systems or networks. Knowledge with SQS,
From C$108K/yr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. As a Frontend Software Engineer on the Operator Core Tooling pod, you'll play a vital role in building the robust services that power our critical operations tooling platform. Your work will directly empower our micromobility operations teams by providing them with intuitive and efficient tools, significantly improving their daily workflows as they manage our fleet. You'll collaborate closely with business leaders, front-end developers, and data scientists across Lyft to achieve this impact. Responsibilities: Help define the roadmap and architecture based on technology and business needs Write well-crafted, well-tested, readable, maintainable code Have a good grasp and ability to explain the various tradeoffs made in decisions Participate in code reviews to ensure code quality and distribute knowledge Lead projects from idea to positive execution Incorporate considerations for business context and failure modes in your work Proactively participate in resolving ongoing incidents Unblock, support, effectively communicate and obtain buy-in across teams to achieve results Share your knowledge by giving brown bags, tech talks, and evangelizing appropriate tech and engineering best practices See the direct impact of your work on the efficiency of our operating teams Experience: 3+ years of software engineering industry Advanced knowledge of JavaScript Experience working with modern JavaScript frameworks, like React Experience working with NodeJS and Express applications Experience working with design systems (e.g. Bootstrap, Salesforce Lightning, GitHub Primer) Good understanding of web performance and how browsers and DOM work Experience with unit, integration, and end-to-end testing Experience designing, building and improving a set of team owned components Culture of investigating and solvin
From C$1.4M/yr
About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu
From C$136K/yr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Our Infrastructure team is passionate about building software to solve problems at massive scale. We do this often, and when we believe our solution is worth sharing with the community, such as Envoy Proxy , we open source our ideas for the benefit of others. As a Infrastructure Engineer at Lyft, you will run our Production Infrastructure by monitoring system availability and take a holistic view of our platform health. You will build software and platforms to automate infrastructure platform operations and management. By measuring and monitoring our operations you will seek opportunities to optimize our systems in order to push our platform forward, anticipating our customers' needs in order to continually improve the platform. You will provide Lyft partner teams with operational support to help them build robust large scale distributed systems. About the Team Data Pipelines is at the heart of all critical data flowing through Lyft supporting hundreds of services that impact millions of drivers and passengers every day. Our team’s mission is to empower Lyft engineers to self-serve in building and maintaining data pipelines as needed to support products that deliver the world’s best transportation experience. We leverage a variety of technologies to store, stream and manage data making it available to our internal customers. Responsibilities: Maintain and analyze metrics from; operating systems; control planes; and applications to assist in fault detection and performance enhancement Design, develop and deploy tooling and systems that continually improve the reliability, scalability and efficiency of our platform Balance feature development speed and reliability with service-level objectives Operate and improve our Infrastructure using industry best practices and tools Participate in design and
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by building high-performance, scalable and reliable machine learning systems? Do you want to help define and build the next generation of AI platforms powering advanced NLP applications? We are looking for a Site Reliability Engineer to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will work closely with many teams to deploy optimized NLP models to production in low latency, high throughput, and high availability environments. You will also get the opportunity to interface with customers and create customized deployments to meet their specific needs. As a Site Reliability Engineer you will: Build self-service systems that automate managing, deploying and operating services. This includes our custom Kubernetes operators that support language model deployments. Automate environment observability and resilience. Enable all developers to troubleshoot and resolve problems. Take steps required to ensure we hit defined SLOs, including pa
From C$172K/yr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. As the Engineering Manager for the Lakehouse Foundation team, you will lead a group of engineers responsible for the foundational data layer that all of Lyft's data systems and emerging AI workloads are built on. The team owns catalog and metadata management, table formats and storage, and the access patterns and gateways through which other engineering teams interact with Lyft's data. As Lyft converges on a unified lakehouse architecture, this team builds and operates the single source of truth that powers analytics, machine learning, experimentation, and every business decision made from data. You will play a key role in shaping the team's technical direction, partnering with peer Data Platform teams on a multi-year platform evolution, and developing engineers who operate with autonomy on systems of significant scale and complexity. Lyft's Infrastructure teams build the foundational systems that the rest of engineering depends on to move fast, ship reliably, and scale efficiently. These are high-leverage roles where the work you and your team do has a multiplicative effect across the company. We're looking for experienced leaders who can balance the discipline of operating critical infrastructure with the curiosity to keep evolving how Lyft builds. Engineering at Lyft is a place where managers and engineers operate with high ownership and strong technical judgment. Our engineers expect their managers to be honest, available, and focused on the work that matters: developing their teams, removing obstacles, and giving people the support they need to do their best work. We build teams that are inclusive, technically rigorous, and have a strong sense of ownership for what they build. Responsibilities: Lead a team responsible for Lyft's foundational data layer, including catalog and metadata management,
About the Role: We're seeking a Director of Product Support Operations to lead the team that runs how Tubi's Tech Org plans, executes, and ships. PSO is a deliberately lean, AI-first team of Technical Program Managers and Customer Experience specialists built around three durable jobs: orchestrating agentic systems, people, and the cross-team programs that land Tubi's most complex initiatives for over 100 million monthly active users. You will own the Tech Org's most complex, company-level programs, spanning live events, platform integrations, monetization, and financial operations, along with the agentic systems that carry the org's operational load. You will also own the customer experience function end to end and run the Tech Org's operating cadences, including planning cycles, executive business reviews, and org-wide communications. You'll work closely with engineering, product, design, data science, and finance leaders as a trusted executive partner. This team consists of builders, not just coordinators. If you're passionate about building with AI rather than coordinating around it, landing complex programs at company scale, and developing world-class talent while staying hands-on yourself, this is the role for you. This is a hybrid role based out of our San Francisco, New York, Los Angeles, or Toronto offices. What You'll Do: Define the multi-year strategy and 12-month roadmap for Product Support Operations, aligning the portfolio with company objectives Lead, develop, and recruit high-performing Technical Program Management and Customer Experience team members, setting a high bar through direct feedback and rigorous performance management Operate as a player-coach: personally design, ship, and maintain agentic workflows and lead high-stakes programs while modeling the builder-TPM standard for the function Own execution of the Tech Org's most complex, company-level programs spanning multiple organizations and external partners, with impact validated by metrics
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About the Role Cohere builds frontier AI for the enterprise. We're hiring a Senior Director of Integrated Marketing to bring our demand engine, product marketing, and field/partner motion under one strategic roof. This is the senior marketing leader most accountable for pipeline. You're the person who turns Cohere's product, brand, and category narrative into qualified enterprise opportunities for our sales team. Sub-functions report to you: Product Marketing, Web Marketing, Growth, and Events & Partnerships. You'll work with the Chief of Staff on planning, budget, and operating cadence, and with the Marketing Operations Director and her analytics team on attribution, forecasting, martech, and the dashboards your pipeline number lives in. The demand engine only runs if you and the analytics group work as one team on measurement, and building that partnership is part of the job from week one. This role is built for someone who has done it before at a high-growth B2B SaaS or infrastructure company, and is ready to do it again at the frontier of AI. What You'll Run Integrated marketing strategy Build and run the integrated mark
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role As the founding Channel Partner Manager for Canada, you will build the partner motion that helps Canadian businesses discover, evaluate, and adopt Ramp through trusted advisors and ecosystem partners. This is a zero-to-one role. You will own the full partner lifecycle, from identifying and engaging Canadian accounting, advisory, private equity, and venture capital partners through signing, enablement, activation, and partner-sourced pipeline. Your early results will help determine how Ramp scales its Canadian channel motion. You will work closely with Sales, Growth, Product, Engineering, and the broader Partnerships organization. The right person is a hands-on builder: equally comfortable opening a new relationship, running a partner meeting, solving an operational blocker, and turning early learnings into a repeatable playbook. What You’ll Do Build Ramp’s Canadian channel strategy and operating rhythm from the ground up. Research, identify, source, qualify and sign net-new Canadian partners. Create joint go-to-market and activation p
From C$1.5K/yr
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world's biggest financial problems. We're looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn't a place for complacency, it's where ambitious people do the best work of their careers. We're a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Software Platform team accelerates developer velocity and increases system reliability by building the foundational platforms and tools that power Robinhood engineering. Within this group, the Kubernetes Compute team focuses on building and operating a highly available, scalable Kubernetes-powered container platform. We ensure that our infrastructure seamlessly supports reliable application deployments, integrates core platform capabilities, and enables multi-region scalability. We are expanding our core container systems to support our next phase of technical growth! As a Senior Software Develope r, you will focus heavily on building, operating, and expanding our container provisioning platforms. You will be responsible for designing resilient container infrastructure and contributing to our technical migration to Amazon EKS to improve platform reliability. In this position, you will collaborate with engineering teams across Robinhood to deliver reliable platform integrations for core capabilities like networking and security. Your work will directly help our infrastructure scale efficiently while maintaining a high standard of safety and system uptime. This role is bas
From C$1.5K/yr
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world's biggest financial problems. We're looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn't a place for complacency, it's where ambitious people do the best work of their careers. We're a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Software Platform team accelerates developer velocity and increases system reliability by building the foundational platforms and tools that power Robinhood engineering. Within this group, the Kubernetes Compute team focuses on building and operating a highly available, scalable Kubernetes-powered container platform. We ensure that our infrastructure seamlessly supports reliable application deployments, integrates core platform capabilities, and enables multi-region scalability. We are expanding our core container systems to support our next phase of technical growth! As a Software Developer, you will focus on building, maintaining, and scaling our container provisioning platforms. Working alongside senior engineers, you will write code to improve our infrastructure capabilities and actively participate in our technical transition to Amazon EKS. In this role, you will collaborate with teams across the organization to ensure robust platform integrations for everyday application needs like security and networking. Your efforts will directly improve system visibility, automation, and reliability across the platform. This role is based in our Toronto office(s), with in-perso
Overview: Qsight is a high-growth division of Guidepoint focused on building data intelligence solutions for the healthcare sector. Qsight leverages proprietary datasets and rigorous analysis of alternative data sources to generate actionable insights for top-tier institutional investors, medical device manufacturers, and pharmaceutical companies. The Qsight team develops market intelligence products designed to be highly relevant, accurate, and scalable – delivering superior insights to a diverse, global client base. We are seeking an experienced, motivated Tehnical Operations Engineer to join our growing team. This is a multiple-hats role focused on SaaS/platform operations and tier-2 support for client-facing systems. You will own the administration and reliability of key tools, troubleshoot and resolve escalations with clear documentation, and build lightweight automation and reporting to reduce manual work as we scale. You will partner closely with Customer Success, Product, and Engineering to proactively monitor, support, and improve critical systems. Through practical, creative problem-solving, you will strengthen reliability, accelerate time to resolution, and increase operational visibility. Day to day, you will triage and resolve client technical questions, manage vendor license administration and renewals, and produce reporting that informs operational decisions. This role is a launchpad toward an SRE/Platform Engineering track as you grow into deeper automation, reliability engineering, and systems design work. This is a hybrid position based out of our Toronto office. What You’ll Do: Platform Support Own routine ops and configuration changes for critical SaaS platforms – Including Tableau, Freshdesk, Datadog, and our own client facing and internal portals Configure and maintain Freshdesk portals, routing, SLAs, permissions, integrations, etc. based on business requirements. Automate manual operations with Python, PowerAutomate, and shell scri
Other cities to consider
More places hiring for this role
Get new operating engineer jobs in Toronto, Canada by email
Daily job updates · Unsubscribe anytime