At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. As a Senior Software Engineer, Data on the Mapping team, you will collaborate with our world-class team of engineers, product managers, and scientists to grow and improve the quality of recommended routes and accuracy of our travel time estimations. You will lead the architecture and long-term technical direction of our offline experimentation tooling and route simulation services — the systems that let Lyft test routing changes safely before they reach production. You'll also build scalable data pipelines for experimentation, analytics, and machine learning models, along with the data governance and observability systems that keep them trustworthy. Your work will enable integration with partner teams and allow stakeholders across Engineering, Data Science, and Product to make data-informed decisions that directly impact Lyft’s growth and profitability. Our technology stack is based on the latest technologies such as AWS, Databricks, Kubernetes and Airflow. You will work with incredibly passionate and talented colleagues from software engineering, machine learning and data science on projects that directly impact millions of riders and drivers. Responsibilities Own core data pipelines end-to-end, building deep subject matter expertise in the systems you manage and defining/managing SLAs for pipelines, services, and datasets to ensure reliability at scale Serve as the technical owner and architectural lead for our offline experimentation platform and route simulation services, setting technical direction, evaluating trade-offs, and ensuring the systems scale with Lyft's routing and mapping ambitions Continuously evolve data models and schemas to meet business and engineering requirements Develop AI tools that support self-service management of data pipelines (ETL) and schema evolution, and perform han
Jobs in Canada
Senior Observability Engineer in Toronto
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current senior observability engineer jobs in Toronto. Filter by work mode, employment type, experience, department, date posted and distance.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. AS A SENIOR SOFTWARE ENGINEER YOU WILL: Drive high-impact initiatives that span our product areas and full tech stack, including golang and Python on the backend and TypeScript/React on the frontend. Own and deliver features across the notebook service, container runtimes, and UI — designing and shipping medium-to-large projects independently, from ambiguous problem statements through production and post-launch. Advance core platform initiatives such as runtime management and patching, environment reproducibility and replication, security and compliance, and observability for notebooks. Extend the product to operate reliably in regulated and air-gapped environments, where security, compliance, and operational rigor are paramount. Promote strong collaboration within a cross-functional team and partner closely with embedded product managers and designers, as well as platform organizations across Snowflake Be a strong contributor to the product vision and drive team planning. Build for scale, reliability, and high performance, and participate in the on-call rotation to keep a Tier-1 production service healthy. Mentor, coach, and empower more junior team members, and raise the engineering bar through high-quality design and code review. OUR IDEAL CANDIDATE WILL HAVE: 7+ years o
From C$1.4M/yr
About the Role: Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission is to engineer resilience from the ground up, enabling our product teams to innovate rapidly while ensuring our users have a stellar experience. We own the availability, latency, performance, and capacity of our platform, and we achieve our goals through a culture of data-driven decision-making, blameless learning, and relentless automation. As a Senior Site Reliability Engineer, you are a hands-on engineer who blends deep software development expertise with a passion for operational excellence. You will be responsible for designing, building, and running the resilient, scalable, and increasingly self-healing systems that power our products. You will apply sound engineering principles to solve our most complex reliability challenges, with a mandate to automate everything, eliminate toil, and write robust, maintainable code. You will be a force multiplier, mentoring other engineers and elevating the site reliability bar for the entire organization. This is a hybrid role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: System Architecture & Design: Design, build, and maintain scalable, highly available, and fault-tolerant distributed systems. Partner with development teams as a reliability consultant, reviewing designs and influencing architectural decisions to ensure new services are built with reliability, observability, and performance as core principles, not afterthoughts. Automation & Software Development: Write robust, performant, and maintainable code to automate operational tasks, and CI/CD pipelines. Build the internal tools, libraries, and frameworks that enable engineering teams to self-service their
From C$190K/yr
About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! About the Team Engineers on this team build our rules-based calculation engine for processing sales commissions. This might sound simple if you have never been exposed to sales compensation plans, it is not. We are low on meetings and high on accountability. Most of the team is in the EST time zone, with a few located in PST and Central as well. We are still evolving many areas of the platform, which means there is meaningful room to improve the design, reliability, and scalability of the systems we build. What you’ll be doing Reporting to the Manager of Data Platform, you will play an important role in the evolution of our Spark-based data platform. You’ll design and build data-rich platform capabilities, contribute to system design discussions, and help ensure our data systems remain reliable, maintainable, and scalable as Forma grows. As a Senior Engineer, Data Platform, you are expected to operate with strong ownership and sound technical judgment. This includes identifying risks in the work you own, surfacing edge cases, asking thoughtful questions, and proposing improvements that strengthen the quality and reliability of the platform. You will: Design, build, and improve Spark-based data pipelines and platform services. Work with complex data models representing sales compensation plans, hierarchies, relationships, and enterprise datasets. Build reliable, deterministic data systems that customers and internal teams can trust. Improve testing, observability, data quality,
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Our Senior Software Engineers lead and mentor engineers, delivering high-value products for our customers and infrastructure that enables our business to scale. Vanta’s team and technology surface are growing quickly, and it’s essential that we invest in the right abstractions and systems to enable us to scale with our business. As a Senior Software Engineer, you’ll be responsible for setting technical direction to enable our product and infrastructure to scale with our business, driving complex projects across our technical stack, and mentoring our talented engineering team. Your past experience will be leveraged to enable and accelerate Vanta’s growth. Our business has found incredible product-market fit and has monetized effectively since the day we signed our first customer. We’re growing at a blistering pace, which presents career-defining opportunities for engineers to accelerate their growth and to contribute to a rapidly-scaling company. Visit our Vanta Engineering Blog to learn more about what our team is working on! The Integrations Platform team mission is to power the world’s largest trust automation ecosystem, enabling any person or agent to build, connect, and automate trust seamlessly. We own Vanta’s integration ecosystem, which currently includes over 400 integrations across Cloud Providers (AWS, Azure, GCP), Identity Providers, Mobile Device Management (MDM), and Human Resources Information System (HRIS). We are focused on developing the Integration Platform. This includes creating shared primitives for authentication, lifecycle, observability, and publishing to ensure all integrations are built on the same fou
$100K – $125K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role At Sentry, Support is an engineering discipline. Our customers are the greatest technical minds in the world—developers at elite enterprises building the future of software—and they deserve answers that go deeper than a knowledge base link. We are architecting the Technical Support engine . We’re looking for a veteran engineer to help us redefine the standard of technical support by combining deep human expertise with autonomous agentic systems. You are a debugger of both code and systems. You will treat support volume as a data signal to build automated resolution paths, ensuring our human engineers only touch the most complex, high-impact architectural puzzles. Sentry Support Engineers aren't just clearing queues; they are Orchestrators . You will engage with our users across GitHub, Discord, and our internal systems, while acting as the Technical Lead for our Agentic Ops. You ensure that when a developer asks a complex question, our systems have the right context and a seamless "Human-in-the-Loop" path to you when deep, nuanced expertise is required. In this role you will Master the Sentry Ecosystem & Support Elite Developers Deep-Dive Debugging: Perform root-cause analysis on complex issues and distributed tracing gaps across polyglot environments. Support the Great Minds: Act as a strategic consultant for senior engineers at our largest enterprise customers, solving high-stakes architectural challenges that push the boundaries of observability. Troubleshoot SDK Implementations: Go deep into the source code of Sentry’s SDKs to help developers instrument complex frameworks and custom environments. Engin
From C$160K/yr
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Streaming Foundations team builds services and operates data pipeline infrastructure to support event streaming, messaging, and analytics use cases. We are looking for a Software Engineer who is passionate about distributed systems, platform engineering, and solving data-intensive problems at scale. In this high-impact role, you will get to work with engineers throughout the organization to build foundational infrastructure that allows Auth0 to scale for years to come. What you’ll be doing Help set the technical direction for the team and influence the engineering roadmap for the Platform’s streaming capabilities Design and lead the implementation of our most complex and critical systems for data-intensive use cases. Research and champion new technologies and architectural patterns to solve strategic challenges and scale the platform. Lead and influence cross-functional initiatives, ensuring technical alignment and successful execution across multiple teams. Improve the operational posture of our systems by designing for observability, reliability, and scalability, and by mentoring others in operational best practices. Coach and mentor senior engineers and act as a technical leader across the engineering organization. Collaborate with different stakeholders like product teams whenever needed. What you’ll bring to our teams 7+ years of software development experience in a fast-paced, agile environment Experience working with Golang or Java is preferred H
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Specialist, Product Management – Attributes & Insights Our Purpose: We work to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion (https://www.mastercard.us/en-us/vision/who-we-are/diversity-inclusion.html) for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team – one that makes better decisions, drives innovation, and delivers better business results. About the Team Mastercard Identity, within the Security Solutions organization, leads the development of products and services that enable global commerce, powers financial inclusion, prevents crime and makes some of the most seamless experiences possible. As part of Mastercard Identity, the Attributes & Insights team drives the development and management of products, programs, and services focused on deterministic attribute ve
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Counsel, Services - Privacy, AI & Data Responsibility Overview Mastercard is seeking a Senior Counsel to serve as a dedicated Canadian privacy and data responsibility counsel supporting Mastercard’s products, services, and go-to-market activities in Canada. This role will provide practical, business-oriented legal advice on Canadian privacy, AI, data protection, consumer protection, and related regulatory requirements, with a primary focus on enabling Canadian product launches and supporting remediation of existing products, services, processes, and controls to meet Canadian legal and regulatory expectations. Role The Senior Counsel will work closely with product, regional, regulatory, commercial, and global privacy stakeholders to help Mastercard bring innovative products and services to the Canadian market in a responsible, compliant, and scalable way. The role will support the identification, prioritization, and implementation of compliance controls for Canadian requirements. You will: Serve as the Canadian Data Protection Officer for our Services business Act as the lead legal advisor on Canadian privacy, AI, data protection, consumer protection, and related regulatory matters. Provide practical legal guidance to product, business, and legal teams on Ca
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Specialist, Attribute Verification Product Management Our Purpose: We work to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion (https://www.mastercard.us/en-us/vision/who-we-are/diversity-inclusion.html) for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team – one that makes better decisions, drives innovation, and delivers better business results. About the position Services within Mastercard is responsible for acquiring, engaging, and retaining customers by managing fraud and risk, enhancing cybersecurity, and improving the digital payments experience. We provide value-added services and leverage expertise, data-driven insights, and execution. Mastercard Identity, within the Security Solutions organization, leads the development of products and services that enable global comme
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Lyft is building the next generation of intelligent agents powered by AI. We're seeking a Senior AI Agentic Engineer to lead this transformation across the enterprise. This is a strategic role for someone who can bridge the gap between AI technology and business value, designing and deploying AI-powered workflows that deliver measurable outcomes. You'll work across IT, people operations, marketing, sales, finance, legal, and procurement to architect AI-first solutions that solve complex business problems. Responsibilities: Architect and implement AI agents and intelligent workflows that transforms complex, multi-step business processes and deliver quantifiable business outcomes Partner with business leaders and stakeholders across multiple departments to identify high-impact AI opportunities that align with organizational objectives Elicit requirements from diverse stakeholders and translate complex business problems into technical solutions Build compelling business cases that clearly articulate ROI, implementation costs, benefits, timelines, and strategic alignment Evaluate, recommend, and implement AI solutions that best fit organizational needs and use cases Design solutions with an AI-first approach, ensuring optimal value delivery and user experience Monitor and measure the performance of AI workflows, using data-driven insights to demonstrate value and drive continuous improvement Design and deliver training programs, workshops, office hours, and enablement materials that help teams understand and adopt agentic solution Create frameworks, best practices, templates, and reusable patterns that accelerate AI adoption and ensure consistent quality Build a community of practice around AI and intelligent agents, empowering others to identify opportunities and contribute to the agentic roadmap Act as
From $96.4K/yr
Level Up Your Career with Zynga! At Zynga, we bring people together through the power of play. As a global leader in interactive entertainment and a proud label of Take-Two Interactive, our games have been downloaded over 6 billion times—connecting players in 175+ countries through fun, strategy, and a little friendly competition. From thrilling casino spins to epic strategy battles, mind-bending puzzles, and social word challenges, our diverse game portfolio has something for everyone. Fan-favorites and latest hits include FarmVille™, Words With Friends™, Zynga Poker™, Game of Thrones Slots Casino™, Wizard of Oz Slots™, Hit it Rich! Slots™, Wonka Slots™, Top Eleven™, Toon Blast™, Empires & Puzzles™, Merge Dragons!™, CSR Racing™, Harry Potter: Puzzles & Spells™, Match Factory™, and Color Block Jam™—plus many more! Founded in 2007 and headquartered in California, our teams span North America, Europe, and Asia, working together to craft unforgettable gaming experiences. Whether you're spinning, strategizing, matching, or competing, Zynga is where fun meets innovation—and where you can take your career to the next level. Join us and be part of the play! Position Overview We are seeking a Senior Product Manager with experience in Ad monetization or Ad Technology systems to shape, execute & distribute our Ad Technology SDK across 50+ Games at Zynga! This role specializes in key aspects of our product lifecycle, building customized ad formats for our games and delivering experiments, readouts and optimizations. This role requires a blend of product management and ad monetization experience. Ideal candidates should have experience working on monetization/payments systems and should be deeply familiar with how games balance engagement with IAA/IAP monetization. What You'll Do Lead experiments for our Advertising platform from experiment design through analysis and rollout -> fine-tuning supply and demand side features that impact how our players see ads
About the Team: Tubi's Internal Tools team is at the forefront of AI integration, developing everything from developer resources to production-grade AI for business operations. We are the group responsible for turning AI from an experiment into an operating capability: training, infrastructure, developer agents, and AI-powered business systems. Engineers operate with high ownership and autonomy, collaborating on shared architectural decisions and AI infrastructure. What You'll Do: Own systems end to end — design them, build them, and support them in production. Lead the projects you own: sequence the work, decide what lands first, and set technical direction for the engineers working with you. Sit with the people who use what you build, and turn what you learn there into a system. Design the service boundaries, contracts and schema evolution that let our platforms grow without breaking the teams depending on them. Make our AI systems dependable in production: evaluation harnesses, human approval steps before an agent acts, retries that handle a model returning something unexpected, and cost tracking that tells you what a task costs before you run it. Build what other engineers build on — agent skills, tool and MCP integrations, shared libraries — and raise the bar through code review, design discussion and mentoring. Spot the platform work nobody has asked for yet, make the case for it, and build it. Your Background: 5+ years of professional experience building and operating production systems, from design through production ownership. A system you designed and can walk us through end to end — where its boundaries sit, what constrained it, and what you chose against. Strong programming proficiency in a statically typed language such as Rust, Go, C++, Java, Kotlin, C#, or TypeScript. Production Rust is a plus rather than a requirement. You have owned a service in production: you wrote the runbooks, you knew what it cost, and you were the one paged when it broke. Expe
From C$108K/yr
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Role Overview As the Okta Research & Design (ORD) Jira Owner & Administrator , nested within the Technical Program Management (TPM) organization, you will define and execute the strategy for how ORD plans, tracks, and reports on engineering work within the full toolchain of Atlassian Jira, Jira Advanced Roadmaps (Plans), Confluence, Rovo, and AI integrations. In this role, you act as a critical bridge between product & technical leadership, agile teams, and cross-functional business units — translating methodology, best practices, and process frameworks into practical workflow improvements that accelerate delivery and reduce manual overhead for a 1,000+ person engineering, product, and design organization. By optimizing tooling configurations, automating reporting, and eliminating fragmented management processes, you will directly accelerate ORD’s developer velocity and strengthen operational efficiency. Core Responsibilities 1. Platform Ownership & Strategic Tooling Direction Strategic Vision: Define, implement, govern an AI-led strategy for utilizing Jira for product planning, tracking, and reporting tooling that supports end-to-end portfolio tracking from high-level corporate objectives down to sprint-level stories. E2E Ecosystem Management: Complete ORD ownership of the toolchain (Jira, Jira advanced Roadmaps, AI integrations, Confluence), overseeing architecture, schema updates, workflow management, screen/notification schemes, cu
From C$1.4M/yr
About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu
Other cities to consider
More places hiring for this role
Get new senior observability engineer jobs in Toronto, Canada by email
Daily job updates · Unsubscribe anytime