About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! Senior Staff Backend Engineer About the Team We build enterprise software that helps organizations optimize sales performance, enabling go-to-market agility. Our engineering organization includes multiple product application teams responsible for delivering core customer-facing capabilities. We are seeking a Senior Staff Backend Engineer to join our application teams and help set technical direction across multiple domains within engineering. You'll work alongside staff, senior, and early-career engineers, and partner closely with engineering leadership to define, evolve, and scale the systems that power enterprise-grade product workflows. This is an opportunity to own complex, multi-domain technical problems and shape product direction beyond a single team. We are low on meetings, high on accountability. Most of the teams are in the EST time zone, but we have a few located in AST, PST, and Central as well. What you'll be doing You will play a pivotal role in shaping the technical direction of our application stack across multiple domains. You will lead development efforts for our most complex initiatives, the kind that span two or more teams or product areas, and serve as a technical benchmark for system design, code quality, and long-term maintainability. You'll operate at the intersection of data modelling, business logic, and enterprise-scale reliability, and your work will often set standards that neighboring teams adopt. This remains a hands-on
Jobs in Canada
Software Engineer Ml Infrastructure Platform in Toronto
155 active opportunities · Updated October 2026
Showing
15 jobs
Explore current software engineer ml infrastructure platform jobs in Toronto. Filter by work mode, employment type, experience, department, date posted and distance.
About the Role We are a small team of AI builders in Paytm Labs. As a Staff AI Platform Engineer, you will work across inference and agentic systems. You will contribute to Paytm's AI inference platform (Pi), serving internal teams and enterprise customers - running our own coding and domain-specific models (voice, vision, risk, fintech workflows) as well as third-party models. You will also architect and build the platform that enables autonomous AI agents to operate safely and reliably in production - the runtime, orchestration, and developer tooling for agents to reason, plan, use tools, and execute complex multi-step workflows, automating both software development and business processes. You will work at the intersection of LLMs, distributed systems, and production fintech infrastructure, helping define how inference and agentic AI are built and deployed across payments, risk, fraud, collections, support, and developer experience.
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! As a Senior Security Engineer you will: Serve as trusted advisor to team’s leadership and partner teams by clearly articulating business risks associated with security issues Lead security operation functions – including vulnerability management, SAST, DAST, detection engineering, and incident response – in CI/CD and cloud-native production environments Integrate security into our applications throughout the software development lifecycle Collaborate with product and development teams, driving the success of larger projects to ensure that software is built and deployed securely without compromising agility and speed Driving and supporting bug bounty program, application security reviews and threat modeling, including code review and dynamic testing Assess and integrate security tools to automate and scale security processes, i.e: evaluate open-source vs vendor solutions Gather and analyze security metrics to address security issues with cross-team dependencies Be a problem solver who is empathetic to developer concerns and will employ constructive and flexible approach to building innovative solutions You may be a good fit if: 5
About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team We are a brand new team, working quickly to accelerate Stripes engineering productivity by effectively deploying LLM agents and tools to automate large swaths of the engineering workflow. We’re a cross-continent team, spanning North America and Europe - filled with high agency engineers with strong product perspectives. Some examples of what we work on: here and here . What you’ll do You will build the next generation of internal AI coding tools and platforms, to massively accelerate Stripe’s engineering productivity safely and with a clear focus on our users’ needs. Who you are We’re looking for someone who meets the minimum requirements to be considered for the role. If you meet these requirements, you are encouraged to apply. The preferred qualifications are a bonus, not a requirement. Minimum requirements We’re looking for someone who has: Strong software engineering skills and ability to write high quality code at high velocity 2-5 years of professional hands-on software development experience, able to write well-factored algorithms and have experience with commonly used data structure and algorithms Hands on experience building tools or platforms used by other engineers or internal users Strong collaboration skills, can work across workstreams within your team and contribute to your peers’ success Customer obsession, ability to articulate and represent customer experience in various forums to drive the right outcome Have t
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. As a PCB Power Design Engineer, you will design and optimize power distribution systems for Tenstorrent’s next-generation AI accelerator cards, balancing performance, area, cost, and reliability. You will contribute across the full power design lifecycle, from component selection and simulation through board bring-up, validation, debugging, and production readiness. Working closely with hardware, firmware/software, thermal, mechanical, and validation teams, you will help deliver robust power architectures for high-performance AI systems. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An electrical engineer with experience designing, analyzing, and debugging power distribution systems for high-performance or high-power products. A hands-on engineer who enjoys moving from schematics and simulations into lab bring-up, testing, troubleshooting, and design validation. A systems-oriented collaborator who can work effectively across hardware design, firmware/software, thermal, mechanical, and validation teams. A detail-oriented problem solver who balances electrical performance, power integr
About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! About the Team Engineers on this team construct our rules-based calculating engine for processing sales commissions. This might sound simple if you have never been exposed to sales comp plans, it is not! We are low on meetings, high on accountability. Most of the team are in EST time zone but we have a few located in PST and Central as well. We are far from maintenance / progressive evolution in many areas, there is a lot of room to make a big impact in the overall design. What you’ll be doing Reporting to the Manager of Data Platform, you will play a critical role in the evolution of our Spark based data platform. You'll lead development efforts for our complex, data-rich platform features while being an example to the team of code quality and thoughtful software design. You will be working on the most challenging code at Forma. As a Staff Engineer, you are expected to operate with a high degree of ownership and trust. This includes proactively identifying architectural risks, surfacing edge cases or constraints others may not see, and advocating for improvements that strengthen the long-term integrity of the system. We value engineers who bring forward thoughtful perspectives - even when they challenge assumptions - and who help the team see around corners. You will: Design and evolve backend services that power product workflows. Architect data models representing hierarchical & graph structures, relationships, and large-scale enterprise datasets. Build
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Join Tenstorrent and help bring next-generation AI accelerator technology from silicon bring-up to production. You’ll work at the forefront of hardware innovation, diagnosing complex issues across chips, systems, firmware, and software while collaborating with some of the brightest engineers in the industry. This role offers the opportunity to solve challenging technical problems, build impactful debug solutions, and directly influence the reliability and performance of cutting-edge AI compute platforms. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A hands-on hardware debug engineer who thrives on solving complex, cross-functional problems at the intersection of silicon, firmware, and software. A curious and analytical problem solver who enjoys digging into failures, identifying root causes, and driving issues from initial discovery through resolution. An engineer with strong post-silicon validation and bring-up experience who is comfortable working in the lab and getting deep into system-level behavior. Someone who enjoys building tools, improving debug methodologies, and creating
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world's biggest financial problems. We're looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn't a place for complacency, it's where ambitious people do the best work of their careers. We're a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Developer Infrastructure org is the engine behind Robinhood's entire engineering organization — a collection of tightly integrated teams whose collective mission is to make every engineer at Robinhood faster, more reliable, and exponentially more productive. The org spans four major teams: DevX (Developer Experience), TestX (Test Infrastructure), Backend Platform, and Mobile Platform. DevX owns Robinhood's monorepo and Bazel-based build infrastructure — the critical layer between a developer writing code and that code being ready to ship — along with the company's full CI/CD pipeline and remote build execution cluster. TestX owns the infrastructure behind Robinhood's entire test experience: the integration test environments, and personal development environments that serve as miniature simulations of the full Robinhood system, giving engineers a safe, isolated space to test their code end-to-end before it ever touches production. Backend Platform and Mobile Platform own the core language runtimes, libraries, IDEs, and developer toolchains across Python, Go, TypeScript, Swift, and Android. Together, these teams share a single north star: leveraging AI and agentic systems
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy behind the next generation of AI computing systems. In this highly visible technical leadership role, you'll drive reliability from architecture through production, partnering across hardware, software, and manufacturing teams to build high-performance AI platforms that set the standard for uptime, durability, and quality. If you're passionate about solving complex engineering challenges and influencing products at scale, you'll have the opportunity to shape technology powering the future of AI. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are You've spent 8+ years in reliability engineering, ideally in high-performance computing, AI hardware, or data center systems. You're comfortable with the statistical side of the job, HALT, HASS, ALT, MTBF, Weibull analysis, and FMEA are all familiar territory. You can work through a technical problem in a thermal lab and then explain the risks and trade-offs clearly to leadership. You're good at bringing people together, mechanical, electrical, thermal, softw
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is looking for a Firmware Engineer working on microcontrollers and SoCs, focused on low-level C/C++ development and board bring-up. You’ll implement and debug firmware, develop boot/power/reset sequences, and use lab tools to diagnose issues across the hardware–software boundary. You’ll collaborate closely with hardware, board, and system software teams while building strong skills in modern embedded platforms, RTOS/Embedded Linux, and automated testing. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An experienced embedded engineer with a strong foundation in computer engineering, electrical engineering, computer science, or a related field, and a track record of building reliable software for real-world systems. A strong problem solver who enjoys working where software meets hardware and can move comfortably between architecture, implementation, debugging, and system-level thinking. Curious by nature and energized by complex challenges, with a willingness to explore new technologies, development approaches, and AI-assisted tools to make engineering more effective. A thou
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's Core Platform team provides the foundational infrastructure that powers all engineering at Vanta. We're expanding upmarket to support enterprise customers, which requires strategic investment in platform systems that ensure security, reliability, and developer productivity at scale. As we expand upmarket to support enterprise and regulated customers, we’re investing heavily in platform capabilities that scale securely while reducing cognitive load for product teams. As the Engineering Manager, Core Platform at Vanta, you'll own the foundational infrastructure that every engineer builds on, ensuring it scales with company growth while remaining fast, simple, and reliable. This team’s ownership spans shared services infrastructure, observability and monitoring, datastore management, and async work systems. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow high-performing platform engineering teams that deliver reliable, scalable infrastructure and operational excellence for Vanta’s products and customers Set technical direction and drive multi-quarter platform initiatives spanning infrastructure reliability, security, scalability, and developer experience across shared systems and services Partner closely with product engineerin
From C$122K/yr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. The People Development team is building the infrastructure for how Lyft grows, develops, and equips its people. As Technical Learning Manager, you will own the strategy and execution of technical learning programs for Lyft's Engineering, Science, Product, and Design organizations — one of our highest-leverage investments in people. This role sits at the intersection of learning design, program management, and functional excellence across the technology organizations. You will inherit an established portfolio of programs for software engineers — including Tech Learning, compliance training, and documentation initiatives — and will be responsible for running, evolving, and expanding them across all of Lyft’s tech orgs. You will report to the Head of People Development and work closely with Functional excellence leadership, HRBPs, and cross-functional partners to ensure our technical learning programs are high-quality, well-adopted, and tied to real business outcomes. Responsibilities: Program Ownership Own the strategy and execution of Lyft's technical learning portfolio, including engineering continuing education programs, documentation improvement initiatives and compliance training Manage mid-cycle programs with active stakeholder relationships and scheduled commitments — ensuring continuity, quality, and follow-through Provide editorial oversight for Lyft's internal technical content initiatives, including the Tech Blog — managing workflow, stakeholder relationships, and publication processes in partnership with engineering contributors Assess the current program portfolio and make recommendations grounded in engineer needs and business priorities Provide engaging learning experiences that empower technologists to do their best work at Lyft Embed AI upskilling in all programs including
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Opportunity: Okta Access Gateway (OAG) Enterprises run on a mix of modern cloud services and mission-critical on-premises systems (such as Oracle E-Business Suite, SAP, PeopleSoft, and custom legacy web apps). Okta Access Gateway (OAG) solves the enterprise hybrid cloud challenge by extending Okta’s cloud identity, Adaptive MFA, and Zero Trust security policies to on-premises and legacy applications without requiring custom code changes or traditional VPNs. As the Engineering Manager for Okta Access Gateway in Toronto, you will lead and grow a team of software engineers building the next generation of our hybrid access and gateway infrastructure. You will partner closely with Product Management, Architecture, Security, and Quality teams to deliver high-throughput, mission-critical security software deployed across multi-cloud and enterprise datacenters globally. What You’ll Do People Leadership & Team Growth Lead, mentor, and empower an engineering team, fostering an inclusive, high-performance, and psychologically safe engineering culture. Drive career progression, goal setting, regular 1:1s, and continuous feedback to help engineers grow their technical and leadership skills. Attract, interview, and hire diverse engineering talent to scale Okta’s engineering presence in Toronto. Delivery & Operational Excellence Own the end-to-end execution and delivery of key product roadmap initiatives, balancing feature velocity, technical debt, and softwar
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team of bold thinkers and sharp problem-solvers who are wired to make an impact. The Ops Platform organization develops internal platforms that replace repetitive manual processes with AI-driven systems. These tools support key areas such as Fraud Operations, Account Operations, Financial Crimes Operations, and Retirement Services. The team works closely with product, data science, and operations partners to deliver reliable systems that improve decision-making and efficiency! As a Software Developer, you will design and build platforms that enable operational teams to investigate and resolve issues more quickly and accurately. You will work with large datasets and signals to create tooling that supports fraud investigation and other operational workflows. You will collaborate with data scientists and machine learning engineers to translate manual processes into automated systems. Your work will focus on improving system reliability, reducing operational effort, and increasing the speed at which new products and features can be supported across Robinhood’s offerings. This role is based in our Toronto, ON office, with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams. What you’ll do You will define technical direction and make architectural decisions for systems that support operational workflows across multiple product lines You will build tools that proces
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world's biggest financial problems. We're looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn't a place for complacency, it's where ambitious people do the best work of their careers. We're a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Software Platform team accelerates developer velocity and increases system reliability by building the foundational platforms and tools that power Robinhood engineering. Within this group, the Kubernetes Compute team focuses on building and operating a highly available, scalable Kubernetes-powered container platform. We ensure that our infrastructure seamlessly supports reliable application deployments, integrates core platform capabilities, and enables multi-region scalability. We are expanding our core container systems to support our next phase of technical growth! As a Software Developer, you will focus on building, maintaining, and scaling our container provisioning platforms. Working alongside senior engineers, you will write code to improve our infrastructure capabilities and actively participate in our technical transition to Amazon EKS. In this role, you will collaborate with teams across the organization to ensure robust platform integrations for everyday application needs like security and networking. Your efforts will directly improve system visibility, automation, and reliability across the platform. This role is based in our Toronto office(s), with in-perso
Related career options
Similar roles with stronger pay
Demand 50/100 · 8 jobs
$1.3M – $1.3M/yr
Salary →Demand 67/100 · 15 jobs
$972.9K – $972.9K/yr
Salary →Demand 51/100 · 10 jobs
$249K – $249K/yr
Salary →Demand 53/100 · 26 jobs
$241K – $241K/yr
Salary →Demand 61/100 · 48 jobs
$240K – $240K/yr
Salary →Demand 67/100 · 38 jobs
$238.8K – $238.8K/yr
Salary →Other cities to consider
More places hiring for this role
Get new software engineer ml infrastructure platform jobs in Toronto, Canada by email
Daily job updates · Unsubscribe anytime