Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is looking for a Firmware Engineer working on microcontrollers and SoCs, focused on low-level C/C++ development and board bring-up. You’ll implement and debug firmware, develop boot/power/reset sequences, and use lab tools to diagnose issues across the hardware–software boundary. You’ll collaborate closely with hardware, board, and system software teams while building strong skills in modern embedded platforms, RTOS/Embedded Linux, and automated testing. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An experienced embedded engineer with a strong foundation in computer engineering, electrical engineering, computer science, or a related field, and a track record of building reliable software for real-world systems. A strong problem solver who enjoys working where software meets hardware and can move comfortably between architecture, implementation, debugging, and system-level thinking. Curious by nature and energized by complex challenges, with a willingness to explore new technologies, development approaches, and AI-assisted tools to make engineering more effective. A thou
Jobs in Canada
Ai Platform And Agentic Engineer in Toronto
363 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai platform and agentic engineer jobs in Toronto. Filter by work mode, employment type, experience, department, date posted and distance.
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's Governance and Compliance platform is the operating layer enterprises trust to run their security programs. We are hiring an engineer who will own the architectural foundation that makes it work for their most complex organizational structures. Program Structure and Trust is a newly chartered group at Vanta with a focused mission: build the enterprise org model that lets customers bring their compliance and security structure — product lines, business units, isolated data environments, cross-cutting audits, scoped approvals, and data residency requirements — natively into the platform. The problems this team solves determine whether Vanta can serve the enterprise customers it's increasingly winning. This is the defining technical role of the group. The Principal Engineer owns the design, phased delivery, and long-term technical direction of Vanta's enterprise org model — a multi-quarter initiative that cuts across the platform and establishes the foundation for how enterprise customers structure, segment, and operate inside Vanta. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as a Principal Engineer at Vanta: Own the design and multi-quarter delivery of Vanta's enterprise org model, including hierarchical product lines and business units, isolated data access and ownership, cross-cutting audit workflows, scoped approvals, and EU and GovCloud data residency support Define and evolve the core abstractions that let the platform absorb structurally diverse, often conflicting enterprise requirements — solving for the general case rather than one-off customer accommodations R
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Join Tenstorrent as a Staff Reliability Engineer and help define the reliability strategy behind the next generation of AI computing systems. In this highly visible technical leadership role, you'll drive reliability from architecture through production, partnering across hardware, software, and manufacturing teams to build high-performance AI platforms that set the standard for uptime, durability, and quality. If you're passionate about solving complex engineering challenges and influencing products at scale, you'll have the opportunity to shape technology powering the future of AI. This role is hybrid, based out of Toronto, Canada. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are You've spent 8+ years in reliability engineering, ideally in high-performance computing, AI hardware, or data center systems. You're comfortable with the statistical side of the job, HALT, HASS, ALT, MTBF, Weibull analysis, and FMEA are all familiar territory. You can work through a technical problem in a thermal lab and then explain the risks and trade-offs clearly to leadership. You're good at bringing people together, mechanical, electrical, thermal, softw
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a Director of Accounting Operations to lead and scale the company’s core accounting operations function. This leader will build a disciplined, accurate, audit-ready, and scalable organization that can support rapid growth, global expansion, increasing transaction complexity, and future public-company requirements. In this role, you will oversee day-to-day accounting operations across AP, AR, payroll accounting, fixed assets, accruals, reconciliations, and close operations while partnering across global finance functions. This role is hybrid, based out of Santa Clara,CA; Austin,TX; Boston, MA; or Toronto, ON. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An experienced accounting operations leader who thrives in fast-paced, high-growth environments and brings strong technical accounting judgment, operational discipline, and a builder’s mindset. Equally comfortable solving detailed accounting issues and designing scalable processes, systems, and controls, with a strong foundation in US GAAP, close, reconciliations, accruals, AP, AR, payroll accounting, fixed assets, internal controls, and audit support. Able to tell wha
From C$140K/yr
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Okta Verify Team Okta Verify Team builds desktop and mobile applications for authentication and authorization across Okta-protected resources. Our mission is to enable customers to access these resources securely. We develop Okta cloud services and client software that allow users to seamlessly login to devices and use Okta authenticators to access applications securely. The Staff Software Engineer in Test Opportunity We seek a passionate and experienced C++ / Linux Staff Software Engineer in Test to join our dynamic team. We are looking for someone who is excited about testing and automation for the Linux Platform and can build for this enterprise setting. Okta Engineering strongly believes in automated testing, UX design, and an iterative process to build high-quality next-generation software. The ideal candidate is passionate about leading quality strategy for large-scale, mission-critical client applications on Linux platform.This is a great opportunity to work in an environment of scale, work in a talented and passionate group, build processes, and be front and center in Okta’s vision of passwordless access to any application. The role gives the candidate an excellent opportunity to learn about interesting problems in security and identity space. It also has a lot of visibility within Okta and has great growth potential. What you’ll be doing Collaborate with the product management, development, and cro
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. As a Senior Software Engineer, Data on the Mapping team, you will collaborate with our world-class team of engineers, product managers, and scientists to grow and improve the quality of recommended routes and accuracy of our travel time estimations. You will lead the architecture and long-term technical direction of our offline experimentation tooling and route simulation services — the systems that let Lyft test routing changes safely before they reach production. You'll also build scalable data pipelines for experimentation, analytics, and machine learning models, along with the data governance and observability systems that keep them trustworthy. Your work will enable integration with partner teams and allow stakeholders across Engineering, Data Science, and Product to make data-informed decisions that directly impact Lyft’s growth and profitability. Our technology stack is based on the latest technologies such as AWS, Databricks, Kubernetes and Airflow. You will work with incredibly passionate and talented colleagues from software engineering, machine learning and data science on projects that directly impact millions of riders and drivers. Responsibilities Own core data pipelines end-to-end, building deep subject matter expertise in the systems you manage and defining/managing SLAs for pipelines, services, and datasets to ensure reliability at scale Serve as the technical owner and architectural lead for our offline experimentation platform and route simulation services, setting technical direction, evaluating trade-offs, and ensuring the systems scale with Lyft's routing and mapping ambitions Continuously evolve data models and schemas to meet business and engineering requirements Develop AI tools that support self-service management of data pipelines (ETL) and schema evolution, and perform han
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by building high-performance, scalable and reliable machine learning systems? Do you want to help define and build the next generation of AI platforms powering advanced NLP applications? We are looking for a Site Reliability Engineer to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will work closely with many teams to deploy optimized NLP models to production in low latency, high throughput, and high availability environments. You will also get the opportunity to interface with customers and create customized deployments to meet their specific needs. As a Site Reliability Engineer you will: Build self-service systems that automate managing, deploying and operating services. This includes our custom Kubernetes operators that support language model deployments. Automate environment observability and resilience. Enable all developers to troubleshoot and resolve problems. Take steps required to ensure we hit defined SLOs, including pa
Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 130 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com . Reddit’s Ads Data Science team is looking for a highly experienced Staff Data Scientist to advance the intelligence powering the advertiser experience on Reddit. In this role, you'll take deep ownership of a high-impact problem space within advertising, specializing in measurement, identity, and signal quality. This is a high-impact, high-autonomy role where you'll influence strategic direction, set a high technical bar, and drive cross-functional initiatives across one or more critical focus areas in the Ads organization. Responsibilities: Design the Future of Ads Identity: Develop/employ probabilistic models for identity resolution. Design the methodology that links on-platform and off-platform actions to maximize addressability while honoring privacy. Advance Lift Methodologies & Experimentation: Own the statistical rigor behind Reddit’s Brand and Conversion Lift products. Innovate experimental design and develop infrastructure that supports large-scale, high-velocity, low-bias testing for advertisers. Maximize Signal for Predictive Performance: Define the strategy for new signal sources. You will mathematically quantify the value of these signals and work with modeling teams to incorporate them into predictive models, directly improving bidding efficiency and ROAS. Define Ground Truth & Evaluation Frameworks: Solve the industry-wide challenge of validating identity and measurement. Design the objective functions and truth sets used to train our models and measure the incremental impact of our identity graph.
Engineering Manager, AI Conversation Platform Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The newly formed Conversation Platform team aims to build a conversation platform for all merchants who use Stripe. We are doing so by (a) automating the easy tasks, and (b) assisting our users in the difficult tasks. Some examples include customizing the Stripe landing page to suggest bespoke integrations, allowing users to command the Stripe API in natural language, and resolving user issues automatically. We are developing RAG based systems on the latest LLMs as well as fine-tuning our own models. We’re an end-to-end team going from ideas to models to shipping in production. What you’ll do Responsibilities Driving an ambitious vision for AI/ML that benefits our users Setting the technical & process direction for the team based on business goals Brainstorm and coordinate product integrations with partner teams Proposing new ideas and building prototypes Be an integral part of a larger ML community internally & externally Hire & develop a world-class team to deliver high-quality ML systems. Coach engineers to help them grow in their careers and maintain a high bar Who you are We are looking for ML Engineering Managers who are passionate about using ML to improve products and delight customers. You have experience leading teams that develop streaming feature pipelines, build ML models, and deploy them to production, even if it involves making substantial changes to backend code. You a
Engineering Manager, AI Conversation Platform Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The newly formed Conversation Platform team aims to build a conversation platform for all merchants who use Stripe. We are doing so by (a) automating the easy tasks, and (b) assisting our users in the difficult tasks. Some examples include customizing the Stripe landing page to suggest bespoke integrations, allowing users to command the Stripe API in natural language, and resolving user issues automatically. We are developing RAG based systems on the latest LLMs as well as fine-tuning our own models. We’re an end-to-end team going from ideas to models to shipping in production. What you’ll do Responsibilities Driving an ambitious vision for AI/ML that benefits our users Setting the technical & process direction for the team based on business goals Brainstorm and coordinate product integrations with partner teams Proposing new ideas and building prototypes Be an integral part of a larger ML community internally & externally Hire & develop a world-class team to deliver high-quality ML systems. Coach engineers to help them grow in their careers and maintain a high bar Who you are We are looking for ML Engineering Managers who are passionate about using ML to improve products and delight customers. You have experience leading teams that develop streaming feature pipelines, build ML models, and deploy them to production, even if it involves making substantial changes to backend code. You a
From C$190K/yr
About Forma.ai: Forma.ai is a Series B startup that's revolutionizing how sales compensation is designed, managed and optimized. We handle billions in annual managed commissions for market leaders like Edmentum, Stryker, and Autodesk. Our growth has been fuelled by our passion for fundamentally changing and shaping how companies use sales intelligence to drive business strategy. We’re welcoming equally driven individuals who are excited about creating something big! About the Team Engineers on this team build our rules-based calculation engine for processing sales commissions. This might sound simple if you have never been exposed to sales compensation plans, it is not. We are low on meetings and high on accountability. Most of the team is in the EST time zone, with a few located in PST and Central as well. We are still evolving many areas of the platform, which means there is meaningful room to improve the design, reliability, and scalability of the systems we build. What you’ll be doing Reporting to the Manager of Data Platform, you will play an important role in the evolution of our Spark-based data platform. You’ll design and build data-rich platform capabilities, contribute to system design discussions, and help ensure our data systems remain reliable, maintainable, and scalable as Forma grows. As a Senior Engineer, Data Platform, you are expected to operate with strong ownership and sound technical judgment. This includes identifying risks in the work you own, surfacing edge cases, asking thoughtful questions, and proposing improvements that strengthen the quality and reliability of the platform. You will: Design, build, and improve Spark-based data pipelines and platform services. Work with complex data models representing sales compensation plans, hierarchies, relationships, and enterprise datasets. Build reliable, deterministic data systems that customers and internal teams can trust. Improve testing, observability, data quality,
From C$1.4M/yr
About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Role Overview We are seeking a Platform Experience and Developer PM to own how developers and enterprise technical teams build on, integrate with, and operate Cohere's model platform. This is a high-leverage role sitting at the intersection of three domains: Managed Services and Models as a Service. Own Cohere's managed service offerings as a product. This is broader than model serving alone. It includes the full range of how enterprises consume and operate Cohere's capabilities, from shared multi-tenant model access to dedicated single-tenant deployments, and from synchronous real-time inference to high-volume asynchronous workloads. You will own the product thinking around deployment models, data residency and regional compliance requirements, self-serve provisioning, and the operational controls that give enterprises confidence in running production workloads on Cohere’s infrastructure. API and SDK. Own the roadmap for how developers build on Cohere. This means setting the direction for our APIs and SDKs, thinking carefully about interface design and ergonomics, and ensuring we ship developer primitives that are stable, well-
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Role We are looking for a Principal Product Manager to lead the strategic evolution of the Auth0 Platform Infrastructure. In this highly visible, high-impact leadership role, you will be the visionary force behind the core infrastructure, consumption models, and platform capabilities that power Auth0’s immense global scale. You will own the multi-year strategy for critical platform pillars, with a heavy emphasis on redefining our rate-limiting and platform consumption frameworks. As a Principal product leader, you will navigate multi-dimensional scaling challenges (including the paradigm shift of autonomous AI agents), influence executive stakeholders, and ensure our platform remains the most resilient, secure, and performant identity solution for the world’s largest enterprises. What You’ll Do Drive the Long-Term Platform & Consumption Strategy: Define and own the multi-year roadmap for Auth0’s core infrastructure. Align deeply technical platform architecture with broader business objectives, Go-To-Market strategies, and overarching revenue goals. Spearhead the Strategic Evolution of Rate Limits & Monetization: Transform rate limits from a protective mechanism into a core business strategy. Pioneer dynamic consumption models, proactive policy-aware observability, and intelligent throttling strategies that align platform costs with customer value, reduce escalations, and unlock new SKUs. Architect Global Cloud Resilience & Scale: Champion ov
From C$108K/yr
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are building and maintaining a highly scalable asynchronous platform that empowers our organization to handle critical business cases. As a software engineering team, our mission is to create robust and innovative solutions that drive the success of our business and deliver unparalleled value to our customers. We adopt Infrastructure as Code practice to automate the provisioning and configuration of our resources, which helps reduce manual configuration and improve consistency. Our team culture is built on collaboration, open communication, and a supportive environment where each member's ideas are valued and contributions are recognized. We believe in the importance of fostering a positive workplace culture that inspires innovation and creativity. Responsibilities: Maintain and analyze metrics from; operating systems; control planes; and applications to assist in fault detection and performance enhancement Design, develop and deploy tooling and systems that continually improve the reliability, scalability and efficiency of our platform Balance feature development speed and reliability with service-level objectives Operate and improve our Infrastructure using industry best practices and tools Participate in design and production readiness reviews, platform management and capacity planning ceremonies with cross-functional teams Document Infrastructure operations process and insights, identify repeatable actions and ruthlessly automate repetitive tasks Participate in our teams on-call rotations, respond to incidents and support other teams mitigate customer impacting events Experience: 5+ years experience working on teams responsible for software development, automation and systems engineering Experience building large-scale infrastructure, distributed systems or networks. Knowledge with SQS,
Other cities to consider
More places hiring for this role
Get new ai platform and agentic engineer jobs in Toronto, Canada by email
Daily job updates · Unsubscribe anytime