Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Evaluation is critical to making progress in scaling intelligence. As models continue to become superhuman in many real-world use cases, we must continue to develop new evaluation techniques that accurately reflect what models are already capable of, as well as set the agenda for what future models should be capable of. In this role, you are responsible for creating these next-generation evaluation methods and infrastructure to measure LLM progress. As a Senior Research Scientist, Model Evaluation, you will: Create ambitious new evaluation benchmarks that push the limits of what our models can accomplish. Work on highly cross-functional teams to translate model feedback into trustworthy, repeatable evaluations. Conduct research to advance the state-of-the-art in LLM evaluation methods, including training LLM judges; refining LLM-based data synthesis pipelines; and improving evaluation efficiency. Build scalable and reusable tools for digging into model performance. You may be a good fit if: You enjoy rapidly building prototypes that demonstrate the boundaries of what LLMs are capable of, and you have developed res
Jobs in Canada
Model Behavior Engineer in Canada
352 active opportunities · Updated October 2026
Showing
15 jobs
Explore current model behavior engineer jobs across Canada. Filter by work mode, employment type, experience, department, date posted and distance.
About the Role: As a Staff Software Engineer on the ML Infrastructure team, you will collaborate closely with the Machine Learning and Product teams to build world-class machine learning inference platforms. These platforms power essential services like personalized recommendations, search, and content understanding across Tubi. A core responsibility of this team is developing and maintaining low-latency ML model serving systems that support Deep Learning, LLM, and Search models. This involves building self-service infrastructure and critical components such as the inference engine, feature store, vector store, and experimentation engine. You will improve the way we deploy and operate our services and even contribute to open-source projects. This role grants the architectural freedom to explore new frameworks, lead critical cross-functional projects, and transform the capabilities of our ML and Product teams. Responsibilities: Design and build scalable, high throughput, and low latency distributed systems using Scala Build reusable components and services that serve various ML applications like Personalization, Search, Ads and Exploration Partner closely with ML engineers to understand their challenges and limitations and develop scalable solutions to address them. Proactively recommend solutions to keep our ML Inference stack state of the art. Take a data driven approach to identifying & optimizing latency, cost, and efficiency of our infra. Lead large scale cross functional refactorings if necessary Mentor other engineers on the team on system design, effective incident management, interviewing, leveraging LLMs for work, etc. Collaborate with ML, Product, and cross functional engineering teams to define the long term vision and architecture for ML Infrastructure at Tubi. Your Background: Experience designing and building scalable, distributed systems in any modern backend language (e.g., Scala, Java, Python, Go, C++); experience with Scala or JVM b
From $290.4K/yr
Scale's LLM post-training platform team builds our internal distributed framework for large language model training. The platform powers MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs. It also serves as the underlying training framework for the data quality evaluation pipeline. Scale is uniquely positioned at the heart of the field of AI as an indispensable provider of training and evaluation data and end-to-end solutions for the ML lifecycle. You will work closely with Scale’s ML teams and researchers to build the foundation platform which supports all our ML research and development works. You will be building and optimizing the platform to enable our next generation LLM training, inference and data curation. If you are excited about shaping the future AI via fundamental innovations, we would love to hear from you! You will: Build, profile and optimize our training and inference framework. Collaborate with ML and research teams to accelerate their research and development, and enable them to develop the next generation of models and data curation. Research and integrate state-of-the-art technologies to optimize our ML system. Ideally you’d have: Passionate about system optimization Experience with multi-node LLM training and inference Experience with developing large-scale distributed ML systems Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc. Strong software engineering skills, proficient in frameworks and tools such as CUDA, Pytorch, transformers, flash attention, etc. Strong written and verbal communication skills to operate in a cross functional team environment. Nice to haves: Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and multimodal, etc. Compensation packages at Scale for eligible roles include base salary, equity,
From $189.6K/yr
Scale’s ML platform (RLXF) team builds our internal distributed framework for large language model training and inference. The platform has been powering MLEs, researchers, data scientists and operators for fast and automatic training and evaluation of LLM's, as well as evaluation of data quality. Scale is uniquely positioned at the heart of the field of AI as an indispensable provider of training and evaluation data and end-to-end solutions for the ML lifecycle. You will work closely across Scale’s ML teams and researchers to build the foundation platform that supports all our ML research and development. You will be building and optimizing the platform to enable our next generation of LLM training, inference and data curation. If you are excited about shaping the future AI via fundamental innovations, we would love to hear from you! You will: Build, profile and optimize our training and inference framework Collaborate with ML teams to accelerate their research and development and enable them to develop the next generation of models and data curation Research and integrate state-of-the-art technologies to optimize our ML system Ideally you’d have: Strong excitement about system optimization Experience with multi-node LLM training and inference Experience with developing large-scale distributed ML systems Strong software engineering skills, proficient in frameworks and tools such as CUDA, Pytorch, transformers, flash attention, etc. Strong written and verbal communication skills and the ability to operate in a cross functional team environment Nice to haves: Demonstrated expertise in post-training methods &/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and multimodal, etc. Compensation packages at Scale for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the positi
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! In September 2026 we raised a $350 million Series E at a $3.5 billion valuation , and we are scaling our engineering and research teams to meet demand. The role Frontier AI data is expensive to make and hard to measure. Every task we deliver is tested against the strongest models, often through many long-running agent rollouts. Your job is to make that process faster, cheaper, and more rigorous with ML and AI You will be one of the early members of ML & Research Engineering at Snorkel. You will study how frontier-grade data is generated and evaluated, form hypotheses, validate them against real production data, and ship the winners at scale. You will shape the discipline's direction, its standards, and the team that grows around it. What you'll work on Efficient agentic evals. Cut the cost of long-horizon agent evaluation with adaptive sampling, statistically grounded early stopping, model cascades, caching, and cheap-first gating. AI model routing. Route every eval and judge call to the cheapest model that clears the quality bar, with fallback, monitoring, and cost attribution. Fine-tuned small models. Fine-tune and serve open-weight models (LoRA and other
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! We’re looking for a Research Scientist to advance how high-quality data and environments for AI agents are created. You’ll build and optimize pipelines that combine real-world data, automated generation, and human expert input. Working with domain experts, academic partners, customers, and our product and engineering teams, you’ll scale these pipelines to target frontier model performance gaps and expand data and environment diversity. Your work will amplify human knowledge and judgement, enabling experts to create and refine data and agentic environments that strengthens Snorkel’s position as the frontier data lab. This role is ideal for someone who wants to advance frontier AI through data and environment creation and enjoys turning research into reusable, scalable systems. Location: San Francisco, New York, OR REMOTE Main Responsibilities Design, implement, and optimize reusable pipelines that combine AI capabilities with expert judgment to accelerate data and agentic environment creation. Design and run rigorous experiments to validate proof-of-concept approaches, measure their impact on data quality, pipeline efficiency, and model performance, and communic
C$200K – C$250K/yr
Senior Machine Learning Developer Location: Montreal, Quebec, Toronto, or Ontario About Numa Numa is building the platform to power AI-native dealerships, rearchitecting automotive service and sales with advanced AI agents that automate customer interactions, streamline operations, and reimagine how dealerships work. Numa integrates AI into every aspect of dealership functions—from rescuing customer calls and voicemails that generate more revenue, to reducing customer resolution times that drive overall customer satisfaction (CSI), to improving dealership team productivity and accountability. Numa has raised $50 million from leading investors (Google, Threshold, Costanoa, Mitsui, and Touring Capital). The Role We’re hiring a Senior Machine Learning Developer to build and ship ML/AI systems that interact with real customers thousands of times a day. Our voice agents book service appointments, rescue missed calls, and route callers through natural conversations. You’ll work across products and platforms. You’ll ship AI features including prompts, agents, tools, and production ML models, while building the evaluations and tooling that help teams ship with confidence. You’ll also contribute to our ML platform, including model serving, LLM infrastructure, and production observability. At Numa, we believe great ML is about more than building bigger models. It’s about knowing whether a change is good enough to ship. Our evaluation-first approach makes that measurable in CI and production. What You’ll Do Build conversational AI systems for phone and SMS that understand customer needs, take action, and know when to act autonomously Develop tooling such as memory, knowledge graphs, and validated customization that help agents reason and adapt to dealership needs Train, evaluate, and deploy ML models using Ray Serve and Dagster for prediction, classification, ranking, and capacity forecasting, and keep them healthy in production Create offline and online evalua
From $170K/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About the Role Snorkel AI is looking for a Head of Talent Acquisition Operations & Insights to build and lead the systems, processes, tools, metrics, and operational infrastructure that power our recruiting organization. This leader will own the design and execution of a modern TA operations function that enables Snorkel AI to scale hiring with speed, quality, consistency, and data-driven decision-making. This is a builder role for someone who has stood up TA operations in a hyper-growth technical startup environment. You should be equally comfortable designing the strategy, implementing the systems, improving process, building reporting, leading recruiting coordination, and partnering with recruiting and business leaders to improve how hiring gets done. We are looking for someone who brings strong operational discipline, deep knowledge of recruiting systems and workflows, and a forward-looking perspective on how AI can modernize talent acquisition. What You’ll Do Build and lead the TA operations function for Snorkel AI, including recruiting systems, tools, workflows, reporting, process, and coordination. Own the recruiting tech stack, including ATS configu
From $130K/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About the Team Marketing at Snorkel is growing rapidly and anchored by high-ownership operators who lead independently, align cross-functionally, and consistently deliver outsized results. We partner across Sales, Product, Research, and the executive team to translate complex AI and data value into differentiated positioning, integrated programs, and field-ready enablement that accelerates growth. The culture is high standards, high autonomy, and high collaboration. About the Role Reporting to the Sr. Director of Product Marketing, the Product Marketing Manager, Frontier Labs will own the GTM execution for our frontier lab business. You will partner closely with Research, FDE, and frontier-facing sales teams to translate technical work into research-credible positioning, repeatable sales plays, and high-quality GTM programs that resonate with research and procurement leaders inside frontier labs. You will keep the frontier asset library (battlecards, technical one-pagers, decks, benchmark and eval narratives) sharp against a fast-moving market, and bring competitive and customer insight into every motion. This is a hands-on role for a product marketer who wants
From $252K/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About Snorkel Snorkel AI is the frontier AI data lab, helping teams build the data and environments behind high-performing frontier and agentic AI. We combine technology with research-driven AI data development to create datasets, benchmarks, evals, and custom solutions for real-world AI systems. Founded out of the Stanford AI Lab in 2019, Snorkel works with leading AI labs and enterprises to move from better data to better outcomes. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About The Role Snorkel is hiring a Head of Security to build and lead our security function end-to-end — infrastructure security, application security, and governance, risk & compliance (GRC). You'll own the security function end-to-end — strategy, team, and execution — and operate as the primary security voice with customers, auditors, and the exec team. You'll report to the CTO. This is a builder's role: you'll take security from its current state to a mature, right-sized function as Snorkel scales, hiring and developing the team as needs grow. Key Responsibilities Security Leadership & Team Building Define Snorkel's overall security strategy,
From $1.3M/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! Accounting Manager, Technical Accounting & Financial Reporting San Francisco, CA (Hybrid) About The Role We are seeking a highly motivated Accounting Manager, Technical Accounting & Financial Reporting to play a key role in building and scaling Snorkel's accounting organization. This position combines technical accounting expertise with operational excellence. You will partner closely with the Controller and Accounting team to own key areas of the monthly close, prepare technical accounting analyses, support financial statement audits, and help establish scalable accounting processes and internal controls. The ideal candidate enjoys solving complex accounting issues while remaining hands-on in day-to-day accounting operations. You thrive in a fast-paced startup environment, embrace ambiguity, and enjoy building processes that support a rapidly scaling business. What You'll Do Month-End Close & General Accounting Support and help lead the monthly, quarterly, and annual close processes. Prepare and review journal entries, reconciliations, accruals, and supporting schedules. Perform detailed balance sheet reconciliations and investigate reconcil
From $125K/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! Counsel, Commercial What You’ll Do Snorkel AI is looking for a Counsel, Commercial to join our growing Legal team, reporting to Snorkel’s Associate General Counsel, Commercial. This is a key role across Snorkel’s commercial contracting function, including vendor agreements and customer agreements. You will serve as a trusted legal partner to business teams across the company, drafting, reviewing, and negotiating a wide range of commercial contracts while helping build scalable processes. You’ll work closely with cross-functional stakeholders, combining strong commercial legal skills with the hands-on, execution-oriented approach required in a fast-growing technology company. This hybrid role is based in San Francisco or New York City. Responsibilities Draft, review, and negotiate a broad range of commercial contracts, including vendor and customer agreements, NDAs, MSAs, SOWs, order forms, and other commercial arrangements. Serve as a trusted legal partner to cross-functional teams on commercial contracts. Manage a high volume of contracts and competing priorities, ensuring timely execution and alignment with business objectives. Review and negotiate commercial
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! The Role We're looking for our founding AI Data Product Manager to own Snorkel's Agentic Data and RL Environments roadmap. In this role, you'll lead the product strategy for a variety of data types (e.g. Agentic Coding, Computer Use). You will shape the roadmap for the datasets Snorkel invests in by understanding the market, incorporating frontier lab needs and collaborating with researchers at Snorkel and our academic partners. This role is highly cross-functional, sitting between Research, GTM and Operations. As a founding member for this role, you will be in charge of setting up the frameworks to build the roadmap, gather data from relevant sources, and share the roadmap with both internal and external stakeholders. What You'll Do Own the "data as a product" roadmap for Snorkel's Agentic and RL Environment focus areas, working x-functionally with research, academic partners, and GTM to define the skills and capabilities for our datasets Shape new "data" product areas and work with academic partners and research leaders to build Snorkel's competitive edge in the market Collaborate cross-functionally to help shape the roadmap and data strategy and influence bu
$190K – $240K/yr
About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About This Role We're looking for a Staff HR Business Partner to build and own the people strategy for Snorkel's Data as a Service (DaaS) organization. This role is hybrid ( 3 days/week in office) in San Francisco, CA . The DaaS org is a delivery-first team that has more than tripled in size over the last six months, with no signs of slowing. They deliver high-quality data operations and AI deployment outcomes for frontier labs and AI teams. This org has a unique composition: forward deployed engineers, technical and operations delivery managers, a supply team managing a workforce comprised of multiple worker types at scale, and others. The people challenges here require an HRBP who has seen this kind of complexity before, such as workforce planning across FTEs and contractors, building a high performance culture rooted in delivery outcomes, and keeping a geographically dispersed, operationally complex team connected to Snorkel's culture. You'll partner directly with our DaaS GM and leadership team, and you'll need to be as comfortable in the operational weeds as you are in strategic conversations. The ideal background is professional services, managed services
C$135K – C$210K/yr
Overview: Guidepoint seeks an experienced Data/AI Engineer as an integral member of the Toronto-based AI team. The Toronto Technology Hub serves as the base of our Data/AI/ML team, dedicated to building a modern data infrastructure for advanced analytics and the development of responsible AI. This strategic investment is integral to Guidepoint’s vision for the future, aiming to develop cutting-edge Generative AI and analytical capabilities that will underpin Guidepoint’s Next-Gen research enablement platform and data products. This role demands exceptional leadership and technical prowess to drive the development of next-generation research enablement platforms and AI-driven data products. You will develop and scale Generative AI-powered systems, including large language model (LLM) applications and research agents, while ensuring the integration of responsible AI and best-in-class MLOps. The Senior AI/ML Engineer will be a primary contributor to building scalable AI/ML capabilities using Databricks and other state-of-the-art tools across all of Guidepoint’s products. Guidepoint’s Technology team thrives on problem-solving and creating happier users. As Guidepoint works to achieve its mission of making individuals, businesses, and the world smarter through personalized knowledge-sharing solutions, the engineering team is taking on challenges to improve our internal application architecture and create new AI-enabled products to optimize the seamless delivery of our services. This is a hybrid position based in Toronto. What You'll Do: Architect and Build Production Systems: Design, build, and operate scalable, low-latency backend services and APIs that serve Generative AI features, from retrieval-augmented generation (RAG) pipelines to complex agentic systems. Own the AI Application Lifecycle: Own the end-to-end lifecycle of AI-powered applications, including system design, development, deployment (CI/CD), monitoring, and optimization
Other cities to consider
More places hiring for this role
Get new model behavior engineer jobs in Canada by email
Daily job updates · Unsubscribe anytime