Jobs in Canada

Senior Engineering Manager Machine Learning Infrastructure Ads in San Francisco

182 active opportunities · Updated October 2026

Explore current senior engineering manager machine learning infrastructure ads jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

57/100

steady · 52 related jobs

Hiring trend

-7.4%

Job postings compared with the previous 30 days

Remote options

11.5%

Share of matching jobs listed as remote

Typical salary

$216K – $216K/yr

Based on 7 salary observations

V
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$260K – $290K/yr

Quick readStrong listing-quality and freshness signals

About VSCO For years, we've helped photographers create their work. Now we're building what comes next. VSCO exists for photographers. Not as a side feature, not as an afterthought, but as the whole point. We build the connected system photographers rely on, and we've spent over a decade earning the trust of a global creative community that takes the craft seriously. Photography is at an inflection point. AI is reshaping what's possible for creative work, and that's where our mission shines. VSCO is building the full photographer's workflow: from creating and editing your work, to delivering it to clients, to running your business. All of it built thoughtfully, with photographers leading the way. We believe the future of photography tools creates more space for creativity, handling the busy work so photographers can focus on the craft. If you care about craft, community, and what technology can unlock for creative people, this is the work. We're a mission-driven and focused company where your work ships quickly, is meaningful, and reaches tens of millions of people worldwide. You'll have a real say in what we build and how we build it. We hire people who don't wait to be asked, naturally connect the dots, care about the quality of what they ship, and believe the best outcomes come from building together. About The Role VSCO is hiring a Senior Staff Engineer to own Reflex, our production GPU image-processing engine, and the path it takes into Studio Pro on iOS and macOS. This is a critical technical leader. You will deepen a stack that is already in customers’ hands: new capabilities (RAW, large-image export, desktop), production quality (color correctness, memory, performance), and the contracts that let product engineers ship on top of the engine without becoming graphics programmers. You will also lead company-wide engineering mentorship and force-multiplication — tools, workflows, and AI-assisted development that make the rest of the team more effective. In This

ReactRestAIGo
V
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$190K – $220K/yr

Quick readStrong listing-quality and freshness signals

About VSCO For years, we've helped photographers create their work. Now we're building what comes next. VSCO exists for photographers. Not as a side feature, not as an afterthought, but as the whole point. We build the connected system photographers rely on, and we've spent over a decade earning the trust of a global creative community that takes the craft seriously. Photography is at an inflection point. AI is reshaping what's possible for creative work, and that's where our mission shines. VSCO is building the full photographer's workflow: from creating and editing your work, to delivering it to clients, to running your business. All of it built thoughtfully, with photographers leading the way. We believe the future of photography tools creates more space for creativity, handling the busy work so photographers can focus on the craft. If you care about craft, community, and what technology can unlock for creative people, this is the work. We're a mission-driven and focused company where your work ships quickly, is meaningful, and reaches tens of millions of people worldwide. You'll have a real say in what we build and how we build it. We hire people who don't wait to be asked, naturally connect the dots, care about the quality of what they ship, and believe the best outcomes come from building together. About The Role VSCO is hiring a Senior Software Engineer to be a primary contributor on Reflex, our production GPU image-processing engine, and significantly guide the path it takes into Studio Pro on iOS and macOS. You will deepen a stack that is already in customers’ hands: new capabilities (RAW, large-image export, desktop), production quality (color correctness, memory, performance), and the contracts that let product engineers ship on top of the engine without becoming graphics programmers. In This Role, You Will Be a primary contributor on the Reflex imaging engine: DAG execution, WGSL operators, GPU resource/memory budgets, and color management (linear workin

ReactAIGoRust
R
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $190.8K/yr

Quick readStrong listing-quality and freshness signals

Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 130 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com . The Media Experience team exists to empower users to discover, consume, and engage with the best media content on Reddit. We do that by making video consumption intuitive and engaging. Because video is a powerful medium for community connection, our goal is to ensure our video experience is best-in-class for every user. Our north star is growing video views and video users . These are users who consistently return to Reddit for the unique, authentic content that fuels our communities. Achieving this level of consistent engagement requires a product flywheel that surfaces the right video features, UI optimizations, and playback improvements at the right moment to accelerate organic engagement. We use data, Quality of Experience (QoE) signals, and rapid product experimentation to identify and resolve any friction preventing a seamless viewing experience. We move with urgency, instrument everything, and treat a missed engagement milestone as a signal to learn from, not a number to ignore. If you want to build iOS video experiences that measurably change the trajectory of tens of millions of users and improve the overall engagement of Reddit itself, this is your team. What you’ll do: 🚀 Drive the Technical Strategy for Media Engagement Lead the iOS architecture for user-facing features that boost video discovery and consumption. You’ll design and build immersive viewing surfaces and intuitive UI/UX flows that directly maximize watch time, user retention, and daily active video consumers. 🔬 Champion for User Experience With

RestAISwiftGo
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$150K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on the modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing Solve challenging problems that arise in providing high performance interactive experience to enable analytics and workflows use cases on top of modern warehouses Build software using the latest developer tools and using programming languages like Rust, Go, GraphQL, Typescript Develop new algorithms and techniques for improving the performance and interactivity for enabling analytics and workflows for the world largest companies Triage product or system issues and debug/track/resolve by analyzing the sources of issues Design and implement new software features to support our fast growing user growth Collaborate with cross-functional groups - infrastructure, design, product, customer support, sales and marketing to build an innovative product capabilities Qualifications We Need 5+ years industry experience building and maintaining high-quality software Experience building and deploying robust and secure web applications in a continuous deployment environment Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong Computer Science fundamentals Qualifications We Want (also, skills you’ll learn!) Experience building software capabilities for analyzing large scale data web applications Data driven aptitude and its application to solve distributed system problems Data model design, and API development experience to enable customer facin

TypeScriptPythonSQLGraphql
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$240K – $270K/yr

Quick readStrong listing-quality and freshness signals

About the Role At Sigma, we’re not just adding AI—we’re building the future of how people work with data. Our platform already lets users explore billions of rows of data in seconds with a spreadsheet-like interface, analyze and present their data in workbooks, and build data apps and workflows. Now we’re pushing further, applying AI to reshape how people build in Sigma, discover insights, and make smarter decisions—fast. That’s where you come in. As an AI/ML Engineer, you’ll join a growing team focused on building the AI foundation that will power Sigma for the future. Your work will become an integral part of the workflow for the thousands of enterprises that run on Sigma. What You’ll Do Partner with product, design, and engineering teams to identify high-impact AI/ML opportunities Prototype and productionize AI systems that feel intuitive but do a lot under the hood—recommendations, natural language interfaces, agentic workflows, and more Develop and scale AI/ML infrastructure that powers both internal tooling and customer-facing features Tackle novel UX problems at the intersection of AI, BI, and apps What You Bring Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related field (required) 10+ years of experience building and deploying production-grade AI/ML systems Deep knowledge of machine learning, deep learning, and applied AI Experience across the full ML lifecycle: data curation, training, deployment, monitoring A track record of building things that ship—whether it’s recommendations, search, machine translation, or something equally complex Experience adapting or training foundation models (language or multimodal) for novel domains Bonus Points (or skills you’ll build here) You've built agents that can plan, reason, and use tools You know your way around cloud infrastructure (AWS, GCP, Azure) You’ve worked in a fast-moving startup or high-growth environment Additional Job details The base salary range for this posit

PythonSQLAWSAzure
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$170K – $240K/yr

Quick readStrong listing-quality and freshness signals

Senior Software Engineer - Observability and Reliability About the Role We are growing the engineering team and looking for engineers who have the chops to build and deliver world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible. What You Will Be Doing Build observability tools and platforms, including: metrics, logging, distributed tracing, dashboarding, alerting, application performance management Build with modern tools and languages like Go, Open Telemetry and Kubernetes Participate in on-call rotation and ensure uptime of services Create runtime tools/processes that optimize cloud triaging and limit downtime Define best practices around making our systems and services measurable Collaborate with peers and stakeholders through design and code reviews to ensure best practices amongst available technologies. We expect successful candidates to be coding a majority of their time Qualifications We Need Strong Computer Science fundamentals 5+ years industry experience building and maintaining high-quality software, especially software other engineers use You apply a product mindset to infrastructure systems and feel accomplished enabling others Desire to be a great teammate and have fun at work Strong sense of craftsmanship, and a healthy academic curiosity Qualifications We Want (also, skills you’ll learn!) Experience building systems for data analytics Distributed systems monitoring and profiling skills Knowledge of cloud application security models Administered cloud service infrastructure (GCP, AWS, Azure) Startup experience Additional Job details Additional Job details The base salary range for this position is $170k - $240k annually. Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience. Base pay is one part of the Total Package that is provided to compensate and recognize e

PythonSQLAWSAzure
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$170K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing You will be responsible for developing elegant and responsive user experience using the latest front-end technologies. You'll own substantial pieces of the product, from design to launch Working with our product, UX design, and backend development teams, you will develop new features and technologies that make our product experience awesome and radically simplify the user experience for non-technical users You will leverage your technical expertise in front-end application development in the creation of novel visualizations for structured and unstructured data and develop new techniques for improving the performance and interactivity of the application Use modern frontend frameworks like React, GraphQL, TypeScript and Node.js Qualifications We Need 5+ years industry experience building and maintaining high-quality software An eye for great design and a passion for building products that provide a great user experience The ability to make the right trade-offs between functionality and delivery speed that supports delivering value to customers, all the while iterating based on feedback and roadmap priorities Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong computer science fundamentals Qualifications We Want (also, skills you’ll learn!) Experience building software capabilities for analyzing large scale data web applications Prior exp

TypeScriptPythonReactNode.js
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$170K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Role Sigma is transforming how businesses run by delivering a high performance platform on the modern data architecture. Hence, we are growing the engineering team and looking for engineers who are excited to solve challenging problems, deliver impactful capabilities throughout our stack to build world-class technology. You will be part of a talented team of engineers with a shared mission to make data easily accessible for all users. What You Will Be Doing Solve challenging problems that arise in providing an interactive experience on data warehouses for data exploration and analysis Build with modern tools and languages like Rust, Go, GraphQL, Node, and Kubernetes Build backend distributed services, new algorithms and modern API to support a cloud application Triage product or system issues and debug/track/resolve by analyzing the sources of issues Design and implement new software features to support our fast growing user base Collaborate with peers and stakeholders through design and code reviews to ensure best practices amongst available technologies Qualifications We Need 5+ years industry experience building and maintaining high-quality software Experience building and deploying robust and secure web applications in a continuous deployment environment Desire to be a great teammate and have fun at work without compromising ownership towards your work Strong sense of craftsmanship, and a healthy academic curiosity to solve challenges at sigma Strong Computer Science fundamentals Qualifications We Want (also, skills you’ll learn!) Data driven aptitude and its application to solve distributed system problems Data model design, and API development experience SQL query optimization and database internals Administered cloud service infrastructure (GCP, AWS, Azure) Prior experience working at high growth company solving technical problems to enable continued success Additional Job details The base salary range for this position is $170k - $240

PythonSQLAWSAzure
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $180K/yr

Quick readStrong listing-quality and freshness signals

About Scale At Scale AI, our mission is to accelerate the development of AI applications. For 8 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including: generative AI, defense applications, and autonomous vehicles. With our recent Series F round, we’re accelerating the abundance of frontier data to pave the road to Artificial General Intelligence (AGI), and building upon our prior model evaluation work with enterprise customers and governments, to deepen our capabilities and offerings for both public and private evaluations. About Data Engine Our Generative AI Data Engine powers the world’s most advanced LLMs and generative models through world-class RLHF (Reinforcement Learning with Human Feedback), human data generation, model evaluation, safety, and alignment. The data we are producing is some of the most important work for how humanity will interact with AI. Our Approach As part of the interview process, you’ll be considered for opportunities across several teams within the GenAI Engineering organization, based on your interests, expertise, and business needs. Potential team placements include Allocation, Growth, Frontier Data, Trust & Safety, Pay, Operator, or Tasking Experience. Together, these teams power Scale’s AI data operations - from building high-impact datasets that push the boundaries of LLM capabilities, to optimizing contributor onboarding and incentives, to safeguarding data integrity through advanced trust, safety, and security measures. They work at the intersection of ML, operations, and analytics to ensure we deliver the highest-quality data at scale. Responsibilities: Design, build, and maintain robust, scalable systems across the full stack, including front-end, back-end, and infrastructure layers Implement high-impact features using modern technologies such as TypeScript, React, Node.js, MongoDB, Elasticsearch, and Temporal Collaborate closely with internal operators (your use

TypeScriptPythonReactNode.js
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $180K/yr

Quick readStrong listing-quality and freshness signals

Scale GP (Scale Generative AI Platform) is an enterprise-grade Generative AI platform providing APIs for knowledge retrieval, inference, evaluation, and more. We are seeking a strong Senior Full-Stack Engineer to help us build, scale, and refine our rapidly growing product. The ideal candidate is deeply grounded in software engineering best practices and experienced in developing and scaling modern web applications end-to-end. You will work across the stack—from React/TypeScript frontends to Python-based backends—while integrating with LLMs and machine learning systems. You will solve complex challenges in scalability, reliability, and product experience while owning significant product areas in a fast-paced environment. What You’ll Do Own major full-stack product areas , driving features from design through production deployment. Build modern frontend experiences using React and TypeScript, ensuring performance, usability, and responsiveness. Develop reliable backend services in Python, working with distributed systems, data pipelines, and ML/LLM components. Integrate with LLMs, vector databases, and AI infrastructure to power intelligent product experiences. Deliver experiments and new features quickly , maintaining high quality and tight feedback loops with customers. Collaborate across product, ML, and infrastructure teams to shape the direction of Scale GP. Adapt quickly —learning new technologies, frameworks, and tools as needed across the stack. Ideal Experience 5+ years of full-time engineering experience , post-graduation. Strong experience developing full-stack applications using React, TypeScript, and Python . Experience scaling or shipping products at high-growth startups . Familiarity with LLMs, vector databases, embeddings, or other modern AI tooling (tinkering or production experience welcome). Proficiency with SQL and modern API development. Experience with Kubernetes , containerization, and microservice architectures. Experience working with at leas

TypeScriptPythonReactSQL
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $1.6M/yr

Quick readStrong listing-quality and freshness signals

About the Team DoorDash’s GenAI Platform team sits within Machine Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and Deliveroo teams safely bring GenAI-powered products, agents, automation, and personalization to production. Our mission is to increase the velocity of business impact from GenAI. A central pillar of that work is running frontier open-weight LLMs and VLMs (such as GLM, Qwen, Kimi, and DeepSeek) ourselves — real-time GPU serving, high-throughput batch inference, and fine-tuning on autoscaling GPUs — delivering large cost and latency wins (for example, a billion embeddings produced roughly 20× cheaper and visual models served roughly 72% cheaper). We also own core platform surfaces including the LLM Gateway, Agent Gateway, evals infrastructure, guardrails, and cost attribution. About the Role You will join a small, high-leverage team building production infrastructure for Generative AI at DoorDash, leading the design and architecture of our open-weights model platform spanning inference and fine-tuning: real-time GPU serving, high-throughput batch inference, and model fine-tuning. You’ll set technical direction across model serving and inference engines, fine-tuning and training pipelines, GPU autoscaling and utilization, batch pipelines, backend services, and observability, and mentor engineers as you go. This role is ideal for a senior engineer who enjoys owning ambiguous, high-impact systems and pushing the cost/performance frontier of GPU inference and fine-tuning in a fast-moving technical area where product needs, model capabilities, vendor ecosystems, and cost/performance tradeoffs are evolving quickly. You’re excited about this opportunity because you will… Lead the design of infrastructure that helps DoorDash teams move GenAI ideas from prototype to production, increasing the velocity of business impact from AI across the company. Own and evolve our open-weights serving stack — real-time GPU endpoints, high-thr

PythonAWSGCPKubernetes
DU
📍 San Francisco, Canada· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team The Consumer Engineering Team is responsible for helping consumers discover and order everything they love globally. Our work spans the entire consumer journey across homepage, search, store discovery, item exploration, checkout and post checkout. We aim to craft a hyper-personalized, delightful and frictionless experience for millions of our customers. About the Role As a Senior Staff Machine Learning Engineer on Core Cx, you will set the personalization (P13n) strategy for the entire consumer shopping journey and bring that strategy to life. You will use our robust data and machine learning infrastructure to implement new ML solutions to make the consumer search experience more relevant, seamless, and delightful across restaurant, grocery, retail and all business at DoorDash . You will modernize the recommendation system leveraging AI. You will demonstrate a strong command of production level machine learning, experience with solving end-user problems, and collaborate well with multi-disciplinary teams. You're excited about this opportunity because you will… Drive the engineering vision, strategy, and execution for an organization of 150+ Grow, build, and nurture impactful business-focused product engineering teams. Scale the team by developing leaders internally and attracting world-class talent Mentor and guide a fast-growing organization in setting the right architectural patterns, working with various vendors in the space, and making judicious investments in the right areas anticipating what the company needs a few years down the road. Partner with Business, Product, and other Engineering teams to transform DoorDash from local commerce to agentic commerce We're excited about you because you have… B.S. or M.S. in Computer Science or equivalent. 10+ years of industry experience developing machine learning models with business impact, and shipping ML solutions to production. Proficiency in using AI coding tools (e.g., Claude Code) in th

AWSGitRestMachine Learning
🔔

Get new senior engineering manager machine learning infrastructure ads jobs in San Francisco, Canada by email

Daily job updates · Unsubscribe anytime