Jobs in Canada

Engineering Architect in San Francisco

188 active opportunities · Updated October 2026

Explore current engineering architect jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

G
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$200K – $270K/yr

Quick readStrong listing-quality and freshness signals

About Glean: Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry’s most advanced enterprise search has evolved into a full-scale Work AI ecosystem, powering intelligent Search, an AI Assistant, and scalable AI agents on one secure, open platform. With over 100 enterprise SaaS connectors, flexible LLM choice, and robust APIs, Glean gives organizations the infrastructure to govern, scale, and customize AI across their entire business - without vendor lock-in or costly implementation cycles. At its core, Glean is redefining how enterprises find, use, and act on knowledge. Its Enterprise Graph and Personal Knowledge Graph map the relationships between people, content, and activity, delivering deeply personalized, context-aware responses for every employee. This foundation powers Glean’s agentic capabilities - AI agents that automate real work across teams by accessing the industry’s broadest range of data: enterprise and world, structured and unstructured, historical and real-time. The result: measurable business impact through faster onboarding, hours of productivity gained each week, and smarter, safer decisions at every level. Recognized by Fast Company as one of the World’s Most Innovative Companies (Top 10, 2025), by CNBC’s Disruptor 50, Bloomberg’s AI Startups to Watch (2026), Forbes AI 50, and Gartner’s Tech Innovators in Agentic AI, Glean continues to accelerate its global impact. With customers across 50+ industries and 1,000+ employees in more than 25 countries, we’re helping the world’s largest organizations make every employee AI-fluent, and turning the superintelligent enterprise from concept into reality. If you’re excited to shape how the world works, you’ll help build systems used daily across Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, and many more - deeply embedded where people get things done. You’ll ship agentic capabilities on an open, extensible stack, with the craf

AWSGCPGitAI
G
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$200K – $270K/yr

Quick readStrong listing-quality and freshness signals

About Glean: Glean is the Work AI platform that helps everyone work smarter with AI. What began as the industry’s most advanced enterprise search has evolved into a full-scale Work AI ecosystem, powering intelligent Search, an AI Assistant, and scalable AI agents on one secure, open platform. With over 100 enterprise SaaS connectors, flexible LLM choice, and robust APIs, Glean gives organizations the infrastructure to govern, scale, and customize AI across their entire business - without vendor lock-in or costly implementation cycles. At its core, Glean is redefining how enterprises find, use, and act on knowledge. Its Enterprise Graph and Personal Knowledge Graph map the relationships between people, content, and activity, delivering deeply personalized, context-aware responses for every employee. This foundation powers Glean’s agentic capabilities - AI agents that automate real work across teams by accessing the industry’s broadest range of data: enterprise and world, structured and unstructured, historical and real-time. The result: measurable business impact through faster onboarding, hours of productivity gained each week, and smarter, safer decisions at every level. Recognized by Fast Company as one of the World’s Most Innovative Companies (Top 10, 2025), by CNBC’s Disruptor 50, Bloomberg’s AI Startups to Watch (2026), Forbes AI 50, and Gartner’s Tech Innovators in Agentic AI, Glean continues to accelerate its global impact. With customers across 50+ industries and 1,000+ employees in more than 25 countries, we’re helping the world’s largest organizations make every employee AI-fluent, and turning the superintelligent enterprise from concept into reality. If you’re excited to shape how the world works, you’ll help build systems used daily across Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, and many more - deeply embedded where people get things done. You’ll ship agentic capabilities on an open, extensible stack, with the craf

AWSGCPGitAI
SA
📍 San Francisco, Canada· Hybrid
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! In September 2026 we raised a $350 million Series E at a $3.5 billion valuation , and we are scaling our engineering and research teams to meet demand. The role Frontier AI data is expensive to make and hard to measure. Every task we deliver is tested against the strongest models, often through many long-running agent rollouts. Your job is to make that process faster, cheaper, and more rigorous with ML and AI You will be one of the early members of ML & Research Engineering at Snorkel. You will study how frontier-grade data is generated and evaluated, form hypotheses, validate them against real production data, and ship the winners at scale. You will shape the discipline's direction, its standards, and the team that grows around it. What you'll work on Efficient agentic evals. Cut the cost of long-horizon agent evaluation with adaptive sampling, statistically grounded early stopping, model cascades, caching, and cheap-first gating. AI model routing. Route every eval and judge call to the cheapest model that clears the quality bar, with fallback, monitoring, and cost attribution. Fine-tuned small models. Fine-tune and serve open-weight models (LoRA and other

PythonMachine LearningAI
A
📍 San Francisco, Canada
✓ High-confidence listingCompany trend -86.2%
Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: Everyone at Airbnb thinks about trust, but our team obsesses over it daily. At the core of trust is safety, and thus we spend a significant amount of our time and energy keeping the community safe. The Trust team is responsible for developing the technology that helps protect our community and platform from fraud while also ensuring our hosts, guests, homes, and experiences meet our high standards. We constantly work to fight against online fraud (such as monetary loss, compromised accounts, spam and scam in messages, fake inventory, etc.) as well as offline fraud (theft, property damage, personal safety, etc.). We also work on onboarding and screening of users, and think about complex topics like identity and reputation to ensure that every interaction with Airbnb helps build trust in us and our community. You'll work side-by-side with talented product managers, data scientists, software engineers, fraud intelligence, and operations teams. Together, you'll design and build ML solutions that have direct, meaningful impact on user trust, business success, and the global Airbnb community. The Difference You Will Make: As a Senior Machine Learning Engineer on the Trust team, you will actively contribute code and ideas that shape the ML systems protecting millions of Airbnb users. You'll own and deliver ML projects end-to-end — from designing and training models to productionizing and operating them at scale, while collaborating closely with cross-functional partners. You'll tackle real-world challenges such as account takeover, fake accounts, payment fraud, and bot detection. Your work

PythonJavaMachine LearningRecruitment
DU
📍 San Francisco, Canada
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team At DoorDash we’re building the industry’s most scalable and reliable delivery network to support our three-sided marketplace of consumers, merchants, and dashers. TPMs on our Engineering teams provide fundamental support to the underlying technology services as well as the engineers building these services. It’s no simple task, but it wouldn’t be interesting if it was! About the Role Platform Services is one of DoorDash’s organizations responsible for extending DoorDash beyond the consumer marketplace. The team owns three distinct business lines: Drive, a white-label delivery API that enables external businesses to access DoorDash’s delivery network; Parcels, which provides last-mile logistics for retailers and merchants; and Enterprise (DoorDash for Business), a b2b product suite powering corporate meal programs and large ordering for organizations of all sizes. We are looking for a Senior Technical Program Manager to help scale Platform Services. You will lead important engineering wide initiatives across Enterprise, Drive, and Parcels, working with Engineering, Product, Operations, and business leaders to deliver dependable, extensible logistics capabilities for external businesses. Programs will span backend services, partner and merchant integrations, operational tooling, and platform interfaces. They require strong technical judgment, cross organizational coordination, stakeholder management, and ownership of results. You will report into the centralized Technical Program Management team in Engineering. You're excited about this opportunity because you will… Drive Platform Services’ Most Strategic Programs - You will partner closely with Engineering and Product leadership, Strategy & Operations, and cross-functional teams to turn Platform Services’ strategy into executable technical programs. The portfolio spans Drive (DoorDash’s white-label delivery offering), Parcels (last-mile logistics for merchants and retailers), and Enterprise (DoorDa

ExcelLogisticsRecruitment
A
📍 San Francisco, Canada· Full-time
✓ High-confidence listingCompany trend -86.2%

From $196K/yr

Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The Data Science team works closely with partners from Design, Product and Engineering across Airbnb’s product portfolio. Our work involves 0-to-1 innovation, applying sophisticated modeling techniques to critical Airbnb problems and finding new ways to leverage science for the good of the Airbnb community. Web prototypers work with Data Scientists to identify key questions during the development of these new solutions and build software prototypes that answer them. The work of DS prototyping covers everything from designing the right data visualization to focused explorations that help solve key design questions to broad explorations that help validate ideas. The Difference You Will Make: This role is focussed on partnering with the Data Science team to develop prototypes that show the potential of new data models targeting critical business and user problems. Step change advances in ML/AI have the potential to transform Airbnb, but doing that well requires not just building the right models, but also creating the right experience -- your work will be critical in enabling us to do that. Sophisticated and impactful models are often complex and hard to understand -- by building effective prototypes you can help close that gap and allow innovation to flourish. Your prototypes will enable leadership to make critical and strategic business decisions by demonstrating technical feasibility, suggesting implementation approaches, enabling user studies, and more. A Typical Day: As a web prototyper, you will create prototypes that enable differentiating and industry-leading produc

JavaScriptTypeScriptJavaReact
HI
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

$107.4K – $155.7K/yr

Quick readStrong listing-quality and freshness signals

Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The HP IQ Design team is dedicated to creating simple, delightful, and transformative user experiences. We're passionate about developing innovative yet familiar interfaces that feel like natural extensions of the human experience. We value collaboration across disciplines, from software engineering to hardware design, and we're excited to work with people who see technology as a tool to help solve real-world problems. As a Product Designer at HP IQ, you'll collaborate with a multidisciplinary team to explore and shape new possibilities in AI-powered digital experiences. The ideal candidate brings intense curiosity, a dedication to craft, and an openness to sharing, developing, and challenging ideas. You love to experiment, engage deeply in the design process, and continuously iterate on concepts until they feel "just right"—and even then, keep exploring ways to improve. As a member of the team, you'll have the opportunity to contribute across the digital design process, develop your craft alongside experienced designers and engineers, and help create new interfaces, interactions, and

JavaScriptJavaRedisGit
HI
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

C$107.4K – C$155.7K/yr

Quick readStrong listing-quality and freshness signals

Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role As a Product Design Engineer at HP IQ, you’ll work at the intersection of creativity, engineering, and design to develop groundbreaking devices that redefine how people interact with technology. You’ll collaborate closely with industrial designers, hardware engineers, and interaction designers to transform ideas into functional prototypes and production-ready designs. From 3D CAD modeling and hands-on prototyping to engineering analysis and design validation, you’ll help solve complex mechanical challenges and drive designs from early concepts toward production. You’ll have the opportunity to take ownership of meaningful engineering work while learning from an experienced, multidisciplinary team and contributing to products that push the boundaries of what’s possible. What You Might Do Design mechanical parts, components, and assemblies for new consumer devices using 3D CAD (NX) Create detailed 2D engineering drawings, define specifications and tolerances, and work closely with overseas vendor partners to ensure accuracy and quality Design and run experiments, applying engineering ana

RedisRestAISEM
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $252K/yr

Quick readStrong listing-quality and freshness signals

About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About Snorkel Snorkel AI is the frontier AI data lab, helping teams build the data and environments behind high-performing frontier and agentic AI. We combine technology with research-driven AI data development to create datasets, benchmarks, evals, and custom solutions for real-world AI systems. Founded out of the Stanford AI Lab in 2019, Snorkel works with leading AI labs and enterprises to move from better data to better outcomes. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About The Role Snorkel is hiring a Head of Security to build and lead our security function end-to-end — infrastructure security, application security, and governance, risk & compliance (GRC). You'll own the security function end-to-end — strategy, team, and execution — and operate as the primary security voice with customers, auditors, and the exec team. You'll report to the CTO. This is a builder's role: you'll take security from its current state to a mature, right-sized function as Snorkel scales, hiring and developing the team as needs grow. Key Responsibilities Security Leadership & Team Building Define Snorkel's overall security strategy,

AWSCI/CDAIGo
SC
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $110K/yr

Quick readStrong listing-quality and freshness signals

The Sigma Commercial Solutions Engineer believes in the power of analytics to transform organizations and uncover new data insights to increase business agility. The Sigma Solutions Engineer will act as the trusted advisor to our prospects and customers working in tandem with Sales, Business Development, Product Management, Customer Success, and Support. This individual is ultimately responsible for managing and delivering on all activities related to the technical presales cycle. This includes presentations, demonstrations, and hands on development of prototypes. You will align closely with our Commercial Sales Representatives focused on sales in new and existing accounts. What You Will Be Doing Understand and uncover business challenges and issues faced by the customer and be able to run targeted discovery sessions or workshops Engage with business users to define, create, and showcase solution prototypes Build and present customized demos for customers, trade shows, and webinars Confidently present and articulate the business value of the Sigma platform to all levels within an organization Deliver product, technical, and security related responses to RFPs/RFIs Participate in product, sales, and relevant technology certifications to acquire, maintain, and grow skill sets Work as a team player by contributing, learning, and sharing new knowledge Be conversant in integration and data migration approaches to help customers develop their data lifecycle and analytics strategy with Sigma Manage multiple customer engagements concurrently Attain quarterly and annual objectives assigned by management Become a Sigma champion and product expert; prospects and customers will look to you for advice and expertise Qualifications We Need Minimum 2 years of analytics, business intelligence, or sales engineering experience Customer relationship skills Strong understanding of database concepts Understanding of advanced spreadsheet concepts Cloud

PythonSQLAIGo
DU
📍 San Francisco, Canada
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team DoorDash is building the world’s most reliable on-demand logistics engine for delivery! We’re looking for machine learning engineer interns to join our fast-growing engineering team to help us develop a 24x7 global infrastructure system that powers DoorDash’s three-sided marketplace of consumers, merchants, and dashers. About the Role As a Machine Learning Engineer intern at Doordash, you’ll work on tackling new challenges in machine learning and artificial intelligence. You’ll conduct research that can be applied across Doordash engineering teams and engage in external collaborations and mentoring, while also performing research in any of the following areas: Auction, Game theory, Recommender systems, Ranking, AdTech, Computer Vision, Causal Inference, and Big data analytics. We offer a 12-week summer internship program in our San Francisco, Sunnyvale, New York, or Seattle offices. You’re excited about this opportunity because you will… Use cutting-edge research in ML/AI, NLP, RecSys, Ranking, Computer Vision, Causal Inference, Ad Tech, Graph analysis to solve real-world problems across discovery, ads,forecasting, fulfillment and search experiences at Doordash. Contribute and execute on research ideas that can be applied and used to improve product experience at Doordash. Collect, analyze, and synthesize findings from data and use these insights to build relevant ML models. Write clean, efficient, and sustainable code We’re excited about you because you… Are working towards a Masters degree in Computer Science, ML, NLP, Statistics, Information Sciences or related field and are graduating between Fall 2027 & Summer 2028 Have a mastery of at least one systems languages (Java, C++, Python, Kotlin, GoLang) or one ML framework (Tensorflow, Pytorch, MLFlow) Have experience in research and in solving analytical problems Are a strong communicator and team player. Have a passion for applied ML and the Doordash product Ideally have

PythonJavaMachine LearningArtificial Intelligence
DU
📍 San Francisco, Canada
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Team DoorDash is building the world’s most reliable on-demand logistics engine for delivery! We’re looking for machine learning engineer interns to join our fast-growing engineering team to help us develop a 24x7 global infrastructure system that powers DoorDash’s three-sided marketplace of consumers, merchants, and dashers. About the Role As a Machine Learning Engineer intern at Doordash, you’ll work on tackling new challenges in machine learning and artificial intelligence. You’ll conduct research that can be applied across Doordash engineering teams and engage in external collaborations and mentoring, while also performing research in any of the following areas: Auction, Game theory, Recommender systems, Ranking, AdTech, Computer Vision, Causal Inference, and Big data analytics. We offer a 12-week summer internship program in our San Francisco, Sunnyvale, New York, or Seattle offices. You’re excited about this opportunity because you will… Use cutting-edge research in ML/AI, NLP, RecSys, Ranking, Computer Vision, Causal Inference, Ad Tech, Graph analysis to solve real-world problems across discovery, ads,forecasting, fulfillment and search experiences at Doordash. Contribute and execute on research ideas that can be applied and used to improve product experience at Doordash. Collect, analyze, and synthesize findings from data and use these insights to build relevant ML models. Write clean, efficient, and sustainable code We’re excited about you because you… Are working towards a PhD degree in Computer Science, ML, NLP, Statistics, Information Sciences or related field and are graduating between Fall 2027 & Summer 2028 Have a mastery of at least one systems languages (Java, C++, Python, Kotlin, GoLang) or one ML framework (Tensorflow, Pytorch, MLFlow) Have experience in research and in solving analytical problems Are a strong communicator and team player. Have a passion for applied ML and the Doordash product Ideally have work

PythonJavaMachine LearningArtificial Intelligence
A
📍 San Francisco, Canada· Full-time
✓ High-confidence listingCompany trend -86.2%

From $132K/yr

Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join AirSupport helps Airbnb employees stay productive wherever and however they work. Within AirSupport, Executive Support provides a dedicated, high-touch technology experience for Airbnb’s senior leaders and their Executive Business Partners. The team combines deep technical expertise, proactive support, and strong cross-functional partnership across office, remote, travel, event, and other approved environments. Executive Support owns the executive technology experience, working closely with engineering, security, infrastructure, AV, workplace, and other technology partners who own the underlying technologies. The Difference You Will Make As a Senior Executive IT Support Engineer, you will be one of the most senior technical individual contributors within AirSupport and a trusted technology partner to Airbnb’s executive population. You will combine hands-on technical expertise with strong judgment, proactive planning, and cross-functional leadership. You will own complex and sensitive executive technology issues through resolution, anticipate risks before they become disruptions, and lead improvements that make the executive experience more reliable and supportable. Your impact will extend beyond the issues you personally resolve. You will identify patterns in what Executive Support is seeing, connect them to broader employee technology priorities, develop recommendations, and influence the teams responsible for the underlying technology. You will help shape Executive Support technical direction, represent the support perspective in major technology decisions, and identify opportuniti

AIRustExcelRecruitment
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $252K/yr

Quick readStrong listing-quality and freshness signals

Scale GP (Scale Generative AI Platform) is an enterprise-grade Generative AI platform that provides APIs for knowledge retrieval, inference, evaluation, and more. We are looking for a strong engineer to join our team and help us build and scale our product in a fast-paced environment. The ideal candidate will have a strong understanding of software engineering principles and practices, as well as experience with large-scale distributed systems. You will be responsible for owning large new areas within our product, working across backend, frontend, and interacting with LLMs and ML models. You will solve hard engineering problems in scalability and reliability. You will: Own large new areas within our product Work across backend, frontend, and interacting with LLMs and ML models Deliver experiments at a high velocity and level of quality to engage our customers Work across the entire product lifecycle from conceptualization through production Be able, and willing, to multi-task and learn new technologies quickly Ideally you'd have: 7+ years of full-time engineering experience, post-graduation Experience scaling products at hyper growth startups Experience tinkering with or productizing LLMs, vector databases, and the other latest AI technologies Proficient in Python or Javascript/Typescript, and SQL Experience with Kubernetes Experience with major cloud providers (AWS, Azure, GCP) Compensation packages at Scale for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position and may be inclusive of several career levels at Scale; it will be determined during the interview process based on work location and additional factors, including job-related skills, experience, qualifications, interview performance, and relevant education or training. Scale employees in eligible roles are also granted equity based compensation, subject to Board of Director approval

JavaScriptTypeScriptPythonJava
SA
📍 San Francisco, Canada· Full-time
✓ High-confidence listing

From $252K/yr

Quick readStrong listing-quality and freshness signals

About Scale AI At Scale, our mission is to develop reliable AI systems for the world's most important decisions. Our products provide the high-quality data and full-stack technologies that power the world's leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. Scale Frontier Data is the organization behind the training and evaluation data that frontier labs depend on. We build the systems, tooling, and expert workflows that turn hard human expertise into signals that models can learn from, across reasoning, coding, agentic tool use, and domain expertise. Reinforcement learning environments are now the center of gravity for that work: the difference between a model that demos well and a model that reliably completes long-horizon work is almost always the quality of the environments and reward signals it was trained against. Responsibilities As a Staff Software Engineer, RL Environments, you'll own the technical foundation for how Scale builds, runs, verifies, and delivers RL environments at scale. An RL environment is a real piece of software: a containerized world with real dependencies, real state, real tools, and a grader that has to be correct even when the agent is creative about breaking it. Building one is a full-stack engineering problem. Building thousands of them reproducibly, cheaply, with trustworthy reward signals and throughput measured in millions of rollouts is a systems problem that very few people have solved. You'll work on both. You'll design the platform: sandboxed execution, environment packaging and versioning, rollout orchestration, trajectory capture, verifier frameworks, and the authoring surfaces that let engineers and domain experts produce environments without reinventing infrastructure each time. And you'll go deep on the environments themselves by instrumenting real applications, designing task suites that expose specific capability gaps, and building graders that

TypeScriptPythonReactAWS
🔔

Get new engineering architect jobs in San Francisco, Canada by email

Daily job updates · Unsubscribe anytime