Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Metronome, now part of Stripe, is the leading usage-based billing platform built for modern software companies. With Metronome, companies can launch products faster, offer any pricing model, and streamline finance workflows without writing code. Our platform computes millions of invoices per billing period and is scaling rapidly to accommodate new customers, saving them hours of development time and manual invoicing and enabling them to use consumption data to better serve their customers. Our customers love our product and approach, and we’re humbled to work with amazing companies like OpenAI, NVIDIA, Confluent, and Anthropic. What you’ll do As a member of our technical support engineering team, you will be on the front lines providing world-class customer service. As part of our engineering organization, you will become an expert on our product and partner closely with Metronome's engineers, customer success, solutions architecture, and growth teams, as well as our customers' developers. Your primary responsibilities will include handling customer escalations through our ticketing system, using internal observability tools to diagnose and scope customer-facing issues, collaborating directly with customers via Slack and other channels, and developing internal tools and documentation to improve the support experience. Since we view every support escalation as an opportunity to learn, both as individuals and as a company, you will play a
Jobs in Canada
Software Engineer Infrastructure in San Francisco
149 active opportunities · Updated October 2026
Showing
15 jobs
Explore current software engineer infrastructure jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
$150K – $210K/yr
Location: San Francisco, CA (Hybrid) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role We're seeking an experienced Senior Optimization Engineer to join our Data Science Team. In this role, you will lead the design, development, and deployment of optimization models that power our software platform across applications including electricity markets, renewable energy, and battery energy storage systems. You will be responsible for developing production-grade optimization engines that solve complex operational and planning problems at scale. This role requires deep expertise in mathematical optimization, strong software engineering skills in Python, and experience building optimization models that integrate with production systems. The ideal candidate has significant experience in the energy industry, particularly electricity markets and battery storage optimization. This position emphasizes technical leadership, ownership of complex optimization projects, and collaboration across engineering, product, and commercial teams to deliver high-impact optimization solutions. Key Res
About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by millions of people, then we want to talk to you! About the Role We are hiring a Firmware Validation & Integration Engineer for our autonomy software team. This is a critical role to build robust and scalable validation for our firmware and systems to ensure reliability at every level. In this role, you will work with our electrical, firmware, and autonomy engineers to build the infrastructure and test suites required to validate the system. This includes designing and implementing our Hardware-in-the-Loop (HIL) simulation environments and automation frameworks from the ground up. You will report to the Autonomy Platform Lead on our Autonomy Platform Team at DoorDash Labs. We expect this role to be hybrid with some time in-office and some time remote. You’re excited about this opportunity because you will… Play an integral role on a small and focused team. Design and build Hardware-in-the-Loop (HIL) systems to simulate vehicle dynamics and sensor data for comprehensive firmware and system-level validation. Develop automated test infrastructure and software tools to exercise multiple embedded platforms throughout our robot system. Interface many layers of our control system including vehicle controls, power management, and motion control to ensure seamless system integration. Implement low-level test sequences and validation algorithms to safely stress-test vehicle components such as batteries, drive-train, and thermal management devices. Collaborate with cross-functional teams to identify edge cases and hardware-software corner cases that impact vehicle safety and performance. We’re excited about you because… BS/MS degree in Computer Science, Robotics, Electrical Engineering, or related technical field. 5+ years of experience in validati
Role Overview We are seeking a Staff Simulation Engineer to build an end-to-end aerial autonomy simulation stack at DoorDash Labs. This is a highly technical, hands-on leadership role focused on defining and implementing the simulation architecture that underpins autonomy development, validation, CI/CD testing, and pilot training. You will operate as the technical authority for simulation: owning core architecture decisions, developing key components yourself, and setting engineering standards. You will build and mentor a small, high-caliber simulation team while remaining deeply involved in implementation and system design. This role is ideal for someone who has built simulation systems from first principles, understands simulator internals deeply, and is excited to create a world-class platform from scratch. Key Responsibilities Architect and implement an end-to-end simulation stack for aerial autonomy at DoorDash Labs.. Develop high-fidelity simulation capabilities, including: Flight dynamics modeling Contact modeling and constraint handling Sensor and perception simulation Autonomy software-in-the-loop (SITL) integration Design and implement scalable simulation infrastructure to support: Regression testing in CI/CD pipelines Continuous validation of flight autonomy and autopilot software stack Mission-level testing and scenario generation Build cloud-deployed simulation systems to enable large-scale parallel testing and pilot training. Partner closely with autonomy, controls, and aircraft teams to ensure simulation fidelity and validation alignment. Establish technical direction, architecture standards, and performance benchmarks for simulation. Mentor and grow a small team of simulation engineers while remaining deeply hands-on. Required Qualifications Master’s or PhD in Computer Science, Electrical Engineering, Mechanical Engineering, Robotics, Aerospace Engineering, or a related field. 10+ years of experience in robotics or physics-based simulation. Deep expe
From $302.4K/yr
Director of Engineering, Physical AI Role Overview The Director of Engineering will report to the General Manager of Physical AI, and will be responsible for leading a multi-disciplinary engineering organization. In this senior leadership role, you will own the execution of the Physical AI Data Engine — the platform powering the next generation of Physical AI/Embodied AI. You will collaborate closely with Operations and GTM to guide product direction and help solve the data bottleneck that stands between today's robotics research and real-world deployment. This role requires significant ownership in a fast-paced environment and you will motivate internal teams to set the pace for business growth. Travel will come into play. Key Responsibilities: Set and drive the technical vision across data collection infrastructure, teleoperation systems, ML training pipelines, model evaluation frameworks, annotation tooling, and research Lead a multidisciplinary engineering organization—spanning engineering managers, software engineers, ML engineers, and ML research scientists—while designing the organizational structure, talent strategy, and culture required to scale rapidly without compromising on quality or strategic alignment Maintain exceptional technical and operational excellence by deeply understanding team deliverables, asking incisive questions, identifying slipping standards early, and knowing precisely when to step in Drive cross-functional alignment across Engineering, Operations, and GTM on platform architecture, release processes, and shared priorities Collaborate with researchers and clients to architect and deliver scalable, production-grade data infrastructure tailored for complex robotics workloads Required Qualifications: Bachelor's degree in Engineering, Robotics, Computer Science, or a related technical field 8+ years of engineering experience in fast-paced environments, including 4+ years direct people management demonstrated history of recruiting, mentorin
From $24K/yr
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About the Role & Team We’re looking for an Engineering Manager to lead the Data Infrastructure team within Statsig Experiment at Amplitude. You will lead a multidisciplinary team of software engineers, data engineers, and data scientists responsible for the systems that power experimentation at scale. The team owns three critical areas: Data ingestion: Collecting and importing experiment exposures, custom events, OpenTelemetry data, and real user monitoring data across SDKs, streaming systems, cloud storage, and customer data warehouses. Data computation: Building distributed computation systems that transform raw data into accurate, timely experiment results. Stats engine: Developing and productionizing the statistical methods that help customers make trustworthy decisions from their experiments. This is not a traditional data engineering management ro
C$45 – C$51/hr
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About the Role HP IQ’s Security Team is creating something the world has never seen before. We are attempting something truly impactful — innovating at the deepest levels of hardware and launching a service that will inspire users to experience computing in an entirely new way. Privacy and security are not just priorities; they are fundamental to our product and essential to our success. What You Might Do We are seeking a software engineering intern ready to take on the challenge of securing users’ devices, users’ data, and HP IQ’s infrastructure. We see security as the key to accomplishing what other companies cannot. If you thrive at balancing exceptional user experiences with strong security, this is the place for you. Essential Qualifications Pursuing a degree (Bachelors or Graduate) in Computer Science or related technical field Strong software engineering skills Ability to thrive in a collaborative environment Interest in privacy & security Interest in product design & development Experience with C, C++, Java or Python Preferred Skills An understanding of security concepts
$195K – $1.2M/yr · Jobiba est.
At Scale, our mission is to develop reliable AI systems for the world's most important decisions. For 10 years, Scale has provided the high-quality data and full-stack technologies that power the world's leading models, and has helped enterprises and governments build, deploy, and oversee AI applications that deliver real impact. We work closely with industry leaders like Meta, Ernst & Young, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force. Scale's internship is not a side project. Interns own real, shipped work on the same roadmaps as full-time engineers, with mentorship from world-class talent and a culture that values ownership, speed, and truth-seeking. Many of our interns return as full-time Scaliens. Example Projects Build reinforcement learning and post-training data pipelines that power frontier model development Develop evaluation infrastructure that measures model reliability for enterprise and public sector customers Ship agentic AI applications and the tooling that makes them observable, testable, and safe to deploy Ship tools that accelerate the growth of new qualified contributors on Scale's platform Build fraud-detection systems that remove bad actors and keep Scale's contributor base safe and trusted Use models to estimate the quality of tasks and contributors, and guarantee quality on requests at large scale Devise advanced matching algorithms that pair contributors to customers for optimal turnaround and accuracy Create optimized and efficient UI/UX tooling, in combination with ML algorithms, for 100k+ contributors completing billions of complex tasks Develop new AI infrastructure products to visualize, query, and explore Scale data Requirements A graduation date in Fall 2027 or Spring 2028 with a Bachelor's degree (or equivalent) in a relevant field (Computer Science, EECS, Computer Engineering, Statistics) Available for a Summer 2027 internship (May/June start dates) in San Franci
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Lyft Infrastructure builds the systems engineers depend on to ship stable, scalable, and efficient services. We're hiring a Senior Technical Program Manager to run cross-functional programs across our infrastructure and data platform teams. This role blends program delivery with product sense: you'll own the roadmap for your area, set priorities, and act as the voice of the customer back into how we build. Responsibilities Run infrastructure programs end to end, from kickoff through delivery Own the roadmap for your platform area: shape the strategy, sequence the work, and make the prioritization calls Drive data platform migration and modernization work, coordinating across engineering, data, and platform teams to keep dependencies and timelines under control Be the voice of the customer: partner with engineering teams across Lyft, surface their pain points, and feed that back into priorities and roadmaps Define success metrics and adoption goals, gather and document customer requirements, and make sure what ships actually solves the problem Build feedback loops with customer teams and turn what you hear into concrete improvements Partner with engineering and infrastructure leads to build plans, call out risks early, and keep stakeholders aligned Own program health: track milestones, surface blockers before they slip, and keep decision-makers in the loop Use your technical background in distributed systems and data infrastructure to ask sharp questions and build plans the team believes in Share in the team's release oncall rotation Experience 5+ years in Technical Program Management or a TPM/PM hybrid role A background in software, data, or systems engineering, enough to go deep with engineers Experience owning a roadmap: setting strategy, prioritizing across competing demands, and defining what suc
$160K – $185K/yr
About the role Every enterprise is racing to build AI-powered apps and agents but speed without the right runtime creates chaos, not transformation. Sigma is the AI Runtime Environment that makes those apps governable, scalable, and real, and the Sr. PMM, Sigma Apps will play a critical role in developing and executing the go-to-market narrative and strategy for Sigma Apps. This role requires someone who understands both sides of the enterprise software conversation: the IT leaders and data engineers who evaluate and govern application infrastructure, and the line-of-business owners who care about outcomes, speed, and usability. You'll translate Sigma's application development capabilities into stories that resonate with both and create the materials the sales team needs to tell those stories in the field. What you'll do Develop and maintain positioning and messaging for Sigma Apps including no-code app building, AI-assisted workflow automation, embedded analytics, and apps built with external coding agents working closely with the Director of Product Marketing, Apps and the Product team. Support new feature launches and capability expansions, coordinating across product, design, marketing, and sales to bring new capabilities to market clearly and effectively. Work closely with Product and Engineering to deeply understand and influence the Sigma Apps roadmap bringing market, customer, and competitive insights that help shape and prioritize it. Partner with sales reps and the Enablement team to understand what's working in the field and use those insights to sharpen messaging, update battlecards, and improve enablement materials. Partner with the Enablement team to build and maintain sales enablement content solution briefs, pitch decks, use case guides, competitive comparisons, and discovery question frameworks that help the field confidently sell Sigma Apps Build and maintain competitive analysis and battlecards for the no-code/low-code, embedded anal
From $134.4K/yr
At Scale, we believe that the next frontier of artificial intelligence is embodied. The Physical AI team is focused on building general AI that can reason and act in the physical world. By leveraging Scale’s massive, industry-leading data infrastructure, we are partnering with frontier labs to build Foundation Models for Physical AI that will redefine the future of automation. To support our rapid hardware-software iteration cycles and ensure a world-class R&D environment, we are looking for a Safety Coordinator / Lab Lead to anchor our physical testing operations. Role Overview As the Safety Coordinator / Lab Lead , you will play a mission-critical role in scaling our physical testing infrastructure safely and efficiently. This is a high-impact position where your highest-priority responsibility will be owning the end-to-end execution of safety audits and incident documentation . Operating at the intersection of cutting-edge AI foundation models and complex robotics hardware, you will ensure our researchers, engineers, and autonomous systems interact in a secure, compliant, and highly organized environment. Core Responsibilities Priority Focus: Safety Audits & Incident Documentation Rigorous Safety Audits: Design, schedule, and execute routine safety audits across all physical testing environments, robot cells, and hardware workspaces to ensure continuous compliance with internal benchmarks and industrial safety standards. Incident & Near-Miss Documentation: Own the end-to-end incident management pipeline. Act as the primary point of contact for documenting, archiving, and analyzing any lab incidents, mechanical anomalies, or near-misses. Root-Cause Analysis (RCA): Lead structured post-incident investigations to identify systematic risks, authoring comprehensive RCA reports and implementing Corrective and Preventive Actions (CAPA). Data-Driven Risk Mitigation: Treat safety data as a core operational asset—tracking safety metrics and audit trends to proa
From $1.6M/yr
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. Software Engineer II, Growth at Amplitude (View all jobs) San Francisco Bay Area About The Role & Team The Growth organization is focused on helping users realize long-term value from Amplitude. The team plays a critical role in that journey by driving expansion within existing accounts. Many of our customers start with just a few users — often in product or data teams. Our job is to unlock value for everyone else. We build features that help new users set up, get activated, leverage AI for frictionless insights, highlight relevant parts of the product they might otherwise miss, and expand usage across teams and departments. This is especially impactful in large, complex organizations, where getting from the first few users to broad adoption is a multiplier for retention and revenue. We work across Amplitude’s core product suite and integrations,
From $170K/yr
Airtable is the no-code app platform that empowers people closest to the work to accelerate their most critical business processes. More than 500,000 organizations, including 80% of the Fortune 100, rely on Airtable to transform how work gets done. Airtable’s mission is to bring the power of computing and software development to everyone. We are developing a powerful and extensible toolkit that our customers can leverage to solve a variety of different problems and workflows. We’ve seen our most sophisticated customers use the product to run global processes across thousands of employees, coordinate precision manufacturing pipelines, and consolidate previously siloed mission-critical data into a single source of truth. The complexity of these use cases requires us to be extremely thoughtful about how we design and implement new functionality in the product and make sure it’s both easy to use and comprehend for our customers and maintainable for us. As a Full-Stack, Backend engineer at Airtable, you will have the opportunity to work with customers to deeply understand their needs and workflows. You will collaborate with cross-functional partners across product management, design, research and data science to create innovative new features that enable our customers to do their best work. You will be responsible for owning and executing the end-to-end implementation of these new features that will contribute to making our toolkit even more powerful and successful. We currently have openings on: The Admin & Governance Team (Full-Stack/BE) ensures Airtable is secure, compliant, and enterprise-ready. It owns key admin capabilities like the Admin Panel, SSO, and audit systems, as well as foundational features like User Groups. This team's mission is to accelerate organizational value for the largest customers with enterprise-first governance and controls. The Omni Capability & Quality Team (Full-Stack/BE) brings the power of AI directly to Airtable end users—
From $1.6M/yr
About the Team The Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company at significant scale and continue to expand the workloads, capabilities, and consumer base we serve. Orchestrating and operating thousands of Spark cluster deployments is a complex distributed system problem which the team invests heavily in runtime optimization, systems architecture, multi-tenant scheduling, and end-user tooling. About the Role As a Software Engineer on Spark Platform, you will execute across the surfaces of our in-house Spark deployment that serves the entire company. The work spans Spark runtime upgrades and performance, multi-tenant scheduling and executor bin-packing on Kubernetes, cluster lifecycle automation, and the observability and incident automation that keep the platform sustainable. You will move between layers as the work demands — picking up the next high-leverage problem regardless of where it sits — and partner closely with the rest of the team and with platform consumers across the company. You must be located in San Francisco, Sunnyvale, Seattle, or New York City for this hybrid position. You will report into the Engineering Manager on our Spark Platform team. You're excited about this opportunity because you will… Build and operate an in-house Spark platform that runs at company-wide scale, spanning runtime, scheduler, reliability, and user-facing tooling. Drive multi-tenant scheduling, executor bin-packing, and cost-aware placement that let a small team serve dozens of consumer teams. Own pieces of cluster lifecycle automation — provisioning, upgrades, capacity changes, and node-failure handling — at a scale where these stop being manual events. Build the observability and incident automation that make the platform debuggable end-to-end and keep on-call sus
About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Role Join Twitch’s Memberships team within the Commerce Engineering organization, where we’re building rewarding experiences allowing creators to make a living doing what they love. We’re the team behind our continuous patronage features including channel Subscriptions, Gifting, and Turbo. As a member of our team, you’ll work alongside our highly engaged and collaborative team to design, build and maintain systems that scale to millions of concurrent users. We actively seek to improve our experiences and are looking for members that are passionate about our end users. Our team is based in San Francisco, CA and Seattle, WA. You Will: Create interactive experiences that are rewarding for both viewers and creators Architect and build robust, scalable applications that can handle millions of concurrent users Participate in Operational Excellence work to maintain and support our live services Collaborate with fellow engineers, product managers and designers to build new products and solutions You Have: 1+ years of professional software development experience with a focus on building scalable systems Excellent proficiency in modern programming languages (Python, Java, Go) and distributed system technologies A track record of building product experiences that users love Sharp proble
Higher-paying openings
Jobs with higher listed pay
Staff Software Engineer, Lyft Business
Lyft · San Francisco, CA
Senior ML Software Engineer, Mapping
Lyft · San Francisco, CA
Staff Software Engineer
Amplitude · San Francisco, Canada
Staff Software Engineer - DevX
Amplitude · San Francisco, Canada
Senior Staff Software Engineer - Pricing and Packaging
Gusto, Inc. · San Francisco, CA
ML Software Engineer, ETA
Lyft · San Francisco, CA
Related career options
Similar roles with stronger pay
Demand 50/100 · 8 jobs
$1.3M – $1.3M/yr
Salary →Demand 55/100 · 20 jobs
$972.9K – $972.9K/yr
Salary →Other cities to consider
More places hiring for this role
Get new software engineer infrastructure jobs in San Francisco, Canada by email
Daily job updates · Unsubscribe anytime