About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role Modal's LLM inference platform delivers frontier performance for open-source models with best-in-class elasticity and developer experience, made in part possible by our custom runtime with GPU memory snapshots and multi-cloud substrate . We're looking for a leader to own the direction and execution of this platform to continue to establish us as the clear market leader, working closely with customers like Cognition, Doordash, Ramp, and many more. You'll be leading a group of highly talented engineers working on our market-leading LLM inference offering, spanning the serving stack, routing infrastructure, internal agentic optimization platform, and the user-facing product surface area. This is a hands-on leadership role — expect to split your time between technical contribution, product shaping and people management depending on what the team needs. You'll set direct
Jobs in United States
Inference Engineering And Product Lead in San Francisco
222 active opportunities · Updated September 2026
Showing
15 jobs
Explore current inference engineering and product lead jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders About the Role We're hiring the first Account Managers at Modal. You'll report to the Regional Director of Account Management and be a founding member of the team. This function does not exist yet. There is no playbook, no territory map, no established motion. You'll own a book of business from day one and build the motion at the same time — from fast-moving AI startups to large enterprise teams running critical infrastructure on Modal. This is a commercial role with a revenue target. You'll be measured on retention and expansion across your accounts. While you won't be delivering the technical recommendations and implementation, the work is technical by nature. Our customers are engineers running GPU workloads, inference, and batch jobs in production, and you need to hold your own in those conversations. The profile we're hiring is a technical account manager. You've worked at companies that are deepl
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal builds AI infrastructure products that developers love. That's how we grew so quickly, and why word of mouth remains one of our most important channels today. In this role, you will primarily create and distribute technical content that is unique, educational, and practical. This content will be the first Modal touchpoint for many of our users. We want to not only showcase the power and developer experience of Modal, but also serve as a trusted resource for them when implementing new AI technologies. In this role, you will: Distill the latest advancements in AI technology and educate developers on how to incorporate them. Give demos/talks about Modal and adjacent tools at developer events. Engage with users in our community, both online (X, LinkedInReddit, Slack) and at in-person events. Build relationships, integrations, and joint marketing activities with o
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal builds the infrastructure that lets engineers run AI workloads without the usual pain. To do this well, we need exceptional people – and that’s where you come in. As the first dedicated GTM recruiter on our Talent team, you’ll own sales, GTM, and other G&A searches end-to-end. You’ll work closely with our Head of Talent, founders, and GTM leads to shape how we hire and help bring in the people who will define what Modal becomes. What you’ll do: Drive full-cycle recruiting for key hires across GTM and G&A functions (sourcing, pitching, guiding interviews, and closing candidates) Partner with GTM leaders to understand the real work and calibrate on what great looks like Help set our hiring bar and how we evaluate talent Execute creative top-of-funnel strategies that resonate with a strong community of experienced GTM talent Deliver a fast, respectful, h
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal is hiring a high-impact Solutions Architect to drive technical strategy across our most strategic enterprise accounts. You will operate as the executive technical counterpart to Enterprise Account Executives, leading complex evaluations, shaping infrastructure modernization roadmaps, and driving multi-product adoption across AI/ML workloads. This role is not demo support. It is a strategic, consultative position requiring strong architectural depth, executive presence, and the ability to influence 7–8 figure infrastructure decisions. You will work directly with CTOs, VPs of Engineering, and ML platform leaders to help them rethink how AI infrastructure should be built and operated. If you thrive in high-velocity technical sales environments and want to shape the infrastructure layer powering modern AI companies, this role is for you. What You’ll Do: Own the t
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal is the cloud platform built for AI. We're used by the world's leading AI labs, startups, and researchers to run compute-intensive workloads: training runs, inference, sandboxed code execution, and more. We're hiring a Community Manager in SF to make Modal a fixture in the AI developer community. You'll bring developers together through meetups, hackathons, and events of our own, and build the kind of community that keeps showing up. You know how to rinse and repeat the process, but always with a creative bend. In this role, you will: Co-host developer meetups with partners in our ecosystem. Find the right speakers, build the relationships, and run the events together. Prior examples: High Performance Inference for Open LLMs , Voice AI Builders Night , RL with Modal and Prime Intellect , FDE Happy Hour . Sponsor hackathons that attract highly technical enginee
AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. About the Role We're seeking a Revenue Operations Manager with a strong track record, a builder's mindset, and a bias for action to join our in-person team in New York or SF. This is a high-impact, hands-on role. You'll own the entire revenue operations function, from top-of-funnel lead routing through deal close and commission administration. You'll work closely with our Head of Finance & People Ops and sales leadership to build the systems, dashboards, and processes that scale our go-to-market motion. What You'll Do: Own the lead routing process from inbound and partnering with marketing to ensure proper attribution Run effective territory management & strategy for Geo based decisioning Support & strategise every aspect of revenue operations in your territory Own the strategy for capacity forecasting, inputs, throughputs & outputs being the conduit back to finance in
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is looking for a Product Designer to help shape the next generation of our platform. You'll join a small design team with a lot of ownership and a high bar for craft, working across the entire Baseten product to solve hard problems for technical users. This role covers the full arc of the work: early exploration through production. You'll establish patterns that show up across the product, evolve our design system, and raise the quality bar as we scale. RESPONSIBILITIES Lead product design end to end, including research, product definition, prototypes, design reviews, specs, and final implementation. Work directly with engineering and product to frame problems, explore solutions, and ship quickly. Collaborate with customers to understand their workflows, validate ideas, and test prototypes. Push the visual and interaction quality of the Baseten product across surfaces. Create interactions and details that make complex technical workflows feel clear, fast, and polished. Evolve our design system and establish patterns that scale across a growing product. Look across the product for opportunities to improve consistency, usability, and overall quality. Get into the code with engineers to polish UI and make sure the shipped experience matches the design. Help shape how the design team works and raise the bar for product quality across the company. REQUIREMENTS A portfolio that demonstrates strong visual and
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Lead large-scale brand campaigns across digital, events, and out-of-home. Partner with engineering and product marketing on major product launches. Turn complex technical ideas into clear, compelling visual communication. Evolve the Baseten visual identity into a brand system that scales. Own projects end-to-end, from concept through launch, collaborating directly with marketing, product, engineering, and leadership. Shape our employer brand and help define how new hires experience Baseten. Raise the bar for craft across every customer touchpoint. ABOUT YOU You have a portfolio of exceptional work with outstanding visual craft and attention to detail. You care deeply about quality and sweat the details. You communicate ideas clearly and thrive in collaborative environments. You know when to build systems, not just execute on assets. You default to ownership and are comfortable leading highly cross-functional projects from concept through launch. You're excited by technical products and know how to make complex ideas feel consumable without oversimplifying them. You thrive in a fast-moving environment with a high bar for quality. You have strong opinions about design and can articulate why something works - and why it doesn't. BONUS Motion design and animation experience Experience designing for developers or highly technical audiences. BENEFITS Competitive compensation, including meaningful equity. 100% covera
About the Team: The OpenAI API team builds the foundation that enables every developer to harness OpenAI’s models safely, reliably, and at scale. We design and operate the systems that power model serving, API access, billing, developer tooling, and enterprise integrations—forming the connective tissue between OpenAI’s research breakthroughs and real-world products. Our mission is to make it effortless for anyone to build with OpenAI technology. We’re responsible for the infrastructure and product layers that allow millions of developers to integrate GPT models, fine-tune behavior, manage data, and deliver transformative experiences to their users. We collaborate across product, research, and engineering teams to ensure that innovation in model capabilities translates directly into value for customers. The API team spans multiple disciplines, including product management, infrastructure engineering, developer experience, and data systems. We care deeply about reliability, scalability, and simplicity—creating tools that let developers focus on their ideas while we handle the complexity of running world-class AI systems. About the Role: We are seeking an experienced Product Manager to define and scale the construction of our data processing, data privacy, billing, and access controls products. You will set strategy and execute on projects like expanding our regional data processing footprint, enabling new inference caching controls in the API or building APIs that make it easier for organizations to manage their spend limits. You will also define the strategy and ship foundational capabilities that ensure customers use OpenAI products securely, privately, and with enterprise-grade controls. This role partners deeply with engineering, security, legal, compliance, finance and leadership to deliver high-trust, enterprise-grade systems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to n
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking an experienced SoC Architect to lead the definition and development of next-generation custom AI silicon for edge deployments. This role will be responsible for shaping the architecture of highly efficient, high-performance SoCs optimized for machine learning inference and on-device intelligence. You will work cross-functionally with internal engineering teams and external ecosystem partners to translate product requirements into scalable silicon solutions, driving execution from concept through delivery. In this role you will: Define the architecture and technical roadmap for custom SoCs targeted for edge applications. Drive system-level tradeoff analysis across compute, memory, interconnect, power, thermal, and cost constraints. Architect energy-efficient ML compute subsystems optimized for inference workloads and real-world deployment environments. Collaborate with internal hardware, software, systems, and product teams to align architecture with platform needs. Partner with external silicon vendors, IP providers, and manufacturing partners to execute development plans. Lead hardware/software co-design efforts to maximize performance per watt and end-to-end system efficiency. Guide implementation teams through microarchitecture, RTL development, validation, and bring-up phases. Operate effectively in agile development environments and help teams deliver against aggressive schedules and milestones. You might thrive in this role if: Proven exper
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Web Designer / Design Engineer to own baseten.co day to day, from concept through production. You'll design and build landing pages, bring product launches to life, run experiments, improve conversion, and keep raising the quality bar on the site. This is a high-ownership role on the brand design team, sitting between brand design and frontend engineering. We want the website to be one of the best expressions of the Baseten brand and one of the best developer-focused sites on the internet, which means treating it as a product that's constantly evolving rather than something we redesign every few years. You'll have a lot of autonomy to ship, while partnering closely with brand, product marketing, engineering, and leadership on bigger launches. RESPONSIBILITIES Own the design and frontend implementation of the Baseten marketing site. Ship landing pages, product pages, launches, and campaigns from concept through production. Improve the site continuously, finding and acting on opportunities to make it better rather than waiting for a redesign cycle. Partner with growth and product marketing to test messaging, layouts, and conversion improvements without sacrificing craft. Build polished interactions, motion, and details that make the site feel distinctly Baseten. Build and evolve the components, patterns, and design systems that let the team move quickly without lowering the bar. Turn complex
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking a Sales Manager to help lead a team of Strategic Account Executives. This person will be responsible for building and coaching a team of high-performing sales reps, driving revenue growth, and partnering closely with product and engineering to bring our cutting-edge AI infrastructure to customers. RESPONSIBILITIES Lead and mentor a team of Strategic Account Executives to consistently exceed pipeline and revenue goals. Help define and execution on go-to-market strategy for our fastest growing customer segment. Collaborate cross-functionally with Marketing, Product, and Engineering to align customer needs with Baseten’s product roadmap. Be deeply engaged with the product, enabling reps to have highly technical conversations with prospects and customers. Foster a culture of accountability, learning, and collaboration within the sales team. REQUIREMENTS 6+ years of closing sales experience, with 2+ years in management leading high-performing teams. Strong technical acumen, ideally with background in AI infrastructure, cloud infrastructure, or developer platforms. Comfortable operating in the weeds with technical products and guiding reps through complex deals. Proven track record of success in Strategic/Enterprise sales environments. Based in San Francisco or New York and open to coming in office at least 3 days per week (Tuesday-Thursday). BENEFITS Competitive compensation, including meaningful equ
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking a Sales Manager to help lead a team of Account Executives within our Startups segment. This hire will be responsible for building and coaching a team of high-performing sales reps, driving revenue growth, and partnering closely with product and engineering to bring our cutting-edge AI infrastructure to customers. RESPONSIBILITIES Lead and mentor a team of Startups Account Executives to consistently exceed pipeline and revenue goals. Help define and execution on go-to-market strategy for our fastest growing customer segment. Hire and scale the team by recruiting, interviewing, and onboarding top talent Collaborate cross-functionally with Marketing, Product, and Engineering to align customer needs with Baseten’s product roadmap. Be deeply engaged with the product, enabling reps to have highly technical conversations with prospects and customers. Foster a culture of accountability, learning, and collaboration within the sales team. REQUIREMENTS 4+ years of closing sales experience, with 2+ years in management leading high-performing teams. Strong technical acumen, ideally with background in AI infrastructure, cloud infrastructure, or developer platforms. Comfortable operating in the weeds with technical products and guiding reps through complex deals. Proven track record of success in high velocity sales environments. Based in San Francisco or New York City and open to coming in office at least 3 d
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for an Enterprise Account Executive to help build and scale Baseten’s Enterprise go-to-market motion. You’ll own new business in verticals where AI adoption is accelerating including enterprise software, financial services, and big tech. This is a high-autonomy role where you’ll contribute directly to our Enterprise GTM strategy—identifying new use cases within your verticals, winning lighthouse customers that become references for their industries, and feeding signal back to product and engineering to shape our roadmap. You’ll work alongside Baseten’s founders, forward-deployed engineering team, and GTM leadership. WHAT YOU’LL DO Own a revenue target and all aspects of the sales cycle from prospecting to close, including outbounding and engaging Tier 1 accounts in your assigned verticals Drive new logo acquisition and strategic expansion within key accounts, prioritizing organizations that can serve as lighthouse customers within their industries Become a trusted advisor to customers—understand their unique infrastructure needs, co-innovate on solutions, and lead with conviction by providing clear recommendations grounded in deep industry expertise Partner with Baseten’s forward-deployed engineers to run technical evaluations, POCs, and architecture reviews that earn customer confidence Collaborate cross-functionally with Product, Engineering, Legal, and Marketing to bring new solutions to marke
Other cities to consider
More places hiring for this role
Get new inference engineering and product lead jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime