About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal is the cloud platform built for AI. We're used by the world's leading AI labs, startups, and researchers to run compute-intensive workloads: training runs, inference, sandboxed code execution, and more. We're hiring a Community Manager in SF to make Modal a fixture in the AI developer community. You'll bring developers together through meetups, hackathons, and events of our own, and build the kind of community that keeps showing up. You know how to rinse and repeat the process, but always with a creative bend. In this role, you will: Co-host developer meetups with partners in our ecosystem. Find the right speakers, build the relationships, and run the events together. Prior examples: High Performance Inference for Open LLMs , Voice AI Builders Night , RL with Modal and Prime Intellect , FDE Happy Hour . Sponsor hackathons that attract highly technical enginee
Jobs in United States
Inference Engineering And Product Lead in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current inference engineering and product lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. About the Role We're seeking a Revenue Operations Manager with a strong track record, a builder's mindset, and a bias for action to join our in-person team in New York or SF. This is a high-impact, hands-on role. You'll own the entire revenue operations function, from top-of-funnel lead routing through deal close and commission administration. You'll work closely with our Head of Finance & People Ops and sales leadership to build the systems, dashboards, and processes that scale our go-to-market motion. What You'll Do: Own the lead routing process from inbound and partnering with marketing to ensure proper attribution Run effective territory management & strategy for Geo based decisioning Support & strategise every aspect of revenue operations in your territory Own the strategy for capacity forecasting, inputs, throughputs & outputs being the conduit back to finance in
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is looking for a Product Designer to help shape the next generation of our platform. You'll join a small design team with a lot of ownership and a high bar for craft, working across the entire Baseten product to solve hard problems for technical users. This role covers the full arc of the work: early exploration through production. You'll establish patterns that show up across the product, evolve our design system, and raise the quality bar as we scale. RESPONSIBILITIES Lead product design end to end, including research, product definition, prototypes, design reviews, specs, and final implementation. Work directly with engineering and product to frame problems, explore solutions, and ship quickly. Collaborate with customers to understand their workflows, validate ideas, and test prototypes. Push the visual and interaction quality of the Baseten product across surfaces. Create interactions and details that make complex technical workflows feel clear, fast, and polished. Evolve our design system and establish patterns that scale across a growing product. Look across the product for opportunities to improve consistency, usability, and overall quality. Get into the code with engineers to polish UI and make sure the shipped experience matches the design. Help shape how the design team works and raise the bar for product quality across the company. REQUIREMENTS A portfolio that demonstrates strong visual and
From $295.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers. We’re looking for an EM to lead a team of exceptional ML infrastructure engineers, build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You Will: Lead strategic planning and roadmap execution of scalable production-ready ML systems including model training, data pipelines, feature engineering and model inference. Own the architecture, establish engineering best practices of scalability, reliability, and cost-effectiveness of ML infrastructure (e.g., training, serving, feature). Work closely with data scientists, ML engineers, platform teams, and product stakeholders to design, implement, and operate robust ML platforms that accelerate model development and deployment. Recruit, mentor, and grow a high-performing team of ML infrastructure engineers. You Have: 5+ years of experienc
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Lead large-scale brand campaigns across digital, events, and out-of-home. Partner with engineering and product marketing on major product launches. Turn complex technical ideas into clear, compelling visual communication. Evolve the Baseten visual identity into a brand system that scales. Own projects end-to-end, from concept through launch, collaborating directly with marketing, product, engineering, and leadership. Shape our employer brand and help define how new hires experience Baseten. Raise the bar for craft across every customer touchpoint. ABOUT YOU You have a portfolio of exceptional work with outstanding visual craft and attention to detail. You care deeply about quality and sweat the details. You communicate ideas clearly and thrive in collaborative environments. You know when to build systems, not just execute on assets. You default to ownership and are comfortable leading highly cross-functional projects from concept through launch. You're excited by technical products and know how to make complex ideas feel consumable without oversimplifying them. You thrive in a fast-moving environment with a high bar for quality. You have strong opinions about design and can articulate why something works - and why it doesn't. BONUS Motion design and animation experience Experience designing for developers or highly technical audiences. BENEFITS Competitive compensation, including meaningful equity. 100% covera
About the Team: The OpenAI API team builds the foundation that enables every developer to harness OpenAI’s models safely, reliably, and at scale. We design and operate the systems that power model serving, API access, billing, developer tooling, and enterprise integrations—forming the connective tissue between OpenAI’s research breakthroughs and real-world products. Our mission is to make it effortless for anyone to build with OpenAI technology. We’re responsible for the infrastructure and product layers that allow millions of developers to integrate GPT models, fine-tune behavior, manage data, and deliver transformative experiences to their users. We collaborate across product, research, and engineering teams to ensure that innovation in model capabilities translates directly into value for customers. The API team spans multiple disciplines, including product management, infrastructure engineering, developer experience, and data systems. We care deeply about reliability, scalability, and simplicity—creating tools that let developers focus on their ideas while we handle the complexity of running world-class AI systems. About the Role: We are seeking an experienced Product Manager to define and scale the construction of our data processing, data privacy, billing, and access controls products. You will set strategy and execute on projects like expanding our regional data processing footprint, enabling new inference caching controls in the API or building APIs that make it easier for organizations to manage their spend limits. You will also define the strategy and ship foundational capabilities that ensure customers use OpenAI products securely, privately, and with enterprise-grade controls. This role partners deeply with engineering, security, legal, compliance, finance and leadership to deliver high-trust, enterprise-grade systems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to n
About the Team OpenAI's Industrial Compute organization builds and operates the infrastructure required to train and serve frontier AI models. The Capacity Planning team connects rapidly changing research and product demand with the compute, networking, storage, power, data center, hardware, and operational resources required to make that demand executable. About the Role We are seeking a Technical Program Manager to build and lead capacity planning across OpenAI's large-scale AI infrastructure. You will translate uncertain workload demand into clear infrastructure requirements, allocation decisions, supply commitments, activation priorities, and long-range capacity strategies. This role sits at the intersection of research, engineering, infrastructure, finance, sourcing, deployment, and operations. You will create the planning models, operating cadences, governance mechanisms, and source-of-truth systems that allow teams to understand what capacity is required, what is available, what is at risk, and what decisions must be made. This is not a finance-only forecasting or reporting role. Success requires technical fluency across the infrastructure stack, strong analytical judgment, and the ability to move consequential decisions forward when requirements, timelines, and supply conditions change quickly. Key Responsibilities Own capacity-planning processes across near-term workload allocation, quarterly execution, and longer-range infrastructure horizons. Translate research, training, inference, and product demand into compute, accelerator, cluster, networking, storage, rack, power, and site requirements. Develop scenarios that make assumptions, confidence levels, constraints, sensitivities, and decision points explicit. Reconcile requested demand against contracted, delivered, installed, activated, and workload-usable capacity. Partner with research and engineering teams to understand workload priorities, technical dependencies, utilization patterns, and changing req
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking an experienced SoC Architect to lead the definition and development of next-generation custom AI silicon for edge deployments. This role will be responsible for shaping the architecture of highly efficient, high-performance SoCs optimized for machine learning inference and on-device intelligence. You will work cross-functionally with internal engineering teams and external ecosystem partners to translate product requirements into scalable silicon solutions, driving execution from concept through delivery. In this role you will: Define the architecture and technical roadmap for custom SoCs targeted for edge applications. Drive system-level tradeoff analysis across compute, memory, interconnect, power, thermal, and cost constraints. Architect energy-efficient ML compute subsystems optimized for inference workloads and real-world deployment environments. Collaborate with internal hardware, software, systems, and product teams to align architecture with platform needs. Partner with external silicon vendors, IP providers, and manufacturing partners to execute development plans. Lead hardware/software co-design efforts to maximize performance per watt and end-to-end system efficiency. Guide implementation teams through microarchitecture, RTL development, validation, and bring-up phases. Operate effectively in agile development environments and help teams deliver against aggressive schedules and milestones. You might thrive in this role if: Proven exper
From $221.4K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. WHY DATA SCIENCE & ANALYTICS? The Data Science & Analytics organization’s mission is to increase our speed, frequency and acumen of making decisions at scale by instilling a data-influenced approach to building products. We cover a wide area of the data spectrum including analytical data engineering, product analytics, experimentation, causal inference, statistical modeling and machine learning. Aligned and partnering with product verticals, we use this extensive toolbelt to discover new opportunities and unmet use cases, influence and shape the product roadmap and prioritization, build data products and measure impact on our community of players and creators. WHY CREATOR SERVICES? At Roblox, the Creator Services team enables unbounded creation through reliable core services and novel AI applications. As a Senior Data Scientist focused on Machine Intelligence , you will bridge the gap between high-tier engineering infrastructure and cutting-edge ML applications. You will be the primary strategic partner to product and engineering leadership, transforming unstructured data into actionable business insights and user-facing products. This is a "zero-to-one" environment. You will be tas
Datadog's integrations are the connective tissue between our platform and the technologies our customers run in the real world. As a Sr. PM on the Agent Integrations team, you will own the vision, prioritization, and execution for 100+ integrations that run directly inside the Datadog Agent from foundational infrastructure (MySQL, Kafka, Kubernetes) to the rapidly growing landscape of self-hosted AI and on-premise enterprise technologies. This is a high-impact, breadth-first role at the intersection of infrastructure observability and the frontier of AI-native workloads. At Datadog, we place value in our office culture; the relationships it builds, the creativity it brings, and the collaboration of being together. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own the Agent Integrations roadmap. Determine which new integrations to build and which existing ones to improve, balancing customer demand, business impact, and engineering capacity across a catalog of 100+ technologies. Drive the expanding AI integration surface. Lead product strategy for self-hosted AI workloads, including LLM inference frameworks (e.g., Hugging Face TGI, BentoML), AI agents, MCP servers, and model orchestration tools, so Datadog customers can monitor every layer of their AI stack. Expand on-prem and hybrid coverage. Prioritize and execute new integrations for on-prem technologies including storage systems, HPC schedulers, network devices, and legacy enterprise platforms where customers run critical workloads. Build observability for ERP systems. Define and drive Datadog's strategy for monitoring enterprise ERP platforms (SAP, Oracle EBS/Fusion, Microsoft Dynamics) covering performance, job execution health, and integration layer telemetry so enterprise customers can observe their ERP stack alongside the rest of their infrastructure. Analyze adoption and customer feedback at scale. Use data from multiple sources to
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Web Designer / Design Engineer to own baseten.co day to day, from concept through production. You'll design and build landing pages, bring product launches to life, run experiments, improve conversion, and keep raising the quality bar on the site. This is a high-ownership role on the brand design team, sitting between brand design and frontend engineering. We want the website to be one of the best expressions of the Baseten brand and one of the best developer-focused sites on the internet, which means treating it as a product that's constantly evolving rather than something we redesign every few years. You'll have a lot of autonomy to ship, while partnering closely with brand, product marketing, engineering, and leadership on bigger launches. RESPONSIBILITIES Own the design and frontend implementation of the Baseten marketing site. Ship landing pages, product pages, launches, and campaigns from concept through production. Improve the site continuously, finding and acting on opportunities to make it better rather than waiting for a redesign cycle. Partner with growth and product marketing to test messaging, layouts, and conversion improvements without sacrificing craft. Build polished interactions, motion, and details that make the site feel distinctly Baseten. Build and evolve the components, patterns, and design systems that let the team move quickly without lowering the bar. Turn complex
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking a Sales Manager to help lead a team of Strategic Account Executives. This person will be responsible for building and coaching a team of high-performing sales reps, driving revenue growth, and partnering closely with product and engineering to bring our cutting-edge AI infrastructure to customers. RESPONSIBILITIES Lead and mentor a team of Strategic Account Executives to consistently exceed pipeline and revenue goals. Help define and execution on go-to-market strategy for our fastest growing customer segment. Collaborate cross-functionally with Marketing, Product, and Engineering to align customer needs with Baseten’s product roadmap. Be deeply engaged with the product, enabling reps to have highly technical conversations with prospects and customers. Foster a culture of accountability, learning, and collaboration within the sales team. REQUIREMENTS 6+ years of closing sales experience, with 2+ years in management leading high-performing teams. Strong technical acumen, ideally with background in AI infrastructure, cloud infrastructure, or developer platforms. Comfortable operating in the weeds with technical products and guiding reps through complex deals. Proven track record of success in Strategic/Enterprise sales environments. Based in San Francisco or New York and open to coming in office at least 3 days per week (Tuesday-Thursday). BENEFITS Competitive compensation, including meaningful equ
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking a Sales Manager to help lead a team of Account Executives within our Startups segment. This hire will be responsible for building and coaching a team of high-performing sales reps, driving revenue growth, and partnering closely with product and engineering to bring our cutting-edge AI infrastructure to customers. RESPONSIBILITIES Lead and mentor a team of Startups Account Executives to consistently exceed pipeline and revenue goals. Help define and execution on go-to-market strategy for our fastest growing customer segment. Hire and scale the team by recruiting, interviewing, and onboarding top talent Collaborate cross-functionally with Marketing, Product, and Engineering to align customer needs with Baseten’s product roadmap. Be deeply engaged with the product, enabling reps to have highly technical conversations with prospects and customers. Foster a culture of accountability, learning, and collaboration within the sales team. REQUIREMENTS 4+ years of closing sales experience, with 2+ years in management leading high-performing teams. Strong technical acumen, ideally with background in AI infrastructure, cloud infrastructure, or developer platforms. Comfortable operating in the weeds with technical products and guiding reps through complex deals. Proven track record of success in high velocity sales environments. Based in San Francisco or New York City and open to coming in office at least 3 d
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for an Enterprise Account Executive to help build and scale Baseten’s Enterprise go-to-market motion. You’ll own new business in verticals where AI adoption is accelerating including enterprise software, financial services, and big tech. This is a high-autonomy role where you’ll contribute directly to our Enterprise GTM strategy—identifying new use cases within your verticals, winning lighthouse customers that become references for their industries, and feeding signal back to product and engineering to shape our roadmap. You’ll work alongside Baseten’s founders, forward-deployed engineering team, and GTM leadership. WHAT YOU’LL DO Own a revenue target and all aspects of the sales cycle from prospecting to close, including outbounding and engaging Tier 1 accounts in your assigned verticals Drive new logo acquisition and strategic expansion within key accounts, prioritizing organizations that can serve as lighthouse customers within their industries Become a trusted advisor to customers—understand their unique infrastructure needs, co-innovate on solutions, and lead with conviction by providing clear recommendations grounded in deep industry expertise Partner with Baseten’s forward-deployed engineers to run technical evaluations, POCs, and architecture reviews that earn customer confidence Collaborate cross-functionally with Product, Engineering, Legal, and Marketing to bring new solutions to marke
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for an Enterprise Account Executive to help build and scale Baseten’s Enterprise go-to-market motion. You’ll own new business in verticals where AI adoption is accelerating including enterprise software, financial services, and big tech. This is a high-autonomy role where you’ll contribute directly to our Enterprise GTM strategy—identifying new use cases within your verticals, winning lighthouse customers that become references for their industries, and feeding signal back to product and engineering to shape our roadmap. You’ll work alongside Baseten’s founders, forward-deployed engineering team, and GTM leadership. WHAT YOU’LL DO Own a revenue target and all aspects of the sales cycle from prospecting to close, including outbounding and engaging Tier 1 accounts in your assigned verticals Drive new logo acquisition and strategic expansion within key accounts, prioritizing organizations that can serve as lighthouse customers within their industries Become a trusted advisor to customers—understand their unique infrastructure needs, co-innovate on solutions, and lead with conviction by providing clear recommendations grounded in deep industry expertise Partner with Baseten’s forward-deployed engineers to run technical evaluations, POCs, and architecture reviews that earn customer confidence Collaborate cross-functionally with Product, Engineering, Legal, and Marketing to bring new solutions to marke
Other cities to consider
More places hiring for this role
Get new inference engineering and product lead jobs in United States by email
Daily job updates · Unsubscribe anytime