About the Team The Applied AI Engineering team partners closely with customers to help them move from experimentation to production with OpenAI’s technologies. We act as trusted technical advisors, working across customer strategy, architecture, deployment, and adoption to help organizations realize meaningful impact from frontier AI. The Startups segment serves fast-moving, high-growth companies that are often building new products, workflows, and businesses directly on top of AI. These customers move quickly, operate with high ambiguity, and expect practical, creative, and technically rigorous partnership. About the Role We are looking for an Applied AI Engineering Manager, Startups to lead and scale the Startups Applied AI Engineering motion. This team helps high-growth startups move quickly from experimentation to production, unlock meaningful usage, and build durable technical partnerships with OpenAI. This leader will operate in a high-velocity customer segment where founders, CTOs, and technical teams expect speed, judgment, and hands-on problem-solving. They will balance team leadership, technical depth, customer prioritization, and cross-functional influence across Sales, Product, Engineering, Research, and broader go-to-market teams. In this role, you will define how OpenAI supports startup customers at scale: identifying where deep technical engagement can unlock outsized impact, building repeatable deployment mechanisms, and ensuring the team can serve a broad and dynamic customer base without losing quality or strategic focus. In this role, you will: Craft and continuously refine the strategic vision and operating model for the Startups Applied AI Engineering team, aligning it with OpenAI’s broader company objectives and the evolving needs of high-growth startup customers. Lead, mentor, and grow a team of high-performing technical ICs supporting startup customers across AI-native, developer-led, and product-led companies. Help startups move from early e
Jobs in United States
Inference Technical Lead in San Francisco
268 active opportunities · Updated October 2026
Showing
15 jobs
Explore current inference technical lead jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI’s mission is to ensure the responsible and widespread adoption of artificial intelligence. In support of that mission, the Marketing team helps deeply understand customer audiences and market dynamics, influence the development of the right products, build sustainable and customer-aligned monetization models, and drive awareness, adoption, and usage across OpenAI’s products and platform. We take a data-driven approach to understand markets, develop monetization strategies, and uncover customer needs that shape product strategy and messaging. We partner closely with Sales, Partnerships, Product, Engineering, Research, Comms, and Design to deliver a cohesive end-to-end customer experience and lead go-to-market efforts for new product launches across channels. About the Role As a Product Marketing Manager, Research & Safety , you will shape how the world understands OpenAI’s frontier research and our approach to developing AI responsibly. You will establish the strategic narrative for some of the company’s most important research themes—including frontier capabilities, safety, preparedness, reasoning, and compute—and translate complex technical work into clear, credible stories for researchers, policymakers, and the broader public, working with cross-functional teams to translate that messaging and reach developers and enterprise. Working closely with Research, Safety Systems, Policy, Communications, and Product teams, you’ll lead the go-to-market strategy and messaging for major research and safety initiatives. You’ll develop narratives that build trust, translating technical concepts into tangible artifacts and messaging, drive industry understanding of emerging evaluation frameworks, and help communicate OpenAI’s unique perspective on responsible frontier AI development. You’ll also identify creative and non-traditional ways to bring these stories to the audiences that matter most. In this role, you will: Establish the strategic narrative f
Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we’re changing the way people think about and interact with personal finance. We’re a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we’re at the forefront. We’re proud to come to work every day knowing that what we do has a direct impact on people’s lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The Role At SoFi, we're building AI capabilities that fundamentally change how our teams operate, innovate, and serve our members. We're looking for a Senior Director, AI Enterprise Transformation to lead enterprise-wide AI adoption by partnering with technology and business leaders to identify, prioritize, and deliver high-impact AI initiatives. This is a highly cross-functional leadership role for someone who combines deep technical understanding of modern AI technologies with exceptional executive influence and change leadership. You will work across Engineering, Product, Enterprise Architecture, Data, Corporate IT, Operations, and business functions to accelerate AI adoption, establish best practices, and drive measurable business outcomes. Success in this role requires someone who can translate emerging AI capabilities into scalable enterprise solutions while inspiring teams across SoFi to embrace new ways of working. You will serve as both a strategic advisor and transformation leader, helping shape SoFi's AI roadmap and ensuring AI investments deliver meaningful impact across the company. What You'll Do Lead enterprise-wide AI transformation initiatives that improve operational efficiency, employee productivity and growth for SoFi. Partner with executive leadership to develop and execute SoFi's AI tr
From $214K/yr
Datadog is seeking a strategic, visionary, and results-oriented Senior Director, Growth Marketing – Organic Growth (SEO/GEO/PLG) to lead our organic acquisition strategy across traditional search engines and emerging AI/LLM platforms. This leader will own the vision, strategy, and operating model responsible for driving measurable growth in organic traffic and inbound pipeline through content programs, off-page authority, and AI/LLM discoverability initiatives. In this role, you will lead a team of organic growth specialists, while partnering closely with Website Experience, Product Marketing, and Engineering teams. You will define Datadog's long-term organic growth strategy, establish investment priorities, and ensure the organization is positioned to win across an increasingly complex discovery ecosystem. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own Datadog’s content-led SEO/GEO/PLG strategy, defining the topics, formats, and ecosystems that drive measurable growth in traffic and pipeline Lead, coach, and develop a high-performing team Identify and prioritize high-impact content opportunities across product areas and influence cross-functional teams to bring that content to life Map and optimize Datadog’s presence across the full ecosystem of LLM-ingested content (e.g., YouTube, Reddit, review sites) to improve AI-driven discoverability Design and execute a comprehensive off-page strategy, including link acquisition, digital PR, and authority-building initiatives Partner closely with the Website Experience team, who owns technical SEO, to ensure content is effectively surfaced, indexed, and performant Create and contribute to high-impact content (e.g., flagship pieces, new formats, or experimental channels), setting the standard for qualit
About the Team OpenAI’s Pricing team sits at the center of product, go-to-market, finance, and strategy. We define how OpenAI packages, prices, and scales access to our products across consumer, SMB, and enterprise customers, turning deeply technical product usage and market signal into company-level decisions. We’re looking for a senior Data Scientist to be the first dedicated data science hire on the Pricing team. This is a rare zero-to-one role with direct exposure to OpenAI’s CFO, Head of Pricing, and senior leaders across Product and GTM. You will help build the analytical foundation for pricing at OpenAI, shape executive decisions, and define what excellent pricing data science looks like. About the Role As a founding Data Scientist for Pricing, you will design the analyses, models, experiments, and decision frameworks that guide pricing strategy across OpenAI’s business. You’ll work side-by-side with the CFO, Head of Pricing, and senior leaders across Product and GTM on ambiguous, high-leverage questions where simple reporting is not enough, translating customer behavior, product usage, revenue outcomes, and market dynamics into clear recommendations. This role combines hands-on technical depth with executive-ready storytelling. You should be excited to build from first principles, operate with high independence, and influence decisions that shape how OpenAI grows and serves customers around the world. In This Role, You Will Serve as a senior analytical partner to the CFO, Head of Pricing, Product, and GTM leaders on pricing and monetization decisions. Build the analytical foundation for pricing across consumer, SMB, and enterprise segments, from exploratory analysis to repeatable decision systems. Design and execute analyses that connect customer behavior, product usage, conversion, retention, revenue outcomes, and pricing strategy. Develop models, algorithms, experiments, and decision frameworks for complex pricing, packaging, discounting, and willingness-t
About the Team The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust. About the Role As a Research Program Manager, you will lead initiatives that test the safety and robustness of OpenAI’s models through creative experimentation and structured evaluation. You’ll coordinate efforts across research and engineering teams to transform ambiguous risks into concrete research programs and influence future model development and deployment. We’re looking for people who are technically savvy, comfortable with ambiguity, and excited about shaping the future of safe AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead programs that explore unexpected model behaviors and identify failure modes. Translate vague or emergent risk signals into clear priorities and actionable research plans. Design and run creative evaluations, experiments, and red-teaming campaigns. Collaborate with research, product, and deployment teams to integrate findings into model training and deployment cycles. Develop repeatable systems for tracking model performance and understanding emerging behavior patterns. You might thrive in this role if you: Have strong experience in technical program management, with excellent organizational and communication skills. Are familiar with large language models, prompt engineering, or model evaluation techniques. Are comfortable managing fast-paced, high-uncertainty projects and shaping them from the ground up. Are creative and resourceful in devising new methods for testing model behavior and performance. Can effectively coordinate across technical and non-technical stakeholders to drive alignment and execution. About OpenAI OpenAI is an AI resear
About the Team OpenAI’s mission is to build safe artificial general intelligence (AGI) that benefits all of humanity. Achieving this requires bringing the world’s most exceptional talent under one roof to push the boundaries of what’s possible. Our Research Recruiting team plays a critical role in this effort. We are an embedded part of the research organization, working side by side with our research staff to deeply understand evolving priorities, build trust, and strategically shape the future of OpenAI’s talent. About the Role You will own and execute long-term talent strategies to identify, engage, and recruit many of the world’s leading and emerging AI researchers, research engineers, and technical scientists working at the frontier of machine learning. This is not a traditional execution-focused recruiting role. You will operate as a strategic partner to OpenAI’s research staff, helping define hiring priorities, shape search strategy, influence candidate evaluation, and guide hiring decisions that directly impact the direction and quality of our frontier-model research and fulfillment of our mission. In this role, you will: Partner directly with research and technical staff to define hiring priorities, shape search strategies, and anticipate future talent needs as technical roadmaps evolve. Proactively identify and cultivate exceptional AI/ML research talent across industry, academia, and emerging labs, often before formal hiring needs exist. Use market insights and candidate signals to influence hiring decisions, leveling, and compensation strategy for highly specialized research roles. Serve as a trusted advisor throughout candidate evaluation and closing — helping leaders calibrate for research excellence, long-term potential, and organizational fit. Collaborate closely with your sourcing partner to execute complex, high-impact searches in ambiguous or rapidly evolving technical domains. You might thrive in this role if you: Significant experience recruitin
About the Team OpenAI’s mission is to ensure that general-purpose artificial intelligence benefits all of humanity. The Payments team works across product, engineering, design, and finance to build the financial infrastructure that makes OpenAI’s products accessible to consumers and enterprises around the world. As AI introduces new ways for people and organizations to work, the team is defining how to support and monetize emerging forms of product usage, from usage-based pricing to agentic work. We’re building the foundational systems that help OpenAI products deliver clear, reliable, and scalable payment experiences while ensuring that this powerful technology is deployed responsibly. About the Role In this role, you’ll lead design for one of OpenAI’s most foundational product areas: the payments and monetization infrastructure that supports our consumer and enterprise products. You’ll partner closely with product, engineering, and cross-functional teams to shape how customers understand, manage, and pay for entirely new kinds of AI usage. Your work will extend beyond traditional checkout and billing. You’ll help define the systems, frameworks, and experiences behind durable pay-as-you-go models, Codex usage, and agentic workflows, translating complex business and technical requirements into intuitive experiences. As a product designer in a highly ambiguous and rapidly evolving space, you’ll influence both product strategy and the underlying infrastructure that OpenAI products depend on. This role is based in our San Francisco HQ. We offer relocation assistance to new employees. In this role, you will: Lead the design direction for foundational payments, billing, and monetization experiences across OpenAI’s consumer and enterprise products. Design and ship high-quality, end-to-end product experiences, from early systems and interaction concepts to high-fidelity prototypes and production-ready designs. Shape the infrastructure and product frameworks that support em
About the Team We bring OpenAI's technology to the world through products like ChatGPT and the OpenAI API. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role OpenAI is looking for an experienced Performance Engineer to help us scale the performance, reliability, and efficiency of our systems. In this role, you'll apply deep technical expertise to optimize infrastructure and application-level performance across mission-critical products like ChatGPT and our developer API. You’ll work cross-functionally with teams building core services, training models, and developing real-time user experiences to push our latency, throughput, and cost-efficiency to the next level. We are looking for engineers who thrive in ambiguous environments, value deep systems understanding, and are motivated by delivering measurable impact. This is a highly technical, individual contributor role focused on root-cause analysis, profiling, instrumentation, and architecture-level performance improvements across our stack. In this role, you will: Analyze and optimize performance across application, middleware, runtime, and infrastructure layers—networking, storage, Python runtime, GPU utilization, and beyond. Develop tooling and metrics that provide deep observability into system performance. Collaborate closely with infra, platform, training, and product teams to identify key performance goals and drive systemic improvements. Influence architecture and design decisions to prioritize latency, throughput, and efficiency at scale. Lead investigations into high-impact performance regressions or scalability issues in production. Drive performance testing strategies and help define SLAs/SLOs around latency and throughput for critical systems. You might thrive in this role if you: Have 7+ years of experience in software engineering with a strong tr
About the Team The Strategic Finance team at OpenAI plays a critical role in shaping the company’s long-term trajectory. We partner closely with Product, Engineering, and Go-To-Market teams to inform high-stakes decisions through rigorous data science and economic modeling. As part of our expanding Data Science function, we’re building a best-in-class Forecasting capability to drive real-time, data-driven decision-making across user growth, revenue, compute infrastructure, and more. We are developing scalable forecasting infrastructure to help us understand and anticipate business dynamics in an increasingly complex, usage-based world. Our models are foundational to planning, pricing, operational efficiency, and growth strategy - supporting key investment decisions and unlocking OpenAI’s full potential. About the Role We’re looking for a senior Machine Learning Data Scientist to lead our forecasting initiatives. You’ll be one of the founding members of the Forecasting pillar within Strategic Finance Data Science, responsible for building and scaling robust, interpretable, and production-ready forecasting systems. Your models will power critical business decisions by predicting core metrics such as DAU/WAU, revenue, LTV, compute consumption, and profitability. This is a highly cross-functional role, requiring technical excellence, strong product intuition, and business acumen. You’ll collaborate with product managers, researchers, engineers, and finance leaders to operationalize forecasting insights, influence company-wide strategy, and build foundational forecasting capabilities at OpenAI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Build statistical and machine learning models to solve forecasting needs across product, finance, infrastructure, and GTM domains. Own the end-to-end modeling lifecycle , including scoping, feature engineerin
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Vice President, Software Engineering Overview Decision Stream is Mastercard's next-generation AI-native decisioning platform, designed to power intelligent, real-time decisions across fraud, authentication, payments, and risk. Built with a startup mindset and enterprise-scale ambition, the platform combines innovation, speed, and engineering excellence to redefine decisioning across Mastercard's global ecosystem. We are seeking a visionary and hands-on VP of Software Engineering to help build and scale the platform. This leader will partner closely with Product, Architecture, AI, and Platform Engineering teams to drive technology strategy, architecture, engineering execution, and organizational growth. What You'll Do Lead Through Technical Excellence • Serve as a senior technology leader and role model for engineering teams. • Drive architecture, design, and technology decisions across distributed systems, streaming, AI, and cloud-native platforms. • Engage deeply with engineers, architects, and product leaders to solve complex technical challenges. • Influence engineering standards, software quality, and operational excellence. Build and Scale the Platform • Help shape and deliver a highly scalable, resilient, and secure decisioning platform operating at Mastercard scale. • Balan
About the team OpenAI’s mission is to build safe artificial general intelligence (AGI) which benefits all of humanity. This long-term undertaking brings the world’s best scientists, engineers, and business professionals into one lab together to accomplish this. In pursuit of this mission, our Go To Market (GTM) team is responsible for helping customers learn how to leverage and deploy our highly capable AI products across their business. The team is made of Sales, Solutions, Support, Marketing, and Partnership professionals that work together to create valuable solutions that will help bring AI to as many users as possible. About the Role Our GTM team is uniquely positioned to help customers realize the transformative potential of advanced AI models for their businesses and end users. As part of the GTM Strategy & Operations team, you’ll play a critical role in guiding the GTM strategy and driving the operational efficiency to accomplish this mission. This role serves as a trusted advisor to GTM leadership—providing data-driven insights, managing core operating cadences, and leading high-impact projects that influence how we engage with customers and scale our business. You’ll collaborate cross-functionally with Finance, Enablement, Data and Growth Strategy teams to align efforts, drive efficiencies, and accelerate growth. This role is a central role, meaning it is not tied to a specific business partner, but rather designing the global processes for the business. In this role, you'll: Design, build, and manage our operating rhythms (Forecasting, Pipeline Council, Big Deal Reviews, and MBRs/QBRs). Conduct strategic analyses to determine trends and identify opportunities for process and strategy optimization Collaborate with GTM leadership and cross-functional stakeholders to develop go-to market strategy and resource plans Lead strategic projects to improve efficiency and effectiveness across the revenue organization. Partner closely with technical teams to impl
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. Product at Baseten Product at Baseten is a nascent function. Our company today has a strong engineering culture, is heavily customer-obsessed, and moves fast. We're building the product function now, and you'd be one of the people who defines it. You'll work directly with our founders and with some of the best systems and AI engineers and you'll set the standard for what product looks like here. PMs at Baseten don't sit above engineers - you earn ownership by being technical, finding the truth in front of customers, building great cross-functional relationships, and just shipping great product experiences. The role Once a model is deployed, keeping it fast, reliable, and economical at scale is where production inference is won or lost. You'll own the surface that makes that happen: how deployments autoscale, how traffic is routed, how the system fails over, and how workloads scale across clusters and regions. You'll own these as products end to end - both how they work under the hood and how customers configure and observe them - and you'll help set and define the roadmap that infrastructure and product teams alike can build towards. This space is largely still evolving - think Cloud Infrastructure in mid-2000s. Your job is to make it 10x easier to reliably scale and serve AI models in production and set the market standard. Impact and outcomes you'll drive You will own how workloads scale and where they land — autosca
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a senior social media manager to own how Baseten shows up on social. Our audience is ML engineers, infrastructure teams, and technical founders, and most of them meet us first on X or LinkedIn, around a model launch, a benchmark, an open source release, or a customer result. This role decides what that first impression is. This is a senior individual contributor role. You set the strategy and you write the posts. Day to day you'll work with product marketing, comms, design, our engineers, and our founders. This role relies on technical credibility. You don't need an engineering background, but you do need to understand what we're claiming and why it matters. We post about latency, throughput, and GPU cost, and we hold ourselves to getting those details right. RESPONSIBILITIES This role is the face of the Baseten brand on our social channels and builds our direct line of communication with the community across X, LinkedIn, YouTube, and the communities where our audience already spends time. Track the conversation across AI and open source, and move quickly when we have something useful to add. Set the social strategy: what we post where, how each account grows, and how we measure it, with a clear point of view on which channels deserve investment and which don't. Translate technical work into posts worth sharing: model launches, benchmark results, open source projects, engineering deep dives, an
Other cities to consider
More places hiring for this role
Get new inference technical lead jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime