Jobs in United States

Inference Technical Lead in San Francisco

268 active opportunities · Updated October 2026

Explore current inference technical lead jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI's Enterprise team builds AI-powered enterprise products and shared platform capabilities that help organizations put advanced AI to work securely and at scale. Our work spans enterprise workflows, agent experiences, integrations, identity, administration, security, governance, and deployment. About the Role As a Technical Program Manager on Enterprise, you will lead the technical strategy and execution behind the products and shared capabilities that make ChatGPT, Codex, and future OpenAI products useful, secure, and scalable for organizations. You will translate customer needs, competitive dynamics, and product priorities into actionable plans, influence architectural direction, and deliver durable capabilities across application, platform, and infrastructure layers. The role requires deep technical fluency, strong product judgment, and the ability to move between hands-on execution and broader enterprise strategy. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Drive technical strategy and execution for enterprise product and AI workflow initiatives, from design through implementation, launch, customer rollout, and iteration. Partner with engineering teams to influence architectural direction, interface definitions, and implementation tradeoffs across full-stack products, APIs, integrations, and shared platform systems. Translate enterprise customer requirements into actionable product priorities across AI-powered workflows, agent experiences, integrations, permissions, data access, evaluations, identity, security, governance, and deployment readiness. Represent the needs of enterprise buyers, IT administrators, security teams, business leaders, developers, and end users in product and technical decisions. Identify adoption barriers, competitive gaps, and opportunities to make OpenAI products easier for organizations

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. Safely delivering increasingly capable AI systems requires scalable technical safeguards, clear ownership of emerging risks, rigorous deployment readiness, and close coordination across research, engineering, product, operations, legal, policy, and external partners. Our Technical Program Managers lead complex, high-stakes initiatives that turn safety commitments into deployed systems and measurable outcomes. We work across model development, infrastructure, product, and operational response to help ensure our technology is deployed responsibly and cannot be used to cause serious real-world harm. About the Role We’re seeking Technical Program Managers to drive complex product, platform, and safety initiatives across ChatGPT, API, enterprise, and related deployment environments. These roles operate at the intersection of technical strategy and execution: you will turn safety and product priorities into actionable plans, influence architectural and operational decisions, and deliver durable capabilities across model, infrastructure, application, and platform layers. Depending on the role, you may enable sensitive or high-impact model deployments, integrate safeguards into cloud and API platforms, prevent violent misuse and other serious harms, improve detection and enforcement systems, create platform solutions for safety or establish new programs as risks evolve. You will partner deeply with engineers, researchers, product managers, and operational teams while communicating technical tradeoffs and program decisions to senior leadership. You bring technical fluency, product judgment, and a strong execution record. You’re comfortable navigating ambiguity, advocating for users and developers, balancing safety with model usefulness, and leading cross-functional work with urgency, rigor, and empathy. Specific focus areas and scope will vary by opening and level. Thi

AWSRestMachine LearningAI
C
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -79.2%

From £100K/yr

Quick readStrong listing-quality and freshness signals

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Cohere has a strong AI narrative, the most deployable product, and a CEO who co-authored the foundational paper of modern AI. The company is positioned to become the transatlantic leader in sovereign AI for enterprise and government, and needs a storytelling engine to deliver that story at scale. This role sits within the team that is building that engine. This is an opportunity to define the narratives that influence how enterprises, developers, and the broader ecosystem understand the future of AI and Cohere’s role in building it. In this cross‑functional, individual‑contributor role, you’ll partner deeply with product, research, engineering, and go‑to‑market teams to translate complex technical work into clear, compelling stories. You’ll help Cohere articulate what we’re building, why it matters, and how our technology is reshaping enterprise AI. You’ll craft narratives that resonate with technical and non‑technical audiences alike, and build programs that elevate Cohere’s product leadership and technical credibility. This is a rare opportunity to build a product and technology communications function at a fast

GitMachine LearningAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team Enterprise Verticals builds role-specific ChatGPT Work experiences for high-value enterprise workflows. We combine product engineering, plugins and skills, connectors, data, evaluations, and customer evidence to turn useful demos into reliable daily work. This opening sits within the Technology vertical inside Enterprise Verticals. The group focuses on repeatable workflows for people at technology companies, beginning with functions such as data and analytics, sales, and design, and carries the shared platform needs—tool integration, permissions, quality measurement, and safe rollout—across those experiences. We work closely with Design, Research, GTM, Security, and platform teams, as well as with customers and design partners. Success means that people can reach a trustworthy first result, understand what the system did, and keep using the workflow—not merely that a prototype exists. About the Role We are looking for an exceptionally experienced, hands-on full-stack engineer to define and build the next generation of AI-powered enterprise workflows. You will take on the hardest and most ambiguous problems in the Technology vertical: translating real customer needs into product direction, designing the systems behind the experience, and personally writing and shipping production-quality code across the stack. You will own the technical direction and end-to-end delivery of products spanning ChatGPT Work surfaces, backend services, plugins, connectors, enterprise data, permissions, and evaluations. You will make foundational architecture and product tradeoffs; establish patterns other engineers can build on; and hold these experiences to a high bar for reliability, security, observability, and customer value. This is an individual-contributor role for an engineer who leads through technical judgment, direct execution, and influence—not people management. You should be equally comfortable working directly with customers, setting direction with senior cro

TypeScriptPythonReactNode.js
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI Finance helps ensure the organization is set up for success in pursuit of its mission of ensuring that artificial general intelligence benefits all of humanity. Within Finance, the Technical Revenue team partners with Product, Legal, GTM, Strategic Finance, Revenue Accounting, Finance Systems, and the company’s GTM Deal Desk organization to enable scalable, audit-ready monetization. We advise on commercial and contract design, establish defensible accounting positions under ASC 606, and provide clear conclusions and handoffs for downstream execution by Revenue Accounting Operations. About the Role We’re looking for a Senior Manager of Technical Revenue Enablement to lead Revenue Recognition deal advisory for OpenAI’s general B2B commercial activity. Reporting to the Head of Technical Revenue, you will serve as the primary Revenue Recognition counterpart to OpenAI’s Deal Desk organization, partnering with Legal, GTM, Product, Pricing, and Strategic Finance throughout the presignature deal lifecycle. You will own intake and triage, advise on non-standard terms, approve arrangements supported by established policy and precedent, and route novel or strategic matters to the appropriate technical owner. You will build a scalable deal-advisory model by translating ASC 606 into practical contract guardrails, approved language, decision trees, review thresholds, and clear service levels. This role requires deep technical accounting judgment, commercial fluency, strong stakeholder influence, and the ability to make timely, risk-based decisions in a fast-moving environment. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead the Revenue Recognition deal-advisory function for general B2B transactions, serving as the primary finance counterpart to OpenAI’s GTM Deal Desk organization. Partner early with GTM Deal Desk, Legal, GTM, Pr

Artificial IntelligenceAIAccountingFinance
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$200K – $240K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Platform org at Sentry is the engine that everything else runs on, spanning developer platform, core infrastructure underpinning the product, SRE, security, and IT. It's a broad, technically complex, and deeply consequential organization. We're looking for a Staff Technical Program Manager who can operate at the intersection of technical depth and strategic execution: someone who thrives in complexity, builds trust with senior engineering leaders, and has a gift for turning ambiguity into clarity and momentum. You'll report to the Head of Technical Program Management and work closely with the VP of Platform Engineering, their staff, and partner teams across Engineering, Product, and Design (EPD). This is a high-visibility role with real influence. You'll work directly with the CTO and senior leaders, shape how the Platform org operates, and help Sentry scale through one of its most important chapters. In this role, you will: Drive strategic execution. Partner with the VP of Platform Engineering and senior EPD leaders to translate priorities into clear, measurable plans. Own sequencing, milestones, and key decisions, and make sure leadership always has reliable visibility into progress and tradeoffs. Build data-driven delivery health. Establish the metrics and dashboards that reflect delivery confidence, risk, and engineering health across the Platform org. Keep planning and reporting high-signal and lightweight, focused on outcomes rather than activity. Manage capacity, dependencies, and risk. Create visibility into resourcing, cross-team dependencies, and constraints so leaders can align investment to the

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Product Marketing team shapes how customers understand, adopt, and realize value from OpenAI’s technology. We work across Product, Research, Sales, Solutions, Partnerships, and Customer Success to bring customer insight into our product strategy and translate technical capabilities into clear, credible stories and solutions. About the Role AI becomes meaningful when it helps people do the work that matters to them. For a finance team, that might mean understanding complex information faster. For a healthcare provider, it might mean navigating clinical workflows more effectively. For a sales or marketing team, it might mean creating entirely new ways to reach and serve customers. We’re looking for a senior product marketing leader to shape how OpenAI serves the business functions and industries where our technology can make a meaningful difference. You’ll define how our models and products meet the needs of teams such as sales, marketing, and finance, as well as industries including financial services, healthcare, and retail. Working closely with Product, Research, Sales, Solutions, and Partnerships, you’ll identify important customer problems, influence product strategy, and build relationships with the ecosystem partners and data providers needed to bring complete solutions to market. You’ll also build and lead the product marketing team responsible for turning these opportunities into durable customer value. You might thrive in this role if you: Have 12+ years of experience in product marketing, industry marketing, solutions marketing, or enterprise go-to-market, ideally across enterprise software, cloud, data, developer, or AI platforms. Have built and led high-performing teams, mentored senior marketers, and know how to create clarity in fast-moving, ambiguous environments. Understand how different industries and business functions evaluate technology, adopt new tools, and define value. Have shaped positioning and go-to-market strategies for c

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with research, software, and external hardware partners to shape the next generation of AI systems, from silicon through full-scale deployments. Our team focuses on understanding and optimizing performance across the full system stack—ensuring that architectural decisions are grounded in rigorous, quantitative analysis of real-world workloads. About the Role We are seeking a Performance Modeling Lead to build and lead a small, high-impact team responsible for answering forward-looking architectural questions across AI infrastructure systems. You will develop modeling frameworks and methodologies to evaluate system-level tradeoffs and guide key design decisions. Your work will directly influence reference architectures, vendor designs, and long-term infrastructure strategy. This role sits at the intersection of AI workloads, system architecture, and quantitative modeling, and requires strong technical judgment, ownership, and the ability to translate complex analysis into clear, actionable guidance. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Build and own a performance modeling framework/toolchain to evaluate AI systems across multiple levels of abstraction. Analyze and quantify architectural tradeoffs across compute, memory, networking, storage, and system topology. Develop performance models to guide decisions on: scale-up vs. scale-out architectures interconnect and network design memory hierarchy and system balance. Translate modeling outputs into clear recommendations for internal teams and external hardware vendors. Influence reference designs and vendor roadmaps through data-driven insights. Partner closely with machine learning, systems, and hardware teams to understand workload characte

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

Technical Program Manager – Applied Product & Platform About the Team The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional experiences to developers, consumers, and businesses alike. Technical Program Managers at OpenAI play a key leadership role in scaling these efforts, partnering deeply with product, engineering, design, and go-to-market teams to bring ambitious ideas to life and ensure clarity and discipline in execution. About the Role We’re seeking a Technical Program Manager to drive complex, large-scale product and platform initiatives across ChatGPT, API, and related surfaces. This role sits at the intersection of technical strategy and execution. You’ll be responsible for translating product vision into actionable plans, shaping interfaces and experiences from both consumer and developer perspectives, and ensuring successful delivery across web, mobile, and platform layers. You’ll bring deep technical fluency, product sense, and a strong execution track record. You’re comfortable managing ambiguity, influencing architectural direction, advocating for developers, and leading cross-functional delivery with discipline and empathy. Location: San Francisco, CA (Hybrid – 3 days/week in-office) In this role, you will: Collaborate with product, engineering, and design to shape roadmaps, technical strategies, and long-term platform direction. Translate product vision into clear, scalable technical execution plans, balancing short-term delivery with long-term platform evolution. Lead end-to-end execution of cross-functional product and platform programs—from planning through launch and iteration. Own timelines, milestones, dependency management, risk mitigation, and accountability across teams. Partner with engineering teams to influence ar

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI's Strategic Sourcing team helps the company scale responsibly, efficiently, and at speed. We partner with leaders across Engineering, Research, IT, Systems, Finance, Legal, and other teams to shape commercial strategies, negotiate critical agreements, and build resilient supplier ecosystems. Our work connects technical strategy, commercial judgment, financial discipline, and execution in support of OpenAI's mission. About the Role OpenAI is seeking a Strategic Technology Negotiations Lead to personally lead some of the company's most complex and consequential technology negotiations. This is not a people manager role. It is a senior strategic IC operator role for an expert negotiator who wants to remain close to the work and personally drive high-impact outcomes. We are especially interested in leaders who have managed teams and are intentionally seeking an individual contributor role where their impact comes through judgment, influence, and direct ownership. Rather than owning a fixed category, you will be deployed against high-priority opportunities where deal complexity, commercial stakes, executive visibility, or time pressure require exceptional negotiation leadership. Your initial focus will include data platforms and infrastructure, including data lake and lakehouse technologies, observability, and enterprise SaaS, with flexibility to work across other strategic technology areas. You will lead negotiations from strategy through execution, aligning decision-makers and driving agreements to closure. Many of these negotiations exist within broader supplier and partner ecosystems. You will look beyond the immediate transaction to account for interconnected cost, equity, revenue, partnership, risk, and long-term strategic implications. Success requires strong economics, sound judgment under pressure, executive credibility, and the ability to bring stakeholders with you through difficult decisions. This role is based in San Francisco, CA. We u

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The OpenAI Audit Team is on a mission to build the future of internal audit from the ground up. Our ambition will be powered by a high-energy, technically exceptional team with the judgment, intellectual curiosity, and creativity to harness the latest advances in AI and design a truly next-generation audit function. As AI reshapes how work is performed across the enterprise, Internal Audit will use AI, automation, and data analytics to identify and assess the most significant and emerging risks across the business, including technology, cybersecurity, finance, compliance, operations, and data. We will build trusted partnerships at every level—from the Board of Directors and senior leadership to the teams delivering on OpenAI’s mission every day. We will operate as both an independent assurance provider and a trusted advisor, bringing an objective and pragmatic perspective to critical decisions. By engaging closely with management while preserving our independence, we will help the business innovate responsibly, move with confidence, and manage risk without creating unnecessary barriers. About the Role As the Cybersecurity & Technology Audit Leader, you will help shape the strategy, methodology, technology, and culture of a new audit function. You will lead complex audits and advisory reviews of cybersecurity, technology, data, and AI-related risks while advising leaders on practical ways to strengthen governance, resilience, and risk management. You are an experienced, hands-on professional who combines deep technical expertise with strong business judgment. You can quickly move between executive-level governance questions and detailed technical analysis, use data to identify and assess risk, and form clear, well-supported conclusions in fast-moving or ambiguous situations. You will have meaningful influence over how the function develops, including how we apply AI, automation, and analytics to audit planning, testing, monitoring, and reporting. T

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the role We are building a higher education researcher motion that helps funded labs adopt OpenAI across core research workflows. This role will develop relationships with principal investigators, researchers, research software engineers, research computing teams, and university technology leaders, then translate high-value use cases into sustained Pro, Codex, API, and ChatGPT Edu usage. This is a hybrid business development and program management role. You will identify and qualify priority labs, design phased access and fellows programs, run pilots, remove technical and institutional blockers, and create repeatable paths from individual researchers to lab- and institution-level adoption. Additionally, this role will launch and manage a community of researchers with events, communications and support. In this role, you will: Build a pipeline of funded labs, research centers, and technical champions at priority R1 universities; qualify opportunities based on workflow value, funding path, influence, and expansion potential. Run discovery with researchers and university stakeholders to understand workflows, data and security needs, procurement constraints, and success criteria. Design and operate phased access and fellows programs, including eligibility, selection, offer mechanics, onboarding, office hours, community programming, and pilot goals. Translate research workflows into effective use of Pro, Codex, API, and ChatGPT Edu in partnership with Solutions, Product, and Education account teams. Manage pilots end to end, remove trust, funding, and technical blockers, and drive measurable activation, retention, and expansion. Turn early usage into product feedback, workflow documentation, peer proof, case studies, and repeatable enablement. Build clear handoffs and expansion paths from researcher to lab to institution; track fellow selection, activation, retained usage, workflow proof, conversion, and expansion. You might thrive in this role if you: Relevant exp

AWSRestAIGo
NR
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -73.9%

From $118K/yr

Quick readStrong listing-quality and freshness signals

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity Are you looking for a high-impact finance role that offers true flexibility? We are seeking a Lead Revenue Accountant to join our team in a fully remote capacity . In this pivotal role, you won’t just be managing spreadsheets; you will be a key partner to our Revenue Accounting Senior Manager, helping to shape and build our revenue processes from the ground up. This is your chance to apply your technical mastery in a fast-paced, modern environment where your contributions directly influence our financial reporting and growth. If you thrive on autonomy and want to skip the commute to focus on meaningful work, this is the role for you. What You'll Do: Lead the Process: Take full ownership of the revenue-related month-end close, ensuring our financial reporting is accurate, timely, and transparent. Technical Consultation: Review customer contracts and non-standard terms to determine the correct revenue treatment, providing clear and concise accounting documentation. Financial Integrity: Prepare and reconcile essential revenue schedules, including deferred revenue and contract assets/liabilities, maintaining a high standard of precision. Process Innovation: Drive continuous improvement by streamlining our Order-to-Cash processes and supporting vital SOX audit controls. Strategic Collaboration: Act as a subject matter expert on special projects and ad-hoc requests, providing the revenue accounting insights needed for senior management to make informed decisio

GitRestAIGo
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $1.3M/yr

Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity We are seeking a technically skilled, customer-focused individual to join our Technical Onboarding team as a Solutions Consultant. In this role you will own the onboarding and adoption of our enterprise customers, taking them from a signed contract to a secure, well-governed, and widely adopted deployment of our platform. You will serve as the trusted technical advisor for each engagement, running discovery, designing the right approach, and leading the hands-on delivery yourself. This spans a range of engagement types, from securing and standing up a customer's enterprise instance to driving broad team adoption and API collaboration at scale. Most of our customers are large, security-conscious organizations in regulated industries, so this is a consultative, high-impact role where you will drive real technical outcomes and directly influence customer success. What You’ll Do Engagement Ownership: Own structured onboarding and adoption engagements end to end, from kickoff and project planning through configuration, testing, closing, and post-engagement hypercare, serving as the single point of ownership for the platf

JavaScriptPythonJavaAWS
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Privacy Engineering team builds the systems and technical foundations that govern how user data is understood, retained, accessed, and used across OpenAI. We partner with Product, Data, Infrastructure, Security, and Legal to translate policy and trust commitments into durable architecture and enforceable controls. Our work spans data inventory and mapping, classification and lineage, retention and deletion, access governance, purpose and usage controls, auditability, and lifecycle automation. We aim to make policy-aligned data handling the default while giving teams clear, reliable primitives for building and operating products at scale. About the Role We are looking for an experienced Software Engineer to drive the architecture and execution of user data governance across OpenAI. You will define technical direction, build shared platforms and controls, and lead cross-functional programs that make data flows discoverable, policies enforceable, and ownership explicit. This role is well suited to a senior engineer who can move between deep systems design and organization-wide influence, turn ambiguous requirements into pragmatic roadmaps, and operate high-trust systems end to end. This position is based in San Francisco. Relocation assistance is available. In this role, you will: Set the technical strategy and architecture for user data governance across data mapping, classification, lineage, retention, deletion, access, and permitted usage. Design and build shared services, APIs, metadata systems, and policy-enforcement mechanisms that make governance controls consistent, scalable, and auditable. Establish reliable inventories of user data, system ownership, data flows, and policy applicability across products, infrastructure, analytics, and research systems. Partner with Product, Data, Infrastructure, Security, and Legal leaders to define decision rights, translate requirements into controls, and drive adoption across teams. Own governance systems

AWSRestAIGo
🔔

Get new inference technical lead jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime