About the Team The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional experiences to developers, consumers, and businesses alike. The ChatGPT Multimodal team works across voice, image generation, and other multimodal experiences to turn frontier research capabilities into reliable products. The team connects product usage and failure patterns with research, evaluation, data, inference, capacity, and external partnerships so that model and product improvements translate into better experiences for users. About the Role We are seeking a Technical Program Manager to build the flywheel that helps ChatGPT multimodal products learn from real-world usage and improve quickly. You will lead programs spanning production-signal mining, evaluation and data pipelines, research-to-production parity, multimodal capacity planning, and complex cross-functional dependencies for voice and image-generation launches. You will work closely with product engineering, research, Human Data, inference and capacity teams, safety partners, and external vendors or product partners. Success requires technical depth, strong systems thinking, comfort with ambiguity, and the ability to turn fragmented or manual work into durable mechanisms that teams adopt. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Build a system for mining production conversations and product signals to identify representative multimodal workflows, user needs, and failure modes. Establish and maintain evaluations for the highest-priority multimodal behaviors and use cases, with clear coverage, quality standards, and ownership. Package production signals into decision-ready data and
Jobs in United States
Ai Outcomes Manager in San Francisco
1,456 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai outcomes manager jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team Compute Foundations builds the software that manages OpenAI’s GPU compute infrastructure across sites, data centers, and infrastructure providers, supporting model training and inference. Our systems turn large, heterogeneous fleets of machines into dependable compute for research and products. We build Kubernetes-based control planes, controllers, services, and APIs that coordinate the lifecycle of machines and clusters. We connect global infrastructure management with the realities of bare-metal systems, giving clients consistent interfaces across differences in hardware, topology, and provider behavior. About the Role You will build distributed systems that provision, configure, and manage compute throughout its lifecycle. Your work will connect global services and Kubernetes controllers with the systems that bring machines online, update them safely, and recover them when something goes wrong. This role combines software architecture with an understanding of how machines and data centers work. You might design a lifecycle API, improve controller performance under high concurrency and provider rate limits, or trace a provisioning failure from an API through reconciliation to network boot or host configuration. You will help these systems remain reliable as the fleet expands across sites and generations of GPU hardware. We value depth in relevant systems and the ability to connect layers. You do not need to arrive as an expert in every component of the stack. In this role, you will: Design, build, and operate Kubernetes-based controllers and distributed services that coordinate infrastructure across sites, isolate failures, and scale as GPU capacity grows. Define APIs and resource models that let clients request and track lifecycle operations through consistent interfaces across hardware platforms and providers. Build provisioning and configuration services that coordinate network boot, hardware management interfaces, and the deployment of firmware,
$150K – $200K/yr
About Taskrabbit: Taskrabbit is a marketplace platform that conveniently connects people with Taskers to handle everyday home to-do’s, such as furniture assembly, handyman work, moving help, and much more. At Taskrabbit, we want to transform lives one task at a time. As a company we celebrate innovation, inclusion and hard work. Our culture is collaborative, pragmatic, and fast-paced. We’re looking for talented, entrepreneurially minded and data-driven people who also have a passion for helping people do what they love. Together with IKEA, we’re creating more opportunities for people to earn a consistent, meaningful income on their own terms by building lasting relationships with clients in communities around the world. Taskrabbit is a hybrid company with employees distributed across the US and EU and a Built In — Best Places to Work (2022, 2023, 2024, 2025) continually ranked across multiple national and regional categories. Join us at Taskrabbit, where your work will be meaningful, your ideas valued, and your potential unleashed! Prior to applying please note: This role is hybrid requiring 2 days in office at our San Francisco hub every Tuesday & Wednesday (located at 130 Sutter St, San Francisco, CA). About the Role: The Architecture Team drives Taskrabbit's Platform Modernization efforts, guiding engineering teams toward our Target State Architecture (TSA). As a Staff Software Architect, you'll serve as an embedded representative of the Architecture team, partnering with engineering teams across the platform to accelerate their modernization work. You'll operate as a subject matter expert on platform modernization within a small team of architects, applying AI-assisted and Spec-Driven Development (SDD) practices to help teams build new components using our Target State Architecture stack of technologies. This role is hands-on and results-oriented: pairing directly with engineers and technical leads, not just advising from the sidelines. You will be: Act
$170K – $225K/yr
About Taskrabbit: Taskrabbit is a marketplace platform that conveniently connects people with Taskers to handle everyday home to-do’s, such as furniture assembly, handyman work, moving help, and much more. At Taskrabbit, we want to transform lives one task at a time. As a company we celebrate innovation, inclusion and hard work. Our culture is collaborative, pragmatic, and fast-paced. We’re looking for talented, entrepreneurially minded and data-driven people who also have a passion for helping people do what they love. Together with IKEA, we’re creating more opportunities for people to earn a consistent, meaningful income on their own terms by building lasting relationships with clients in communities around the world. Taskrabbit is a hybrid company with employees distributed across the US and EU and a Built In — Best Places to Work (2022, 2023, 2024, 2025) continually ranked across multiple national and regional categories. Join us at Taskrabbit, where your work will be meaningful, your ideas valued, and your potential unleashed! Prior to applying please note: W e are currently unable to provide visa sponsorship for this position (including H-1B, OPT, F1, CPT or other employment-based visas). Candidates must be legally authorized to work in the United States without employer sponsorship now or in the future. This role is hybrid requiring 2 days in office at our San Francisco hub every Tuesday & Wednesday (located at 130 Sutter St, San Francisco, CA). About the Role Data Science plays a crucial role in driving impact at Taskrabbit. We are seeking a highly skilled and motivated Staff Data Scientist to join us, working closely with cross-functional teams from Product teams and occasionally Commercial Operations to provide data-driven insights and solutions that enhance our products and accelerate growth while minimizing marketplace losses. What you will work on Be a strategic thought partner with stakeholders in Product and occasionally Commercia
About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models. We’re looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience. You’ll transform complex operational data into actionable insights that directly influence how OpenAI allocates and scales one of the world’s largest AI compute environments. Key Responsibilities Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency. Develop forecasting models for inference demand across products, regions, and model families. Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities. Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies. Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs. Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions. Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps. Communicate technical findings clearly to both engineering teams and executive leadership. Qualifications MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience). 5+ years of experience working in the infrastructure data science space. Strong ex
About the Team The Finance Platform & Technology team builds and scales the systems and data architecture that power OpenAI’s core financial operations. We enable business agility, compliance, and operational excellence across procure-to-pay, quote-to-cash, supply chain, financial planning, and asset management. We partner with Procurement, Accounting, Tax, Legal, Security, Data, and Engineering to modernize workflows through thoughtful platform design, reliable integrations, scalable automation, and trusted data. About the Role As a Business Systems Lead for Procure-to-Pay, you will be a hands-on engineer who designs, builds, and operates the integrations and first-party applications that power OpenAI’s procurement workflows. You will translate business needs into secure, scalable software, APIs, data flows, and automation across Oracle Fusion, Zip, and connected platforms. You will build the future of buying at OpenAI using OpenAI’s own technology, from guided intake and approval experiences to supplier onboarding, purchasing, receiving, invoicing, and downstream financial data flows. You will own the technical roadmap and support model for these capabilities, improving today’s platforms while deciding where to integrate, configure, or build as OpenAI scales. Your core strength will be software and integration engineering. You will personally write code, troubleshoot cross-system failures, and take solutions through testing, deployment, and production support. You will also make targeted functional configurations in procurement platforms and partner with functional specialists on deeper process and module design. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and operate integrations across Oracle Fusion, Zip, and connected systems using APIs, events, messaging, and batch interfaces where appropriate. Build first-party
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry's issue platform processes billions of events every day to help millions of developers find and fix bugs. The hard part is deciding which events point to a problem, which belong together, and what a developer needs to know to investigate. As a Senior Software Engineer on the Issue Detection team, you'll design, build, and operate the systems that make those decisions. You'll work on real-time processing pipelines, monitors, and analysis systems that detect problems and turn them into issues. The work combines distributed systems with product engineering. Choices about detection accuracy and processing latency affect which problems developers see and how soon they can act. You'll help shape how developers monitor their applications and how Sentry groups related events into issues. You'll also build the context developers and AI agents need to investigate what went wrong. Keeping these systems reliable and fast as Sentry grows is part of the job, alongside making the issues they produce more useful. In this role you will Build and scale features on a product surface handling billions of events daily, where both query latency and correctness are immediately visible to users. Own the design and delivery of substantial projects end to end, scoping alongside product and design, making the technical calls within your scope, shipping, and instrumenting what you ship so the team can measure it. You will contribute to meaningful technical product decisions : grouping quality, search performance, migrations and backfills against enormous datasets, and making the surface work well for both humans and agents. Champi
From $130K/yr
About Mixpanel Mixpanel is the leading product intelligence and analytics platform, trusted by more than 29,000 companies to help understand how people use the products they build. By combining powerful analytics with AI that knows your business, Mixpanel helps teams see what’s working, diagnose what’s not, and decide what to build next. Learn more at mixpanel.com . About the Team The Revenue Marketing team at Mixpanel is responsible for pipeline generation across paid, website, and product-led channels. This role sits within that team as its dedicated engineering owner — accountable for the technical systems that power how people discover, evaluate, and request Mixpanel. You are the only engineer dedicated to this surface, which means you set its technical direction rather than execute against someone else's. You work directly with our marketing teams and partner with Growth Engineering on shared systems. Website design and content are handled by our web design team; your focus is the technical infrastructure, integrations, and systems that sit underneath. About the Role As our Marketing Systems Engineer, you own the technical layer of our marketing engine: the integrations and infrastructure that power our marketing website. The website is our primary lead capture system that turns traffic into pipeline. You own the systems that connect our marketing surface to Salesforce, Customer.io/Hubspot, and our broader GTM stack. You set the technical roadmap for this surface, move fast, measure impact, and treat reliability as a first-class concern rather than a cleanup task.You will also collaborate with Growth Engineering on shared infrastructure including handraiser routing and tracking. Responsibilities Own the technical health of the Mixpanel marketing website: page speed, WCAG compliance, Google Tag Manager, technical SEO, and third-party integrations including Qualified, Optimizely, and TrustArc. Own the integrations between the marketing website and our GTM stack t
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role you will Own technical delivery across multiple deployments from first prototype to stable production Build full-stack systems that deliver customer value and sharpen how we learn Embed closely with customer teams, understand their needs, and guide adoption of what you build Scope work, sequence delivery, and remove blockers early Make trade-offs between scope, speed, and quality; adjust plans to protect delivery Contribute directly in the code when progress or clarity depends on it Codify working patterns into tools, playbooks, or building blocks that others can use Share field feedback that helps Research and Product understand where the models succeed and where they can improve Keep teams moving through clarity and follow-through You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work Have scoped and delivered complex systems in fast-moving or ambiguous environments Write and review production-grade code across frontend and backend using Python, JavaScript,
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. The Fraud Data team at Plaid builds the machine learning systems that power Plaid’s fraud detection products, leveraging insights from across Plaid’s network to help identify and stop fraud before it happens. Our team works across the full data science and machine learning lifecycle—from discovering new signals and experimenting with models to deploying and optimizing them in production. We continuously learn from real-world model performance and customer feedback to improve our systems and develop new ways to protect customers and consumers from evolving fraud threats. As a Senior Machine Learning Engineer on Plaid's Fraud Data team, you will develop models that improve fraud detection for our customers. You will identify predictive patterns in Plaid's network data and lead projects from initial experiments through model deployment and ongoing improvement. Investigate fraud patterns and model errors to identify new signals, improve detection, and expand coverage across customers and use cases. Develop training datasets and predictive features, addressing challenges such as incomplete labels, class imbalance, data leakage, and changing fraud behavior. Design, train, and tune model
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. About the Team At Embedded Insights, we find the best machine learning opportunities for external products and internal systems, and collaborate with cross-functional partners to bring them to life. We are a central team of Machine Learning Engineers and Data Scientists. We embed with partner teams to build and apply machine learning models that improve internal decision-making and power the Plaid product suite. About the Role You will be the first Data Scientist on the Embedded Insights team, part of Plaid’s Data organization. You will establish the analytics and metrics backbone for a team supporting a diverse set of internal and external products. You will help drive better decision-making, support machine learning model development, and contribute directly to the health of the Plaid network and the quality of Plaid’s products. Your day-to-day work will include: Analyzing entities across the Plaid network to understand behavior and identify opportunities, anomalies, and risks. Creating foundational metrics, dashboards, and monitoring systems that provide a clear view of network health and machine learning model performance. Evaluating the value and performance of machine learni
$220K – $450K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry is looking for an Engineering Manager to lead our Growth team (internally called "Value Discovery") and help expand our product-led business by helping customers discover and get the most out of our platform. At Sentry, we do growth with an explicit bias for customer value. Our experiments are rooted in strong product sense and empathy for developers, not conversion pressure. We never surprise developers with new charges or push features they don't need. In a binary tradeoff between doing good by the customer or good by the business, the customer always wins. You'll collaborate closely with several areas of the business including Product, BizOps, Data, Design, and GTM to identify the moments in a developer's workflow where Sentry can deliver more value, make that value easier to find, and support that your approach works with data. In this role you will You'll lead a team of engineers, set its priorities, and be accountable for the metrics it moves. This is a new role, and the roadmap is yours to define and execute in partnership with design and data teams. Lead work across Sentry's self-serve funnel — signup and onboarding, first-run activation moments in-product, trial and upgrade paths, and pricing and billing surfaces in partnership with the Billing team — largely in our Python/Django and TypeScript/React codebase Evolve our experimentation foundations, in close partnership with our Data team — flagging, instrumentation, and readouts are real but far from finished Run a fast, disciplined experimentation loop where hypotheses are argued before they're built Hold conversion and activation alongside ch
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking a Mechanical Engineer to design, build, and own the mechanical side of our robotic actuator dynamometer and test infrastructure. You will create the test stands, couplings, fixtures, load paths, guarding, and serviceable lab hardware that enable rigorous characterization of robotic actuators. This role combines precision mechanical design with hands-on lab work. You will take robotic actuator test infrastructure from requirements and analysis through CAD, fabrication, assembly, commissioning, and iteration, partnering closely with electrical and software engineers to deliver safe, flexible, high-uptime test cells. In this role, you will Own the mechanical architecture of dynamometer and actuator test cells, including frames, bases, load paths, alignment, guarding, and serviceability. Design dynamometer structures, robotic actuator fixtures, load-motor mounts, couplings, shafts, bearings, adapters, and torque-reaction hardware. Translate robotic actuator test requirements into robust mechanical systems for torque, speed, thermal, durability, backdrive, efficiency, and failure testing. Perform first-principles analysis and simulation for stiffness, strength, fatigue, vibration, thermal growth, critical speed, and safety factors. Create precise, repeatable alignment strategies that protect test articles, load machines, sensors, and couplings. Design modular fixturing that supports rapid changeover across actuator and motor variants without compromising measurement quality. Work closely with electrical engineers on cable routing
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We’re looking for a Robotics Control Systems Engineer to take on a foundational role within our robotics team. You’ll help architect, implement, tune, and verify the control infrastructure that enables intelligent, reliable, and responsive robot behavior. This is a deeply hands-on role focused on real-time systems, actuation, dynamics, low-level hardware interaction, and whole-robot performance. You’ll spend significant time working directly with robots onsite: debugging behavior, tuning subsystems, running experiments, and providing feedback across mechanical, electrical, and software teams. This role is based in San Francisco, CA, and requires in-person 5 days a week. In this role, you will: Design and implement real-time control algorithms for robotic systems, including motion control, feedback loops, state estimation, actuator control, and subsystem tuning. Define the control architecture from low-level actuators and hardware interfaces through whole-robot behavior and policy. Identify and characterize actuator, hardware, and software parameters through rigorous experimentation, testing, commissioning, and verification. Work with machine learning engineers to implement reinforcement learning models. Collaborate across mechanical, electrical, and software teams to integrate control logic with sensing and actuation hardware. Help inform the mechanical and electrical design to maximize capability and flexibility. Create the control system architecture; determine the correct level of abstraction from actuators all the way up to whole-robot
From $92K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About tvScientific We are the first CTV advertising platform purpose-built for performance marketers. For game developers and publishers, we bridge the gap between massive TV reach and granular User Acquisition (UA) metrics. Built by ad-tech veterans, our platform combines media buying, optimization, and MMP attribution to help gaming brands automate CTV campaigns, drive app installs, and maximize Return on Ad Spend (ROAS). Join the tvScientific team as an Account Manager (Gaming), where you'll lead strategic client relationships for gaming and app clients, drive revenue growth, and ensure client success on our cutting-edge platform. As an Account Manager on our team, you'll be responsible for managing a portfolio of key client accounts, developing and executing strategic account plans, and driving revenue growth through upsell, cross-sell, and
Other cities to consider
More places hiring for this role
Get new ai outcomes manager jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime