About the Team : We build core first-party app experiences in ChatGPT and Codex, define the primitives for high-quality app third-party app experiences, and collaborate with best-in-class partners across consumer and enterprise categories to bring delightful experiences to our customers. About the Role: We are hiring a Product Manager to shape and scale the app ecosystem across ChatGPT and Codex. This person will own both 1P product experiences and partner-led launches, translating user needs, model capabilities, platform constraints, developer and partner requirements, and enterprise controls into products that feel delightful, reliable, and safe. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Develop the strategy and roadmap for ChatGPT and Codex app ecosystem experiences across consumer and enterprise use cases. Build and ship high-quality first-party app experiences that demonstrate the best of what apps can do inside ChatGPT and Codex. Collaborate with best-in-class partners to create app experiences that solve real user and business workflows. Define the product foundations and quality standards needed for a trusted app ecosystem. Lead cross-functional execution across engineering, design, research, partnerships, GTM, legal, privacy, security, support, and data/evals. Use customer, user, partner, and model-behavior insights to prioritize the roadmap and improve post-launch performance. You might thrive in this role if you: Have built and scaled consumer or enterprise product experiences with strong product taste and measurable user impact. Have built app platforms, marketplaces, partner ecosystems. Are technically fluent enough to reason about MCPs, APIs, SDKs, and model/product constraints. Can move fluidly between strategy, product, partner judgment, and operational execution. Communicate crisply in writing and bring clarity to ambiguous, fast-moving product areas. Care deeply about user trust, safe
Jobs in United States
Quality Lab Supervisor in United States
6,553 active opportunities · Updated October 2026
Showing
15 jobs
Explore current quality lab supervisor jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE OPPORTUNITY We are looking for Senior Software Engineers to join our team. This is a specialized, high-impact role sitting at the intersection of high-performance computing (HPC) and Large Language Model (LLM) engineering. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work. RESPONSIBILITIES Benchmarking : Evaluate, run and automate standard LLM quality benchmarks (GSM8K, MMLU) alongside custom performance suites for specific workloads (e.g., long-context window, KV cache reuse, disaggregated serving). DevEx Improvement : Develop and maintain internal GPU-enabled development environments (similar to GitHub Codespaces). You will ensure the team has seamless, high-performance "dev machines" optimized for model experimentation. Tool Development : Build and contribute to open-source tools such as InferenceMAX and genai-bench to automate model evaluation, benchmarking and analysis. System Profiling : Use profilers like PyTorch Profiler, NVIDIA Nsight Systems and py-spy to collect performance profiles, identify bottlenecks, and debug the compute/networking stack. Monitoring & Observability : Develop real-time dashboards and alerts to monitor system health, model startup times, and runtime performance. Continuous Integration : Auto
ABOUT THE TEAM The Scaled/Digital Customer Success team helps customers realize value from Mural through education, digital programs, and learning experiences at scale. This team plays a critical role in making high-quality customer education more personalized, accessible, and scalable. This role sits within our Scaled and Digital Customer Success team and reports to the Director of Customer Success. YOUR MISSION As a Customer Enablement Manager, your mission is to bring our customer enablement strategy to life through high-quality content, personalized learning experiences, and scalable programs. You will translate customer needs and product capabilities into clear, outcome-based learning journeys that help teams work better together and achieve real-world success with our platform. You will build and scale modern customer education programs that help customers adopt priority use cases, reduce time-to-value, and increase product confidence. WHAT YOU'LL DO Execute scalable education initiatives including self-guided, live, cohort-based, and peer-led learning experiences. Design outcome-based learning journeys and adaptive paths tailored to priority roles, use cases, and adoption milestones. Develop and maintain high-quality reusable content such as courses, guides, videos, templates, and interactive simulations. Partner cross-functionally with Product, Marketing, and CS to align education with launches and identify education gaps. Deliver internal enablement to CS Teams globally to ensure consistency in Customer Journey, Strategy Execution, and customer outcomes. Track and measure learning engagement, proficiency, and product adoption to continuously improve program effectiveness. WHAT YOU'LL BRING 3+ years of experience in customer education, instructional design, or scaled Customer Success, ideally in B2B SaaS. Proven ability to design modern learning experiences using microlearning, scenarios, and social learning principles. Strong content creation skills with th
From $131K/yr
The product operations team enables product excellence at scale at Datadog. With over thirty standalone products on a single platform, product operations drives consistency, efficiency, and quality across the organization. As a Product Operations Manager, you will help the product organization do its best work. You'll build and scale the processes, frameworks, and cross-functional workflows that enable PMs and their partners to ship effectively. This includes touchpoints with inbound, delivery, outbound workstreams in product development. This is a role for someone who thrives in the space between teams: driving alignment, removing friction, and ensuring that the operational foundations of the product org can keep pace with Datadog's growth. At Datadog, we place value in our office culture, the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own and improve the full Product Development Life Cycle . Own and drive operational excellence for product launches. This includes coordinating across Product, Engineering, Marketing, Sales, Customer Success, and Support to ensure we are ready to bring quality products and support to our customers. Build AI-powered workflows and agents that automate and accelerate operational work. From customer feedback triage to launch tracking to cross-functional reporting, you'll experiment with and deploy AI tools and custom agents to help the product org move faster and smarter. Build and maintain a single source of truth for launches giving cross-functional partners real-time visibility into what's shipping, when, and what's needed from each team. Drive centralized product knowledge and enablement , ensuring PMs and stakeholders have the tools, resources, templates, and context they need to succeed and making it easily discoverable. Analyze and quantify product signals. Query pro
From $192K/yr
As Engineering Manager for Threat Detection, you will lead a high-performing team that powers Datadog's detection program. Threat Detection is the organization responsible for keeping Datadog ahead of an evolving threat environment: closing coverage gaps faster, raising the bar on signal quality, and shipping detections that hold up under the scale and complexity of cloud-native infrastructure. Your team will combine direct detection expertise, platform engineering, and applied AI to ship detections at a pace and scale traditional rule-writing alone cannot match. Examples of what your team will work on include detection-authoring agents, the detection platform that powers every rule in production, coverage analysis, alert triage and response automation, and the evaluation infrastructure that holds these systems to a high bar of fidelity. Detection authorship is a shared responsibility across the organization, and your team will contribute both by building the systems that scale our authoring capacity and by writing detections directly when their domain expertise is the right tool. You will partner closely with our Security Incident & Response Team (SIRT), Cyber Threat Intelligence (CTI), AI Engineering teams, and Datadog's broader Security organization. This is a high-impact leadership role: you will grow a team of security and software engineers responsible for building and executing our detection and AI strategy. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the strategy, roadmap, and execution of Datadog Security's shift to AI-accelerated detection and response. Drive development of high-fidelity detections as a shared responsibility across the organization, ensuring your team's systems and direct contributions raise the bar on coverage and
About the Team At OpenAI, the User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving company priorities. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are looking for a senior program manager to build the safety, quality, and risk operations supporting a new category of consumer devices. You will translate ambiguous product risks and evolving requirements into practical operating models, workflows, escalation paths, launch-readiness plans, and cross-functional decision-making. This is a foundational role: the systems you build will shape how OpenAI launches, monitors, and improves a new category of consumer devices safely at scale. You will help establish how potential safety incidents, product-quality concerns, sensitive customer escalations, privacy-sensitive issues, and other emerging device risks are identified, investigated, resolved, and incorporated into product and operational improvements. You will turn incomplete requirements into practical workflows, decision rights, launch plans, quality controls, measurement, and durable ownership. The role centers on program building, operational judgment, and execution. We welcome candidates from product safety, quality assurance, regulatory operations, technical program management, healthcare, medical devices, aerospace, consumer technology, and other environments involving complex products or regulated risks. Direct consumer-hardware experience is helpful but not required. The strongest candidates learn unfam
About the Team OpenAI’s Central Procurement – Hardware Operations team is building scalable, AI-enabled operating foundations for a rapidly growing hardware footprint. We help the business move quickly while maintaining the data quality and controls needed to manage financial and operational risk responsibly. As OpenAI scales across R&D, new-product introduction, manufacturing, and third-party custody models, we are building the foundations to absorb complexity without adding unnecessary friction. That means establishing practical standards, strengthening discipline where it matters, using AI thoughtfully, and continuously improving how teams manage hardware operations. About the Role We are hiring an Asset Compliance Program Lead to establish the cross-business framework that gives OpenAI reliable lifecycle visibility over hardware-related financial assets and capital equipment. You will define the controls, systems roadmap, and evidence practices needed to manage those assets within OpenAI’s financial-control scope. This is a senior individual-contributor role in Central Procurement – Hardware Operations within Finance. Distinct from sourcing and transaction execution, you will partner with hardware business units, Accounting, Financial Risk Management, Procurement, contract manufacturers, and other third-party custodians to deploy practical operating mechanisms. These mechanisms will connect how assets are purchased, built, received, moved, held, verified, and retired with the ownership, data, reporting, and evidence needed for financial governance. The role covers those assets regardless of location or custody, including manufacturing equipment and tooling, supplier- and contract-manufacturer-held assets, leased assets, and other third-party-held equipment. You will stay close to the operational details—how assets move, where records diverge, which controls are not working, and where evidence is incomplete—and use that view to strengthen processes, accountab
About the Role The Mechanical Commissioning Project Engineer owns mechanical commissioning planning, readiness, quality coordination, and test execution oversight for the project. This role ensures mechanical systems are installed, inspected, started, balanced, controlled, and tested in a way that supports reliable integrated facility performance. Reports to the Commissioning Project Lead and partners closely with mechanical contractors, equipment vendors, design/engineering teams, the electrical commissioning lead, controls stakeholders, and vendor field/test engineers. Key Responsibilities Develop and maintain the mechanical commissioning scope, readiness criteria, inspection strategy, and discipline test execution plan. Review mechanical design packages, specifications, submittals, method statements, controls narratives, sequence assumptions, and testing requirements for commissionability and risk. Coordinate mechanical QA/QC inspections with contractors and vendor field/test engineers, including installation checks, pre-functional readiness, deficiency capture, and closeout tracking. Own mechanical commissioning procedure development and review, including equipment startup, functional testing, controls verification, balancing prerequisites, failure mode validation, and integrated systems testing inputs. Coordinate with equipment vendors on factory/site acceptance requirements, startup support, test prerequisites, documentation packages, and vendor participation during critical tests. Support readiness and execution for cooling, ventilation, hydronic, pumping, heat rejection, controls, and other project-specific mechanical systems. Lead discipline-level review of mechanical test results, deficiencies, corrective actions, retest requirements, trend logs, and acceptance evidence. Maintain mechanical commissioning dashboards and status inputs for the Commissioning Project Lead, including risk items, resource needs, test readiness, and issue aging. Partner with the E
About the Role The Electrical Commissioning Lead owns electrical commissioning planning, readiness, quality coordination, and test execution oversight for the project. This role ensures electrical systems are inspectable, testable, safe to energize, and ready for coordinated functional and integrated systems testing. Reports to the Commissioning Project Lead and partners closely with electrical contractors, equipment vendors, design/engineering teams, the mechanical commissioning lead, controls stakeholders, and vendor field/test engineers. Key Responsibilities Develop and maintain the electrical commissioning scope, readiness criteria, inspection strategy, and test execution plan. Review electrical design packages, specifications, submittals, method statements, sequence assumptions, and testing requirements for commissionability and risk. Coordinate electrical QA/QC inspections with contractors and vendor field/test engineers, including installation checks, pre-functional readiness, deficiency capture, and closeout tracking. Own electrical commissioning procedure development and review, including startup, energization readiness, functional testing, failure mode validation, controls interfaces, and integrated systems testing inputs. Coordinate with equipment vendors on factory/site acceptance requirements, startup support, test prerequisites, documentation packages, and vendor participation during critical tests. Support energization planning, switching and safety coordination, access sequencing, temporary conditions, witness points, and hold points. Lead discipline-level review of electrical test results, deficiencies, corrective actions, retest requirements, and acceptance evidence. Maintain electrical commissioning dashboards and status inputs for the Commissioning Project Lead, including risk items, resource needs, test readiness, and issue aging. Partner with the Mechanical Commissioning Lead on cross-discipline dependencies, including controls, life safety int
About the Team The Recursive Self-Improvement (RSI) team works across research, engineering, product, and infrastructure to build AI systems that accelerate and ultimately conduct high-quality research at OpenAI. We work to automate real research workflows and improve research productivity by building systems and feedback loops, designing evaluations, and training models to develop missing capabilities. Our work spans the full lifecycle of model training, evaluation, and deployment to help researchers move faster and tackle increasingly ambitious problems. About the Role We’re hiring research scientists , research engineers , and AI systems engineers to work on automating research at OpenAI. This role is based in San Francisco, CA. In this role, you will: Design evaluations for research judgment, hypothesis generation and testing, and long-horizon experiment execution. Turn real research workflows and model failures into data and evaluation flywheels. Improve model research capabilities through agent harnesses, synthetic data, RL environments, and model training. Build and maintain safe, reliable integrations between our models and OpenAI’s research infrastructure. Develop research agents, experiment-orchestration systems, and sandboxed runtimes that support real research workflows. Create metrics and economic models to understand RSI’s current and future effects on research productivity, model capabilities, and the safety of internal deployments. This is a high-ownership role for researchers and engineers who thrive in ambiguity, move fluidly between research and implementation, and turn emerging opportunities into rigorous, reliable, scalable results. You might thrive in this role if you: Have research or engineering experience across LLM training, model evaluations, agent systems, synthetic data, research infrastructure, or large-scale distributed systems. Are a strong generalist who can move between open-ended research and practical implementation, turning ambig
About the Team The Support Automation team at OpenAI scales the organization by applying cutting-edge AI models to real-world challenges, automating and enhancing work across the organization. From customer operations to engineering, we develop an ecosystem of automation products that empower our colleagues and drive impact. We're passionate about crafting products that serve those around us, blending rapid prototyping with a focus on long-term quality and reliability. By creating reusable solutions, we create patterns that can be applied across diverse domains within OpenAI. TLDR: this team leverages OpenAI technology to improve OpenAI, and you’ll have the opportunity to leverage the full extent of our tech (both public and pre-released) to accomplish this mission. About the Role We’re looking for a Backend Software Engineer with experience working in ML/LLM-heavy domains to help to design and build an evals infrastructure that measures the quality of OpenAI’s support automation. This is a deeply technical and highly cross-functional role where you’ll build robust systems and backend services that serve as the foundation for how knowledge is created, accessed, and applied across OpenAI. The role will especially focus on working closely with Data Science and Research partners to design and build evals at scale. In this role, you will: Design eval pipelines that are reliable, reproducible, and extendable Build the infrastructure for continuous eval monitoring frameworks (regression/drift monitoring, building robust golden datasets) along with feedback loops that ultimately strengthen support automation Design, build, and maintain backend services and APIs to support intelligent automation and knowledge systems Integrate and structure data across internal platforms, transforming it into formats optimized for use by downstream systems and AI workflows. Collaborate closely with data, research, and engineering teams to integrate OpenAI models into high-leverage workflows
About the Team GTM Growth Engineering builds AI-native products and systems that help OpenAI's go-to-market and B2B marketing organizations operate with greater speed, focus, and leverage. Our mandate is revenue leverage: products tied directly to pipeline quality, customer engagement, seller and marketer productivity, and the speed at which OpenAI can bring its technology to customers. We build the infrastructure and user experiences behind high-impact GTM workflows, including customer context, prioritization, routing, campaign execution, review surfaces, feedback loops, and measurement. Our work combines product craft, applied AI, reliable systems, and thoughtful operational design. About the Role We’re looking for a product-minded Software Engineer to build AI-powered products and full-stack experiences for GTM Growth Engineering. You will own meaningful product slices end to end, from user experience and frontend implementation to backend APIs, integrations, data models, instrumentation, and launch readiness. This is a role for engineers who want to build products that do real work in production. You will partner with Product, Design, Data Science, Sales, B2B Marketing, and operations teams to understand high-value workflows and ship systems that improve customer engagement, pipeline, conversion, and team productivity. The role is ideal for a strong product engineer who can move between product craft, systems engineering, applied AI, and measurable business outcomes. You should be excited to build from ambiguous problem statements, ship quickly, and improve products based on real user feedback. What You'll Do Build AI-powered products and workflows that help sales and B2B marketing teams identify opportunities, coordinate work, and engage customers more effectively. Own full-stack product experiences from prototype through launch, instrumentation, iteration, and production hardening. Design intuitive user journeys that combine polished interfaces, reliable servi
$272K – $320K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role Build the most advanced AI Meeting Notes product — and expand it into broader “AI data capture” features that help teams turn conversations into durable context, tasks, and knowledge. Our mission is to 10x the rate of business context & data that enters Notion — optimized for agents — so teams get superhuman memory across workstreams and customers. Notion workspaces that use AI Meeting Notes already enter 6x more data on a daily basis, so we’re well on our way. What You'll Achieve Ship end-to-end product experiences across capture → transcript → summary → follow-ups (full-stack ownership). Make meeting & data capture feel effortless and magical (e.g., speaker identification via audio waveforms, richer in-meeting UX, smarter organization). Improve summary quality that teams trust: structure, factuality, and citations that make downstream agents and humans more capable. Raise the bar on reliability & observability across the pipeline (SLOs, debugging workflows, incident response) for realtime systems. Build agentic meeting workflows that turn discussions into tasks, follow-ups, and organized knowledge — so “w
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Mission The Global Revenue Enablement (GRE) organization exists to equip every audience in our ecosystem — field sellers, technical specialists, and partners — with the clarity, content, and skills they need to execute with speed and consistency. This Program Manager sits within the PMO + Data team and owns the design, execution, and continuous improvement of Snowflake's Authorized Training Partner (ATP) program. You will build and scale a certified partner and instructor network that extends Snowflake's educational reach into markets and segments the core team cannot serve directly — while protecting brand integrity and learner experience at every touchpoint. You are the DRI for training partner program success metrics. You'll design the operating model, set quality standards, and continuously optimize how the program runs. The first 6 months will include a heavier focus on learning the team's operations and designing workflow improvements; afterward, the role shifts more toward partner program ownership and strategic growth. Key Responsibilities Partner Program Design & Ownership Own execution and continuous improvement of the Authorized Training Partner (ATP) program: manage partner onboarding, certification, quality standards, and performance management within e
From $17/hr
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Looking for a career with purpose, stability, and opportunities to grow? Join CVS Health's Mail Order Pharmacy team and play an important role in helping patients receive the medications they need. This full-time opportunity offers guaranteed 40 hours per week, no direct customer interaction, competitive benefits, and the chance to advance your career in as little as six months based on performance and skill development. No Pharmacy Experience? No Problem. This is an excellent opportunity to begin a career in pharmacy and healthcare operations. Comprehensive training is provided, and your recruiter can guide you through the pharmacy technician licensing process. Starting Pay: $17.00/hour Additional Shift Differential: Earn an extra $1.00/hour for all hours worked on this shift. Why Join CVS Health? We invest in our colleagues and provide industry-leading benefits, including: 10 paid company holidays Up to 16 paid vacation days accrued during your first year Medical, dental, vision, life, and disability insurance Up to 30% employee discount at CVS stores Tuition reimbursement programs Paid parental leave 401(k) with company match Career advancement opportunities Comprehensive training and development Learn more about our benefits at www.BenefitMoments.com . About th
Other cities to consider
More places hiring for this role
Get new quality lab supervisor jobs in United States by email
Daily job updates · Unsubscribe anytime