Jobs in United States

Model Designer in San Francisco

839 active opportunities · Updated October 2026

Explore current model designer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team The People Technology team builds and operates the systems that support how OpenAI hires, develops, and supports its people. The team brings together People Systems and People Innovation Labs, a product engineering group focused on rethinking how we find and retain exceptional talent and help employees do their best work. People Systems owns the company’s core people-technology ecosystem, including platforms such as Workday and Ashby. People Innovation Labs builds new employee and People Team experiences on top of that foundation, including OpenHouse, our internal employee hub, and AI-powered products and automations. Together, we are working toward a model in which our enterprise systems provide reliable data, controls, and core business logic, while employees and managers can complete more of their work through simple, integrated, and AI-native experiences. About the Role We are looking for a People Systems Lead to manage the People Systems team and shape how our core systems evolve. You will be responsible for the reliability and effectiveness of our current environment while helping us move beyond the constraints of traditional enterprise software. This includes stabilizing and improving platforms such as Workday and Ashby, designing the integrations that connect them to the broader technology ecosystem, and partnering with People Innovation Labs to surface workflows through OpenHouse, Slack, and AI-powered experiences. This role requires someone who is comfortable moving between strategy, technical design, and team leadership. You should understand People systems deeply, be able to work through integration and architecture decisions with engineers, and translate complex organizational needs into scalable solutions. You will also manage vendor relationships, develop the People Systems team, and drive alignment across People, Engineering, Finance, Security, Legal, and other partners. This role could be a fit for someone who has grown up in People S

AWSRestAIGo
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As an AI Accelerator Systems Software Technical Program manager at OpenAI, you will help bring our chips/system hardware roadmap to life, navigating an array of technical and partnership challenges. We’re looking for people excited to push the frontiers of computing by navigating technical explorations and are passionate about building. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Manage the end-to-end software development from design to implementation for our AI acceleration systems, working across technical, cross-functional and external stakeholders Lead planning and scheduling of AI system software designs with our strategic partners and vendors Coordinate and lead internal resources and communication for efficient interaction with partners and vendors. You might thrive in this role if you: Have experience as a software technical program manager for data center system products (server, GPU, TPU, networking, storage and so on) taking products from concept to volume in a data center environment ensuring the systems scale with high quality Know end-to-end software development program management techniques from concept, design, production, deployment into the data center Want to help design some of the world’s largest supercomputing systems, working at the edge of complex hardware challenges Enjoy working with and enabling world-clas

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team At OpenAI, the User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving company priorities. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are looking for a senior program manager to build the safety, quality, and risk operations supporting a new category of consumer devices. You will translate ambiguous product risks and evolving requirements into practical operating models, workflows, escalation paths, launch-readiness plans, and cross-functional decision-making. This is a foundational role: the systems you build will shape how OpenAI launches, monitors, and improves a new category of consumer devices safely at scale. You will help establish how potential safety incidents, product-quality concerns, sensitive customer escalations, privacy-sensitive issues, and other emerging device risks are identified, investigated, resolved, and incorporated into product and operational improvements. You will turn incomplete requirements into practical workflows, decision rights, launch plans, quality controls, measurement, and durable ownership. The role centers on program building, operational judgment, and execution. We welcome candidates from product safety, quality assurance, regulatory operations, technical program management, healthcare, medical devices, aerospace, consumer technology, and other environments involving complex products or regulated risks. Direct consumer-hardware experience is helpful but not required. The strongest candidates learn unfam

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team GTM Growth Engineering builds AI-native products that help OpenAI's go-to-market and B2B marketing organizations scale with greater speed, intelligence, and operational effectiveness. We apply OpenAI models to real business workflows and build the systems that make those applications useful and dependable: customer context, agent behavior, feedback, evaluation, experimentation, and appropriate human oversight. Our work brings together software engineering, applied AI, product, data, and GTM operations. We measure success through the quality of customer engagement, pipeline, conversion, and the effectiveness of our sales and marketing teams. About the Role We're looking for an Applied AI Engineer to build production systems that help AI-powered go-to-market workflows improve over time. You will connect agent behavior, customer and operator feedback, evaluation, experimentation, and business outcomes to make these systems more effective, reliable, and responsive to evolving customer needs. This is a deeply technical, cross-functional role with end-to-end ownership of the agent improvement loop: understand production behavior, identify failure modes, improve how the system decides or acts, and validate the resulting impact. You will partner with Engineering, Product, Data Science, Sales, and B2B Marketing to turn real-world signals into safer, more effective agent behavior and measurable improvements in customer engagement, conversion, qualified pipeline, and team productivity. In this role, you will: Own the production improvement loop across agent behavior, customer and operator feedback, evaluation, experimentation, and verified business outcomes. Instrument agent workflows so model interactions, tool use, decisions, failures, human edits, and downstream outcomes can be understood in context. Define meaningful quality standards, representative evaluation datasets, regression coverage, and production monitoring for real GTM workflows. Investigate why a

PythonAWSRestAI
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team OpenAI is at the center of some of the highest-impact multimodal work in AI. ChatGPT serves a massive global audience, and enables diverse interactions via text, speech, and visuals. As interactive surfaces grow, models also need to adapt to emerging harm, understand user intent and situational context, and respond appropriately. The Chat and Multimodal Safety team is responsible for ensuring that OpenAI’s increasingly multimodal models and products behave safely across these experiences. We develop the research, training methods, and evaluations needed to make these experiences safe. Our work sits at the frontier of responsibly deploying powerful AI, in close partnership with Personal AGI, io, model training, and product teams. About the Role As a Researcher on the Chat and Multimodal Safety team, you will help shape how frontier models perceive and reason the world, and translate that understanding into safe behavior. We’re looking for people who combine deep technical expertise with strong safety judgment. Strong candidates often bridge perception and language: they may have built vision-language models, worked on modality fusion or image encoders, developed multimodal post-training or evaluations, or advanced safety for image, video, or audio systems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Define and advance multimodal safety research for text, vision, and audio, connecting perception and semantic understanding to safe model behavior. Build training and evaluation methods for VLMs, including post-training, safety evals, and interventions that help models respond safely and appropriately in varied contexts. Collaborate closely with Personal AGI, Consumer Devices, and product/model teams to translate research into safer ambient, embedded, and personalized multimodal experiences. You might thrive in this role if you:

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team The Growth team drives user and revenue growth across ChatGPT’s consumer and business segments as well as other OpenAI products worldwide. We operate across the full funnel - from awareness and acquisition through activation, retention, and expansion - using a combination of global performance marketing, AI-powered workflows, in-product optimization, insights, experimentation, and creative ops engineering. About the Role We are hiring a Lifecycle Lead to build the company-wide owned-channel capability that helps teams reach users with relevant, timely, and trustworthy experiences. This senior, hands-on leader will set the lifecycle strategy, partner with Engineering to build the orchestration and deployment platform, and establish the operating model that allows teams across the company to launch and improve evergreen programs safely at scale. You will sit at the intersection of platform, product, and campaign strategy. You will partner with Engineering, Product, Data Science, and Analytics on the underlying systems, and with Product Marketing Managers and other client teams to design journeys that help new, active, and returning users reach value and build durable habits. In this role, you will: Partner with product to set the company-wide vision, roadmap, and operating model for lifecycle and owned-channel engagement. Partner with Engineering, Product, Data Science, and Analytics to shape the tooling and infrastructure for identity, audiences, eligibility, consent, triggers, orchestration, decisioning, frequency, experimentation, localization, quality assurance, and observability. Define scalable deployment workflows—including self-service and centrally supported paths, intake, templates, approvals, governance, service levels, and incident response—so teams across the company can launch safely and efficiently. Partner with Product Marketing Managers and other client teams to translate audience, product, and business goals into evergreen journey stra

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving areas of work. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are seeking a Device Safety & Risk Operations Specialist to build the safety operating model for a new category of consumer hardware. This is a senior individual-contributor role for someone who can turn emerging product risks and incomplete requirements into practical workflows, controls, launch plans, and durable systems. You will define how product-safety incidents, critical escalations, regulated cases, and privacy-sensitive issues should be identified, investigated, escalated, resolved, and learned from. You will also establish operational requirements for case management, data access, decision logging, quality assurance, monitoring, and cross-functional response. You will stand up priority workflows through launch and early operations, then help transition them into durable homes across USRO and partner teams. The right person combines deep operational judgment with strong technical and hardware product fluency. They can move from executive-level risk framing to detailed workflow design, tabletop exercises, launch readiness, frontline guidance, and post-launch improvement. Location / work model: San Francisco, CA; hybrid, 3 days/week in-office. Please note: This role may involve exposure to sensitive or concerning material. Strong discretion, judgment, and resilience are essential. In This Role, You Will:

AWSRestAIGo
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team The CoT Monitorability team at OpenAI studies whether and when the chain-of-thought of frontier reasoning models is monitorable enough to support scalable oversight. We study how to measure monitorability , which training mechanisms affect monitorability, and speculative methods to improve monitorability. While we mostly focus on CoT monitorability at the moment, we care more generally about any form of monitorability, auditing methods, and improving alignment. We were the first to show that chain-of-thought monitoring can be a practical additional safety mechanism, and today our monitoring systems are actively used on OpenAI’s largest RL training runs to detect misbehavior. The issues we surface are then used to help improve our reward functions, environments, etc (without directly training against a CoT monitor). Our work sits in Alignment and intersects with model training, alignment evaluations, monitoring, and frontier-risk research.We care most about monitorability where the stakes are high, and about preserving useful oversight signals as models become more capable. About the Role We’re looking for a researcher with strong empirical ML expertise and a deep interest in model behavior, alignment, or interpretability. Direct chain-of-thought interpretability experience is welcome but not required; strong candidates may come from broader interpretability, alignment, model training, or investigative model-behavior work. As a researcher on the Alignment team, you will design and run experiments that improve our understanding of model monitorability. You will investigate how training interventions across the model-development pipeline influence whether reasoning remains legible, build evaluations that make those questions measurable, and help translate findings into practical oversight and training recommendations. You may also help develop new monitoring models or methods and apply them to OpenAI’s largest training runs. This role is especially well

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team OpenAI’s Industrial Compute organization is building the infrastructure required to support the next generation of frontier AI systems. Through a combination of strategic partnerships and self-built data center campuses, we are scaling the power, cooling, electrical, mechanical, and controls infrastructure needed to deliver compute at unprecedented scale. The Commissioning organization is responsible for ensuring this infrastructure is safely tested, validated, integrated, and transitioned into reliable operations. For our self-build campuses, the team operates through a hybrid delivery model: OpenAI provides commissioning leadership, discipline ownership, governance, and project integration, while commissioning partners provide field and test engineering capacity to support inspections, startup, testing, and turnover. About the Role We are seeking a Commissioning Project Lead to own the commissioning strategy and execution for a large-scale, self-build data center project. You will lead the overall commissioning program from early construction planning through startup, functional testing, integrated systems testing, and final turnover. You will establish the commissioning execution plan, integrate commissioning activities into the master project schedule, coordinate multidisciplinary readiness, and lead the vendor commissioning partners providing field and test engineering capacity. This role serves as the primary commissioning interface to project leadership, construction management, contractors, equipment vendors, operations, and commissioning partners. You will be responsible for creating clarity across organizations, identifying readiness and schedule risks early, and ensuring the facility progresses through testing and turnover against clearly defined acceptance criteria. The role will initially support planning and coordination in a hybrid capacity and transition to full-time onsite presence as construction, inspections, startup, testing, and t

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving areas of work. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are seeking a Device Safety & Risk Operations Specialist to build the safety operating model for a new category of consumer hardware. This is a senior individual-contributor role for someone who can turn emerging product risks and incomplete requirements into practical workflows, controls, launch plans, and durable systems. You will define how product-safety incidents, critical escalations, regulated cases, and privacy-sensitive issues should be identified, investigated, escalated, resolved, and learned from. You will also establish operational requirements for case management, data access, decision logging, quality assurance, monitoring, and cross-functional response. You will stand up priority workflows through launch and early operations, then help transition them into durable homes across USRO and partner teams. The right person combines deep operational judgment with strong technical and hardware product fluency. They can move from executive-level risk framing to detailed workflow design, tabletop exercises, launch readiness, frontline guidance, and post-launch improvement. Location / work model: San Francisco, CA; hybrid, 3 days/week in-office. Please note: This role may involve exposure to sensitive or concerning material. Strong discretion, judgment, and resilience are essential. In This Role, You Will:

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team OpenAI’s Education team is building products that advance how people learn with AI. The team works across higher education institutions, K-12 districts, and country-level partnerships, including applied research on how AI affects learning and cognitive outcomes. The team owns owns ChatGPT Edu, ChatGPT for Teachers, and related product/research work. The team partners closely with go-to-market, research, Consumer Learning, and model teams to turn education-specific insights into product experiences that can improve ChatGPT more broadly. Some of our recent work: New Education Plugins for ChatGPT Work and Codex New tools for understanding AI and learning outcomes Education for countries Advancements in higher education Early product work - Introducing Study Mode About the Role We’re looking for a hands-on Tech Lead Manager to lead and manage a team of senior full-stack engineers building AI-native learning experiences in ChatGPT. This person will combine technical execution, product judgment, and people leadership: they will write and ship code, manage engineers, and help shape the product direction for how students and Educators use AI. In This Role, You Will Lead and manage a team of three senior full-stack engineers. Build product experiences for ChatGPT Education, ChatGPT for Teachers, and AI-native learning workflows. Partner with research teams on field studies, randomized control trials, classifiers, data pipelines, and cognitive-outcome measurement. Collaborate with Consumer Learning and model teams to translate education insights into broader ChatGPT behavior and product improvements. Drive execution across product, engineering, research, go-to-market, and partner teams. Help define product strategy, priorities, and delivery plans for a new product pod. You Might Thrive In This Role If You Have several years of direct people-management experience with engineers. Are still highly technical and comfortable doing IC engineering work. Have strong pr

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Role We're looking for a Head of Competitive Intelligence to build and lead a world-class competitive intelligence function. This person will transform market signals into strategic advantage by developing the frameworks, analysis, and insights that inform our product strategy, go-to-market approach, and executive decision-making. This is not a research role. It's a highly cross-functional strategy role that sits at the intersection of Product, GTM, Research, and Leadership. You'll establish the systems, operating cadence, and analytical rigor that help the company understand where we win, where we're vulnerable, and how the market is evolving. What You'll Do Build and own the company's competitive intelligence strategy and operating model. Develop and maintain comprehensive competitive assessments, including Harvey Ball analyses, feature matrices, product comparisons, pricing analyses, and market landscape reviews. Create executive-ready competitive insights that influence product strategy, roadmap prioritization, pricing, and investment decisions. Build best-in-class objection handling and competitive messaging for Sales, Marketing, Customer Success, and Partnerships. Monitor competitor product launches, model releases, acquisitions, partnerships, pricing changes, funding, and GTM motions, synthesizing complex information into clear strategic recommendations. Establish repeatable processes and tooling for collecting, validating, and distributing competitive intelligence across the company. Partner closely with Product Management to inform roadmap decisions and identify areas of strategic differentiation. Collaborate with Marketing to sharpen positioning and messaging based on competitive dynamics. Support executive leadership with board-ready competitive analyses and market briefings. Build dashboards and reporting that track competitive movements and emerging industry trends. Develop frameworks that evaluate competitors across capabilities, enterprise r

AWSRestMachine LearningAI
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team OpenAI is building AI systems that can help professionals perform complex, high-value work with greater speed, rigor, and creativity. Investment banking is one of the most demanding environments for knowledge work: bankers must synthesize fragmented information, exercise judgment under pressure, and produce precise, defensible models, analyses, and client materials. Our team works across Research, Product, Engineering, and Go-to-Market to make OpenAI's models genuinely useful for these workflows. We translate real professional work into product requirements, evaluations, training signals, and repeatable customer solutions. We care not only whether a model can generate an answer, but whether it can deliver accurate, defensible work that experienced bankers can trust and use. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. About the Role We are looking for a Subject Matter Expert in Investment Banking to help define what excellent AI-assisted banking work looks like and turn that standard into better models and products. You will bring deep, current knowledge of how investment banking work is actually performed, including company and industry research, financial analysis and modeling, valuation, diligence, transaction execution, and the creation and review of client materials. You will use that expertise to design realistic tasks and evaluations, create and assess high-quality reference work, diagnose model failures, and help our technical teams improve model behavior and product experiences. This is a hands-on individual-contributor role for someone who enjoys both doing the work and explaining what makes it good. You should be comfortable moving between an Excel model, a presentation, a source document, an evaluation rubric, a product prototype, and a conversation with researchers or customers. You will help us distinguish outputs that merely look pl

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team The Safety Systems team is responsible for various safety work to ensure our best models can be safely deployed to the real world to benefit the society and is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency. The Model Safety Research team aims to fundamentally advance our capabilities for precisely implementing robust, safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to enforce nuanced safety policies without trading off helpfulness and capabilities, how to make the model robust to adversaries, how to address privacy and security risks, and how to make the model trustworthy in safety-critical domains. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. About the Role OpenAI is seeking a senior researcher with passion for AI safety and experience in safety research. Your role will set directions for research to enable and empower safe AGI and work on research projects to make our AI systems safer, more aligned and more robust to adversarial or malicious use cases. You will play a critical role in shaping how a safe AI system should look like in the future at OpenAI, making a significant impact on our mission to build and deploy safe AGI. In this role, you will: Conduct state-of-the-art research on AI safety topics such as RLHF, adversarial training, robustness, and more. Implement new methods in OpenAI’s core model training and launch safety improvements in OpenAI’s products. Set the research directions and strategies to make our AI systems safer, more aligned and more robust. Coordinate and collaborate with cross-functional team

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.1%

About the Team The Post-Training Frontiers team is responsible for training the frontier agents OpenAI ships to the world (GPT-Next). We train the flagship agentic models behind Codex, ChatGPT, and the API through large-scale reinforcement learning. The team’s work spans four areas. First, execution and science: working with teams across OpenAI to decide what can go into the final model and how, using scientific experiments and evals that are representative of the final pipeline so issues can be recognized early. Second, RL scaling: executing the final large-scale reinforcement learning run, making sure GPUs are used efficiently and training stays healthy. Third, research: improving horizontal capabilities like instruction following, factuality, memory, and multi-agent behavior, where the team’s broad visibility helps identify cross-cutting improvements across teams and domains. Fourth, engineering: maintaining the infrastructure stack and internal tools to ensure that both the final run and all integrations go as smoothly as possible and that the systems are easy to work with. About the Role This role focuses on keeping our frontier RL training runs fast, reliable, and unblocked. You will work across engineering and infrastructure problems as they emerge, from scaling and orchestration issues to inference bottlenecks, numerical problems, and hardware failures, as well as supporting large horizontal integrations in the big run, like multi-agent capabilities or memory. This is a role for a strong generalist who quickly learns anything needed for the task, has high attention to detail, debugs deeply, and is motivated by fixing the highest-impact problem in front of the team. In this role, you will: Keep large-scale async RL training runs moving by jumping into the most urgent engineering and infrastructure problems. Debug issues across training systems, inference, orchestration, scaling, and distributed infrastructure. Improve the reliability and efficiency of RL trai

AWSRestAIGo
🔔

Get new model designer jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime