About the Team The Synthetic RL team develops reinforcement learning methods that leverage synthetic data, environments, and feedback to train and evaluate frontier AI models. The team explores approaches such as self-play, simulators, and other synthetic evaluations to push model capability, generalization, and alignment beyond what is possible with the current prevailing methodology. About the Role As a Research Scientist on the Synthetic RL team, you will develop novel reinforcement learning techniques that use synthetic environments and feedback to improve large-scale models. You’ll work closely with other researchers to design experiments, analyze learning dynamics, and translate research insights into training approaches used in production systems. We’re looking for researchers who enjoy working on open-ended problems, value fast iteration, and want their work to directly shape how frontier models are trained. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Research and develop reinforcement learning algorithms Design and run experiments to study training dynamics and model behavior at scale Collaborate with engineers and researchers to integrate successful approaches into model training pipelines You might thrive in this role if you: Have a strong background in reinforcement learning, machine learning research, or related fields Have strong engineering and statistical analysis skills Enjoy exploring new problem spaces where data, objectives, and evaluation are imperfect or evolving Are motivated by seeing research ideas influence real-world AI systems About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an ex
Jobs in United States
Aws in San Francisco
866 active opportunities · Updated October 2026
Showing
15 jobs
Explore current aws jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency. The Model Policy team aligns model behavior with desired human values and norms. We co-design policy with models and for models by driving rapid policy taxonomy iteration based on data and defining evaluation criteria for foundational models’ ability to reason about safety. Key focus areas include: catastrophic risk, mental health, teen safety and multimodal safety. About the Role Providing access to frontier AI systems raises complex questions around dual-use science and catastrophic risk. How should models respond to requests involving chemical synthesis, biological experimentation, or pathogen research? Where is the boundary between legitimate scientific inquiry and information that could enable misuse? How do we design policies that meaningfully reduce risk without unnecessarily restricting beneficial research? This is a senior role in which you’ll help shape policy creation and development at OpenAI for addressing biological and chemical risks. You will develop structured policy frameworks and taxonomies to guide safe model behavior. This role sits at the intersection of biosecurity expertise, AI safety research, and policy design. You will help ensure that frontier AI systems can support beneficial life sciences research, such as drug discovery, public health, and biosafety, while reducing the risk that these capabilities could be misused. Our relevant publications: Preparedness framework Preparing for future AI capabilities in biology Safety evaluations hub OpenAI GPT5 System Card Evaluating Fairness in ChatGPT Improving Model Safety Behavior with Rule-Based Rewards OpenAI Model Spec Your Responsibilities: Design and maintain model policies governing chemical and biological risk, defining how models should safely handle dual-use scenarios. Develop structured taxonomi
About the Team OpenAI’s mission is to build safe artificial general intelligence (AGI) that benefits all of humanity. Achieving this requires bringing the world’s most exceptional talent under one roof to push the boundaries of what’s possible. Our Research Recruiting team plays a critical role in this effort. We are an embedded part of the research organization, working side by side with our research staff to deeply understand evolving priorities, build trust, and strategically shape the future of OpenAI’s talent. About the Role You will own and execute long-term talent strategies to identify, engage, and recruit many of the world’s leading and emerging AI researchers, research engineers, and technical scientists working at the frontier of machine learning. This is not a traditional execution-focused recruiting role. You will operate as a strategic partner to OpenAI’s research staff, helping define hiring priorities, shape search strategy, influence candidate evaluation, and guide hiring decisions that directly impact the direction and quality of our frontier-model research and fulfillment of our mission. In this role, you will: Partner directly with research and technical staff to define hiring priorities, shape search strategies, and anticipate future talent needs as technical roadmaps evolve. Proactively identify and cultivate exceptional AI/ML research talent across industry, academia, and emerging labs, often before formal hiring needs exist. Use market insights and candidate signals to influence hiring decisions, leveling, and compensation strategy for highly specialized research roles. Serve as a trusted advisor throughout candidate evaluation and closing — helping leaders calibrate for research excellence, long-term potential, and organizational fit. Collaborate closely with your sourcing partner to execute complex, high-impact searches in ambiguous or rapidly evolving technical domains. You might thrive in this role if you: Significant experience recruitin
About the Team The Support team is central to ensuring that our customers' experience with our products is nothing short of exceptional. We resolve complex issues, provide technical guidance, and support customers in maximizing value and adoption from deploying our products. We work closely with Sales, Technical Success, Product, Engineering and others to deliver the best possible experience to our customers at scale. OpenAI's customers represent a range of diverse backgrounds and maturity, from early-stage startups to established global enterprises. Given OpenAI’s breakneck shipping cadence and growth – and the expectation that it will only accelerate – our ability to architect automation systems and agentic workflows for scale is central to our ability to maintain exceptional support quality in the face of AGI. About the Role As a Support Vendor Manager, you will own the health, performance, and long-term scalability of multiple support partner and vendor relationships. This is a vendor leadership role first and foremost: you will drive commercial and operational accountability (SLAs, QBRs, escalation paths, remediation plans), while also building the operating model that enables support to scale without linear headcount growth. You’ll collaborate closely with User Operations teams (e.g., Trust & Safety, Fraud & Risk), Systems/Tooling, Data partners, and Product/PM stakeholders as we launch new workflow and launch and scale new programs. You’ll be responsible for: End-to-end vendor leadership: Own day-to-day oversight, relationship health, and executive-level accountability for multiple support vendors/BPOs. Performance management & remediation: Define and manage SLA/KPI performance expectations, run WBRs/QBRs, identify performance gaps, and drive structured turnaround plans with clear owners and timelines. Escalation and risk management: Serve as the primary escalation point for vendor issues, including incident response, surge events, quality regress
About the Team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role OpenAI is developing custom silicon to power the next generation of frontier AI models. We’re looking for experienced Design Verification (DV) Engineers to ensure functional correctness and robust design for our cutting-edge ML accelerators. You will play a key role in verifying complex hardware systems—ranging from individual IP blocks to subsystems and full SoC—working closely with architecture, RTL, software, and systems teams to deliver reliable silicon at scale. In this role you will: Own the verification of one or more of: custom IP blocks, subsystems (compute, interconnect, memory, etc.), or full-chip SoC-level functionality. Define verification plans based on architecture and microarchitecture specs. Develop constrained-random, directed, and system-level testbenches using SystemVerilog/UVM or equivalent methodologies. Build and maintain stimulus generators, checkers, monitors, and scoreboards to ensure high coverage and correctness. Drive bug triage, root cause analysis, and work closely with design teams on resolution. Contribute to regression infrastructure, coverage analysis, and closure for both block- and top-level environments. You might thrive in this role if you have: BS/MS in EE/CE/CS or equivalent with 3+ years of experience in hardware verification. Proven success verifying complex IP or SoC designs in industry-standard flows Proficient in SystemVerilog, UVM, and common simulation and debug tools (e.g., VCS, Questa, Verdi). Strong knowledge
About the Team Like every team at OpenAI, the Marketing team contributes to our broader mission of ensuring responsible and widespread adoption of artificial intelligence. With that aim in mind, we are responsible for developing and executing strategies that drive awareness, engagement, and usage for OpenAI’s products and platform amongst our core audiences. Our focus extends beyond just promoting product features; we aim to provide valuable insights and resources that help our users make the most out of AI technologies. The Marketing Insights & Analytics team aims to understand the markets, audiences, behaviors, attitudes and needs to shape outreach strategies and inform product and business decisions. We’re the strategic partner that discovers insights, specifically: foundational understanding, marketing strategy, thought leadership, measurement and tracking. About the Role You will be the Market Research Lead that translates insights into actionable marketing plans and executions, drive business results, and orient the team towards strategic thinking to drive brand outcomes. This is critical as we hope to reach millions of users and businesses worldwide. This role can be based in San Francisco, CA or NYC, NY utilizing a hybrid work model (3 days per week in-office). In this role, you will: Play an integral part of Marketing and be a strategic partner to brand, product marketing, creative, product and beyond. Design, execute, and deliver high-impact research using mixed methods. Work on a diverse portfolio of business challenges - how do we reach the next billion people, how do we grow internationally, how do we position our suite of products, and how do we ship joy to our customers. Translate business questions into research plans. Translate data into clear, compelling narratives. Develop deep empathy for developers, builders, and technical decision-makers, understanding their workflows, motivations, and pain points across the lifecycle Shape go-to-market str
About the Team The Codex team is responsible for building state-of-the-art AI systems that can write code, reason about software, and act as intelligent agents for developers and non-developers alike. Our mission is to push the frontier of code generation and agentic reasoning, and deploy these capabilities in real-world products such as ChatGPT and the API, as well as in next-generation tools specifically designed for agentic coding. We operate across research, engineering, product, and infrastructure—owning the full lifecycle of experimentation, deployment, and iteration on novel coding capabilities. About the Role As a Performance & Systems Engineer on the Codex team, you will be responsible for whole-system optimization across a complex, evolving stack. Codex spans LLM inference, cloud orchestration, agentic work management, and multiple product surfaces. Your job will be to identify and land high-leverage changes—across infrastructure, modeling, and product layers—that make Codex agents significantly faster and cheaper to serve. We’re looking for generalists who thrive in ambiguity and love chasing performance bottlenecks to ground. This is a high-ownership role where your work will directly improve the experience of millions of users. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Hunt down and address inefficiencies across the Codex system stack, from agent behavior to LLM inference to container orchestration, and beyond. Build tooling to measure, profile, and optimize system performance at scale. Collaborate with researchers and engineers to land high-ROI changes that improve latency and cost. You might thrive in this role if you: Have experience operating across both ML systems and cloud infrastructure. Enjoy diving into messy, ambiguous problems and emerging with clear wins. Think holistically about performance, balancing spee
About the Team Tax and Trade at OpenAI shapes business strategy by embedding critical tax, export control, customs, and cross-border considerations into how the company builds, sources, scales, and operates in support of the mission. We combine deep expertise with practical systems thinking to look around corners, identify emerging risks and opportunities early, and help teams make smarter decisions at the point where strategy becomes execution. Across procurement, hardware operations, manufacturing, logistics, finance, legal, supplier onboarding, and operator workflows, we build robust, scalable support services leveraging cutting-edge technology—including governed AI and automation—to make complex regulated work more durable, more efficient, and easier to scale. About the Role We’re hiring a Senior Manager, Export Controls to lead OpenAI’s export controls strategy and operating model. This is a senior role with broad scope across advanced computing, semiconductors, software, hardware, manufacturing, and high technology partnerships. You will refine how OpenAI classifies controlled technology, software, and hardware, structures access-controlled environments, manages licensing and supplier commitments, and scales export-control operations in a way that supports the company’s pace of innovation. You will also shape how OpenAI applies AI and agentic workflows to policy-heavy operational work, building systems that make complex rules easier to navigate and easier to execute. In this role, you will: Refine the strategy and operating model for OpenAI’s export controls program across advanced computing, semiconductors, software, hardware, manufacturing, and high technology partnerships. Own export classification and licensing strategy for controlled technical data, software, hardware, and research environments. Lead the design and operation of compliant controlled environments and related governance processes. Partner with Research and Infrastructure to support efficient
About the Role We are seeking a Cloud Infrastructure Engineer to help design and evolve the platforms that power OpenAI’s products. In this role, you will be a hands-on technical leader, driving the architecture, scalability, reliability, and security of critical infrastructure systems. You will help define how we build and operate infrastructure at the next order of magnitude, while influencing technical direction across teams. This role is both deeply technical and highly strategic, requiring strong ownership, sound judgment, and the ability to partner effectively across engineering, product, and research organizations. In this role, you will: Design and build scalable, reliable, and secure infrastructure platforms that power OpenAI products Evolve cloud infrastructure abstractions that enable rapid product development across teams Architect systems to support significant growth, performance, and operational complexity Improve server orchestration, networking, distributed systems reliability, and infrastructure security posture Influence technical direction and infrastructure strategy across multiple teams Partner closely with product, research, and engineering teams to align infrastructure with evolving needs Own operational excellence, including participation in on-call rotations, incident response, and production readiness Mentor engineers and raise the overall technical bar of the organization Contribute to a culture of high ownership, low ego, and thoughtful collaboration You might thrive in this role if you: 8+ years of experience building and operating large-scale infrastructure systems Deep expertise in Kubernetes and container orchestration at scale Strong experience designing cloud abstractions and platform infrastructure (AWS, GCP, Azure, or similar) Proven track record of leading complex technical initiatives across teams Experience operating highly reliable, secure, and scalable distributed systems Security engineering experience or security backgroun
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions tailored to the demands of advanced AI workloads. We work across the full stack—from silicon to system integration—partnering closely with internal teams and external vendors to define and deliver next-generation AI infrastructure. Our team focuses on defining scalable, high-performance system architectures and reference designs that balance performance, cost, and operational efficiency across rapidly evolving technologies. About the Role We are seeking a 3P Architect to define and drive rack- and cluster-level reference designs in collaboration with external partners. This role is responsible for translating workload requirements and system-level goals into concrete architectures, aligning partners on critical design attributes, and ensuring vendor roadmaps meet our infrastructure needs. You will work closely with performance modeling and internal architecture teams to evaluate tradeoffs, while owning the end-to-end definition and execution of third-party system designs. This includes identifying gaps in current technologies, driving vendor development, and shaping future infrastructure capabilities. This role requires strong system intuition, cross-functional leadership, and the ability to operate effectively across internal teams and external ecosystems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Define rack- and cluster-level reference architectures for AI infrastructure deployments. Translate workload requirements into clear system design specifications and partner deliverables. Collaborate with performance modeling teams to evaluate architectural tradeoffs and system behaviors. Align internal stakeholders and external partners on critical system attributes (performance, cost, power, reliability, scalability). Identify gaps in current technology offerings and dr
About The Team Our mission is to bring OpenAI products to life for every customer. Demo Experience equips customer-facing teams with the experiences, systems, and confidence to make frontier capabilities tangible, relevant, and trustworthy. OpenAI’s products and customer needs are evolving rapidly. Demo Experience closes the gap between a frontier capability and a credible customer experience—making new capabilities understandable, demonstrable, and reusable quickly at scale. Working across Product, Engineering, Marketing, Operations, and GTM, we turn recurring customer needs into reusable capabilities and raise the standard for every customer conversation. About The Role Demo Experience Engineers work at the intersection of product engineering, technical storytelling, and GTM execution. You will own ambiguous, high-leverage problems end to end—from building agentic prototypes to creating the infrastructure and self-service tools that make them reliable and reusable. Your work will help customer-facing teams move faster, reduce avoidable failures, and translate frontier product capabilities into clear customer value. You will also turn recurring patterns from customer-facing work into product feedback, launch-readiness improvements, and scalable systems. Success in this role means teams can demonstrate new capabilities sooner and with greater confidence. Recurring requests become reusable capabilities instead of one-off work. Demo experiences are accurate, reliable, and safe. Insights from customer-facing work improve product and readiness decisions. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In This Role, You Will Own end-to-end demo readiness for new and priority product capabilities, including environments, integrations, synthetic data, evaluations, reliability checks, and fallback paths. Build compelling prototypes, LLM agents, and reference flows that mak
About the Team The Platform Analytics team builds the systems OpenAI researchers use to understand the quality and behavior of the models we train including what models are doing, why they behave in a particular way, and how that behavior changes across experiments. Neptune is a core part of this work. It ingests, stores, queries, and visualizes large volumes of metrics from pretraining, post-training, and reinforcement learning. Hundreds of researchers depend on these systems in their daily work to compare experiments, debug unexpected behavior, and decide what to try next. Our scope is broader than metrics. We also build platforms that help researchers analyze samples, traces, evaluation results, and other structured or unstructured data through dashboards, APIs, and increasingly agent-driven workflows. These systems need to remain fast, reliable, and understandable as the scale and complexity of research change quickly. We are not trying to become a consulting team that builds a separate solution for every research project. We work directly with researchers to understand recurring problems, then turn them into reusable infrastructure and platform capabilities that many teams can build on. About the Role We’re looking for a hands-on experienced software engineer who can take ownership of a critical system and drive it from problem definition through production adoption. This person should be able to own a platform such as CacheHouse end to end: define its technical direction, design its data model and storage architecture, integrate it with several research dashboards and workflows, guide one or two engineers, and ensure the system works reliably for its users. The right candidate should already bring the technical judgment, ownership, and execution expected at this level. The primary learning curve should be OpenAI’s stack and research problem space, not learning how to lead a complex engineering effort or deliver a production system. You will work directly with
About the Team The Personalization-Memory team, within OpenAI's broader Personal AGI organization, is focused on developing agents that can learn from prior interactions in order to become more helpful and efficient over time. We build general-purpose memory and personalization capabilities that transfer across ChatGPT and other agentic products, and we collaborate with applied engineering on the product surfaces that allow users to interact with memory. About the Role As a Research Engineer / Research Scientist on the Personalization-Memory team, your work will span memory architecture, post-training, and developing long-horizon tasks for training and evaluations. We're looking for individuals who have a background in reinforcement learning research, are able to iterate quickly, and who can convert scientific rigor and long-term research into realized product impact. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own and pursue a research agenda for improving long-horizon memory and personalization in frontier models. Build robust evaluations for tracking modeling improvements. Design, implement, test, and debug code across our research stack. Collaborate closely with the research and product teams to influence the shape of technical solutions in the product. You might thrive in this role if you: Love being on the cutting edge of RL and frontier model research. Value principled approaches and research craftsmanship. Are passionate about long-horizon tasks, memory, and turning your research into product impact. Are comfortable diving into a large ML codebase to debug. Thrive in a fast-paced, dynamic, and technically complex environment. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI syst
About the Team The Agent Enablement team works across engineering, product, design, and research to bring our technology to the world. We seek to learn from deployment and broadly distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. We aim to make our innovative tools globally accessible, transcending geographic, economic, and platform barriers. Our commitment is to facilitate the use of AI to enhance lives, supported by rigorous insights into how people use our products. About the Role We are looking for experienced full-stack engineers to join our new Agent Enablement team. Our goal is to design and grow an open ecosystem of agent-enabled sites and services. This is a wide-ranging role: you’ll build new user and agent identity protocols, user experiences to control and observe agents across web, desktop, and mobile, and much more. We will rely on you to drive our technical decisions while also steering our product and partnership direction, optimizing for both short-term impact and long-term success of the ecosystem. We value engineers who are impact-driven, autonomous, and adept at removing barriers to forward progress. In this role, you will: Design the primitives and protocols for an open agent ecosystem, enabling our users’ agents to make the best use of sites and services across the internet. Build a next-generation user experience to observe and control agents, across web, desktop, and mobile. Evolve our approach to token consumption across subscriptions and API customers. Execute on fast-paced projects in collaboration with research, design, data science and other product engineering teams. Work closely with our strategic customers and partners to grow the ecosystem. You might thrive in this role if you: Have strong full-stack engineering skills and experience shipping customer-facing products from concept to production. You’re comfortable working across frontend, backend, APIs, data models, and product desig
About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for engineers to build the infrastructure that powers Codex agents in production. This role focuses on the systems that let models safely execute code, interact with tools, complete long-running tasks, and operate reliably and efficiently at scale. You’ll design and operate the infrastructure behind sandboxed execution, orchestration, stateful workflows, app-server and SDK boundaries, and model rollouts. You’ll work at the intersection of distributed systems, developer tooling, and AI, building primitives that make Codex faster, safer, more reliable, and easier for the rest of the organization to build on. What You’ll Do Design and build execution environments for AI agents, including sandboxing, isolation, and reproducibility. Develop systems for agent orchestration across multi-step, tool-using workflows. Build infrastructure for running, testing, and debugging code generated by models. Create state and memory systems that allow agents to persist context across long-running tasks. Optimize tokens, latency, reliability, and cost across Codex’s production fleet. Support model rollouts, capacity planning, and the core tradeoffs between quality, speed, and economics to manage a fleet of frontier agents at scale. Build shared platform capabilities that unblock product teams, partner teams, and open source Codex. Yo
Other cities to consider
More places hiring for this role
Get new aws jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime