Jobs in United States

Aws And Tooling Platform Lead in San Francisco

866 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will build the low-level device runtime that turns compiled programs into efficient, functional and performant execution on OpenAI’s custom AI accelerator. This software will schedule kernel launches, manage device memory and address spaces, coordinate synchronization, and expose reliable abstractions to higher-level runtimes and frameworks. You will work at the boundary of software and hardware, partnering with compiler, kernel, architecture, verification, and silicon teams to define interfaces and validate behavior. You will also use and improve event-based, cycle-accurate simulation to develop runtime capabilities before silicon is available, diagnose performance and correctness issues, and guide hardware-software co-design. In this role, you will: Design and implement the low-level device runtime for OpenAI custom silicon. Build kernel-launch scheduling, command submission, queueing, dependency tracking, and completion handling. Manage device memory spaces, allocation, virtual-to-physical mappings, data movement, and lifetime across concurrent workloads. Implement synchronization primitives, events, barriers, streams, and ordering guarantees that are correct and efficient. Define clean interfaces between the runtime, drivers, firmware, compiler-generated code, kernels, and higher-level execution systems. Use event-based, cycle-accurate simulators to develop, validate, debug, and performance-tune runtime behavior before and after silicon availability. Di

AWSRestAIC++
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -82%

About Team Our Robotics team is focused on unlocking general-purpose robotics and advancing toward AGI-level intelligence in dynamic, real-world environments. Working across the full model and systems stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the physical constraints of real-world systems to improve people’s lives. About the Role We are looking for a TPM to drive development and integration of a range of sensor systems for robotics. This role will drive cross-functional alignment across requirements, engineering design, integration, validation, manufacturing, supply chain, and release processes, helping turn complex sensor-system needs into clear plans, decisions, and milestones. Location and in-person expectations: This role is based in San Francisco, CA and requires in-person presence 4 days a week. In this role you will: Drive requirements alignment across engineering design, integration, testing, and validation for camera modules, LiDAR, IMUs, RADAR, proximity sensors, audio components and the systems they interact with. Coordinate the integration of modules including electrical, mechanical, harnessing, and software interfaces with the full robotic system with deep understanding of timelines to drive the respective PCBAs, enclosures, build and test fixtures, connectors and cables. Establish effective cadences for technical reviews, BOM readiness, change management, production releases, approvals, and decision tracking. Align harnesses, fasteners, assembly fixtures, test fixtures, and documentation so cross-functional teams can execute against a clear plan. Lead validation planning around functional, reliability, NVH failure modes, including testing needs, schedules, and exit criteria. Partner with manufacturing and supply chain to manage handoffs, lead times, dependencies, and production readiness. Drive tradeoff decisions across cost, qua

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -82%

About Team Our Robotics team is focused on unlocking general-purpose robotics and advancing toward AGI-level intelligence in dynamic, real-world environments. Working across the full model and systems stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the physical constraints of real-world systems to improve people’s lives. About the Role We are looking for a Technical Program Manager to own actuator development and integration from system goals through production readiness. The actuator program spans mechanical, electrical, firmware, harnessing, controls, test, reliability, manufacturing, and supply chain, and needs a TPM who can turn cross-functional decisions into clear scope, executable milestones, and timely decisions. In this role, you will help the team converge on the right technical plan, surface risks early, and deliver reliable actuator systems on schedule. This role is based in San Francisco, CA and requires in-person presence 4 days a week. In this role you will: Drive actuator programs end-to-end, aligning scope, milestones, interfaces, dependencies, and exit criteria across engineering teams. Drive scope lock and technical convergence for sprints, MVPs, and stretch goals while connecting component decisions to system performance. Coordinate actuator development across motors, gears, sensing, electronics, and firmware, and align the electrical and mechanical interfaces that connect actuators to the broader robot. Lead validation planning from early prototypes through engineering validation, reliability testing, and production readiness. Drive tradeoff decisions across cost, quality, performance, schedule, and lead time by collaborating cross-functionally and quantifying impact to meet program deliverables. Establish effective mechanisms for technical reviews, change control, design releases, decision tracking, and manufacturing readiness.

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We’re looking for a Product Manufacturing Engineer to drive manufacturing strategy and execution for next-generation AI hardware within the Silicon & Systems team, with a particular focus on PCBA manufacturing, assembly, process development, and production readiness. You will work closely with design engineering, systems engineering, operations, TPMs, contract manufacturers, suppliers, and other external partners to ensure new hardware products successfully transition from concept and prototype builds through NPI and high-volume production. You’ll own critical manufacturing initiatives, identify and resolve production risks, and help establish the processes and controls required to deliver complex hardware at scale. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Drive manufacturing and quality initiatives to ensure product success from concept and early development through NPI, production launch, and scale. Lead manufacturing process development for next-generation AI hardware systems, partnering closely with design engineering, systems engineering, operations, TPMs, suppliers, and manufacturing partners. Own PCBA manufacturing development and production readiness, including assembly processes, process validation, manufacturing test, rework, yield improvement, and successful transition into volume production. Establish NPI manufact

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. Role Overview We are seeking an experienced ASIC Package Signal Integrity / Power Integrity Engineer to drive electrical architecture, modeling, optimization, and validation for the most advanced AI/HPC silicon and package design. This role focuses on high-speed SerDes and memory channel architecture, advanced 2.5D/3D package SI/PI, substrate to package co-design, power-delivery-network optimization, electromagnetic modeling, and simulation-to-measurement correlation. The ideal candidate has strong hands-on experience with high-speed channel and PDN analysis across ASIC packages, interposers, substrates, and power-delivery structures, and can translate simulation results into practical design requirements for interposers and package substrate design optimization. The engineer will work closely with ASIC, package, system, mechanical, thermal, power, and silicon validation teams from early architecture and feasibility studies through production bring-up. In this role you will Own SI/PI architecture and analysis for advanced AI ASIC packages from early feasibility studies through production. Develop and optimize high-speed electrical channels for 200G/400G SerDes, PCIe, HBM, DDR, and chiplet/die-to-die interfaces. Perform package, interposer and substrate modeling using 2D/3D electromagnetic solvers. Define and optimize package stack-ups, transmission-line structures, via transitions, breakout structures, return paths, ground shielding, bump maps, and ball maps based on SI/PI re

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI's Enterprise team builds AI-powered enterprise products and shared platform capabilities that help organizations put advanced AI to work securely and at scale. Our work spans enterprise workflows, agent experiences, integrations, identity, administration, security, governance, and deployment. About the Role As a Technical Program Manager on Enterprise, you will lead the technical strategy and execution behind the products and shared capabilities that make ChatGPT, Codex, and future OpenAI products useful, secure, and scalable for organizations. You will translate customer needs, competitive dynamics, and product priorities into actionable plans, influence architectural direction, and deliver durable capabilities across application, platform, and infrastructure layers. The role requires deep technical fluency, strong product judgment, and the ability to move between hands-on execution and broader enterprise strategy. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Drive technical strategy and execution for enterprise product and AI workflow initiatives, from design through implementation, launch, customer rollout, and iteration. Partner with engineering teams to influence architectural direction, interface definitions, and implementation tradeoffs across full-stack products, APIs, integrations, and shared platform systems. Translate enterprise customer requirements into actionable product priorities across AI-powered workflows, agent experiences, integrations, permissions, data access, evaluations, identity, security, governance, and deployment readiness. Represent the needs of enterprise buyers, IT administrators, security teams, business leaders, developers, and end users in product and technical decisions. Identify adoption barriers, competitive gaps, and opportunities to make OpenAI products easier for organizations

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As an AI Accelerator Systems Software Technical Program manager at OpenAI, you will help bring our chips/system hardware roadmap to life, navigating an array of technical and partnership challenges. We’re looking for people excited to push the frontiers of computing by navigating technical explorations and are passionate about building. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Manage the end-to-end software development from design to implementation for our AI acceleration systems, working across technical, cross-functional and external stakeholders Lead planning and scheduling of AI system software designs with our strategic partners and vendors Coordinate and lead internal resources and communication for efficient interaction with partners and vendors. You might thrive in this role if you: Have experience as a software technical program manager for data center system products (server, GPU, TPU, networking, storage and so on) taking products from concept to volume in a data center environment ensuring the systems scale with high quality Know end-to-end software development program management techniques from concept, design, production, deployment into the data center Want to help design some of the world’s largest supercomputing systems, working at the edge of complex hardware challenges Enjoy working with and enabling world-clas

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Role We are looking for a senior partner success and learning professional to activate cloud partners to co-sell and deploy customer solutions at scale. This role leads the enablement strategy and execution that equip partner teams to turn joint priorities into measurable customer and business outcomes. As a trusted advisor to strategic partners, you will turn shared business priorities into practical enablement programs that help partner sales, technical, customer success, and leadership teams identify joint opportunities, co-sell effectively, and deploy customer solutions successfully. You will own the partner enablement relationship for an assigned portfolio and work across internal teams to ensure programs are coordinated, scalable, and connected to measurable partner, customer, and business outcomes. This is not a traditional training delivery role or a quota-carrying partner account manager position. It is a consultative, partner-facing role for someone who can shape strategy, build relationships, translate complex AI capabilities into effective learning experiences, lead persuasive enablement sessions, and guide strategic partners toward sustained success. In this role, you will: Own strategic partner enablement and success Serve as the primary enablement and success lead for a portfolio of strategic cloud platforms. Develop a deep understanding of each partner’s business, priorities, audiences, existing learning infrastructure, and strategic relationship with our company. Build partner-specific enablement strategies aligned to shared goals, priorities, and the broader partnership strategy. Establish trusted relationships with partner leaders, alliance teams, enablement stakeholders, sales organizations, technical communities, and post-sales teams. Lead ongoing planning conversations, enablement reviews, stakeholder check-ins, and executive updates that keep partners and internal teams aligned. Identify gaps, risks, opportunities, and emerging partn

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The User Operations team (Support) is central to ensuring that our customers' experience with our products is nothing short of exceptional. We resolve complex issues, provide technical guidance, and support customers in maximizing value and adoption from deploying our products. We work closely with Sales, Technical Success, Product, Engineering and others to deliver the best possible experience to our customers at scale. OpenAI's customers represent a range of diverse backgrounds and maturity, from early-stage startups to established global enterprises. About the Role OpenAI’s top-of-funnel is evolving into a signal-driven, data-orchestrated engine powered by AI-native workflows, enriched customer intelligence, and increasingly interconnected systems. We are looking for a Technical Systems Program Manager to help drive the strategy, alignment, prioritization, and execution of these systems across Salesforce and adjacent platforms. This role will partner closely with Engineering, Revenue Operations, Data, and Enterprise Platform teams to operationalize scalable CRM workflows. The ideal candidate combines strong technical fluency, operational systems thinking, and stakeholder leadership with the ability to drive execution across highly cross-functional environments. This role is based in San Francisco, CA, and the team works a hybrid schedule (Monday - Wednesday in office). We will offer relocation assistance if needed. In this role, you will: Own and drive CRM systems initiatives across support Partner with senior business stakeholders and engineering teams to identify operational pain points, define scalable solutions, evaluate technical tradeoffs, and execute across Salesforce and adjacent platforms Align cross-functional stakeholders including Revenue Operations, Support Delivery, Risk, Engineering, Data, Legal, and Enterprise Platform teams while driving prioritization and execution Help operationalize AI-native workflows by defining and building s

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team is responsible for building the next generation of AI-native silicon while working closely with software and research partners to co-design hardware tightly integrated with AI models. In addition to delivering production-grade silicon for OpenAI’s supercomputing infrastructure, the team also creates custom design tools and methodologies that accelerate innovation and enable hardware optimized specifically for AI. About the Role We're looking for an Optical Interconnect System Engineer to design, qualify, and deploy scalable optical connectivity for large-scale AI infrastructure. This role spans fiber-system architecture, optical-mechanical integration, validation, reliability, deployment, and serviceability. You will work with optical, mechanical, electrical, networking, manufacturing, reliability, and data-center teams to translate system needs into practical interconnect solutions. This is a hands-on role for someone who can connect design decisions with installation, qualification, troubleshooting, and long-term operational performance. In this role, you will: Define optical interconnect architectures and requirements across hardware platforms and rack-level systems. Design high-density fiber systems for performance, density, reliability, installation, and serviceability. Lead optical-mechanical integration and cross-functional design reviews. Develop test and qualification plans for optical components, modules, switching platforms, and integrated systems. Own optical loss budgets, routing guidelines, handling requirements, and serviceability criteria. Support system bring-up, deployment, troubleshooting, failure analysis, and reliability improvement. Create reusable design guidelines, interface requirements, and qualification methods. You might thrive in this role if you have: Core experience Experience desi

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI is a frontier AI research and deployment company. Frontier research is at the center of how we advance our mission, with researchers, engineers, product leaders, operators, and many other teams working together to turn new capabilities into systems that benefit humanity. The People team helps OpenAI attract, engage, and support the exceptional talent this work requires. Employer Brand sits at the intersection of Research, Recruiting, Communications, Marketing, and Brand. Its mandate is to continue to establish OpenAI in the talent market as the unique and leading frontier research lab—not simply another technology company—and make our distinctive research environment, mission, culture, and opportunity for impact relevant and tangible to every priority talent audience. About the Role We’re hiring an Employer Brand Manager to build and scale the strategy, narrative, and operating system that shape how priority talent understands OpenAI. This senior individual contributor will anchor our employer brand in OpenAI’s identity as a frontier research lab and translate an evidence-backed “why OpenAI / why now” narrative into campaigns, researcher and employee stories, recruiter and hiring manager enablement, and candidate experiences. This role is especially important as OpenAI competes for exceptional talent across research, engineering, product, and other mission-critical functions in a fast-moving field where external perceptions can be incomplete or change quickly. You will develop a clear, credible talent narrative for priority audiences—anchored in frontier research and substantiated by individual agency, world-class infrastructure, research-to-product translation, deployment scale, and a willingness to answer hard questions candidly. This role is responsible for ensuring the external brand resembles our culture and ethos internally, therefore must remain immersed in various OpenAI research and applied branches. This role is based in San Francisco

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About Team Our Robotics team is focused on unlocking general-purpose robotics and advancing toward AGI-level intelligence in dynamic, real-world environments. Working across the full model and systems stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the physical constraints of real-world systems to improve people’s lives. About Role We are building a supply-chain organization capable of supporting our transition from laboratory-scale development to factory-scale operations. We are looking for an Inventory Manager to build and operate our inventory function from the ground up. This is a highly hands-on role. In the near term, you may start with an empty room and be responsible for determining what racks, shelving, bins, labels, scanners, workflows, and systems are needed to turn it into a functional stockroom. You will receive material, organize inventory, perform counts, move parts between buildings, resolve discrepancies, and establish the processes others will eventually follow. As we grow, you will have the opportunity to develop this foundation into a full-scale, multi-site factory inventory operation. You may hire and manage contractors or onsite inventory administrators, but this role will initially have a significant individual-contributor component and will remain accountable for day-to-day execution. This role owns inventory and internal logistics. It does not own purchasing, production planning, or inbound and outbound supplier logistics. This role is based in San Francisco, CA and requires in-person presence 5 days a week. In this role you will: Build inventory operations from the ground up across various OpenAI facilities. Design and set up stockrooms, receiving areas, and material-storage locations, including selecting racks, shelving, bins, carts, labeling equipment, scanners, and other infrastructure. Personally execute core inventory

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The CoT Monitorability team at OpenAI studies whether and when the chain-of-thought of frontier reasoning models is monitorable enough to support scalable oversight. We study how to measure monitorability , which training mechanisms affect monitorability, and speculative methods to improve monitorability. While we mostly focus on CoT monitorability at the moment, we care more generally about any form of monitorability, auditing methods, and improving alignment. We were the first to show that chain-of-thought monitoring can be a practical additional safety mechanism, and today our monitoring systems are actively used on OpenAI’s largest RL training runs to detect misbehavior. The issues we surface are then used to help improve our reward functions, environments, etc (without directly training against a CoT monitor). Our work sits in Alignment and intersects with model training, alignment evaluations, monitoring, and frontier-risk research.We care most about monitorability where the stakes are high, and about preserving useful oversight signals as models become more capable. About the Role We’re looking for a researcher with strong empirical ML expertise and a deep interest in model behavior, alignment, or interpretability. Direct chain-of-thought interpretability experience is welcome but not required; strong candidates may come from broader interpretability, alignment, model training, or investigative model-behavior work. As a researcher on the Alignment team, you will design and run experiments that improve our understanding of model monitorability. You will investigate how training interventions across the model-development pipeline influence whether reasoning remains legible, build evaluations that make those questions measurable, and help translate findings into practical oversight and training recommendations. You may also help develop new monitoring models or methods and apply them to OpenAI’s largest training runs. This role is especially well

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Business organization works with customers and partners on some of our most complex and consequential opportunities. These efforts require rigorous strategy, strong operating leadership, and coordinated execution across commercial, product, technical, deployment, and go-to-market teams. About the Role We are hiring a Business Lead to serve as the operating leader for a strategically important initiative anchored in a major partnership. This person will turn an ambitious, cross-functional mandate into a clear strategy, operating plan, decision structure, and set of measurable outcomes. This role combines strategy and operations, product judgment, commercial skills, and the ability to get things done. You will identify the most promising product and customer opportunities, develop a point of view on how our products should work together, shape the commercial approach, and personally drive execution across both organizations. This is a hands-on role for someone who wants to own outcomes, not just coordinate work. You will structure ambiguous problems, establish priorities, build trusted relationships with internal and external stakeholders, and work across product, engineering, deployment, and go-to-market teams to remove blockers and deliver results. Success means establishing a durable operating model for the initiative, improving decision velocity, translating strategy into coordinated execution, launching a joint go-to-market motion, landing an initial cohort of customers, and delivering a high-quality enterprise deployment. In this role, you will: Own the initiative’s integrated strategy and operating plan, including priorities, desired outcomes, metrics, owners, dependencies, and decision points Structure complex and ambiguous business problems, develop fact-based recommendations, and translate them into clear choices and executable plans Define success measures and build operating reviews that surface progress, risks, tradeoffs, and requi

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Personal AGI team is responsible for training and improving pre-trained models to be deployed into ChatGPT, the API, and potential future products. In the Model Experience team, we shape the default character and behavior of ChatGPT: how the model communicates, responds to users, uses its capabilities, and behaves across different contexts and languages. Our goal is to make every interaction with ChatGPT thoughtful, helpful, and trustworthy. We take an opinionated view of what good human–AI interaction should look like, then turn that vision into real model behavior through human data, evaluations, reward models, and post-training. Our work sits at the intersection of research, product, and model design. We partner closely with teams across OpenAI to conduct research and ensure our models are thoughtful, safe, reliable to serve millions of users. About the Role As a Research Engineer / Scientist, you will research and develop improvements to our models. Our team works in research areas combining reinforcement learning and products. We're looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models. An ideal candidate is passionate about product-driven research and the quality of human-AI interaction. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own and pursue a research agenda to improve model capability and performance. Collaborate closely with the other research and product teams, allowing customers to optimize their own models. Build robust evaluations for tracking modeling improvements. Design, implement, test, and debug code across our research stack. You might thrive in this role if you: Have a deep understanding of machine learning and machine learning applications. Have good judgment about model behavior and can communicate this judgment effec

AWSRestMachine LearningAI
🔔

Get new aws and tooling platform lead jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime