Jobs in United States

Aws And Tooling Platform Lead in San Francisco

866 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Safety Systems team is dedicated to ensuring the safety, robustness, and reliability of AI models and their deployment in the real world. Building on the many years of our practical alignment work and applied safety efforts, Safety Systems addresses emerging safety issues and develops new fundamental solutions to enable the safe deployment of our most advanced models and future AGI, to make AI that is beneficial and trustworthy. Learn more about OpenAI’s approach to safety About the Role As an Analytics Engineer in Safety Systems, you will play a pivotal role in building a data-centric culture, enhancing decision-making processes, and driving strategic initiatives through analytics. You will partner closely with Engineering, Research, and Data Science to develop and maintain canonical data sources and source-of-truth dashboards that enable both people and AI agents across the organization to derive trustworthy, actionable insights. You will own the consumption layer for safety metrics: defining intuitive, reliable ways for stakeholders across Safety Systems, partner teams, and leadership to understand the safety of our products, answer safety-related questions independently, and inform product decisions and company strategy. Most importantly, you will be a core member of the Safety Systems team, collaborating with researchers and engineers to advance our goals of safe, robust, and reliable AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design and maintain canonical datasets that serve as sources of truth for safety metrics. Develop and refine data products such as dashboards, reports, agent-enabled workflows, and machine-readable interfaces that empower stakeholders to extract and analyze data independently. Work closely with stakeholders in Engineering, Research, and Data Science to understand their decision-making n

SQLAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s mission is to ensure that general-purpose artificial intelligence benefits all of humanity. Our Communications team’s ethos is to support OpenAI’s mission and goals by clearly and authentically explaining our technology, values, and approach to safely building powerful AI. About the Role OpenAI is seeking an experienced communications professional to join our Platform & Research Communications team. This role will work closely with the Research Communications Lead and partner deeply with safety researchers, alignment researchers, and cross-functional teams to shape how OpenAI’s safety research is understood by researchers, journalists, policymakers, and the broader public. This position is responsible for developing and executing external communications strategies around OpenAI’s safety research—from alignment and evaluations to broader work that helps advance the safe development and deployment of increasingly capable AI systems. The ideal candidate brings strong science or technical fluency, excellent storytelling instincts, and experience helping researchers communicate complex work with clarity, accuracy, and nuance. You will partner closely with research leadership, individual researchers, policy, product, safety, legal, and cross-functional communications teams. This role requires both strategic judgment and hands-on execution in a fast-moving environment where research, public understanding, and high-stakes safety narratives intersect. This role is based in San Francisco, CA and follows a hybrid schedule (three days per week in office). Relocation assistance is available. In this role, you will: Shape Safety Research Narratives Develop clear, credible external narratives around OpenAI’s safety research, including alignment, evaluations, preparedness, interpretability, and other areas connected to the safe development of frontier AI. Translate complex technical work into accessible stories without oversimplifying, overstating impact

ReactAWSRestMachine Learning
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s API Multicloud team is responsible for extending OpenAI’s API platform into strategic cloud environments, starting with AWS . The team’s mission is to distribute OpenAI’s API broadly and safely by enabling key API technologies in AWS-native environments, in close partnership with Amazon and internal teams across Codex, Research, Safety Systems, and Applied. The team is focused on bringing core developer and enterprise capabilities into cloud-native environments, including AWS-hosted Codex, model customization / post-training as a service, and new stateful runtime environments for agentic workloads. This work sits at the intersection of production ML systems, developer platforms, model behavior, and large-scale infrastructure. About the Role We’re hiring Machine Learning Engineers to build and improve the AI systems that help strategic partners adapt OpenAI models to important use cases in cloud-native environments. This role spans post-training workflows, evaluation, data pipelines, model behavior, and API/infrastructure integration. You’ll work at the boundary between partner needs and core ML systems: helping teams understand what is and isn’t working, diagnosing issues in training and evaluation workflows, and turning those learnings into improvements to the underlying platform. You should enjoy working with external technical partners, extracting the real goal from messy requests, and pushing back or reframing when the requested experiment is not the highest-leverage path. You’ll collaborate closely with Research, Applied, Safety Systems, infrastructure teams, and external technical partners to solve ambiguous model-performance problems. When you succeed, strategic partners and internal teams will be able to improve model behavior with confidence, driving measurable product improvements while the systems behind that work become more reliable, scalable, and effective over time. In this role, you will Partner with strategic customers and in

PythonAWSKubernetesRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s mission is to build safe artificial general intelligence (AGI) which benefits all of humanity. This long-term undertaking brings the world’s best scientists, engineers, and business professionals into one lab together to accomplish this. In pursuit of this mission, our Go To Market (GTM) team is responsible for helping customers learn how to leverage and deploy our highly capable AI products across their business. The team is made of Sales, Solutions, Support, Marketing, and Partnership professionals that work together to create valuable solutions that will help bring AI to as many users as possible. About the Role Our GTM team is uniquely positioned to help customers realize the transformative potential of advanced AI models for their businesses and end users. As part of the GTM Strategy & Operations team, you’ll play a critical role in guiding the GTM strategy and driving the operational efficiency to accomplish this mission. This role serves as a trusted advisor to GTM leadership—providing data-driven insights, managing core operating cadences, and leading high-impact projects that influence how we engage with customers and scale our business. You’ll collaborate cross-functionally with Finance, Enablement, Data and Growth Strategy teams to align efforts, drive efficiencies, and accelerate growth. This role is a central role, meaning it is not tied to a specific business partner, but rather designing the global processes for the business. In this role, you'll: Design, build, and manage our operating rhythms (Forecasting, Pipeline Council, Big Deal Reviews, and MBRs/QBRs). Conduct strategic analyses to determine trends and identify opportunities for process and strategy optimization Collaborate with GTM leadership and cross-functional stakeholders to develop go-to market strategy and resource plans Lead strategic projects to improve efficiency and effectiveness across the revenue organization. Partner closely with technical teams to impl

SQLAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Education team is building products and experiences that help learners, educators, and institutions benefit from AI in ways that are rigorous, useful, and grounded in real learning outcomes. The work spans both consumer and B2B education, with close collaboration across engineering, learning science, design, data, and research. This team sits in a highly strategic investment area for OpenAI, with strong opportunities to shape how product ideas flow across consumer and institution-facing experiences. Some of our recent work: New Education Plugins for ChatGPT Work and Codex New tools for understanding AI and learning outcomes Education for countries Advancements in higher education Early product work - Introducing Study Mode About the role We’re looking for a product-minded Full Stack Engineer to help build OpenAI’s education products from the ground up. You’ll own end-to-end development across the stack, from early concepting and prototyping through production launch and iteration. This is an opportunity to work on a highly strategic, early-stage product area where engineering judgment, product sense, and customer empathy all matter. You’ll partner closely with leaders across the education org, including learning scientists, researchers, designers, and cross-functional partners, to turn emerging ideas into durable product experiences for schools, universities, and other education stakeholders. In this role, you will: Build and ship product experiences across the full stack for OpenAI’s education offerings Own projects end-to-end, from ideation and technical design through implementation, launch, and iteration Work closely with learning scientists and researchers to translate learning goals and evidence into product decisions Collaborate with design, data, and cross-functional partners to build thoughtful, high-quality user experiences Help define the engineering foundation for a growing education pod, including patterns, systems, and technical

TypeScriptPythonReactSQL
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Premium team owns some of the highest-leverage customer-facing levers in ChatGPT’s consumer revenue business, spanning the paid customer journey: helping users understand the value of paid plans, convert with confidence, and continue finding lasting value in their subscription. Our work is highly cross-functional, partnering with Product, Data Science, Design, FinEng, Finance, Legal, Support, and Marketing to improve free-to-paid conversion, renewal, customer lifetime value, and revenue while keeping the experience trustworthy, scalable, and low-friction. In This Role, You Will: Lead and scale an engineering team responsible for some of ChatGPT’s most important subscription and monetization experiences. Own the technical execution for Premium customer experiences across plan merchandising, paywalls, upgrade flows, checkout UX, plan management, renewals, downgrades, and cancellation. Partner with Product and Data Science to run high-quality experiments across upgrade, trial, renewal, downgrade, and cancellation flows. Improve key subscription metrics including conversion, renewal, churn, ARPU, and lifetime value. Build reliable customer-facing Premium experiences for purchase, plan management, renewal, downgrade, cancellation, and access-related states at scale. Partner closely with FinEng and other platform teams to evolve the billing, payments, and entitlement capabilities that power Premium experiences. Collaborate closely with Product, Design, Data Science, Finance, Legal, Support, and Marketing on monetization strategy and execution. You Might Thrive in This Role If You: Have 5+ years of engineering management experience, Have strong technical expertise in backend, frontend, or full-stack development, with experience building growth-oriented features. Have a track record of improving conversion, retention, or monetization through experimentation and data-driven product engineering. Are experienced with subscription products, plan merchandising

SQLAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for applied AI engineers to help bring Codex agents from impressive demos to dependable tools. This role is about improving agent performance on real software engineering tasks and closing the gap between research capability and real-world usefulness. You’ll work closely with research, infrastructure, and product to ensure agents are not just powerful, but useful, steerable, and reliable in practice. The job is not only to improve model behavior in isolation, but to turn those improvements into measurable gains in solve rate, usefulness, and economic value for users. What You’ll Do Design and iterate on agent behaviors across real-world coding tasks and long-horizon workflows. Work closely with research to develop and run evals to measure agent performance, regressions, failure modes, and edge cases. Improve performance through prompting, tool-use strategies, context construction, and model-facing experimentation. Analyze failures in production and systematically improve robustness and reliability. Build feedback loops and data systems that get better real-task data into evaluation and research. Work with product teams to shape user-facing agent experiences and the interfaces the agent depends on. Help define what “good” looks like for agents completing complex tasks end-to-end. You Might Be a Good Fit If You Ha

PythonAWSRestMachine Learning
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team We’re hiring software engineers to make OpenAI’s networking teams more productive. These teams build and operate the high-performance networking systems that support OpenAI’s training and inference infrastructure at frontier scale. About the Role We’re looking for someone who cares deeply about the developer experience of engineers working on complex infrastructure systems — especially around build systems, test architecture, release pipelines, and reliable development workflows. This role will be embedded with OpenAI’s networking team: making it faster, safer, and easier for engineers to build, test, validate, and ship changes across multi-server, networked, and hardware-adjacent environments. In this role you will: Improve development workflows for engineers building and operating OpenAI’s networking systems Design and improve continuous deployment, release, and validation pipelines Build and maintain test harnesses for multi-server, networked, and hardware-backed environments Improve iteration speed across C++, Python, and build-system-heavy codebases Partner with engineers to identify friction in CI, testing, debugging, and deployment workflows Drive testing and reliability strategy for infrastructure components that support large-scale training and inference workloads Work closely with centralized developer experience teams while staying deeply embedded with the networking engineers closest to the systems You might thrive in this role if: You are motivated by helping other engineers move faster and with more confidence You have experience with CI/CD, release pipelines, testing infrastructure, or build systems You are comfortable moving between C++, Python, and build systems such as CMake, Bazel, or Blaze You enjoy building test harnesses, automation, and workflow improvements for complex systems You do not need to be a networking expert, but you are excited to learn enough about the domain to make the team meaningfully more effective When you see

PythonAWSCI/CDRest
B
📍 San Francisco, California, United States· Full-time
✓ High-confidence listing

$192K – $240K/yr

Quick readStrong listing-quality and freshness signals

Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Engineering at Brex Engineering at Brex is about building systems that scale with speed and intention. Our teams span Software, Data, Security, and IT, and operate with high autonomy and deep collaboration. We tackle hard technical problems, own our outcomes, and push for excellence at every level — from architecture to deployment. It’s an environment where engineering is a craft, and builders become leaders. What you’ll do As a Senior Software Engineer, Infrastructure (Release Engineering) at Brex, you will design, build, and operate the core systems that power Brex’s release, observability, and incident management processes. You will partner closely with product, platform, and operations teams to ensure releases are safe, fast, and reliable, and that our infrastructure scales securely as Brex grows. Where you’ll work This role will be based in our San Francisco office. We are a hybrid environment that combines the energy and connect

PythonJavaSQLAWS
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models. We’re looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience. You’ll transform complex operational data into actionable insights that directly influence how OpenAI allocates and scales one of the world’s largest AI compute environments. Key Responsibilities Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency. Develop forecasting models for inference demand across products, regions, and model families. Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities. Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies. Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs. Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions. Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps. Communicate technical findings clearly to both engineering teams and executive leadership. Qualifications MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience). 5+ years of experience working in the infrastructure data science space. Strong ex

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team The Applied organization brings OpenAI’s most advanced technology to the world through products like ChatGPT and the APIs that power a growing ecosystem of developer and enterprise applications. Data Engineering builds and operates the trustworthy, secure, and reliable data systems that power decisions across OpenAI. About the Role We’re looking for a Data Engineering Manager to lead the Growth & Revenue data engineering team. This leader will own the data strategy and execution for the data subject areas spanning growth accounting across all product surfaces, product partnerships, checkout, billing, payments, revenue, and monetization, helping OpenAI understand how people adopt, engage with, and pay for our products. You will partner closely with several Data Science, Business, and Engineering partners to connect product behavior to trustworthy subscriber, payment, and revenue measurement. In this role, you will: Build, manage, and grow a high-performing, inclusive team across the Growth & Revenue data subject areas. Define the data strategy for all the data subject areas you own. Deliver durable, well-modeled data products that connect product behavior, subscription state, checkout events, payment outcomes, and revenue. Establish trusted metric definitions and data quality standards so product, growth, finance, and executive leaders can make fast, consistent decisions. Partner with Data Science and Product teams to support experimentation, causal measurement, funnel analysis, and scalable self-serve analytics. Partner with Finance and Financial Engineering to ensure analytical revenue views reconcile to financial truth and production billing systems. Raise operational excellence for critical pipelines, including reliability, observability, privacy, governance, and incident response. Set a clear roadmap, make principled tradeoffs, and communicate progress and risk across technical and business stakeholders. You might thrive in this role if yo

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Legal team helps advance our mission by tackling novel legal issues in AI. Our team brings together professionals across technology, privacy, intellectual property, corporate, employment, tax, regulatory, and litigation. Our regulatory compliance work turns legal requirements into practical programs that support responsible AI development and deployment. About the Role As a Legal Program Manager focused on regulatory compliance, you will build and manage cross-functional programs that translate counsel’s guidance into practical, sustainable operations. Your initial focus may include content moderation and/or frontier AI governance, with the mix shaped by team priorities and your strengths. You will partner with internal and external counsel, other legal program managers, and technical and business teams to coordinate implementation, evidence collection, reporting, and ongoing compliance. You’ll build repeatable systems that scale across regulations, products, and jurisdictions, helping teams navigate emerging requirements with clarity and sound judgment. This full-time role is based in San Francisco, CA, or New York, NY. In this role, you will: Lead regulatory compliance programs end to end: define scope, owners, milestones, dependencies, risks, and escalation paths, and drive execution with counsel and cross-functional partners. Translate counsel’s regulatory guidance into repeatable workflows, controls, and documentation. Depending on your portfolio, this may include content moderation disclosures, transparency reporting, reporting and appeals workflows, or frontier AI model launch readiness, evaluation and risk-management evidence, and incident reporting. Build strong partnerships across Product, Engineering, User Operations, Governance, Risk and Compliance (GRC), Global Affairs, Communications, and Go-to-Market to align program priorities and deliverables. Support regulatory inquiries, audits and investigations with counsel, organizing ev

SQLAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI Consumer Devices is building the next generation of products that bring powerful AI into people’s everyday lives. Guided by OpenAI’s mission to ensure AGI benefits all of humanity, our team combines world-class researchers, engineers, designers, and operators who care deeply about creating useful, intuitive, and responsible technology. You’ll have the opportunity to work alongside exceptional people on ambitious, zero-to-one challenges at the intersection of hardware, software, and AI. This is a chance to help define an entirely new category of products—and shape how people experience AI in the future. The Systems Integration team is critical in this mission, turning complex hardware-software development into reliable product signals. We validate complete device experiences across software, cloud services, connectivity, accessories, and real-world operating environments, combining hands-on system testing, structured test development, hardware-in-the-loop environments, diagnostics, and automation to uncover issues that component-level testing alone cannot reveal. About the Role As a Systems Test Engineer, End-to-End Validation , you will design and execute end-to-end testing for complex device experiences spanning hardware, software, connectivity, cloud services, and accessories. You’ll translate product behavior and real-world use cases into structured, reproducible test procedures and build test environments that allow failures to be reliably reproduced and diagnosed. You’ll also identify opportunities to automate repetitive or high-value scenarios, working with engineers to turn complex manual workflows into scalable validation systems. Because this is a new category of devices, you’ll have the opportunity to build the end-to-end validation foundation early—shaping test coverage, environments, and workflows from prototype through launch. We’re looking for someone who combines strong systems thinking, hands-on testing skills, technical curiosi

PythonAWSRestAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s People team hires, engages, and retains world-class talent to safely build and deploy AGI that benefits all of humanity. The People Analytics team helps leaders make rigorous, evidence-based talent decisions and ensures that the systems supporting those decisions are valid, reliable, fair, and accountable. About the Role As a People Data Scientist focused on AI fairness and bias testing, you will help establish how OpenAI evaluates AI-assisted People systems and high-impact talent processes. You will design and conduct rigorous assessments to identify, measure, and mitigate potential bias across the lifecycle of models, agents, decision-support tools, and automated workflows. Your work will span the entire employee life-cycle, such as hiring, performance, promotion, employee development, workforce planning, etc. You will evaluate both technical systems and the broader human-AI decision processes in which they operate, examining not only model performance but also data quality, measurement validity, differential outcomes, human oversight, and unintended consequences. We’re looking for an experienced data scientist or applied researcher who can translate complex fairness questions into defensible evaluation strategies, scalable testing infrastructure, and clear recommendations for technical teams and senior leaders. This role is preferred to be based in San Francisco, CA. In this role, you will: Define and lead fairness and bias-testing strategies for AI-assisted People processes, models, agents, and decision-support systems from development through deployment and ongoing monitoring. Design rigorous algorithmic audits and validation studies, including adverse-impact analysis, subgroup and intersectional evaluation, error-rate analysis, calibration, measurement invariance, reliability, criterion-related validity, and sensitivity testing. Identify the appropriate fairness criteria for each use case, evaluate tradeoffs among competing definitions

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s People team hires, engages, and retains world-class talent to safely build and deploy AGI that benefits all of humanity. The People Analytics team helps leaders make better, evidence-based talent decisions. About the Role As a People Research Scientist, you will bring deep expertise in research design, measurement, experimentation, and applied data science to OpenAI’s most important People programs. You will design studies, evaluate people processes, and help leaders better empower employees, strengthen organizational systems, and deliver exceptional employee experiences. This is a high-ownership individual contributor role combining hands-on research, methodological leadership, and scalable people science capabilities. We’re looking for an experienced researcher who can turn ambiguous People questions into rigorous designs, validated insights, and actionable recommendations. This role is based in San Francisco, CA or Mountain View, CA, with occasional travel to our San Francisco office. What You’ll Do: Design rigorous research and evaluation strategies for recruiting, organizational health, manager effectiveness, employee experience, and talent outcomes. Apply advanced statistical modeling, machine learning, and research methods to inform program design, evaluate effectiveness, and quantify business impact. Partner with People Operations, data engineering, and people systems teams to define data requirements, improve data quality, establish documentation standards, and ensure research datasets are governed, reproducible, and privacy-preserving. Build scalable people science infrastructure, including self-service agentic tools, automated validation workflows, reusable research datasets and analytical pipelines. Develop research playbooks that establish rigorous standards for study design, measurement, validation, and documentation, enabling high-quality, repeatable, and scalable research across the organization. Communicate findings through c

PythonSQLAWSRest
🔔

Get new aws and tooling platform lead jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime