Jobs in United States

Aws And Tooling Platform Lead in San Francisco

866 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Agent Post-Training team creates the frontier agents OpenAI ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that can operate computers, collaborate with people and other agents, and expand what people and organizations can imagine, attempt, and achieve. We define what the next generation of agents should be able to do, build the training signal that teaches those abilities, and run the experiments that make them real. Our work spans coding, tool use, computer use, multi-agent coordination, long-horizon execution, factuality, instruction following, calibrated reasoning, and taste. Our team is where new model capabilities get made. We build the data, environments, graders, training methods, and feedback loops that shape what OpenAI's next agents can do, then carry those capabilities through major training runs and into the products people use. About the Role As a researcher working on Frontier Evals & Environments, you will help build north star model environments to drive progress towards safe AGI/ASI. Your work will directly guide the research programs of the most ambitious training runs happening at OpenAI. Some prior open-sourced evaluations built by researchers in this role include GDPval , SWE-bench Verified , MLE-bench , PaperBench , and SWE-Lancer . If you are interested in feeling firsthand the fast progress of our models, and steering them towards good outcomes, this is the role for you. You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measure whether it worked, and ship improvements into products used by real people. This is a high-agency role for people who want their work to land directly in frontier models. In this role, you might Create ambitious RL environments to push our models to their limits, and measure frontie

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The B2B Marketing team contributes to OpenAI's broader mission of ensuring responsible and widespread adoption of artificial intelligence. We are responsible for developing and executing strategies that drive awareness, engagement, and usage for OpenAI's products and platform. We take a data-driven approach to understand our customers' needs and challenges, ensuring that their voices are reflected in product development and messaging. We then partner closely with Product, Engineering, Growth, Sales, Research, Comms, and Design teams to create a cohesive customer experience across all our channels. About the Role We're looking for an enterprise Product Marketing Manager to own the administrator and governance experience across ChatGPT Work and Codex. IT administrators, security teams, and deployment owners are often the difference between an AI product being available and an organization actually putting it to work. This role will define how OpenAI earns the trust of the people responsible for approving, deploying, managing, and expanding AI inside their organizations. You will translate a fast-moving set of enterprise capabilities into a clear, practical story covering setup, access, controls, governance, security, usage, and value. You will own the administrator audience as a dedicated product marketing discipline, partnering closely with Product, Security, Sales, Customer Success, Growth, Customer Education, and other Marketing teams. Your job is not simply to announce new features. It is to help administrators understand what has changed, what they can control, how to deploy responsibly, and why expanding access is the right decision for their organization. In this role, you will: Own the IT administrator and governance audience: Build a deep understanding of the needs, decisions, objections, and workflows of workspace administrators, IT leaders, security teams, and enterprise deployment owners. Define the enterprise administration and governance s

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Role We’re looking for a senior Strategic Finance Lead to help shape and execute OpenAI’s financial strategy across some of the company’s most important long-term investments and growth priorities. This high-impact role sits at the intersection of corporate finance, capital markets, treasury, and cross-functional execution. In this role, you will: Shape financial strategy for major long-term investments, capital structure decisions, and broader strategic decision-making. Build rigorous financial models and scalable frameworks to evaluate strategic initiatives, funding options, tradeoffs, and risk. Structure and help execute complex strategic initiatives in partnership with cross-functional teams. Translate ambiguous technical and business inputs into clear recommendations and decision-ready materials for executives, the board, and other senior stakeholders. Lead high-priority cross-functional workstreams from concept through execution, bringing strong judgment, ownership, and communication in fast-moving situations. Help build an AI-native finance organization by applying AI to improve workflows, decision-making, and execution across finance processes. You might thrive in this role if you have: 12+ years of experience across strategic finance, corporate finance, capital markets, investment banking, infrastructure finance, or related fields. A strong understanding of capital structure, financing strategy, debt markets, and large-scale investment evaluation. Experience operating in high-growth, high-complexity environments with the ability to navigate ambiguity and move quickly. Exceptional financial modeling, strategic thinking, and problem-solving capabilities. Strong executive presence with the ability to communicate clearly across investors, executives, and technical stakeholders. A high-agency operating style that combines strategic perspective with a willingness to build directly. Excitement around using AI as a core operating advantage, with a mindset

AWSRestAIGo
N
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -87.4%

$185K – $220K/yr

Quick readStrong listing-quality and freshness signals

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: We are seeking a strategic and technically fluent Lead, IT Audit to join our Finance team reporting to the Head of Internal Audit. This is a broad, high-impact role spanning both IT SOX compliance and operational IT audits. You will help establish and elevate our technology controls program end to end — owning the IT SOX lifecycle, designing the IT general and application controls framework, embedding AI and automation into how we test and monitor controls, and delivering value-added operational IT and cybersecurity audits that strengthen how the company builds and runs its systems. You will partner with leaders across Engineering, Security, IT, Finance, and the business to ensure sound technology controls are built into how the company operates as we scale. This role is ideal for someone who thinks like a builder, not just an auditor — someone who can translate complex control and security requirements into practical, scalable processes in a fast-moving SaaS environment with modern cloud architecture and complex data flows. This role can be based in either San Francisco or New York City. We work from our offices on M

AWSAzureGCPCI/CD
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Personality & Model Behavior team, within OpenAI’s broader Personal AGI team conducts research on how to shape personalities and guide the behavior of models. We think about topics such as emotional intelligence, reasoning, and how models interact thoughtfully with users. We’re particularly interested in understanding how individual users want ChatGPT to behave, and creating personalized models that feel uniquely tailored to each user. We integrate this research into ChatGPT and other OpenAI products that are used by hundreds of millions of users. About the Role We're looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models, and in areas like reinforcement learning and reward modeling. An ideal candidate is passionate about product-driven research. In this role, you will: Conduct research around personalization, personality, and model behavior by leveraging and developing tools such as synthetic data, reward modeling, and reinforcement learning. Build robust evaluations and model training pipelines to facilitate our research. Innovate new post-training methods. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. You might thrive in this role if you: Have a deep understanding of machine learning and its applications. Have prior knowledge in training and optimizing models and building evaluations. Are willing to dive into large ML codebases to debug issues. Thrive in dynamic and technically complex environments. Have a track record of delivering innovative, out-of-the-box solutions to address real-world constraints. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through o

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight: Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful autonomy (for example future versions of auto-review ). About the Role We’re looking for strong executors with excellent judgment, comfort with ambiguity, and an understanding of frontier model research. You don’t need prior safety or alignment experience, we also welcome people that recently realized that alignment and safety is a critical area to contribute to. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Train and evaluate frontier models to reduce harmful or misaligned agent actions, forming clear hypotheses and executing independently through ambiguity. Mine incidents and build scalable measurement, data-processing, and evaluation systems that turn real failures into repeatable safety signals. Collaborate closely with post-training, capabilities, oversight, and pre-training partners to ship research-backed mitigations into large-scale training and agent systems. You might thrive in this role if you: Have demonstrated strength in research engineering, ML en

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight : Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful agent autonomy (for example future versions of auto-review ). About the Role This role focuses on oversight and system-level mitigations that enable increasingly capable agents to operate safely and autonomously in real environments. We prioritize building oversight systems that are used in practice today, both internally and externally (see our recent work on action monitoring for codex and former code review ). We also study longer-term questions about how increasingly capable agentis systems can be supervised, constrained, and corrected. We’re looking for a safety&security minded researcher or engineer who can reason rigorously about security boundaries and agent behavior, then build and test practical mitigations. A background in AI control or security is welcome but not required. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and evaluate system-level controls for agent actions like agent-based review. Plan how they fit in a broader syste

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking a Mechanical Engineer to design, build, and own the mechanical side of our robotic actuator dynamometer and test infrastructure. You will create the test stands, couplings, fixtures, load paths, guarding, and serviceable lab hardware that enable rigorous characterization of robotic actuators. This role combines precision mechanical design with hands-on lab work. You will take robotic actuator test infrastructure from requirements and analysis through CAD, fabrication, assembly, commissioning, and iteration, partnering closely with electrical and software engineers to deliver safe, flexible, high-uptime test cells. In this role, you will Own the mechanical architecture of dynamometer and actuator test cells, including frames, bases, load paths, alignment, guarding, and serviceability. Design dynamometer structures, robotic actuator fixtures, load-motor mounts, couplings, shafts, bearings, adapters, and torque-reaction hardware. Translate robotic actuator test requirements into robust mechanical systems for torque, speed, thermal, durability, backdrive, efficiency, and failure testing. Perform first-principles analysis and simulation for stiffness, strength, fatigue, vibration, thermal growth, critical speed, and safety factors. Create precise, repeatable alignment strategies that protect test articles, load machines, sensors, and couplings. Design modular fixturing that supports rapid changeover across actuator and motor variants without compromising measurement quality. Work closely with electrical engineers on cable routing

ReactAWSRestAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -82%

About the Team The Future of Computing Research team is an applied research team within OpenAI’s Consumer Devices group. We study how AI systems perceive people and their surroundings, and we turn that research into capabilities for future products. Our work spans machine learning, sensing, and hardware, with a focus on building systems that work beyond controlled environments. About the Role We’re looking for a machine learning engineer to help shape how future AI systems understand the physical world and the people in it. The role focuses on multimodal perception and authentication, bringing together signals from cameras, microphones, and other sensors. You’ll work with specialized perception models and larger multimodal models, and partner with hardware, firmware, software, and product teams to bring new research into real-world systems. This role is based in San Francisco. We work in the office three days per week and offer relocation assistance. In this role, you will: Research and develop multimodal perception and authentication methods across visual, audio, and other sensing signals. Explore how specialized perception models and larger multimodal models can work together. Design data, training, and evaluation approaches that improve performance in real-world conditions. Study model behavior, robustness, and failure modes across sensing, data, and deployment environments. Integrate and validate new capabilities in real-time or resource-constrained systems. Work with hardware, firmware, software, and product teams to turn research into working systems. You might thrive in this role if you: Have a strong background in computer vision, audio or speech machine learning, multimodal learning, or sensing. Have experience developing specialized machine learning models, larger multimodal models, or both. Have brought research ideas into practical systems, prototypes, or products. Know how to design experiments, build evaluations, and investigate model behavior. Have wo

PythonAWSRestMachine Learning
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The GTM Data Science team partners with Go-to-Market, Technical Success, Product, Engineering, RevOps, and Strategic Finance to build the shared intelligence layer for OpenAI's B2B business. The team turns product usage, customer behavior, revenue, field activity, and customer feedback into rigorous insight products that help leaders and field teams understand where customers are succeeding, where adoption is blocked, and what actions will accelerate durable growth. We are building systems that make customer intelligence proactive: surfacing risk, expansion potential, product gaps, and repeatable playbooks before they show up as escalations or missed opportunities. About the Role As the Applied Data Science & Insights Lead for GTM Intelligence Solutions and Technical Success, you will be a hands-on technical leader responsible for shaping how OpenAI measures, understands, and improves customer adoption across our B2B products. You will build AI/ML-powered intelligence products that connect account health, product usage, customer lifecycle, support tier, qualitative sentiment, commercial context, and field actions into a practical operating system for GTM and Technical Success. This role will build the data science foundation for Technical Success: defining the metrics, models, operating insights, and decision systems that help the team scale customer adoption and expansion with rigor. You will also be expected to build and lead a small mighty team over time: setting direction, hiring and developing talent, creating operating cadences, and holding a high bar for technical rigor and business impact. You will lead the development of models, metrics, and decision systems that recommend what GTM and Technical Success teams should do next, explain why, and measure whether those interventions worked. Your work will help customers move from pilots to production, deepen usage across products, identify high-value use cases, reduce churn risk, and create a f

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Platform team powers how millions of developers and enterprises build with our models. We provide APIs and agentic solutions used by global startups and fortune 500s. We work closely with product, engineering, design, and go-to-market to build a world-class platform that pushes the frontier of AI capabilities. About the Role As a Data Scientist on the Platform team, you will drive a data-driven culture for OpenAI’s API and B2B solutions. You’ll define the metrics that matter for developer success and enterprise value, measure the impact of new models and features, and partner with PMs and engineers to improve model quality, reliability, latency, and cost. Your work will shape how thousands of products adopt agentic AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will Embed with the Platform product team as a trusted partner, uncovering ways to improve developer experience, reliability, and usage growth Define north-star metrics across the developer funnel (activation, retention, growth), as well as latency/cost guardrails for new features and models Design and interpret A/B tests and controlled rollouts (e.g., new model versions, pricing/limits, new API features, new B2B products) Build source-of-truth dashboards and self-serve data tools for product, engineering, and go-to-market teams Translate product learnings into actionable feedback for Research (e.g., failure modes, eval gaps, model response quality) You might thrive in this role if you have 5+ years in a quantitative role in ambiguous, high-growth environments (platforms, APIs, or B2B products a plus) Depth in SQL and Python, with a track record proposing, designing, and running rigorous experiments Experience defining and operationalizing metrics from scratch (including reliability/latency/cost and safety) Strong cross-functional communication with PMs, enginee

PythonSQLAWSRest
C
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -79.2%

From £189.8K/yr

Quick readStrong listing-quality and freshness signals

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Adoption of AI is moving rapidly from pilots to production. Enterprises are investing in generative-AI programs, creating a once-in-a-decade market inflection point. At Cohere, we’re committed to transforming the enterprise landscape through strategic partnerships and innovative AI solutions. As a Partner Development Manager for the Enterprise Market, you’ll be at the heart of our mission to empower businesses across the region with cutting-edge AI technologies. This role is perfect for someone who thrives in a dynamic, cross-cultural environment and is passionate about driving growth through collaboration. You’ll work closely with leading enterprises, system integrators, and technology partners to co-create value, expand market reach, and deliver impactful AI solutions tailored to the unique needs of businesses. Your work will directly contribute to our success. You’ll have the opportunity to shape partnerships that drive digital transformation, foster innovation, and position Cohere as a leader in the AI ecosystem. If you’re a strategic thinker, a relationship builder, and eager to make a meaningful impact in on

AWSGitAIGo
P
📍 San Francisco, CA, United States· Full-time· Remote
✓ High-confidence listingCompany trend -86.3%

From $92K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About tvScientific We are the first CTV advertising platform purpose-built for performance marketers. For game developers and publishers, we bridge the gap between massive TV reach and granular User Acquisition (UA) metrics. Built by ad-tech veterans, our platform combines media buying, optimization, and MMP attribution to help gaming brands automate CTV campaigns, drive app installs, and maximize Return on Ad Spend (ROAS). Join the tvScientific team as an Account Manager (Gaming), where you'll lead strategic client relationships for gaming and app clients, drive revenue growth, and ensure client success on our cutting-edge platform. As an Account Manager on our team, you'll be responsible for managing a portfolio of key client accounts, developing and executing strategic account plans, and driving revenue growth through upsell, cross-sell, and

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

We are hiring a Security Software Engineer to design and implement the hardware-backed security foundations used across OpenAI’s device ecosystem. A central focus of this role is hardening the boundary between our policy systems and the HSMs that protect sensitive cryptographic keys. This boundary determines which operations may be performed, what may be signed, which policies must be satisfied, and how changes to trusted software and policy are authorized. You will develop security-critical software and firmware within, or immediately adjacent to, an HSM trust boundary. Depending on your background, this may include HSM trusted applications, firmware services, cryptographic mechanisms, device drivers, PKCS#11 components, secure-provisioning protocols, or signing-policy enforcement systems. This is a hands-on software-engineering role. You will be expected to design systems, write and review production code, debug across hardware and software boundaries, and carry projects from initial requirements through deployment. It is not an HSM administration, PKI operations, compliance, or architecture-only position. In This Role, You Will Design and implement security-critical software and firmware for HSMs, secure elements, trusted execution environments, and hardware roots of trust. Build and harden the policy-to-HSM boundary responsible for authorizing certificate issuance and cryptographic signing operations. Develop HSM trusted applications, firmware components, host interfaces, device drivers, SDKs, or cryptographic service integrations. Implement or extend cryptographic interfaces such as PKCS#11, OpenSSL providers or engines, platform key-storage APIs, or comparable hardware-security interfaces. Build firmware and software that cryptographically enforces key generation, provisioning, usage, rotation, recovery, and destruction policies. Design and implement HSM-backed certificate authority, code-signing, key-management, and device-identity systems. Develop end-to-end

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team Critical Harm Operations sits within User Safety & Risk Operations and builds enforcement systems for the most severe user-safety risks—including Frontier Risk and Material Harm—that are accurate, fast, defensible, and scalable. The Cyber vertical turns policy into reviewer standards, calibrated judgment, quality systems, escalation paths, and safe automation. About the Role We’re looking for a senior cybersecurity and cyber intelligence practitioner who is also an operations leader, strategist, and people manager. You’ll bring sound cyber judgment, strong analytical and automation instincts, and the ability to stay close to the work while setting direction, managing and developing people, aligning partners, building programs from 0→1, and translating ambiguity into durable operating systems. This role combines senior individual-contributor depth with direct people-management accountability. You’ll personally shape complex cyber judgments, operating models, SOPs, quality systems, and automation strategies while building, coaching, and holding accountable a high-performing team. You’ll also lead through influence across FTEs, BPO partners, and cross-functional teams. Baseline operational health—including quality, queue coverage, productivity, and SLA adherence—is table stakes; success is additionally measured by durable improvements to decision quality, stakeholder alignment, reviewer capability, and the systems that run the operation. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Set the strategic direction for Cyber Operations and translate high-level organizational priorities into a clear operating model, roadmaps, SOPs, escalation paths, quality goals, and safe-access strategies. Serve as the senior cybersecurity and cyber intelligence expert for complex and high-risk decisions across ChatGPT, API, Codex, agents, and

AWSRestAIGo
🔔

Get new aws and tooling platform lead jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime