Jobs in United States

Aws And Tooling Platform Lead in United States

2,026 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

S
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.6%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About The Role The Security Team is responsible for securing all things Sentry: our customers, our code, and everything in between. We are a small but growing team with broad scope, high trust, and the autonomy to tackle hard security problems with creativity and an engineering mindset. We work at a company with a strong developer culture, building a product that millions of developers genuinely love and rely on. That context shapes everything about how we operate. As a Security Engineer on this team, you'll work across application and platform security domains. You'll contribute to the practices that keep Sentry secure as we grow: security reviews, threat modeling, vulnerability management, and embedding secure coding practices into an engineering organization that cares about doing things right. You'll partner closely with product and engineering teams to influence how features are designed and built from the start. You will work as a technical collaborator who helps make the secure path the obvious one. As Sentry expands our agentic product capabilities and development practices, you'll also find yourself at the frontier of a new set of security challenges. In this role, you will Support and help mature Sentry's security review program. From secure code review, to architecture review, and threat modeling. You'll help build the processes, tooling, and culture which make security a natural part of how we ship and operate. Contribute to mature vulnerability management practices. Intake, triage, prioritization, remediation tracking, and support of our bug bounty and responsible disclosure program. Advocate for secure-by-desig

TypeScriptPythonAWSAzure
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Network Engineering team within IT and Security advances the mission of deploying artificial general intelligence (AGI) for the benefit of all by delivering secure, scalable, and resilient network services. We build and operate the connectivity that supports OpenAI’s offices, labs, campuses, cloud environments, people, and devices. By combining strong network fundamentals with security, reliability, automation, and user-centered design, we enable impactful AI research, corporate operations, and product innovation. About the Role As a Network Engineer at OpenAI, you will design, operate, and continuously improve the global networks that connect our offices, labs, campuses, PoPs, cloud environments, people, and devices. The role spans strategic platform engineering and responsive production operations: you will shape architecture, standards, roadmaps, lifecycle plans, and automation while supporting incidents, escalations, and time-sensitive delivery. Operational signals will inform what we stabilize, simplify, standardize, or automate next. We work backward from user needs, investigate root causes, own outcomes end-to-end, and move quickly without compromising security. We are looking for a versatile engineer who can make pragmatic reliability and security tradeoffs, communicate clearly, and turn recurring operational work into durable platforms, tooling, and standards. You will partner across IT, Security, AppEng, Research, Applied, workplace teams, carriers, and vendors. In this role, you will: Design, implement, and operate secure, scalable enterprise networks across offices, labs, campuses, PoPs, cloud connectivity, and hybrid environments. Set strategic direction for network services through architecture, standards, roadmaps, lifecycle planning, capacity strategy, and measurable reliability outcomes. Own production operations, including on-call, incident response, escalations, and time-sensitive delivery, while protecting user experience,

PythonAWSAzureCI/CD
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI's mission is to ensure that artificial general intelligence benefits all of humanity. The Consumer Devices team is building a new generation of AI-powered products that seamlessly integrate hardware and software to create intuitive, transformative experiences. We bring together experts across embedded systems, machine learning, hardware, design, and product engineering to develop products at the intersection of AI and consumer technology. About the Role OpenAI is seeking a System Performance Engineer to profile, benchmark, and optimize performance across our embedded hardware products. In this role, you will work across operating systems, applications, camera and vision, graphics, and platform teams to define product KPIs, build performance tooling, and drive optimizations from early lab characterization through product launch and real-world usage. You will help establish the performance standards that shape the user experience of our products, ensuring they remain responsive, efficient, and reliable throughout their lifecycle. This role requires deep expertise in embedded or high-performance systems, strong operating systems fundamentals, and hands-on experience debugging under tight latency, power, and memory constraints. This role is based in San Francisco, CA. We use a hybrid work model of four days per week in the office and one day working remotely. Relocation assistance is available for new hires. In this role, you will: Develop system performance benchmarks, methodologies, and policies to evaluate end-to-end product behavior. Profile and analyze performance across key product use cases and workloads using custom and industry-standard profiling tools. Partner closely with engineering teams to identify bottlenecks and drive performance optimizations across the software stack. Define high-level product KPIs and establish measurement frameworks to measure launch readiness and monitor performance throughout the product lifecycle. Measure, re

PythonAWSLinuxRest
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s User Operations team ensures that our customers’ experience is nothing short of exceptional. We resolve complex issues, provide guidance, and support customers in maximizing value with our products. Within User Operations, Ads Safety Operations helps OpenAI grow advertising in a way that is safe, trusted, and sustainable for advertisers, users and the business. The Ads Safety Support team sits within Ads Safety Operations and owns the advertiser-facing support journey: helping advertisers navigate ads products, policies, reviews, and enforcement decisions while turning frontline insights into better guidance, tooling, and product experiences. About the Role We’re looking for a Senior Advertiser Support Specialist to help build and scale advertiser support within Ads Safety Operations. You’ll work directly with advertisers through support tickets, escalations, owning high-impact questions related to ads policies, review outcomes, product behavior, and operational processes. Many cases will be novel or ambiguous, and you’ll be expected to investigate carefully, communicate clearly, and drive them to resolution in a fast-moving product environment. As OpenAI’s advertising ecosystem grows, your work will help advertisers operate successfully at scale while serving as a critical bridge between frontline support and the teams responsible for product quality, policy integrity, and platform safety. Beyond resolving individual cases, you’ll help define what world-class support looks like in an AI-native environment. You’ll turn support interactions into operational intelligence, identifying recurring friction, gaps in guidance, and emerging issues to improve support at scale, and protect the quality and safety of OpenAI’s ad ecosystem. You’ll partner closely with Engineering, Product, Policy, Legal and Go To Market partners to improve systems, reduce bugs, and elevate the customer experience, while leveraging automation, agents, and our own AI technol

PythonAWSRestAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $243.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer on our Release Engineering team, you will build the systems that hundreds of engineers across Roblox use every day to safely ship code to tens of millions of concurrent players. Your work will empower teams to ship bold, high-impact changes quickly and confidently on every device Roblox runs on. If you enjoy building developer-facing infrastructure where reliability and blast radius directly impact end-users, you will be right at home on our growing team. You Will: Design and develop backend services and automation that power our release and experimentation process across desktop, mobile, console, VR, and servers Work in C++ engine code to improve telemetry reporting, enhance feature rollout and automatic abort capabilities, and extend release functionality Build progressive client and server rollout, regression detection, and automated rollback systems that keep our weekly multi-platform releases safe at scale Work directly with engineering customers to turn pain points into durable, flexible, and safe-by-default tooling You Have: 5+ years building backend services with C#, Python, TypeScript, or similar Familiar with and comfortable working with C++ Familiar

TypeScriptPythonAWSGit
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $196.8K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox’s daily active users growing at a record pace, we are seeking a senior data infrastructure engineer to join our new Data Insights team. Our team owns the data tooling that empowers Roblox builders to independently make informed and timely data-driven decisions. As an engineer on the team, you’ll work on the platforms behind tools like Superset, Hex, and Python notebooks, which provide critical insights into the health of our business to users at every level of the company. We tackle diverse challenges in data engineering, infrastructure, and analytics, to deliver the insights our customers need. You will collaborate closely with engineers across our data ecosystem to shape the future of product analytics at Roblox. This role offers the chance to be a founding team member and help define both the technical direction and the long-term shape of the product area from the ground up. This role is a great fit for you if you are proficient in designing and scale robust data infrastructure and applications and have a zeal for developing inspiring, easily maintainable, and reusable code. Join our team and make a significant impact at Roblox. You Will: Architect and deliver a high-pe

TypeScriptPythonReactSQL
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. A majority of our users interact with our products in languages other than English, and our products must work seamlessly across languages, regions, and cultures. The Internationalization team builds the infrastructure that enables OpenAI products to ship globally by default. We develop the systems that power localization, international product launches, and high-quality global user experiences across all OpenAI products. About the Role As a Senior Software Engineer on the Internationalization team, you will build the systems that power localization and international product launches at OpenAI. You’ll work on the platform that manages product content, translation workflows, and localization infrastructure across our products. This role sits at the intersection of AI systems, developer platforms, and product infrastructure. In this role, you will Build and scale OpenAI’s localization, content, and experimentation platform used across OpenAI product teams, including open-source components: Develop AI-powered translation pipelines combined with human-in-the-loop review workflows. Design systems that reliably deliver localized product content across web and mobile apps. Build tools that enable linguists and localization teams to review and improve translations. Develop developer tooling that simplifies localization and internationalization workflows. Build and maintain internationalization libraries used across OpenAI products: Design systems that correctly handle numbers, currencies, dates, and pluralization across locales. Improve support for multilingual interfaces and right-to-left languages. Partner with product teams to improve the international readiness of new features. You might thrive in this role if you Have strong software engineering experience building backend or full-stack systems. Have familiarity with Java, React, MySQL, and cloud infrastructure p

JavaReactSQLMySQL
P
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Team Description: Plaid is evolving into an AI-first company, and Intelligent Tooling sits at the center of that transformation. The Intelligent Tooling team is being built from the ground up, and our mission is to establish the technical foundations, operating model, and internal platforms that embed AI deeply into Plaid’s coding tools, internal systems, and the entire software development lifecycle. When we are successful, engineers across Plaid will delegate lower-leverage work to AI agents, move faster with confidence, and spend more of their time designing and inventing for customers. Intelligent Tooling owns the platforms and systems that make this possible - from AI coding integrations and SDLC agents to the internal tools that power Plaid’s operations. Role Description: As a Staff Software Engineer on the Intelligent Tooling team, you will build and operate internal systems that directly impact how engineers across Plaid do their work, and own the technical direction for major parts of that surface. This is a hands-on role with significant ownership, where success is measured by real adoption, reliability, and improvements to developer experience. You will work on AI-powered tooling, interna

AWSCI/CDRestAI
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -82%

$230K – $385K/yr

Quick readStrong listing-quality and freshness signals

About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products, including next-generation ads experiences, that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers, advertisers, and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, and Research to bring new ad experiences into real-world systems across OpenAI surfaces at global scale, including thoughtfully integrating them into the core ChatGPT experience. About the Role We’re looking for an experienced Software Engineer to help build the creative rendering and presentation layer of OpenAI’s ads ecosystem. This is a foundational role responsible for defining how ads are structured, rendered, and delivered across different surfaces, platforms, and media types. You’ll work across the full technical stack to build infrastructure and tooling for new ad formats, including text, image, video, native, conversational, and interactive experiences. You will help ensure these formats render reliably, perform efficiently, and feel natural within the core ChatGPT experience. You’ll collaborate deeply with Product, Design, and Research to create ads experiences that are useful, high-quality, privacy-preserving, and aligned with OpenAI’s standards for safety and user trust. In this role, you will: Design, build, and

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products, including next-generation ads experiences, that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers, advertisers, and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, and Research to bring new ad experiences into real-world systems across OpenAI surfaces at global scale, including thoughtfully integrating them into the core ChatGPT experience. About the Role We’re looking for an Android Engineer to help build the native Android experiences and client-side systems that power how ads are structured, rendered, and delivered across OpenAI’s ads ecosystem. This role sits within the Ads Formats team, which owns the creative rendering and presentation layer for next-generation ads experiences across different surfaces and media types. You’ll help build the Android infrastructure and tooling that support formats such as text, image, video, native, conversational, and interactive ads, while ensuring they render reliably and perform efficiently across platforms. You’ll work closely with backend, Product, Design, Research, and Safety partners to shape Android architecture that supports user-first, privacy-preserving monetization experiences within ChatGPT and OpenAI’s broader mobile ecosystem. In th

AWSRestAIKotlin
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team pAGI Infra team builds and operates the systems that make large-scale model training and evaluation reliable, efficient, and easy to run. Our work spans distributed training infrastructure, inference and grading platforms, compute scheduling, and research tooling. We partner closely with researchers and engineering teams to turn new research needs into dependable infrastructure, improve GPU efficiency, and shorten the path from an experiment to a validated model. About the Role We’re looking for an AI Systems Engineer to help scale the infrastructure behind our training and evaluation workflows. You’ll own projects from identifying bottlenecks and designing solutions through deployment and operation. The work combines distributed systems engineering, performance optimization, and close collaboration with researchers. You might build a shared grading service, improve resource allocation across workloads, or bring a new training stack into production — directly improving how quickly and reliably research moves forward. In this role, you will: Build and operate infrastructure for large-scale training and evaluation, improving reliability, throughput, and resource efficiency. Develop shared inference and grading platforms with automated capacity management, health monitoring, and visibility into performance. Improve compute scheduling and resource allocation to reduce idle GPU time and help workloads recover quickly from failures. Diagnose bottlenecks across training, inference, and orchestration, and work across teams to improve end-to-end performance. Build self-service tools, automated validation, and observability that help researchers launch experiments, diagnose issues, and compare results with less manual intervention. You might thrive in this role if you: Are excited about the potential of personal AGI and want to build the infrastructure that enables it. Have strong software engineering fundamentals and experience building or operating large-scal

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -82%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI's People team helps hire, develop, and support the people building safe and beneficial AGI. Within that team, People Systems builds the technical foundation that enables our HR, recruiting, payroll, benefits, and performance operations to scale with quality, speed, and rigor. We work at the intersection of HR systems, software engineering, and internal tooling. Our goal is not just to keep core systems running, but to build durable technical leverage for the company. About the Role We're hiring a Workday Engineer to help design, build, and operate the systems that power critical people workflows at OpenAI. This is a highly technical role for someone who combines strong Workday expertise with real engineering fluency. You'll build reliable integrations, improve system architecture, automate complex workflows, and help connect Workday to internal tools, external platforms, and emerging AI-driven systems. You should be comfortable going beyond configuration work. We're looking for someone who can reason through ambiguous systems problems, write and debug technical solutions, work effectively in Git-based environments, and use modern developer workflows, including CLI-driven tooling, to build and operate with speed and discipline. You'll partner closely with cross-functional teams across People, Finance, Security, and Engineering, including our People Innovations team, to build systems that are secure, scalable, and practical. Some work will involve improving mature production infrastructure; some will involve building entirely new workflows and capabilities from scratch. In this role, you will Design, build, and maintain Workday integrations, applications, and workflow automations across domains such as payroll, benefits, recruiting, performance, and case management Improve the reliability, quality, and scalability of People systems through strong engineering, testing, and operational practices Build technical solutions that connect Workday with i

AWSGitRestAI
O
📍 United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but operational in how we execute, and we support every product and research effort at OpenAI. Our tenets include prioritizing for impact, enabling researchers and developers, preparing for future transformative technologies, and fostering a strong, collaborative security culture. About the Role OpenAI is seeking a Security Software Engineer to join the Infrastructure Security (InfraSec) team. InfraSec safeguards the core of OpenAI’s research and production environments—GPU supercomputing clusters, multi-cloud infrastructure, datacenters, networking, storage, and the critical services that power our frontier AI models. Our charter spans everything from bare-metal hardware and firmware to Kubernetes clusters, service meshes, and the data pathways that carry highly sensitive model weights and user data. As a Security Software Engineer, you will design and build critical foundational services, such as authentication systems, egress/ingress proxies, access brokers, and key management platforms, that demand high standards of reliability, scalability, and software craftsmanship. These systems form the security backbone of OpenAI’s supercomputing environment and must remain robust under intense scale and adversarial pressure. In this role, you will: Architect and implement production-grade security services (e.g., auth services, access brokers, secure proxies, key-management infrastructure) that provide strong guarantees across hardware, operating systems, Kubernetes, networks, and CI/CD. Partner with infrastructure and research engineers to embed security into high-performance compute clusters, enabling rapid model training and deployment without compromising protection. Develop automation and detection tooling to continuously identif

PythonAWSAzureGCP
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI's Research Data Team exists to accelerate the evaluation, safety and capabilities of our models and products. Made up of technical operators and software engineers, we design the methods in which we acquire and create data. About the Role As a Research Program Manager (RPM), Data Acquisition, you will partner with research, engineering, and operations to design and implement pragmatic solutions for acquiring data. You will be a key interface between our research roadmap and external data offerings. This role is based in our San Francisco HQ and will be part of a team of RPMs pushing the frontier of data acquisition. In this role, you will: Partner deeply with research: Work with researchers to scope data needs, define success criteria, and translate priorities into clear execution plans. Shape the data acquisition pipeline: Identify, evaluate, and advance high impact data opportunities - balancing research value, feasibility, quality, and responsible execution. Unblock yourself: Move work forward even when the path is unclear — using technical judgement, creative problem solving, and scrappy execution to make progress while longer-term solutions are still forming. Build lightweight systems and visibility: Use SQL, Python, dashboards, and simple tooling to track performance, quality, and blockers. Drive technical roadmaps: Collaborate with engineers to enhance data platforms, resolve blockers, and ensure security best practices such as access management. Scale your impact: Equip vendors and internal teams with the context, standards, and operating rhythms needed to focus on the most important problems. You’ll thrive in this role if you: Are proficient in SQL and Python for analysing datasets, querying databases, building dashboards, and generating actionable insights. Are comfortable using APIs, automation, and AI tools such as Codex to accelerate workflows, remove manual overhead, and upskill quickly in unfamiliar technical areas.Experience sou

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI's Human Data Team creates custom data solutions driving groundbreaking research. Our work enhances and evaluates our flagship models and products like ChatGPT, GPT-5, and Sora, and contributes to safety initiatives through collaboration with our Preparedness and Safety Systems teams. About the Role As a Research Program Manager (RPM) in the Human Data team you will partner with research and engineering to design and implement pragmatic solutions for collecting high-quality data. You will be a key interface between our research roadmap, external vendors, AI trainers, and the Human Data engineering team. This role is based in our San Francisco HQ. In this role, you will: Collaborate with Research: Partner with researchers to scope data collection needs, define success metrics, and establish quality measurement frameworks. Design & Execute Data Collection Campaigns: Translate research needs into actionable plans and accelerate execution by leveraging existing tooling and iterating to reach the desired outcome. In many cases, you will need to implement scrappy new solutions while partnering with engineering to design robust/scalable solutions. Unblock Yourself: You must be deeply uncomfortable with the idea of sitting around waiting for external dependencies, and have the technical acumen and drive to figure out how to achieve at least partial success in the interim. Optimize Systems & Processes: Build and optimize dashboards to track campaign performance, leveraging SQL and Python for data analysis and actionable insights. Drive Technical Roadmaps: Collaborate with engineers to enhance data platforms, resolve blockers, and ensure security best practices such as access management. Scale Your Impact : Advise and empower program managers and vendors to drive day-to-day execution so that you can focus on addressing high priority opportunities. You might thrive in this role if you: Are proficient in SQL and Python for data analysis, including q

PythonSQLAWSRest
🔔

Get new aws and tooling platform lead jobs in United States by email

Daily job updates · Unsubscribe anytime