Jobs in United States

Back End Td Reliability Lab Manager in United States

490 active opportunities · Updated October 2026

Explore current back end td reliability lab manager jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team The Technical Success team is responsible for ensuring the safe and effective deployment of ChatGPT and OpenAI API applications for developers and enterprises. We act as a trusted advisor and thought partner for our customers, ensuring developers and enterprises maximize value from our models and products. The Ecosystem team builds and scales third-party integrations around ChatGPT and Codex. We work across product, engineering, partnerships, safety, legal, policy, and go-to-market to create integrations that are useful, trustworthy, and technically sound. About the role We’re looking for an Applied Deployment Engineer to help build and deepen ChatGPT integrations with third-party messaging platforms. The initial focus will be on partner work as well as future opportunities with other messaging apps around the world. This is a hands-on, partner-facing engineering role. You’ll work closely with external product and engineering teams to improve existing integrations, expand capabilities such as image generation, increase retention, and translate partner needs back into OpenAI’s product and engineering roadmap. You should be comfortable writing code daily, making changes across platform or monorepo surfaces when needed, and moving fluidly between technical discovery, product tradeoffs, prototyping, debugging, and executive-level communication. This role is best suited for a strong product-minded engineer with excellent communication skills, sound technical judgment, and the ability to work across cultures, time zones, and organizations. This role is based in our San Francisco office. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Partner with messenger app product and engineering teams to design ChatGPT-powered experiences that fit naturally into chat, group, contact, and app-tab surfaces. Deepen existing partner integrations, with an initial focus on improving user ex

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s API Multicloud team is responsible for extending OpenAI’s API platform into strategic cloud environments, starting with AWS . The team’s mission is to distribute OpenAI’s API broadly and safely by enabling key API technologies in AWS-native environments, in close partnership with Amazon and internal teams across Codex, Research, Safety Systems, and Applied. The team is focused on bringing core developer and enterprise capabilities into cloud-native environments, including AWS-hosted Codex, model customization / post-training as a service, and new stateful runtime environments for agentic workloads. This work sits at the intersection of production ML systems, developer platforms, model behavior, and large-scale infrastructure. About the Role We’re hiring Machine Learning Engineers to build and improve the AI systems that help strategic partners adapt OpenAI models to important use cases in cloud-native environments. This role spans post-training workflows, evaluation, data pipelines, model behavior, and API/infrastructure integration. You’ll work at the boundary between partner needs and core ML systems: helping teams understand what is and isn’t working, diagnosing issues in training and evaluation workflows, and turning those learnings into improvements to the underlying platform. You should enjoy working with external technical partners, extracting the real goal from messy requests, and pushing back or reframing when the requested experiment is not the highest-leverage path. You’ll collaborate closely with Research, Applied, Safety Systems, infrastructure teams, and external technical partners to solve ambiguous model-performance problems. When you succeed, strategic partners and internal teams will be able to improve model behavior with confidence, driving measurable product improvements while the systems behind that work become more reliable, scalable, and effective over time. In this role, you will Partner with strategic customers and in

PythonAWSKubernetesRest
O
📍 Seattle, Washington, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Statsig team within OpenAI owns the experimentation, rollout, dynamic configuration, and analytics infrastructure that sits on the launch path for OpenAI products. Our systems help teams ship safely, evaluate product and model changes in production, and make high-confidence decisions from real-world usage. Statsig began as an independent company built around experimentation, feature management, and product analytics at scale. After Statsig joined OpenAI, the team began the next chapter: bringing that platform expertise and infrastructure into OpenAI as the experimentation and rollout foundation for every product we ship. This is infrastructure with a very direct product consequence. Teams working on ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and shared platform systems depend on Statsig to evaluate configurations, move traffic safely, ingest experiment data, serve analytics, and roll changes forward or back when production reality demands it. We are at a critical point in the platform journey. Adoption is accelerating quickly across OpenAI, and the systems that were already important are becoming load-bearing for how the company launches. The infrastructure needs to stay fast under sharply increasing evaluation volume, reliable when more services depend on it, observable enough to debug quickly, and efficient enough to support OpenAI-wide scale. Recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency across important services. The next phase is to make those gains systematic: a platform that can absorb rapidly growing product velocity while preserving low latency, data quality, operational safety, and developer trust. Based out of OpenAI's Bellevue office, we are a close-knit team that values in-person collaboration, technical depth, operational ownership, and building infrastructure that

VueAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. About the Role The Safety Measurement Product Manager owns OpenAI's approach to measuring harm and safeguard efficacy in production, including driving the strategy for our suite of safety measurement platforms and products used across the company. You will partner closely with our safety research and engineering teams to determine what we measure, where we measure it, and how we measure it, feeding those insights directly into critical leadership decisions and back into our safety work. You will also represent the company's topline safety metric as well as prioritize incoming requests from partner teams to expand our safety measurement platform to more use cases. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Partner closely with data science, research, engineering, policy teams, and other stakeholders to craft a vision for understanding safety outcomes and prevalence on our platforms. Define strategic priorities and product roadmaps focused on improving safety measurement approaches will scaling our measurement platform to more use cases, products, and cross-functional team needs. Establish repeatable processes to integrate cutting-edge AI safety research into OpenAI’s safety m

AWSRestAIGo
O
📍 Seattle, Washington, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions. Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After Statsig joined OpenAI, the team began the next chapter: bringing that deep product expertise, customer intuition, and mature platform infrastructure into OpenAI as the experimentation and rollout platform for every product we ship. Today, we support teams across ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and the shared infrastructure that connects them. These teams rely on Statsig to safely introduce new capabilities, compare product and model behavior, measure impact, and roll changes forward or back with confidence. We are at a defining moment in the platform journey. OpenAI has the data, product surface area, and pace of innovation to learn faster than almost any organization in the world, but that potential only becomes real if teams can experiment responsibly, measure clearly, and roll out changes safely. Adoption of the platform is accelerating rapidly across the company, and recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency for important services. Based out of OpenAI’s Bellevue office, we are a close-knit team that values in-person collaboration, urgency, craft, and impact. We build for other builders, and the best version of this team is one where every OpenAI product team can move faster because the experimentation and rollout layer is dependable, fast, and easy to use. About the Role We are l

VueAWSRestAI
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. About the Role We’re looking for a Senior Developer Advocate to join our Developer Relations team and help developers everywhere build better APIs, faster. You’ll serve as a trusted technical voice for Postman by creating compelling content, building community, and championing developer needs back into the product. This is a high-visibility, high-impact role. You won’t just talk about Postman, you’ll be hands-on with APIs, AI agents, and emerging protocols like MCP , translating complex workflows into accessible, inspiring content that reaches millions of developers. What You’ll Do Create technical content at scale: Write tutorials, blog posts, videos, sample projects, and livestreams that demonstrate API best practices using Postman and adjacent tools. Represent Postman publicly: Speak at conferences, meetups, podcasts, and webinars. Be a recognizable and credible technical voice in the API and developer tooling ecosystem. Build and nurture community: Engage with developers across forums, Discord, GitHub, social media, and in-person events. Lead or support the global Agents & APIs Meetup program and similar community initiativ

TypeScriptPythonCI/CDGit
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Senior Software Engineer - Customer Experience Engineering Our customers are very happy with our technology and are quickly expanding their use cases, only limited by their own imaginations. The Customer Experience Engineering Team aims to enable customers to push these limits, and when customers run into roadblocks or issues, enable them and our internal technical resources to resolve those issues and get back to seeing success with Snowflake. Snowflake’s Customer Experience Engineering Team is expanding! We are looking for a Senior Software Engineer to join our team who likes building intelligent data applications that give insight into diagnostic data and technical content that is relevant to issues customers may be facing. Our Customer Experience Senior Software Engineers enjoy developing features into the Snowflake product that aim to greatly reduce friction points that block customers and regularly require help from Snowflake technical experts. They are excited by the challenges presented by extracting intelligence from a wide variety of data sources, including structured and unstructured data. As a Sr. Software Engineer, you will: Lead and drive projects that span our stack, including Java and Python services hosted in Kubernetes and Snowpark Container Services. Prom

PythonJavaKubernetesRest
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the role We're hiring someone to build the playbook for how we drive practitioner adoption of new capabilities, and bring new practitioners into the platform. You'll pick a segment (data engineers, ML engineers, analytics engineers), pair it with a use case, and figure out what it takes to get practitioners from awareness to activation. That means developing real knowledge of your audience: what they care about, where they spend time, what resonates, and what falls flat. Then working cross functionally to assemble the right assets, sequence them into an exciting journey that leads to adoption, and getting your campaign in front of the right people across owned, paid, in-product, and partner channels. What you're building is a closed loop. Each campaign is instrumented so you can see what actually worked. Did practitioners activate? Did usage stick? How quickly did they find value? Each round of learning makes the next one sharper. The insights you generate don't stay with you; they feed back to marketing, product and account teams to drive broader adoption across organizations. If you want to own the full arc from choosing the audience to designing the playbook to shipping the campaign to building the feedback loops that make it all compound over time, this is that ro

SQLAIGoRust
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is seeking an Enablement Business Partner to serve as the Chief Activation Officer for our Solution Engineering team. Embedded within key leadership teams, you will accelerate field capability tied to business outcomes, owning the full spectrum of how SEs learn, from structured programs to in-the-moment performance support. You translate business problems into capability solutions, designing systems that make the field more capable when and where it matters most. IN THIS ROLE YOU WILL GET TO: Act as the strategic enablement partner for SE leadership, diagnosing skill gaps tied to business outcomes, prescribing targeted interventions, and measuring whether capability actually changed Curate and quality-gate enablement content: validate AI outputs for accuracy and audience fit, maintain a source of truth for your domains (theater and product), and feed evidence-based quality signals back to PM and PMM Diagnose field needs through signals (Gong, Salesforce, theater-level data analysis, leadership conversations, steer co themes, PM roadmap changes, etc), then classify and prioritize gaps with data Design activation experiences, picking the right modality (hands-on or self-serve lab, AI skill, peer learning, coaching nudge), and measuring behavior change, not just comp

PythonSQLRestAI
S
📍 Menlo Park, California, United States· Full-time
✓ High-confidence listingCompany trend -92.9%

$190K – $238K/yr

Quick readStrong listing-quality and freshness signals

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is about empowering enterprises to achieve their full potential — and people too. With a culture that's all in on impact, innovation, and collaboration, Snowflake is the sweet spot for building big, moving fast, and taking technology — and careers — to the next level. Snowflake is growing fast and we're scaling our team to help enable and accelerate our growth. We're passionate about our people, our customers, our values and our culture! We're also looking for people with a growth mindset and the pragmatic insight to solve for today while building for the future. And as a Snowflake employee, you will be accountable for supporting and enabling diversity and belonging. Do you want to shape how the world's best builders — data engineers, ML engineers, and app developers — experience Snowflake? We're looking for a technically sharp, builder-obsessed product marketer to own go-to-market strategy for Snowflake's developer-facing product areas. This person will work at the intersection of product, engineering, and field teams to craft narratives that resonate with hands-on builders — from individual contributors writing code to technical architects evaluating platforms. You'll be Snowflake's voice to the developer community and the developer community's voice back into S

PythonSQLGitAI
🔔

Get new back end td reliability lab manager jobs in United States by email

Daily job updates · Unsubscribe anytime