About the team Online Data builds and operates Habitat, the single product surface of Online Data and the system of record for OpenAI’s online user data. As OpenAI’s scale and product requirements evolve, Habitat is becoming a full-stack, one-size-fits-most database platform with end-to-end ownership of: Provisioning and developer experience APIs and guardrails Scaling, performance, and reliability Data movement, caching, routing, and placement Privacy enforcement and access control Change Data Capture (CDC) as a first-class primitive The foundation for future storage backends You’ll work on the core online database platform behind OpenAI’s products, building and operating Habitat services that handle high-QPS, latency-sensitive workloads across regions. You’ll partner closely with internal platform and product teams to ship safe, reliable systems, then push them to be faster and more cost-efficient through better caching, routing, observability, and operational tooling. This is a critical role for engineers who like owning hard distributed-systems problems end to end and sweating the details from p99 latency to production operations at massive scale. In this role, you will Design and build core abstractions spanning storage, caching, routing, CDC, and privacy enforcement Own a major surface area end to end, from product and API design to operational excellence Improve latency, correctness, and cost efficiency for real production workloads at massive scale Build strong instrumentation, debugging workflows, and developer-first tooling Collaborate closely with internal product and infrastructure teams to understand requirements and ship pragmatic solutions Participate in an on-call rotation and raise the bar on reliability while aggressively improving performance and usability You might thrive in this role if you have A strong track record building and operating high-scale backend or data-intensive distributed systems in production Excellent systems judgment and the a
Jobiba hiring network
Reliability Engineer in Seattle
10 active opportunities · Updated for October 2026
Fresh results
10 shown
Explore current reliability engineer jobs in Seattle. Use filters to narrow by work mode, employment type, experience and date posted.
About the Team The ChatGPT Library team is building the place where people can save, organize, rediscover, and build on the content they create with ChatGPT. Our goal is to make ChatGPT more useful over time by helping users seamlessly return to important files, images, conversations, and other content across their devices. The team works at the intersection of product engineering, design, and AI research to create intuitive, personalized experiences that make users’ content easy to find and act on. On Android, we are focused on delivering fast, reliable, and deeply native experiences that put a user’s evolving body of work at their fingertips. About the Role We are looking for a senior Android engineer to help build the future of ChatGPT Library on mobile. You will own high-impact product experiences across the Android stack, shaping how millions of people save, organize, discover, and interact with their content in ChatGPT. In this role, you will: Build and ship new Android experiences that help users easily access, organize, and build on the content they create with ChatGPT. Own features end to end—from early product exploration and technical design through implementation, experimentation, launch, and iteration. Develop scalable, maintainable foundations that allow Libraries experiences to evolve quickly as new AI capabilities emerge. Improve the architecture, performance, reliability, and responsiveness of content-rich experiences across a wide range of Android devices. Thoughtfully integrate Android platform capabilities to create experiences that feel intuitive and native to mobile. Partner closely with product, design, research, data science, and engineering teams to translate emerging AI capabilities into useful, polished products. Help establish technical direction and raise the quality bar for Android development across the team. How We Work We care deeply about building products that are intuitive, useful, and trustworthy. We move quickly from ideas to wo
About the Team The Statsig team within OpenAI owns the experimentation, rollout, dynamic configuration, and analytics infrastructure that sits on the launch path for OpenAI products. Our systems help teams ship safely, evaluate product and model changes in production, and make high-confidence decisions from real-world usage. Statsig began as an independent company built around experimentation, feature management, and product analytics at scale. After Statsig joined OpenAI, the team began the next chapter: bringing that platform expertise and infrastructure into OpenAI as the experimentation and rollout foundation for every product we ship. This is infrastructure with a very direct product consequence. Teams working on ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and shared platform systems depend on Statsig to evaluate configurations, move traffic safely, ingest experiment data, serve analytics, and roll changes forward or back when production reality demands it. We are at a critical point in the platform journey. Adoption is accelerating quickly across OpenAI, and the systems that were already important are becoming load-bearing for how the company launches. The infrastructure needs to stay fast under sharply increasing evaluation volume, reliable when more services depend on it, observable enough to debug quickly, and efficient enough to support OpenAI-wide scale. Recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency across important services. The next phase is to make those gains systematic: a platform that can absorb rapidly growing product velocity while preserving low latency, data quality, operational safety, and developer trust. Based out of OpenAI's Bellevue office, we are a close-knit team that values in-person collaboration, technical depth, operational ownership, and building infrastructure that
About the Team The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions. Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After Statsig joined OpenAI, the team began the next chapter: bringing that deep product expertise, customer intuition, and mature platform infrastructure into OpenAI as the experimentation and rollout platform for every product we ship. Today, we support teams across ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and the shared infrastructure that connects them. These teams rely on Statsig to safely introduce new capabilities, compare product and model behavior, measure impact, and roll changes forward or back with confidence. We are at a defining moment in the platform journey. OpenAI has the data, product surface area, and pace of innovation to learn faster than almost any organization in the world, but that potential only becomes real if teams can experiment responsibly, measure clearly, and roll out changes safely. Adoption of the platform is accelerating rapidly across the company, and recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency for important services. Based out of OpenAI’s Bellevue office, we are a close-knit team that values in-person collaboration, urgency, craft, and impact. We build for other builders, and the best version of this team is one where every OpenAI product team can move faster because the experimentation and rollout layer is dependable, fast, and easy to use. About the Role We are l
About the team The ChatGPT Library team is building the persistence layer that helps people and organizations collaborate with increasingly capable AI systems. ChatGPT Library provides users with a durable place for their files, context, and creations, forming a substrate between humans and agents—a foundation for AI that can remember, retrieve, and act on the right context over time. The team sits within the Personalization organization and partners closely with product, design, research, model, and infrastructure teams. Its work helps models become a useful second brain for individuals—for example, answering questions such as “What did I work on last week?”—and supports enterprise use cases that turn company knowledge into accessible, actionable context. About the role We are looking for a hands-on engineering manager to lead a team of approximately 9–10 engineers and own execution across the product roadmap. This is a technical leadership role for someone who can set a high product bar, build a strong team, shape architecture, and contribute directly when needed. You will form and translate an ambitious product vision into a focused roadmap and help the team ship reliable, intuitive experiences at the intersection of AI, knowledge, and collaboration. The ideal candidate leads by example. You are comfortable moving between people leadership, product decisions, technical design, and code. Founder or early-stage startup experience is especially valuable because the role requires urgency, sound judgment under ambiguity, and a willingness to operate across boundaries. In this role you will • Lead, coach, and grow a team of approximately 9–10 engineers while creating clarity, accountability, and a healthy execution rhythm. • Own the engineering roadmap for Library, working with product and design partners to prioritize the highest-impact problems. • Set a high bar for product quality, technical excellence, reliability, privacy, and user trust. • Provide hands-on techni
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Engineering at Brex Engineering at Brex is about building systems that scale with speed and intention. Our teams span Software, Data, Security, and IT, and operate with high autonomy and deep collaboration. We tackle hard technical problems, own our outcomes, and push for excellence at every level — from architecture to deployment. It’s an environment where engineering is a craft, and builders become leaders. What you’ll do As a Staff Software Engineer in Banking, you will help shape the technical direction of one of Brex’s most strategic and complex product areas. The Banking org is both a product and platform org, it owns the Brex Business Account product and AP offerings like Bill Pay and Vendors, and the underlying money movement platform and partner integrations that power those experiences. In this role, you’ll work across customer-facing product surfaces and core financial infrastructure, driving architecture, reliability, and
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. About the role You will work on core enterprise platform systems, focusing on services that ensure Synthesia is secure, reliable, and scalable for our largest customers. You will contribute to our new suite of APIs that will unlock our market leading Avatar technology to be utilised by external creative tools. We are an AI native company and use the most powerful assistant tools on a daily basis to increase our speed of development, automate repetitive tasks and widen the scope of what the team can worked on. This includes Claude and Cursor. You will have ownership of projects that span months and multiple teams, requiring you to break down complex, ambiguous problems into clear steps that can be delivered and validated iteratively. Engineers within Synthesia are empowered to contribute heavily to product discussions and build with both user and business considerations in mind. Impact is key to everything we do here. You will work closely with product, security, legal, and infrastructure partners, and will be expected to translate business and regulatory requirements into scalable technical solutions. You will evaluate your work through system health and reliability metrics, leveraging observability and monitori
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s mission is to unlock financial freedom for everyone by making money movement and access to financial data simple and secure. As a Software Engineer, you will design and build the systems that power how millions of people connect to their finances. You will work across the stack, from reliable backend services and APIs to intuitive applications that bring those systems to life. You will collaborate with engineers, product managers, and designers to ship products that make financial services more accessible and transparent. At Plaid, engineers take ownership early, grow quickly, and see their work reach millions of users. Responsibilities: Design & Development: Build and maintain backend services with a focus on performance, reliability and scalability. Collaboration: Work closely with product managers and other stakeholders to define and implement new features that meet product and customer needs. Code Quality: Write clean, maintainable and efficient code. Testing & Debugging: Develop automated tests to ensure the quality and reliability of the codebase. Troubleshoot and resolve issues. Engage in hands-on coding and architectural design, setting and maintaining high technical standard
About the Team The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions. Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After joining OpenAI, the team began its next chapter: bringing deep product expertise, customer intuition, and mature platform infrastructure into the product development system used by every OpenAI team. Today, teams across ChatGPT, Codex, model measurement, consumer monetization, business subscriptions, developer products, and shared infrastructure rely on Statsig to safely introduce new capabilities, measure impact, and roll changes forward or back with confidence. We are at a defining moment as adoption accelerates and the platform becomes a company-wide standard. About the Role We are looking for an Engineering Manager, Statsig Product to lead the product engineering organization responsible for Statsig’s post-acquisition journey at OpenAI. You will define how experimentation, rollout, configuration, and analytics become a simple, reliable, and trusted part of how every OpenAI product team ships. You will set strategy across multiple product and platform workstreams, build the organization and leadership structure needed for the next phase, and establish the operating model for a platform that serves teams across the company. The right leader can operate across product strategy, technical architecture, organizational design, developer experience, reliability, and executive alignment. You will help preserve what made Statsig strong while integrating it deeply into how OpenAI launches, measures, learns, and makes product decisions. In this role, you will: Build, lead,
About the Team The Statsig team at OpenAI builds and operates the experimentation platform that powers product development, measurement, and decision-making across the company. We partner closely with product, engineering, and infrastructure teams to ensure experiments are trustworthy, statistically rigorous, and scalable to the needs of frontier AI products. Our mission is to help teams make better decisions through reliable experimentation. We care deeply about statistical correctness, pragmatic solutions, and building systems that researchers and engineers can trust at massive scale. The team operates at the intersection of experimentation methodology, data infrastructure, causal inference, and product analytics. We are looking for experienced experimentation experts who want to shape the future of experimentation in the AI era. About the Role We are hiring a Staff-level Data Scientist to help lead the evolution of OpenAI’s core experimentation platform. This role is focused on improving the statistical rigor, reliability, and practical usability of experimentation across the company. You’ll work on some of the hardest problems in online experimentation: sample ratio mismatch detection, variance reduction, bias mitigation, metric design, triggered analysis, heterogeneous treatment effects, sequential testing, and experimentation in complex ML systems. You’ll also help translate advanced statistical concepts into pragmatic systems and product experiences that teams can actually use. This is a highly technical individual contributor role with significant influence across methodology, platform architecture, and experimentation best practices. The ideal candidate combines deep statistical expertise with strong systems intuition and hands-on experience building or operating experimentation platforms at scale. In this role, you will: Drive the statistical direction and technical strategy for OpenAI’s experimentation platform Design and improve experimentation methodolo
Get new reliability engineer jobs in Seattle by email
Daily job updates · Unsubscribe anytime