Jobiba hiring network

Back End Td Reliability Lab Manager Jobs

1,684 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current back end td reliability lab manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

T
Twitch
📍 New York City• Full-time• From $151.2K/yr
1mo ago

About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Role The Community team builds products that allow creators to build and grow communities on Twitch. Our products span across creator and viewer journeys. As a Senior Product Manager on the Discovery team, you will own the experience that helps viewers find creators worth coming back for, across search, browse, and serendipitous discovery in the live feed. Your work will focus on helping viewers form their first creator connection and every connection after it, because that bond is what turns an occasional visit into a daily habit. You will be the voice of the viewer for recommendations: defining what viewers want, surfacing what they are not seeing but should, owning the discovery roadmap, and partnering with our ML teams to turn those needs into better ranking and relevance. You will define and execute a product strategy grounded in customer needs, working closely with applied science, design, and engineering to ship and measure your way there. Our team is based in San Francisco, CA but you can work from San Francisco, CA; New York, NY; Irvine, CA; Los Angeles, CA or Seattle, WA. You Will: Own the discovery experience across search, browse, and the live feed, helping viewers find creators worth returning for and form their first connection and every one after it. Define what viewers want fro

sqlrestai
View job →
S
Stripe
📍 New York San Francisco• Full-time
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies, from the world’s largest enterprises to the most ambitious startups, use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Global Transformation Team is a small, high-leverage team with a global remit. We lead the programs that transform how ProServ works, so our users realize the full value of Stripe faster. That means making ProServ AI-native, building one prescriptive methodology that scales worldwide, and creating the knowledge systems that make every engagement smarter than the last. We care about craft, consistency, and landing change, not just shipping it. Everything we build is global by default, AI-native by design, and built to stick. If we do this well, the model we prove here becomes the blueprint for how Stripe transforms itself everywhere. What you’ll do As AI Product Manager for ProServ, you will own the roadmap that makes ProServ AI-native — sitting at the intersection of AI strategy, engineering execution, and organisational change. You will operate with significant autonomy and a high degree of ambiguity. The ability to influence without authority, build momentum x-geo, and translate strategic intent into real-world delivery is essential. Own the AI product roadmap for ProServ Deeply understand the needs of ProServ users by spending time with our teams, observing patterns across their feedback, and interrogating our data. We are users first, every roadmap decision should trace back to making their experience better Translate this into an AI strategy including a prioritised, executable product roadmap Work closely with the ProServ leadersh

aigorust
View job →
PE
1mo ago

About Us: Paytm is India’s leading mobile payments and financial services distribution company. A pioneer of the mobile QR payments revolution, Paytm builds technology that enables small businesses and consumers to participate in the digital economy. Our mission is to serve half a billion Indians and bring them into the mainstream economy through technology. About the Role: We are looking for a strategic and data-driven Growth leader to own and scale growth for the Paytm for Business app. The role will be responsible for driving activation, engagement, retention and winback, with a strong focus on improving service-led product journeys and building a scalable growth engine. The ideal candidate will combine strong product and growth thinking with deep expertise in analytics, experimentation and lifecycle marketing. This is a high-ownership role that will work closely with Product, Engineering, Data, Marketing and Business teams to identify growth opportunities, design interventions and drive measurable outcomes at scale. Key Responsibilities: - Own the growth charter for the Paytm for Business app, covering activation, engagement, retention and win back. - Identify opportunities across key service product journeys and build strategies to improve user adoption, usage and retention. - Build a data-led growth engine using user analytics, cohort analysis, funnel metrics and behavioral insights. - Drive a strong culture of experimentation and rapid iteration, including hypothesis creation, A/B testing and measurement of business impact. - Own and scale lifecycle marketing / CLM initiatives across user segments and stages of the customer journey. - Leverage AI and Agentic capabilities to automate growth workflows, personalization, customer lifecycle management and decision-making. - Define and track growth KPIs, identify gaps and translate insights into actionable product and business initiatives. - Work cross-functionally with Product, Engineering, Data Science, Ma

BA
Bolna AI
📍 India• Full-time
1mo ago

About Bolna Bolna is Voice AI infrastructure built for India. We help businesses deploy intelligent voice agents that can call, converse, and convert in any language, at scale. From collections to customer support to sales, our agents handle millions of conversations so humans don't have to. We're a YC F25 company, backed by General Catalyst, with 1,050+ paying customers and growing fast. Our team of ~25 is based in Bengaluru. The Role You're the person who makes our voice agents actually sound good and actually work for real customers. As an AI Solutions Engineer, you'll sit at the intersection of our product and our customers. Your primary job is to design, write, and iterate on the prompts and tools that power Bolna's voice agents, making them smarter, more natural, and more effective for each use case. You'll work closely with customers to understand their goals, build agent flows, test conversations, and fix what breaks. No heavy coding required-if you can vibe-code a basic script or write a solid system prompt, you're qualified. What You'll Do Write, test, and iterate on system prompts for voice agents across industries like D2C, fintech, healthcare, and logistics Listen to real call recordings, identify where agents fail, and fix them Build conversation flows and call pathways for new customer deployments Help onboard new customers-understand their use case, set up their agent, and get it live Maintain a growing library of prompts, templates, and best practices across Bolna's verticals Red-team agents-try to break them, find edge cases, and make them bulletproof Work with vernacular inputs: test agents in Hindi, Hinglish, and regional languages Feed insights back to product and engineering-you'll see what customers need before anyone else does What We're Looking For Must-have: You've spent serious time prompting ChatGPT, Claude, or similar LLMs-not just casually, but to actually build or solve something You're obsessive about language-you notice when a senten

O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking an Actuator Gear Design Engineer to lead the development of custom gears and gear stages for advanced robotic systems. You will own actuator development from early architecture and concept generation through prototype validation and system integration, partnering closely with mechanical, electrical, controls, firmware, and reliability teams. You will partner with external suppliers and internal manufacturing to create full gearbox assemblies. This role focuses on the design, integration, and validation of precision gearing, including broader knowledge around motor electromagnetics, transmission types, sensing, structural components, and thermal architectures. You will help drive actuator development across the full engineering lifecycle while establishing scalable design, test, and integration practices for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will: Lead the architecture, design, and integration of custom robotic actuator gearing. Define actuator requirements and system-level trade studies around torque density, bandwidth, efficiency, thermal performance, back drivability, inertia, reliability, manufacturability, and cost. Design precision electromechanical assemblies with strong attention to tolerances, alignment, load paths, thermal expansion, sealing, wear, and serviceability. Drive actuator integration into robotic systems, partnering closely with controls, firmware, electrical, and robotics software teams to optimize closed-lo

awsrestai
View job →
O
1mo ago

About the Team The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions. Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After joining OpenAI, the team began its next chapter: bringing deep product expertise, customer intuition, and mature platform infrastructure into the product development system used by every OpenAI team. Today, teams across ChatGPT, Codex, model measurement, consumer monetization, business subscriptions, developer products, and shared infrastructure rely on Statsig to safely introduce new capabilities, measure impact, and roll changes forward or back with confidence. We are at a defining moment as adoption accelerates and the platform becomes a company-wide standard. About the Role We are looking for an Engineering Manager, Statsig Product to lead the product engineering organization responsible for Statsig’s post-acquisition journey at OpenAI. You will define how experimentation, rollout, configuration, and analytics become a simple, reliable, and trusted part of how every OpenAI product team ships. You will set strategy across multiple product and platform workstreams, build the organization and leadership structure needed for the next phase, and establish the operating model for a platform that serves teams across the company. The right leader can operate across product strategy, technical architecture, organizational design, developer experience, reliability, and executive alignment. You will help preserve what made Statsig strong while integrating it deeply into how OpenAI launches, measures, learns, and makes product decisions. In this role, you will: Build, lead,

awsrestai
View job →
O
1mo ago

About the team The Technical Success team is responsible for ensuring the safe and effective deployment of ChatGPT and OpenAI API applications for developers and enterprises. We act as a trusted advisor and thought partner for our customers, ensuring developers and enterprises maximize value from our models and products. The Ecosystem team builds and scales third-party integrations around ChatGPT and Codex. We work across product, engineering, partnerships, safety, legal, policy, and go-to-market to create integrations that are useful, trustworthy, and technically sound. About the role We’re looking for an Applied Deployment Engineer to help build and deepen ChatGPT integrations with third-party messaging platforms. The initial focus will be on partner work as well as future opportunities with other messaging apps around the world. This is a hands-on, partner-facing engineering role. You’ll work closely with external product and engineering teams to improve existing integrations, expand capabilities such as image generation, increase retention, and translate partner needs back into OpenAI’s product and engineering roadmap. You should be comfortable writing code daily, making changes across platform or monorepo surfaces when needed, and moving fluidly between technical discovery, product tradeoffs, prototyping, debugging, and executive-level communication. This role is best suited for a strong product-minded engineer with excellent communication skills, sound technical judgment, and the ability to work across cultures, time zones, and organizations. This role is based in our San Francisco office. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Partner with messenger app product and engineering teams to design ChatGPT-powered experiences that fit naturally into chat, group, contact, and app-tab surfaces. Deepen existing partner integrations, with an initial focus on improving user ex

awsrestai
View job →
O
1mo ago

About the Team OpenAI’s API Multicloud team is responsible for extending OpenAI’s API platform into strategic cloud environments, starting with AWS . The team’s mission is to distribute OpenAI’s API broadly and safely by enabling key API technologies in AWS-native environments, in close partnership with Amazon and internal teams across Codex, Research, Safety Systems, and Applied. The team is focused on bringing core developer and enterprise capabilities into cloud-native environments, including AWS-hosted Codex, model customization / post-training as a service, and new stateful runtime environments for agentic workloads. This work sits at the intersection of production ML systems, developer platforms, model behavior, and large-scale infrastructure. About the Role We’re hiring Machine Learning Engineers to build and improve the AI systems that help strategic partners adapt OpenAI models to important use cases in cloud-native environments. This role spans post-training workflows, evaluation, data pipelines, model behavior, and API/infrastructure integration. You’ll work at the boundary between partner needs and core ML systems: helping teams understand what is and isn’t working, diagnosing issues in training and evaluation workflows, and turning those learnings into improvements to the underlying platform. You should enjoy working with external technical partners, extracting the real goal from messy requests, and pushing back or reframing when the requested experiment is not the highest-leverage path. You’ll collaborate closely with Research, Applied, Safety Systems, infrastructure teams, and external technical partners to solve ambiguous model-performance problems. When you succeed, strategic partners and internal teams will be able to improve model behavior with confidence, driving measurable product improvements while the systems behind that work become more reliable, scalable, and effective over time. In this role, you will Partner with strategic customers and in

pythonawskubernetes
View job →

About the Team The Statsig team within OpenAI owns the experimentation, rollout, dynamic configuration, and analytics infrastructure that sits on the launch path for OpenAI products. Our systems help teams ship safely, evaluate product and model changes in production, and make high-confidence decisions from real-world usage. Statsig began as an independent company built around experimentation, feature management, and product analytics at scale. After Statsig joined OpenAI, the team began the next chapter: bringing that platform expertise and infrastructure into OpenAI as the experimentation and rollout foundation for every product we ship. This is infrastructure with a very direct product consequence. Teams working on ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and shared platform systems depend on Statsig to evaluate configurations, move traffic safely, ingest experiment data, serve analytics, and roll changes forward or back when production reality demands it. We are at a critical point in the platform journey. Adoption is accelerating quickly across OpenAI, and the systems that were already important are becoming load-bearing for how the company launches. The infrastructure needs to stay fast under sharply increasing evaluation volume, reliable when more services depend on it, observable enough to debug quickly, and efficient enough to support OpenAI-wide scale. Recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency across important services. The next phase is to make those gains systematic: a platform that can absorb rapidly growing product velocity while preserving low latency, data quality, operational safety, and developer trust. Based out of OpenAI's Bellevue office, we are a close-knit team that values in-person collaboration, technical depth, operational ownership, and building infrastructure that

vueawsrest
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. About the Role The Safety Measurement Product Manager owns OpenAI's approach to measuring harm and safeguard efficacy in production, including driving the strategy for our suite of safety measurement platforms and products used across the company. You will partner closely with our safety research and engineering teams to determine what we measure, where we measure it, and how we measure it, feeding those insights directly into critical leadership decisions and back into our safety work. You will also represent the company's topline safety metric as well as prioritize incoming requests from partner teams to expand our safety measurement platform to more use cases. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Partner closely with data science, research, engineering, policy teams, and other stakeholders to craft a vision for understanding safety outcomes and prevalence on our platforms. Define strategic priorities and product roadmaps focused on improving safety measurement approaches will scaling our measurement platform to more use cases, products, and cross-functional team needs. Establish repeatable processes to integrate cutting-edge AI safety research into OpenAI’s safety m

awsrestai
View job →

About the Team The Statsig team within OpenAI builds the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Our work sits on the critical path for how product, engineering, research, and go-to-market teams learn from real-world usage and make high-confidence decisions. Statsig began as an independent company focused on helping builders move faster through trustworthy experimentation and feature management. After Statsig joined OpenAI, the team began the next chapter: bringing that deep product expertise, customer intuition, and mature platform infrastructure into OpenAI as the experimentation and rollout platform for every product we ship. Today, we support teams across ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and the shared infrastructure that connects them. These teams rely on Statsig to safely introduce new capabilities, compare product and model behavior, measure impact, and roll changes forward or back with confidence. We are at a defining moment in the platform journey. OpenAI has the data, product surface area, and pace of innovation to learn faster than almost any organization in the world, but that potential only becomes real if teams can experiment responsibly, measure clearly, and roll out changes safely. Adoption of the platform is accelerating rapidly across the company, and recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency for important services. Based out of OpenAI’s Bellevue office, we are a close-knit team that values in-person collaboration, urgency, craft, and impact. We build for other builders, and the best version of this team is one where every OpenAI product team can move faster because the experimentation and rollout layer is dependable, fast, and easy to use. About the Role We are l

vueawsrest
View job →
P
Postman
📍 San Francisco• Full-time
1mo ago

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. About the Role We’re looking for a Senior Developer Advocate to join our Developer Relations team and help developers everywhere build better APIs, faster. You’ll serve as a trusted technical voice for Postman by creating compelling content, building community, and championing developer needs back into the product. This is a high-visibility, high-impact role. You won’t just talk about Postman, you’ll be hands-on with APIs, AI agents, and emerging protocols like MCP , translating complex workflows into accessible, inspiring content that reaches millions of developers. What You’ll Do Create technical content at scale: Write tutorials, blog posts, videos, sample projects, and livestreams that demonstrate API best practices using Postman and adjacent tools. Represent Postman publicly: Speak at conferences, meetups, podcasts, and webinars. Be a recognizable and credible technical voice in the API and developer tooling ecosystem. Build and nurture community: Engage with developers across forums, Discord, GitHub, social media, and in-person events. Lead or support the global Agents & APIs Meetup program and similar community initiativ

typescriptpythonci/cd
View job →

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. The Role You'll be an engineer who builds AI agents in production, sitting close to the customers who depend on them. This is a full-stack engineering job with an unusually short distance between your code and someone's actual workday. You'll write Go and Python, design schemas, build evals, and present your solution to a senior executive at an enterprise - often in the same week. What You'll Do Design and ship production agents. You'll build agents that are mission-critical from day one: embedded in Teams, Slack, intranets, voice lines, and email, taking real actions against SAP, ServiceNow, Workday, and a long tail of systems nobody has heard of. These run at enterprise volume under enterprise scrutiny. Own the full lifecycle. Discovery, build, eval, launch, and the unglamorous months afterward where an agent goes from good to genuinely reliable. Work directly with the people whose problem it is. You'll sit with leaders at global enterprises, extract the process from their heads, and decide what should be an agent, what should be a workflow, and what should stay human. Push your work back into the platform. The best patterns you find in the field become part of Ema's core product

pythonaigo
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Senior Software Engineer - Customer Experience Engineering Our customers are very happy with our technology and are quickly expanding their use cases, only limited by their own imaginations. The Customer Experience Engineering Team aims to enable customers to push these limits, and when customers run into roadblocks or issues, enable them and our internal technical resources to resolve those issues and get back to seeing success with Snowflake. Snowflake’s Customer Experience Engineering Team is expanding! We are looking for a Senior Software Engineer to join our team who likes building intelligent data applications that give insight into diagnostic data and technical content that is relevant to issues customers may be facing. Our Customer Experience Senior Software Engineers enjoy developing features into the Snowflake product that aim to greatly reduce friction points that block customers and regularly require help from Snowflake technical experts. They are excited by the challenges presented by extracting intelligence from a wide variety of data sources, including structured and unstructured data. As a Sr. Software Engineer, you will: Lead and drive projects that span our stack, including Java and Python services hosted in Kubernetes and Snowpark Container Services. Prom

pythonjavakubernetes
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the role We're hiring someone to build the playbook for how we drive practitioner adoption of new capabilities, and bring new practitioners into the platform. You'll pick a segment (data engineers, ML engineers, analytics engineers), pair it with a use case, and figure out what it takes to get practitioners from awareness to activation. That means developing real knowledge of your audience: what they care about, where they spend time, what resonates, and what falls flat. Then working cross functionally to assemble the right assets, sequence them into an exciting journey that leads to adoption, and getting your campaign in front of the right people across owned, paid, in-product, and partner channels. What you're building is a closed loop. Each campaign is instrumented so you can see what actually worked. Did practitioners activate? Did usage stick? How quickly did they find value? Each round of learning makes the next one sharper. The insights you generate don't stay with you; they feed back to marketing, product and account teams to drive broader adoption across organizations. If you want to own the full arc from choosing the audience to designing the playbook to shipping the campaign to building the feedback loops that make it all compound over time, this is that ro

🔔

Get new back end td reliability lab manager jobs by email

Daily job updates · Unsubscribe anytime