Jobiba hiring network

Reliability Engineer Jobs

2,028 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

L
Lyft
📍 San Francisco• Full-time
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. The Lyft Business Product Platform team builds the systems and experiences that power Lyft's B2B products — enabling companies, organizations, and their employees to seamlessly access Lyft's transportation network. We sit at the intersection of product and platform, owning both the customer-facing features and the underlying infrastructure that makes them reliable at scale. Our work directly impacts how businesses integrate with Lyft, how admins manage their programs, and how millions of riders get where they need to go. Responsibilities: Drive architecture and technical design for systems that are highly available, scalable, and built to last — not just for today's requirements but for where the product is heading Own features end-to-end: from shaping the technical spec and design through to production rollout and operational health Think critically about how AI capabilities can be incorporated into Lyft Business products to improve the experience for business admins and riders — and bring that perspective into roadmap and architecture conversations Make well-reasoned trade-off decisions and communicate them clearly to peers, leads, and cross-functional partners Write clean, well-tested, maintainable code and hold a high bar for the same in code reviews Partner across engineering, product, and design to align on direction and get buy-in on technical approaches Proactively engage in incident response, contributing both to resolution and to long-term reliability improvements Grow the team's technical culture through design reviews, tech talks, and mentorship Experience: 5+ years of software engineering experience, with a track record of designing and shipping production systems at scale Strong system design instincts — you can reason through distributed systems trade-offs, identify failure modes, and

sqlawsazure
View job →
D
1mo ago

Role Description As a Software Engineer on the Metadata team, you’ll build and operate the large-scale distributed databases that every Dropbox service depends on. Metadata systems are mission-critical, in the live path for all user operations and must meet stringent requirements for latency, durability, and transactional consistency. You’ll design and evolve the core infrastructure that manages Dropbox’s databases at scale, enabling fast, reliable access to data for millions of users and hundreds of internal services. This work spans distributed systems, replication, caching, and transactional database systems. You’ll collaborate closely with engineers across Infrastructure and Product teams to ensure the metadata layer meets business needs and continues to scale with Dropbox’s growth. This is an opportunity to leverage your expertise in distributed systems and grow into broader technical leadership. Our Engineering Career Framework is viewable by anyone outside the company and describes what’s expected for our engineers at each of our career levels. Check out our blog post on this topic and more here . Responsibilities Design and maintain distributed database systems providing low-latency, strongly consistent data access Implement and optimize replication, consensus, and caching mechanisms to meet availability and performance goals Operate production systems, including participating in the on-call rotation, ensuring high availability and data durability Collaborate with infrastructure and product teams to assess current and future use cases and requirements, supporting the development of a mid- to long-term roadmap that reflects these needs Contribute to system design reviews, postmortems, and reliability improvements Write high-quality, efficient code in Go and Rust for performance-critical systems On-call work may be necessary occasionally to help address bugs, outages, or other operational issues, with the goal of maintaining a stable and high-quality experienc

REMOTEsqlmysqlredis
View job →
D
Dropbox
📍 Poland• Full-time• Remote
1mo ago

Role Description As an Infrastructure Engineer, your role will be crucial in shaping and constructing the robust systems that not only support our current flagship products but also lay the groundwork for the next wave of engineering innovations. From optimizing user experiences across various projects to ensuring seamless scalability and data integrity, you'll be at the forefront of shaping the technological backbone of our platform. Collaborating closely with cross-functional teams, you'll leverage your expertise to tackle audacious challenges and push the boundaries of what's possible. Your contributions will directly impact millions of users, as every line of code you write furthers our mission to revolutionize the way people work and collaborate. Join us in redefining the future, where your passion for building scalable, reliable systems will drive meaningful change on a global scale. Our Engineering Career Framework is viewable by anyone outside the company and describes what’s expected for our engineers at each of our career levels. Check out our blog post on this topic and more here . Responsibilities Build infrastructure capable of managing metadata for hundreds of billions of files, handling hundreds of petabytes of user data, and facilitating millions of concurrent connections. Assist in expanding Dropbox's role as the data-fabric, linking hundreds of millions of applications, devices, and services worldwide, while spearheading efforts to improve interoperability and adaptability across various ecosystems. M easur e and optimiz e Dropbox's analytics platform to maintain its status as one of the most advanced in the industry for extracting meaningful insights from vast data volumes. Collaborat e with cross-functional teams to innovate and implement solutions that enhance the performance, reliability, and security of Dropbox's infrastructure, ensuring a seamless experience for users worldwide. On-call work may be necessary occasionally to help address bugs,

REMOTEpythonjavagit
View job →

Who We Are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies — from the world's largest enterprises to the most ambitious startups — use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the Team The Core Change Management group is responsible for the systems that let every Stripe engineer ship code, configuration, and infrastructure changes safely and at high velocity. You will be embedded primarily on the Service Deployments team — the owners of Stripe's end-to-end code deployment platform — with regular collaboration with the Resource Automation and Feature Deployments teams. What Makes This Role Compelling You own the foundation of how Stripe ships software. The deployment platform sits in the critical path of every engineer's workflow at Stripe. The decisions you make affect thousands of deploys per day across hundreds of services, directly determining how fast and safely Stripe's product evolves. Technically rich, architecturally active. The team is executing several concurrent platform transformations: containerizing host-based services at scale, adding intelligent multi-service deploy pipelines, extending real-time anomaly detection to earlier stages of traffic shifts, and rebuilding deployment event infrastructure on top of a durable message bus. This is not maintenance work — the architecture is in motion. Broad surface area, real ownership. You will span the full stack from container scheduling and deployment orchestration business logic to the developer-facing internal platform UI. The problems are multi-layered: reliability, developer experience, performance, and safety all at once. Your judgment prevents incidents.

awsazurekubernetes
View job →
S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world's largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Revenue and Financial Automation Sub-org within Revenue and Financial Automation: Billing. The Revenue and Financial Automation team at Stripe builds software tools that accelerate the economic and technological growth of global businesses by helping them operationalize their commercial relationships with customers. Our offerings include a billing platform, SaaS analytics, data services, and finance automation products that our customers creatively combine to support various revenue models. Team Matching: exact team matching for one of the subteams will begin during final stages. Please note we may also consider you for different orgs based on your experience, location, etc. More information on our team matching process can be found here. What you'll do We're looking for engineers who want to build the distributed systems, APIs, and backend services that power Stripe's revenue and billing platform. You'll focus on API design and performance, service reliability, and distributed systems challenges, while collaborating across the stack to ship complete solutions for millions of businesses. Responsibilities Scope, architect, and lead technical projects to build and scale distributed backend systems and APIs Design, build, and maintain reliable, high-performance APIs and backend services Own service reliability, including setting performance targets and driving improvements across the stack Debug production issues across distributed services

S
Stripe
📍 Taipei• Full-time
1mo ago

About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe Terminal helps Stripe users extend their online presence into the physical world. The Terminal team’s mission is to make it as easy for businesses to accept in-person payments as the Stripe API has done for online payments. With Terminal, businesses can unlock in-person payments use cases that are right for their business model—whether it’s creating a flagship retail experience, extending their website to a pop-up store, or enabling a mobile point-of-sale at their next event. What we are looking for: As an Android BSP Engineer, you will be responsible for the kernel and driver level system development, which includes building, troubleshooting, and writing automated tests for the Android system on our embedded payments platforms. This team works closely with partner teams throughout the hardware and software product lifecycle, from hardware manufacturing to Android app teams. We also work with external vendors on part selection and initial hardware bring-up. What you’ll do: Bring up new devices and lead debugging and performance tuning exercises that span multiple hardware/firmware/software teams. Design, implement, and maintain drivers and Android services that operate efficiently in a constrained environment and meet the reliability and security requirements of the industry. Own the definition of one or more work streams focused on hardware bring-up, peripheral drivers and communication, and power and performance management and opti

javaaikotlin
View job →
S
Stripe
📍 South San Francisco• Full-time• $212K – $318K/yr
1mo ago

Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Design state-of-the-art ML models and large-scale ML systems for underwriting and portfolio management for Stripe Capital based on ML principles, domain knowledge, risk, regulatory and engineering constraints. Design systems to speed up the time from idea to deployment of new models. Experiment and iterate on ML models (using tools including PyTorch and TensorFlow) to achieve key business goals and drive efficiency. Develop pipelines and automated processes to train and evaluate models in offline and online environments. Integrate ML models into production systems and ensure their scalability and reliability. Collaborate with product and strategy partners to propose, prioritize, and implement new product features. Engage with the latest developments in ML/AI and take calculated risks in transforming innovative ML ideas into productionized solutions. Who you are Minimum requirements Must have a Bachelor's degree or foreign equivalent in Computer Science, Machine Learning, Mathematics, Physics, Statistics, or a related field, plus two (2) years of experience in Building and shipping ML systems in production. Must have two (2) years of experience in each of the following: ML algorithms and model architectures; Designing, training and evaluating machine learning models; Productionizing and deploying machine learning models at scale; Orchestrating data pipelines and leveraging large-s

machine learningaigo
View job →
S
Stripe
📍 Singapore• Full-time
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The Payments engineering powers the processing of hundreds of billions of dollars in card payments annually, enabling users to accept card payments from Asia Pacific, Europe, and Middle East regions. As the Engineering Manager for the team, you’ll have responsibility for expanding the reach of Stripe’s global payments network, delivering the best-in-class experience for card payments in these regions. In this role, you’ll be making some of the most significant strategic, product, and technical decisions for card payments at Stripe. The systems this team builds enables our users to create a single integration with Stripe, then take payments across an ever-expanding set of card networks and grow their businesses globally. What you’ll do Responsibilities Work with other Stripe leaders to author Stripe’s card payments strategy Develop engineers on the team, helping them advance in their careers Empower the engineering team to achieve a high level of technical productivity, reliability and simplicity Manage processes to help the team do its best work and interface effectively with the rest of Stripe Contribute to engineering-wide initiatives as a member of Stripe’s engineering management team Recruit great engineers, in collaboration with Stripe’s recruiting team You may be a fit for this role if: You thrive on a high level of autonomy and responsibility You encourage a healthy work environment that’s both supportive and challenging You’re ex

restaigo
View job →
S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The APAC & EMEA Cards team powers the processing of hundreds of billions of dollars in card payments annually, enabling users to accept card payments from Asia Pacific, Europe, and Middle East regions. As the Engineering Manager for the team, you’ll have responsibility for expanding the reach of Stripe’s global payments network, delivering the best-in-class experience for card payments in these regions. In this role, you’ll be making some of the most significant strategic, product, and technical decisions for card payments at Stripe. The systems this team builds enables our users to create a single integration with Stripe, then take payments across an ever-expanding set of card networks and grow their businesses globally. What you’ll do Responsibilities Work with other Stripe leaders to author Stripe’s card payments strategy in APAC & EMEA Develop engineers on the team, helping them advance in their careers Empower the engineering team to achieve a high level of technical productivity, reliability and simplicity Manage processes to help the team do its best work and interface effectively with the rest of Stripe Contribute to engineering-wide initiatives as a member of Stripe’s engineering management team Recruit great engineers, in collaboration with Stripe’s recruiting team You may be a fit for this role if: You thrive on a high level of autonomy and responsibility You encourage a healthy work environment that’s both supportive and

restaigo
View job →
S
Stripe
📍 San Francisco• Full-time
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies, from the world's largest enterprises to the most ambitious startups, use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Stripe's Infrastructure organization builds and operates the shared platforms that make it possible for engineering teams to build, deploy, run, and scale reliable products. The organization owns foundational capabilities across service networking and discovery, safe feature and configuration changes, data storage and orchestration, and batch computing. Our platforms support a global engineering audience and power systems at significant scale. This includes the control plane for secure service-to-service communication across tens of thousands of microservices, service and feature deployment infrastructure that helps teams introduce and recover from changes safely, infrastructure for running Stripe's containerized services, and data platforms for storage, pipeline orchestration, and large-scale processing. Our batch compute platforms support technologies including Apache Airflow, Apache Spark, Apache Iceberg, Apache Hadoop, and Apache Celeborn. As an Engineering Manager, you will lead or strongly influence teams working on business-critical infrastructure with broad internal adoption. You will balance reliability, scalability, security, developer experience, and operational rigor while helping Stripe move quickly with confidence. You will partner closely with product, security, data, support, and adjacent infrastructure leaders to set direction, resolve cross-team dependencies, and turn recurring needs into durable platform capabilities. Wh

pythonjavamicroservices
View job →

MDM TechnologyMDM Technology, a company affiliated with Dun & Bradstreet, provides enterprise-grade master data management (MDM) solutions that help organizations create a single, trusted view of business entities. By cleansing, matching, linking, and enriching data, anchored by the global standard D‑U‑N‑S® Number business identifier, MDM Technology enables accurate identity resolution across systems to support analytics, compliance, and AI workflows across industries. The Senior Quality Assurance Engineer ensures our software meets user and product needs. The role creates and maintains project test plans, defines and tracks quality assurance metrics, collects and analyzes data for software process evaluation and improvements, and integrates them into business processes combining deep technical expertise in API and Integrations testing and Automation. This role must show a passion for AI enabled Software Quality practices Leveraging modern AI tools to improve testing efficiency, accelerate release cycles, and enhance software reliability across complex enterprise integration.

BA
Bolna AI
📍 Bengaluru• Full-time
1mo ago

At Bolna, we’re building tools that change the way teams leverage Voice AI. We’re looking for a Founding Machine Learning Engineer to own the end-to-end lifecycle of building, evaluating, deploying, and improving models that power millions of production conversations. This is a high-impact, high ownership role where you won’t just work on Bolna’s ML stack—you’ll help build the foundation it scales on. Our team includes IIT alumni with experience at Bain, Atlassian, Uber, Zomato, and LinkedIn, and is backed by leading investors. Responsibilities: Build the data engine - Design pipelines to source and clean conversational voice data across Indian languages, accents, and telephony conditions. Fine-tune models that ship - Fine tune and train models to improve accuracy, speed, and reliability across different use-cases. Define what "good" means - Build evaluation datasets and benchmarks for transcription accuracy, voice naturalness, interruption handling, latency, and end-to-end conversation quality. Set up human-in-the-loop pipelines to capture subjective quality at scale. Ship to production - Work with the engineering team to deploy models into a latency-sensitive, high-volume system. Monitor performance in the wild, debug regressions, and iterate fast. Required Skills: 3+ years of hands-on ML experience with deep practical real-world experience in training models. Strong Python and PyTorch fundamentals with exposure in distributed training, and modern fine-tuning techniques (LoRA, QLoRA, DPO, RLHF, etc.). Training data as a first-class problem. Experience designing data pipelines from collection, cleaning, labeling, deduplication, augmentation and treating data quality as a core engineering discipline. Rigorous about evaluation. You know that "looks good in a demo" is not a benchmark. You build the evals before you trust the model. Speech model experience is a plus with real-time / streaming inference experience where you would have contributed to latency optimization

pythonmachine learningai
View job →

About the Team Enterprise Verticals builds role-specific ChatGPT Work experiences for high-value enterprise workflows. We combine product engineering, plugins and skills, connectors, data, evaluations, and customer evidence to turn useful demos into reliable daily work. This opening sits within the Technology vertical inside Enterprise Verticals. The group focuses on repeatable workflows for people at technology companies, beginning with functions such as data and analytics, sales, and design, and carries the shared platform needs—tool integration, permissions, quality measurement, and safe rollout—across those experiences. We work closely with Design, Research, GTM, Security, and platform teams, as well as with customers and design partners. Success means that people can reach a trustworthy first result, understand what the system did, and keep using the workflow—not merely that a prototype exists. About the Role We are looking for an exceptionally experienced, hands-on full-stack engineer to define and build the next generation of AI-powered enterprise workflows. You will take on the hardest and most ambiguous problems in the Technology vertical: translating real customer needs into product direction, designing the systems behind the experience, and personally writing and shipping production-quality code across the stack. You will own the technical direction and end-to-end delivery of products spanning ChatGPT Work surfaces, backend services, plugins, connectors, enterprise data, permissions, and evaluations. You will make foundational architecture and product tradeoffs; establish patterns other engineers can build on; and hold these experiences to a high bar for reliability, security, observability, and customer value. This is an individual-contributor role for an engineer who leads through technical judgment, direct execution, and influence—not people management. You should be equally comfortable working directly with customers, setting direction with senior cro

typescriptpythonreact
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking an Actuator Gear Design Engineer to lead the development of custom gears and gear stages for advanced robotic systems. You will own actuator development from early architecture and concept generation through prototype validation and system integration, partnering closely with mechanical, electrical, controls, firmware, and reliability teams. You will partner with external suppliers and internal manufacturing to create full gearbox assemblies. This role focuses on the design, integration, and validation of precision gearing, including broader knowledge around motor electromagnetics, transmission types, sensing, structural components, and thermal architectures. You will help drive actuator development across the full engineering lifecycle while establishing scalable design, test, and integration practices for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will: Lead the architecture, design, and integration of custom robotic actuator gearing. Define actuator requirements and system-level trade studies around torque density, bandwidth, efficiency, thermal performance, back drivability, inertia, reliability, manufacturability, and cost. Design precision electromechanical assemblies with strong attention to tolerances, alignment, load paths, thermal expansion, sealing, wear, and serviceability. Drive actuator integration into robotic systems, partnering closely with controls, firmware, electrical, and robotics software teams to optimize closed-lo

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking a senior Actuator Electromagnetic Design Engineer to lead the development of custom electromechanical actuators for advanced robotic systems. You will own actuator development from early architecture and concept generation through prototype validation and system integration, partnering closely with mechanical, electrical, controls, firmware, reliability, and manufacturing teams. This role focuses on the design, integration, and validation of precision electromechanical systems, including motors, transmissions, sensing, structural components, and thermal architectures. You will help drive actuator development across the full engineering lifecycle while establishing scalable design, test, and integration practices for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will: Lead the architecture, design, and integration of custom robotic actuators, including the design, simulation, integration and sourcing of custom electromagnetic components. Define actuator requirements and system-level trade studies around torque density, bandwidth, efficiency, thermal performance, inertia, reliability, manufacturability, and cost. Design precision electromechanical assemblies with strong attention to tolerances, alignment, load paths, thermal expansion, sealing, wear, and serviceability. Drive actuator integration into robotic systems, partnering closely with controls, firmware, electrical, and robotics software teams to optimize closed-loop performance. Devel

awsrestai
View job →
🔔

Get new reliability engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More reliability engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.

Top cities for Reliability Engineer

City links are canonicalized and require at least 20 current jobs.

Countries hiring Reliability Engineer

Country links use the same curated canonical inventory as Jobiba sitemaps.