Jobiba hiring network

Reliability Engineer Jobs

2,049 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

TI
THG Ingenuity
📍 Manchester, United Kingdom• Full-time
16 days ago

About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. Database Platform Manager Company: THG Ingenuity Location: Manchester (Head Office) Reports to: Director of Data Role Overview We are looking for a strong technical leader to run our database platform. Our estate runs on Google Cloud Platform following a recent migration to self-managed services, and there is a real opportunity here to shape the next phase. How we consolidate and modernise the estate, where managed and cloud-native services earn their place, which new technologies are worth adopting, and how the platform scales with the business are all live questions. You will lead the thinking on them and work with stakeholders across the business to agree and deliver the roadmap. You will lead our DBA team, who own every database across the group regardless of the application running on it and who run the platform as a 24/7 service. You will work hand in hand with our Data Reliability Engineering team, whose Principal Engineer is your peer and whose focus is automation, fleet reliability and SLOs. This role is weighted toward hands-on technical depth. You should be as credible in a design review or an incident bridge as you are in a planning session with senior stakeholders. Key Responsibilities Technical direction The technical roadmap for the estate. We want someone genuinely interested in emerging database technologies who evaluates them on merit and can articulate the pros and c

pythonsqlpostgresql
View job →

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Workforce Identity Cloud Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities, such as reducing costs and doing more for your customers. If you like to be challenged and have a passion for solving large-scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate is someone who exemplifies the ethics of, “If you have to do something more than once, automate it” and who can rapidly self-educate on new concepts and tools. Position Overview: The Staff Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses on architecting and managing reliable, scalable, and secure Kubernetes-based platforms on AWS, ensuring high availability and performance while optimising costs and automation. The ideal candidate will have hands-on experience with AWS infrastructure, Kubernetes platform creation, Helm charts, Karpenter scaling, and Istio service mesh. Key Responsibilities: Kubernetes Platform Creation: Design, implement, and maintain highly available, scalable, and fault-tolerant Kubernetes platforms. Ensure clusters are optimised for production workloads, providing high resilience and operational efficiency. AWS Infrastructure Management: Build, manage, and optimise AWS cloud infrastructure, including EKS, ECS, S3, VPCS, RDS, IAM, and more. I

pythonawsdocker
View job →

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Workforce Identity Cloud Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities, such as reducing costs and doing more for your customers. If you like to be challenged and have a passion for solving large-scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate is someone who exemplifies the ethics of, “If you have to do something more than once, automate it” and who can rapidly self-educate on new concepts and tools. Position Overview: The Staff Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses on architecting and managing reliable, scalable, and secure Kubernetes-based platforms on AWS, ensuring high availability and performance while optimising costs and automation. The ideal candidate will have hands-on experience with AWS infrastructure, Kubernetes platform creation, Helm charts, Karpenter scaling, and Istio service mesh. Key Responsibilities: Kubernetes Platform Creation: Design, implement, and maintain highly available, scalable, and fault-tolerant Kubernetes platforms. Ensure clusters are optimised for production workloads, providing high resilience and operational efficiency. AWS Infrastructure Management: Build, manage, and optimise AWS cloud infrastructure, including EKS, ECS, S3, VPCS, RDS, IAM, and more. I

pythonawsdocker
View job →
S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Stripe's Financial Connections group is building technology that expands the scope of problems we tackle beyond card payments. We aspire to enable individuals and businesses to securely share their financial data to improve their product experiences. With more high-quality customer financial data, both Stripe and our users can offer better products to their customers. We have ambitious goals to enable all global businesses to run their entire financial life cycle on Stripe, and we want you to be part of our story. Technical Operations roles in Financial Connections are a dynamic and critical component of the organization's success. As Financial Connections TechOps Manager, you will lead a team of Integration Reliability Engineers who sit at the intersection of product and platform engineers, financial partners, and AI automation. Your role is to ensure that nothing is lost in translation: setting the standard for how the team operates, where it invests, and how the operational backbone of Stripe's open banking ecosystem stays reliable at scale. What you'll do Lead integration reliability and partner operations Own the health of Financial Connections' banking integrations at the team level—define what "reliable" looks like across the partner ecosystem, set the operational bar, and hold the team to it Serve as the management escalation point for complex partner issues, coordinating across financial partners, Stripe engineering, and product lea

aigorust
View job →
DU
16 days ago

About the Team DoorDash Labs, established in 2018, serves as the innovation hub for DoorDash, focusing on developing automation and robotics solutions to enhance last-mile logistics. The team's mission is to create technologies that support and augment human networks, aiming to improve efficiency for Dashers, merchants, and consumers alike. We’re ruthlessly focused on business impact. We are a highly senior team composed of former pioneers from a variety of different robotics industries. As of 2025, DoorDash has completed 10B lifetime deliveries. We’re focused on how to do the next 10B even better. About the Role We are seeking a highly motivated Senior/Staff Test Engineer to join our team. This individual will play a key role in the development and validation of our unmanned platforms at the system and component levels. The ideal candidate has a strong background in test development, test execution, and root cause analysis with a proven track record of collaboratively managing risk throughout a fast paced development process. You’re excited about this opportunity because you will… Run and monitor tests within our facility as well as at outside test labs. Collaborate with a tight knit team to identify and understand test failures. Be hands-on in developing test methods and equipment to uncover failures before they happen in the field. Find clarity through root cause analysis of lab and field failures and suggest design changes to prevent them. Use your creativity to create novel and scaled tests for autonomous systems. We’re excited about you because you have… A bachelors or advanced degree in a relevant engineering discipline. Mastery of test equipment such as environmental chambers, vibration tables, water testers, DAQs, etc.. Ability to bring order to complex test and development programs via clear technical communication and documentation. Experience designing and building testers and equipment. Ability to write Python scripts to automate

pythonawsgit
View job →
🔔

Get new reliability engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More reliability engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.

Countries hiring Reliability Engineer

Country links use the same curated canonical inventory as Jobiba sitemaps.