Jobiba hiring network

Reliability Engineer Jobs

2,028 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

PE
1mo ago

About Meesho Meesho is India's fastest-growing internet commerce company, on a mission to democratize e-commerce for everyone. We serve millions of customers and over 1.75 million sellers through technology-driven innovation, building the scalable systems that power Meesho's most critical surfaces — Search, Recommendations, Personalized Ranking, Logistics, Fraud Detection, and Image Match. The AI Platform sits at the heart of this. It serves a peak of 1M+ real-time deep-learning model inferences per second on ordinary days, scaling 3x+ on sale days — with the reliability that scale demands. The team works at the frontier of applied AI and infrastructure — multi-region inference, novel embedding-search algorithms, and optimized open-weight LLM models — squeezing out every bit of computation and passing the cost savings straight back to customers. About the Role We are looking for an experienced Engineering Manager – AI Engineering to lead the development of scalable AI platforms and infrastructure while managing high-performing engineering teams. You will drive the design, delivery, and optimization of production-grade AI systems powering AI use cases across Meesho.

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary We are seeking a Software Development Engineer In Test, to apply advanced software engineering skills to improve quality across the development lifecycle through scalable test frameworks, developer tooling, and reusable automation solutions. This is a senior individual contributor role for a developer who specializes in testing, testability, and quality engineering. This role combines hands-on engineering with technical leadership across teams. The individual will influence architecture, strengthen automation strategy, improve developer feedback loops, and help establish consistent quality engineering practices that scale across products and platforms. The role carries strong Software Development Engineer in Test expectations, with an emphasis on building engineering solutions that improve product quality, platform reliability, and development velocity. Primary Responsibilities Design, develop, and evolve scalable test frameworks, automation libraries, and developer-facing quality tools Understand how the broader software ecosystem works together and define quality engineering approaches that align to platform strategy Partner with Product, Architecture, and Engineering teams to define test strategy, coverage goals, and acceptance criteria Build reusable utilities, harnesses, mocks, stubs, and data solutions that improve system

typescriptjavaai
View job →
C
Cvshealth
📍 Work From Home, United States• Remote
10 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Locations: CVS Health follows a hybrid work model providing office-based colleagues the ability to flex between working in the office and working from home based on the work you need to accomplish. Various locations available and will be confirmed during conversations with the recruiter. Position Summary Designs and defines the technical architecture and infrastructure required for digital solutions. Utilizing Natural Language Processing, Deep Machine Learning, and Algorithms & Data Structures Writes code, develops software components, and implements complex functionalities according to project requirements. Collaborates with other members of the development team and stakeholders to make high-level architectural decisions, proposes design patterns, and ensures scalability, performance, and maintainability of digital solutions. Leverages advanced programming skills to design and implement complex features, optimize performance, and ensure code efficiency. Integrates various software components or systems, ensuring seamless communication and interoperability between different parts of the digital solution. Writes and executes comprehensive test cases, conducts code reviews, performs debugging, and troubleshoots issues to ensure the reliability, stability, and high quality of digital solutions. Participates in agile or other development

REMOTEawsmachine learning
View job →
O
10 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology anywhere, on any device or app. Our Workforce and Customer Identity Clouds enable secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not looking for someone who checks every single box we’re looking for lifelong learners and people who can make us better with their unique experiences. Join our team! We’re building a world where Identity belongs to you. About Technology Data and Intelligence at Okta At Okta, the Technology Data and Intelligence (TDI) team drives internal efficiency through secure, scalable, and innovative systems. TDI partners with teams across the company to build and support the infrastructure, automation, and enterprise applications that keep operations running smoothly. Focused on enabling productivity and aligning technology with business goals, TDI plays a vital role in both day-to-day operations and long-term strategic growth. The Senior Software Engineer Opportunity We are looking for a Senior Software Engineer to join our growing team in TDI and help scale our internal business solutions with a sharp focus on security, reliability, scalability, and intelligent automation. You will be responsible for designing and developing customizati

javascriptpythonjava
View job →
Z
10 days ago

Level Up Your Career with Zynga! At Zynga, we bring people together through the power of play. As a global leader in interactive entertainment and a proud label of Take-Two Interactive, our games have been downloaded over 6 billion times—connecting players in 175+ countries through fun, strategy, and a little friendly competition. From thrilling casino spins to epic strategy battles, mind-bending puzzles, and social word challenges, our diverse game portfolio has something for everyone. Fan-favorites and latest hits include FarmVille™, Words With Friends™, Zynga Poker™, Game of Thrones Slots Casino™, Wizard of Oz Slots™, Hit it Rich! Slots™, Wonka Slots™, Top Eleven™, Toon Blast™, Empires & Puzzles™, Merge Dragons!™, CSR Racing™, Harry Potter: Puzzles & Spells™, Match Factory™, and Color Block Jam™—plus many more! Founded in 2007 and headquartered in California, our teams span North America, Europe, and Asia, working together to craft unforgettable gaming experiences. Whether you're spinning, strategizing, matching, or competing, Zynga is where fun meets innovation—and where you can take your career to the next level. Join us and be part of the play! What You’ll Do : Develop new and innovative features played by millions of players using Java, C#, C++, Python, javascript. Follow engineering best practices towards ensuring performance, reliability, and measurability Work on large problems and break it up for others to implement. Strong Analytical, programming and debugging skills Perform Design and Code reviews. Be responsible for the Live game health Closely work with other functions like PM, UI/UX, Art, QA Mentor Junior Engineers. Constantly look for opportunities to improve the game performance. Take a hands-on approach in the development of prototypes quickly What you bring: Masters or bachelor’s degree in Computer Science, Engineering or equivalent 8+ years professional experience working in C#, C++, Javascript, Android, IOS, React, Java Solid fundamenta

javascriptpythonjava
View job →
TI
10 days ago

About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. About the Role As Maintenance Lead, you’ll play a central role in keeping our fulfilment operation running safely, efficiently and reliably. Working closely with the Engineering and Facilities Site Manager, you’ll lead the site’s planned and reactive maintenance activity while shaping the technical agenda and building a culture of continuous improvement. This is a high-impact leadership role for an experienced engineer who enjoys balancing the demands of a fast-moving operation with the opportunity to develop long-term reliability strategies. You’ll use modern maintenance systems and techniques to improve equipment availability, reduce recurring failures and help our engineering teams perform at their best. What you’ll be doing Support the development and delivery of the site’s maintenance strategy. Lead a reliability-focused approach using TPM, RCA, RCM and FMEA. Deliver monthly improvement plans that increase equipment reliability and availability. Use CMMS data and failure analysis to identify root causes and deliver lasting solutions. Coordinate the response to live plant issues, ensuring risks are managed and actions are followed through. Own effective weekly and monthly maintenance planning, including preventative maintenance compliance and backlog management. Monitor, analyse and report weekly and monthly KPIs, ensuring results are shared with key stakeholders. Coordinate planned maintenance outages and engineer

auditingrecruitment
View job →

Here’s a summary of the role: As a Staff Ruby on Rails Engineer at Diligent, you will set technical direction for secure, scalable, and high-performing SaaS applications and services that power our governance platform. This role is ideal for you if you thrive on solving the hardest technical problems, shaping architecture across multiple teams, and driving how the organization builds software — including how we responsibly adopt AI into our engineering practices and products. You will own critical systems end-to-end, partner closely with Product, Security, DevOps, and other engineering leaders to shape technical roadmaps, and mentor engineers across the department. A key part of this role is helping define our AI strategy: embedding AI responsibly into our systems and workflows, advising on where AI tools and capabilities can meaningfully improve delivery, and raising the bar on how the team uses them safely and effectively. Here’s a breakdown of what you’ll do (not all of it, just the important stuff): Champion the design, delivery, and evolution of secure, scalable Ruby on Rails applications and services, driving architecture decisions across multiple teams and codebases, with responsibility for scalability, reliability, and the underlying infrastructure required to run them effectively. Set technical direction for major projects and platform initiatives, from solution design and prototyping through implementation and production ownership. Own and evolve the technical roadmap by identifying, shaping, and driving new technical initiatives and investments. Develop and evolve web applications and services with a strong focus on scalability, maintainability, reliability, and long-term platform health. Identify systemic pain points across services and propose pragmatic architectural improvements, including decomposition of monoliths and evolution toward service-oriented or microservices patterns where it adds valu

sqlawsgit
View job →

Job Details: Job Description: The world is transforming - and so is Intel. Intel is a company of bold and curious inventors and problem solvers who create some of the most astounding technology advancements and experiences in the world. With a legacy of relentless innovation and a commitment to bring smart, connected devices to every person on Earth, our diverse and brilliant teams are continually searching for tomorrow's technology and revel in the challenge that changing the world for the better brings. We work every single day to design and manufacture silicon products that empower people's digital lives. Come join us and do something wonderful. Join us as a Manufacturing Failure Analysis Engineer, where you'll play a pivotal role in ensuring the quality and reliability of Intel's next-generation silicon, product, package, platforms, and board process technologies. You will identify failures, conduct root cause analyses, and develop innovative methodologies to enhance performance and yield. By leveraging your expertise, you'll contribute to driving manufacturing innovation and supporting the delivery of world-class products. The ideal candidate brings hands-on experience in probing, strong understanding of device behavior, and the ability to independently execute complex debug workflows. Business group You will be part of Intel's manufacturing organization, a cornerstone of our operations that focuses on advancing semiconductor process technologies and ensuring product excellence. This team is dedicated to delivering high-quality solutions that support Intel's broader mission to create groundbreaking technologies that power the future. Key Responsibilities: • Operate general failure analysis tools in the labs, such as TEM, SEM, optical scopes. • Perform electrical probing an

SEMrecruitment
View job →

Job Details: Job Description: As a Packaging Module Development Engineer, you will play a pivotal role in the development and optimization of Intel's assembly media and collaterals. Your daily work will involve designing innovative solutions to improve media and collateral reliability, manufacturability, and cost efficiency. Your contributions will directly support Intel's cutting-edge assembly packaging technology roadmap, driving advancements in semiconductor manufacturing processes. Business Group You will be part of Intel Corporation's Media Development organization, a team dedicated to advancing packaging technologies and manufacturing processes. This group operates at the forefront of innovation, enabling high-performance products and driving Intel's leadership in the semiconductor industry. By joining this collaborative environment, you will contribute to Intel's broader goals of delivering transformative technology solutions. Key Responsibilities Develop and optimize media and collaterals to meet quality, reliability, cost, yield, and productivity targets. Design and implement new media and collateral solutions, leveraging statistical methods like Design of Experiments (DOE) and Statistical Process Control (SPC). Conduct evaluations of media and collaterals under simulated field conditions, including tests for heat, humidity, temperature cycling, and dynamic forces. Lead initiatives to identify and mitigate media and collateral quality and reliability risks, utilizing innovative tools and methods. Provide technical consultation on media and collateral challenges and deliver solutions that align with operational and technology milestones. Document technical improvements and innovations through research papers and presentations. <p styl

AutoCADrecruitment
View job →

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and establish teams with the most thoughtful people in the world. NVIDIA GH200 superchip provides performance and productivity required for strong scaling for HPC and generative AI workload. Scale out is inherent to design of this massive superchip. We are looking for expert engineers to come and help design rack level solutions for next generation scaling AI supercomputing platforms. We are looking for a strong technical architect to own end to end manageability architecture for these products in data centers. You will work with various component leads internally and externally, drive customer use cases, align architecture with customer requirements and release best products to market. Join us at the forefront of technological advancement. What you’ll be doing: Drive server management for large clusters and data centers deploying GPUs and Grace solution from Nvidia. Work with data center architects and cloud customers to narrow down on requirements for implementation to ensure speed of light product development. Work with internal teams to make sure requirements are designed and implemented in right way with each firmware and software module Collaborate with other leads to design & build data center health management workflow. Drive reliability and optimization in firmware architecture from a data center view point. Work closely with cluster bring up team and resolve is

pythongitai
View job →
M
11 days ago

Our Purpose Mastercard powers economies and empowers people in 200&#43; countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Software Engineer Job Overview: Responsible for the analysis, design, development, testing, and delivery of secure, scalable software solutions. Define requirements for new applications and customization adhering to Mastercard standards, processes, and best practices. Develop, customize, and test applications to integrate to Mastercard specifications. Provide leadership, mentoring, and technical training to other team members. Major Accountabilities • Plan, design, architect, and develop secure, scalable, and maintainable technical solutions and alternatives to meet business requirements in adherence with Mastercard standards, processes, and best practices • Lead day-to-day system development and maintenance activities of the team to meet service level agreements (SLAs) and create solutions with a high level of innovation, cost effectiveness, quality, reliability, and faster time to market. • Accountable for the full systems development life cycle including creating high-quality requirements documents, use cases, designs, and other technical artifacts including but not limited to detailed test strategies, performance benchmarking, release rollout and deployment plans, contingency/back-out plans, feasibility studies, cost and time analysis, and detailed estimates. • Design, develop, test, dep

javadockergit
View job →
M
Mural
📍 Canada Remote• Full-time• Remote
11 days ago

ABOUT THE TEAM At Mural, we're changing how teams collaborate, think, and make decisions together. The Growth & Engagement team helps customers discover value through intuitive product experiences, rapid experimentation, and emerging AI capabilities that make collaboration more effective. We own key parts of the customer journey - from acquisition and onboarding to activation, engagement, retention, and monetization - and use data, experimentation, and customer insights to continuously improve those experiences. We're a cross-functional team of engineers, product managers, designers, and data partners who move quickly, measure outcomes, and learn from every experiment. We value curiosity, collaboration, and thoughtful technical craftsmanship, building systems that enable rapid iteration while maintaining a high bar for quality, reliability, and scalability. YOUR MISSION As a Senior Software Engineer on the Growth & Engagement team, you'll build the products, platforms, and experimentation capabilities that help customers discover value faster and achieve more with our product. You'll partner closely with Product, Design, Data, and other Engineering teams to identify opportunities, validate ideas through experimentation, and deliver experiences that improve customer outcomes and drive sustainable business growth. Beyond shipping features, you'll help shape the technical direction of the team, mentor other engineers, and establish engineering practices that enable us to experiment confidently and move quickly. WHAT YOU'LL DO Design, build, and maintain customer-facing features that improve acquisition, activation, engagement, retention, and monetization. Collaborate closely with Product, Design, Data, Marketing, and Customer Success to identify growth opportunities and translate customer feedback, product analytics, and experimentation to guide product decisions. Design, implement, and analyze experiments using feature flags, A/B testing, and analytics to vali

REMOTEtypescriptreactnode.js
View job →
HI
HP IQ
📍 San Francisco• $149.9K – $270K/yr
11 days ago

Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role As a Senior Platform Engineer at HP IQ, you will help build and evolve the infrastructure, tooling, and shared platform capabilities that enable our engineering teams to develop and operate reliable, secure, and scalable services across cloud and edge environments . You will work closely with application, services, AI/ML, and security teams to improve developer velocity, production readiness, reliability, and operational efficiency across a heterogeneous infrastructure footprint. What You Might Do Design, build, and maintain shared infrastructure and platform capabilities across cloud and edge environments. Build automation and self-service tooling that improves engineering velocity and operational consistency. Develop and maintain Infrastructure-as-Code, deployment workflows, and environment provisioning. Partner with engineering teams on production readiness, including reliability, security, observability, scalability, and recovery. Improve monitoring, alerting, incident response, and operational tooling across distributed environments. Automate repetitive operational t

pythonkubernetesai
View job →
U
11 days ago

About Ubiquiti At Ubiquiti Inc., we create technology platforms for Businesses, Smart Homes, and Internet Service Providers, driven by our goal to connect everyone, everywhere. To date, Ubiquiti has shipped over 100 million devices worldwide, from ISP networking products to next generation of IT solutions. Our growth is made possible by the dedicated team of hundreds behind the scenes. From software developers and product managers to designers and strategists, Team UI is driven to achieve our common goal: Rethinking IT. At Ubiquiti, you’ll heighten your potential and broaden your horizons - all while shaping the future of connectivity. Responsibilities Lead hardware circuit design, schematic capture, and layout review for next-generation, high-density, and high-throughput networking platforms. Evaluate and integrate advanced switching architectures, management subsystems, and cutting-edge high-speed interconnect technologies. Drive hardware architecture design, system bring-up, high-speed signal integrity (SI) validation, and root-cause failure analysis. Partner with mechanical, thermal, and power engineering teams to address challenges related to high power density, thermal dissipation, and system-level reliability. Own BOM structure and support factory deployment to ensure seamless transition of high-layer-count PCBAs from NPI to mass production. Collaborate with cross-functional software, firmware, QA, and compliance teams throughout the entire product lifecycle. Q ualifications Bachelor’s degree or above in Electrical Engineering or a related discipline. 5+ years of hands-on experience in high-complexity system-level hardware design, ideally focused on enterprise-grade networking or high-performance infrastructure equipment. Deep technical understanding of high-speed Ethernet design, high-speed differential signals (advanced SerDes, PCIe, multi-gigabit/ultra-high-speed interfaces), and high-density PCB design rules. Practical experience with complex power d

G
11 days ago

About Graphcore Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors and data center hardware that provide the specialized processing power needed to advance AI while improving the efficiency required for broad adoption. As part of SoftBank Group, Graphcore belongs to a family of companies developing transformative technologies. Our AI Engineering Campus in Austin plays an important role in building the hardware platforms that support the next generation of AI systems. The Opportunity We are looking for a recent graduate or early-career engineer to join the Hardware Platform Development team as a Graduate Systems Engineer. You will contribute to the design, integration, validation, and performance analysis of advanced AI compute platforms. You will work with experienced hardware, firmware, software, mechanical, thermal, and systems engineers throughout the development lifecycle. The role combines subsystem engineering, hands-on laboratory work, test automation, data analysis, troubleshooting, and clear technical documentation. Start: September, 2027 Location: Austin, Texas, USA What You Will Do Contribute to the design, integration, and testing of CPU and high-speed input and output subsystems for advanced compute platforms. Take ownership of defined engineering tasks from requirements and test planning through execution, analysis, and technical review. Develop system-level validation plans, procedures, scripts, and tools that improve test coverage, repeatability, data collection, and analysis. Evaluate platform performance, power, signal behavior, reliability, and interoperability using laboratory measurements and system data. Investigate emerging input and output technologies, including PCIe 6.0 and 800G Ethernet, and assess their use in advanced computing systems. Support platform power, cooling, and energy-efficiency investigations, including liquid-cooling systems for high-performance processor

pythonartificial intelligenceai
View job →
🔔

Get new reliability engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More reliability engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.

Top cities for Reliability Engineer

City links are canonicalized and require at least 20 current jobs.

Countries hiring Reliability Engineer

Country links use the same curated canonical inventory as Jobiba sitemaps.