Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Exa team and lead the charge in redefining enterprise storage by unifying block, file, and object protocols across hybrid-cloud environments. You will combine deep technical expertise in distributed systems with hands-on people leadership to guide architectural decisions and mentor high-impact engineers. This is a unique opportunity to build new engineering teams from the ground up and drive industry-leading innovation alongside Product and Architecture partners. Your work will directly impact how customers consume, scale, and operate mission-critical storage infrastructure. WHAT YOU'LL DO Drive End-to-End System Architecture: Lead the architectural evolution and end-to-end delivery of high-performance, resilient storage systems from initial design concepts to high-quality shipped products. Optimize for Modern Data Workloads: Design and implement robust algorithms and concurrent platform solutions engineered for modern data pipelines, AI infrastructure, distributed computing, and enterprise analytics. Resolve Complex System Engineering Challenges: Apply deep root-cause analysis and system-level insight to solve multi-threaded, high-concurrency performance and reliability issues across Linux platform internals. Cross-Functional Ownership & Leadership: Collaborate across product management, validation, and support teams to align technical roadmaps, establish architectural standards, and drive enterprise
Jobs in India
Reliability Engineer in India
376 active opportunities · Updated October 2026
Showing
15 jobs
Explore current reliability engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Drive the quality strategy for our innovative enterprise storage platform, ensuring zero-downtime resilience across physical hardware and cloud environments like Cloud Block Store and CloudSnap. In this engineering leadership role, you will scale systems testing, feature interoperability, and test automation for mission-critical global applications. Partnering directly with cross-functional development, support, and escalation teams, you will champion a customer-first quality model. This position elevates overall product reliability while shaping how cutting-edge software resilience is delivered at scale. WHAT YOU'LL DO Define & Execute Quality Strategy: Own end-to-end system test designs with a focus on large-scale feature interoperability to guarantee zero-downtime performance across enterprise and cloud environments. Build High-Impact Automation & Tooling: Design and deploy automated test workflows and triage tooling to accelerate defect detection, drastically reducing execution friction across thousands of automated test suites. Simulate Real-World Customer Workflows: Replicate complex customer deployment architectures to validate real-world fault tolerance and overall resilience against failure domains. Drive Root-Cause Resolution: Partner directly with escalation and support engineering teams to analyze and resolve complex defects, utilizing customer feedback loops to eliminate quality gaps. Lead Agi
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. SHOULD YOU ACCEPT THIS CHALLENGE... In this role as an Engineering Manager, you will lead a team of engineers located in Bengaluru, India. You will focus on driving and shaping the direction of our observability software and enabling product engineers to deliver high-quality, reliable software to our customers. You will provide technical leadership and direction, mentor engineers on your team, and collaborate with product, engineering, and cross-functional stakeholders to deliver successful outcomes. The successful candidate must understand the dynamics of global R&D, possess deep knowledge of local culture, and have the ability to champion Pure values and leadership attributes. This role requires the ability to lead and influence multiple stakeholders across cross-functional teams and drive alignment across complex, distributed engineering environments. The team will help build and evolve observability capabilities that provide actionable insights into the health, performance, capacity, and reliability of Pure's products and infrastructure. WHAT YOU'LL NEED TO BRING TO THIS ROLE... 12+ years of combined experience as a software developer and manager 3+ years of technical management experience while staying hands-on 7+ years of hands-on software development experience Strong exposure to one or more of the following areas: distributed systems, systems programming, observability/telemetry, data platforms, or solving prob
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Drive the mission-critical quality strategy for the industry’s most innovative, high-performance storage array platform. In this pivotal engineering leadership role, you will scale systems testing, feature interoperability, and test automation to ensure zero-downtime resilience for global enterprise applications. Partnering directly with cross-functional development, support, and escalation engineering teams, you will champion a customer-first quality model across physical hardware and cloud-native environments (Cloud Block Store, CloudSnap). This position elevates product reliability and shapes how cutting-edge software resilience is delivered at scale. WHAT YOU'LL DO Define & Execute Quality Strategy: Ownership of end-to-end system test designs, focusing on feature interoperability at scale to guarantee zero-downtime performance across enterprise and cloud environments. Build High-Impact Automation & Tooling: Design and deploy automated test workflows and triage tooling to accelerate defect detection, drastically reducing execution friction across thousands of automated test suites. Real-World Customer Simulation: Replicate complex customer deployment architectures and enterprise application workflows to validate real-world resilience, fault tolerance, and resilience against failure domains. Root-Cause Resolution & Continuous Improvement: Partner directly with escalation and support teams to reproduc
We are fueled by a moral imperative to advance mankind, and it all begins with our people, our product, and our purpose. Passion isn’t something we turn on and off; it’s woven into everything we do. If you thrive in high-challenge environments, are inspired by exceptional teammates, and are driven to grow beyond what you thought possible, MX is where you belong. Come build the future with us. Join an award-winning company that isn’t just shaping the financial industry, but transforming it in ways that create meaningful, lasting impact for millions of people. Senior Cloud Platform Engineer At MX, we’re on a mission to empower the world to be financially strong. We power the financial experiences behind thousands of banks, credit unions, and fintechs, helping millions of people better understand, manage, and improve their financial lives. Our platform sits at the intersection of financial data, cloud-scale infrastructure, and trust, and it has to work, every time, at massive scale. We’re entering a pivotal phase of our technology evolution: moving from legacy on-prem infrastructure to a modern, cloud-native platform designed for resilience, security, and developer velocity. It’s a chance to build the next generation of MX’s platform thoughtfully, reliably, and with long-term impact in mind. If you enjoy working on complex systems, care deeply about uptime and reliability, and want your work to directly power the financial well-being of millions, this is where you’ll do the most meaningful work of your career. What You’ll Do Drive Cloud Migration: Drive the end-to-end migration of production workloads from on-premise data centers to GCP. Architect for Reliability: Design and implement production-grade Kubernetes environments and GCP architectures that prioritize 99.99+% availability. Operational Excellence: Improve incident triage and improved recovery timeby implementing cloud-aware diagnostics and automated recovery patterns. Empower Developers: Reduce friction in th
₹2.6Cr – ₹3.4Cr/yr
Are you looking for an opportunity to help solve one of today's biggest business challenges? AI is changing the pace of business, and organizations everywhere are struggling to help their workforce, partners, and customers keep up. At Litmos, we're building the Learning Acceleration Platform that helps organizations build human capability faster—and we're looking for people who are passionate about making a meaningful impact for customers while growing alongside a collaborative, people-first team. Litmos is the Learning Acceleration Platform that helps organizations build capability faster, adapt at the speed business changes, and scale learning to anyone, anywhere. Combining an intuitive platform, AI-powered capabilities, trusted content, expert services, and a broad ecosystem of integrations, Litmos helps organizations accelerate workforce productivity, improve customer adoption and retention, enable high-performing partners, and reduce organizational risk through continuous learning. Organizations such as Hewlett Packard Enterprise, Graco, Sabre, and Russell Mineral Equipment trust Litmos to accelerate learning across their workforce, partners, and customers. Today, more than 11K customers with 30 million learners across 150 countries and 37 languages use Litmos to build the capabilities their organizations need to succeed. Backed by Francisco Partners, one of the world's leading technology investment firms, we're investing in the future of learning—and the people who are building it. Learn more at www.litmos.com . We are looking for a Senior Full Stack Engineer to build and evolve our platform using .NET, React, and SQL Server-based systems. You will work across services, APIs, and front-end systems with a focus on scalability, reliability, and maintainability. You will operate within a globally distributed Agile team (Scrum and ShapeUp), where engineers own delivery, from design through deployment, while collaboratin
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: The Automation Engineer is responsible for designing, developing, and maintaining security automation solutions that enhance detection, response, workflow efficiency, and operational consistency across Operational Security. Working under the Automation Lead, this role builds high-quality SOAR playbooks, integrations, scripts, AI-assisted workflows, and orchestration pipelines to reduce manual workloads and support the Autonomic Security Operations (ASO) model. What you'll be doing: Core Responsibilities Automation Engineering & Development Develop SOAR playbooks, workflows, and automations for alert triage, enrichment, containment, and remediation. Build scalable, reusable automation components, scripts, and integrations. Implement high-quality scripting using Python, PowerShell, and REST APIs. Ensure appropriate version control, QA, testing, and documentation of automation artefacts. Maintain reliability of automations by monitoring performance, exceptions, and system behaviour. Platform Integration & Tooling Engineering Integrate SOAR with SIEM, EDR, TIP, cloud-native secur
About the Role At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale. If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk. Responsibilities Design and build fleet automation systems that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention. Build AI Infrastructure Agents that automate deployment, root-cause failures, incident triage, and autonomous remediation. Develop Fleet Intelligence platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers. Build software that maximizes GPU availability, utilization, performance, and reliability across thousands of accelerators. Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads. Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations. Continuously improve deployment velocity, reliability, and operational efficiency through automation. Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure. Requirements 3+ years building distributed systems, infrastructure platforms, or large-scale backend software. Strong software engineering skills in Python, Go, or Rust . Experience building platforms, automation systems, or developer infrastructure. Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies. Strong systems thinking with the ability to understand problems across hardw
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge Senior Managers of Engineering at OneTrust will make long-term strategic and technical contributions. These individuals set strategic goals for the team, hire engineers, and prioritize projects. You'll be involved technically, too. Developing new products, identifying requirements, and executing with excellence. Your Mission Drive strategic planning and execution while developing key technologies that will enhance OneTrust's long-term, proprietary strategic position. Create new concepts from initial design all the way to market release. You Are Experienced overseeing end-to end-development activities while monitoring reliability and performance of all internal systems and suggesting improvements when required. You will ensure compliance with security regulations while managing software development projects by setting requirements, goals, and timelines. Designing strategies for future development projects based on the company’s overall objectives and resource avail
Who We Are Addepar is a global data and AI platform empowering investment professionals to turn complex financial information into actionable intelligence. Addepar unifies portfolio, market and client data in a total portfolio view and delivers AI-powered insights within investment and client workflows. More than 1,400 firms in nearly 60 countries use Addepar to manage and advise on nearly $9 trillion in assets. Its open platform integrates with nearly 650 software, data and consulting partners to power end-to-end investment operations across firms of all sizes and complexity. Addepar supports clients worldwide with offices in New York City, Salt Lake City, London, Edinburgh, Pune, Dubai, Geneva, Singapore and São Paulo. The Role We are currently seeking a Staff Software Engineer, Infrastructure to join the AI Platform team that powers seamless insights and interaction through natural language and data intelligence across our AI products. As a Staff Software Engineer, you’ll architect, build, and operate the backend and platform systems that power AI Platform. You’ll work across service design, distributed systems, cloud infrastructure, event-driven processing, observability, CI/CD, and production reliability, helping shape the technical direction of a platform that supports scalable, client-facing AI experiences. This role requires a strong software engineering foundation combined with deep infrastructure and systems thinking. We are looking for an engineer who can write high-quality production code, make sound architectural tradeoffs, and own platform capabilities end-to-end — not someone focused only on scripting, cloud configuration, or infrastructure tooling in isolation. You will collaborate closely with frontend, product, and AI/ML engineers to deliver reliable, secure, and scalable systems that align with Addepar’s standards of performance, resilience, and trust. Applicants must have legal authorization to work in the country where this role is based o
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Staff Software Engineer to lead technical direction across a major feature area or system domain at OneTrust. Staff Engineers here own outcomes, not just designs; they decide how ambiguous, cross-cutting problems get solved when no existing playbook applies, and their judgment carries weight across teams they don't formally manage. Your Mission Technical Leadership & Architecture Lead architecture and design for systems with significant scope and blast radius, ensuring decisions hold up under real growth, compliance, and reliability constraints; not just initial requirements. Paying attention to application performance Exercise judgment on where AI-assisted tooling accelerates delivery and where deeper human design thinking is required Cross-Team Collaboration Partner with Product, UX, and other engineering teams early, shaping problems before solutions are locked in. Build working relationships and technical credibility beyond your immediate team. Quality & Standards Set engineering practices for code revie
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Staff Software Engineer to lead technical direction across a major feature area or system domain at OneTrust. Staff Engineers here own outcomes, not just designs; they decide how ambiguous, cross-cutting problems get solved when no existing playbook applies, and their judgment carries weight across teams they don't formally manage. Your Mission Technical Leadership & Architecture Lead architecture and design for systems with significant scope and blast radius, ensuring decisions hold up under real growth, compliance, and reliability constraints; not just initial requirements. Paying attention to application performance Exercise judgment on where AI-assisted tooling accelerates delivery and where deeper human design thinking is required Cross-Team Collaboration Partner with Product, UX, and other engineering teams early, shaping problems before solutions are locked in. Build working relationships and technical credibility beyond your immediate team. Quality & Standards Set engineering practices for code revie
Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before it touches production. That founding instinct still defines how we work. We're accurate and evidence-driven, relentless about clarity, and we'd rather be certain than comfortable, building a groundbreaking platform that transforms how teams run and secure networks across every major cloud and vendor environment. Global leaders like Goldman Sachs, PayPal, S&P Global, IBM, and Dell trust Forward, alongside fast-growing enterprises and government agencies, realizing an average of $14.2 million in annual benefits, according to IDC. Backed by top-tier investors, including A. Capital, Andreessen Horowitz, Goldman Sachs, MSD Partners, Omega Venture Partners, Section 32, and Threshold Ventures, and headquartered in Santa Clara, we're most proud of our team: curious people who'd rather build what doesn't exist than accept how things have always been done. Forward is currently seeking a Senior Backend Software Engineer to work as part of our Platforms team. You will play a critical role in designing, developing, and scaling the core backend services and infrastructure that support our SaaS and on-prem deployments. Your contributions will have a direct impact on the stability, performance, and scalability of our platform, helping to ensure an exceptional experience for our customers. Responsibilities: Platform development: Contribute to the design and development of storage systems, job scheduling systems, data ingestion frameworks, monitoring frameworks etc to ensure high system performance and availability. Feature development: Build and maintain backend frameworks that support essential platform features Scalability & Reliability: Develop scalable, high-performing
Forward is transforming how the world’s most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry’s first network digital twin — a mathematically precise model of the production network that gives IT teams unmatched visibility, verification, and agility across every major cloud and vendor environment. Our customers include global leaders such as Goldman Sachs, PayPal, S&P Global, IBM, and Dell, as well as fast-growing enterprises and government agencies. According to IDC, Forward customers realize an average of $14.2 million in annual benefits through improved efficiency and security. Backed by world-class investors including Andreessen Horowitz, Goldman Sachs, MSD Partners, and Threshold Ventures, Forward offers a people-centric, innovative culture where brilliant minds are shaping the future of network reliability, security, and AI-ready operations. Forward is currently seeking experienced Java developers to work as part of our Network team. Responsibilities Help bring the best ideas from the software development world into the networking industry. Contribute to our code base, systems and software architecture as a member of our engineering team. Help create and optimize network device models for different device vendors and protocols. Help create infrastructure needed to configure, collect and test network devices. Work with peers who are experts in Networking, Distributed Systems, Big Data and Search. Requirements 5+ years of work experience in software development 3+ years of work experience with Java BS in Computer Science or related degree Solid software engineering experience with large code bases Basic understanding of networking and TCP/IP. Strong verbal and written communication skills. Nice to haves Working knowledge of how switches, routers, firewalls or load balancers work. Experience working with networking protocols such as BGP/OSPF/IS-IS, IPv4/IPv6, MPLS, VLAN, VXLAN, etc. This position is a re
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO As Database Support Engineer, you’ll support various critical database platforms across Development, QA, UAT, and Production environments. The role partners closely with application teams, application support, and database engineers and operates within a Follow‑the‑Sun model to ensure availability, performance, and reliability of database services. Key responsibilities include: • Provide operational support for enterprise database platforms in both on-prem private cloud and public cloud • Monitor database health, capacity, performance, and availability, and respond to alerts, diagnose issues, and perform timely remediation • Perform routine maintenance activities (patching, upgrades, housekeeping etc) • Troubleshoot database‑related incidents and collaborate on root cause analysis • Work closely with application owners, application support teams, and DB Engineers • Provide guidance on database best practices and operational standards • Participate in cross‑team problem resolution and continuous improvement initiatives • Contribute to design, implementation and testing of automation and self service capabilities of DB platforms • Drive continuous improvement, identifying opportunities to reduce toil and increase platform efficiency. • Participate in a Follow‑the‑Sun operating model, including shift‑based coverage and handoffs WHAT’S REQUIRED • Bachelor’s degr
Other cities to consider
More places hiring for this role
Get new reliability engineer jobs in India by email
Daily job updates · Unsubscribe anytime