The job purpose of a Shift Engineer - ThermalTools is to oversee the efficient operation, maintenance, and optimization of thermal tools and equipment in manufacturing or industrial settings. They ensure production meets quality standards, adhere to safety regulations, lead teams, troubleshoot issues, and drive continuous improvement to enhance productivity and reliability. Source: Adani Group | Job ID: 51372
Jobs in India
Reliability Engineer in India
376 active opportunities · Updated October 2026
Showing
15 jobs
Explore current reliability engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
At Bolna, we’re building tools that change how businesses leverage voice AI. We’re looking for a Software Engineer to build reliable, scalable systems that power millions of production conversations across languages, industries, and telephony environments. This is a high-impact, high-ownership role where you’ll work on core platform problems across distributed systems, real-time communication, developer infrastructure, and customer-facing products. Our team includes IIT alumni with experience at Bain, Atlassian, Uber, Zomato, and LinkedIn, and is backed by leading investors. Responsibilities Build systems that operate at scale: Design and build backend services that support high-volume, real-time voice AI conversations with strong reliability, performance, and fault tolerance. Own features end to end: Take problems from product requirements and technical design through implementation, testing, deployment, monitoring, and iteration. Improve platform reliability: Build systems that are observable, resilient, and easy to debug. Identify bottlenecks, reduce failure rates, and improve system availability. Work on real-time infrastructure: Solve problems across telephony, streaming audio, webhooks, queues, scheduling, concurrency, and low-latency communication. Build for developers and customers: Improve APIs, SDKs, integrations, dashboards, and internal tools that make the Bolna platform easier to use and operate. Raise the engineering bar: Contribute to technical design reviews, code quality, testing standards, documentation, incident response, and engineering best practices. Required Skills Strong engineering fundamentals: Solid understanding of data structures, algorithms, databases, networking, operating systems, and distributed systems. Backend development experience: 2+ years of experience building and operating production backend systems using Python, Go, Java, Node.js, or a similar language. Production ownership: Experience shipping software to production and own
Position: Engineering Manager - Database Job Location: Noida Role Overview We are seeking a Database Engineering Manager (Individual Contributor) with deep expertise in MySQL and strong working knowledge of MongoDB, PostgreSQL, and Cassandra. This role combines hands-on database administration and optimization with strategic ownership of database reliability, automation, and cloud adoption. The candidate will lead by example—driving technical excellence, influencing best practices, and partnering cross-functionally with DevOps, SRE, and product engineering teams to deliver highly available, secure, and scalable database platforms. Key Responsibilities 1. End-to-End Ownership of MySQL databases in production & staging—availability, performance, and reliability. 2. Architect, manage, and support MongoDB, PostgreSQL, and Cassandra clusters for scale and resilience. 3. Define and enforce backup, recovery, HA, and DR strategies across all critical database platforms. 4. Drive database performance engineering—tuning queries, optimizing schemas, indexing, and partitioning for high-volume workloads. 5. Own replication, clustering, and failover architectures ensuring business continuity. 6. Champion automation & AI-driven operations—design self-healing scripts, predictive scaling, and proactive monitoring solutions. Collaborate with Cloud/DevOps teams on AWS database services (RDS, Aurora, DynamoDB, EC2, S3) to optimize cost, security, and performance. 7. Establish monitoring dashboards & alerting mechanisms for slow queries, replication lag, deadlocks, and capacity planning. Ensure compliance & security standards—encryption, auditing, and regulatory requirements. 8. Lead incident management & on-call rotations, ensuring rapid response and minimal MTTR. 9. Act as a strategic technical partner, contributing to database roadmaps, automation strategy, and adoption of AI-driven DBA practices. Required Skills & Experience 1. 6–10 years of p
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a Staff Software Engineer for our Frontier Security AI team. Snowflake's Frontier Security AI teams develop production-grade LLM applications, intelligent agents, AI infrastructure, and evaluation systems for enterprise customers — products that must meet a high bar for quality, security, reliability, and efficiency while operating over sensitive data at large scale. In this role, you will lead the design and development of our Agentic Harness and agent evaluation platform, working across product, infrastructure, applied AI, security, and modeling teams to take new capabilities from prototype to dependable customer value. AS A STAFF SOFTWARE ENGINEER AT SNOWFLAKE, YOU WILL: Architect and build the Agentic Harness that executes complex, multi-step AI workflows across models, tools, data, and services. Design stable interfaces for tool execution, context construction, state management, memory, permissions, retries, fallbacks, and human review. Own agent quality end to end by building evaluation harnesses, representative datasets, automated graders, experiment pipelines, and release gates. Convert ambiguous reports such as "the agent feels worse" into measurable failure modes, reproducible tests, and durable fixes. Analyze production agent trajectories to identif
About the Role We are looking for a Director of Engineering – AI & Full-Stack SaaS to lead the engineering strategy and execution for a highly scalable, enterprise-grade SaaS platform. This role is ideal for a strong engineering leader who combines deep software engineering expertise, enterprise SaaS experience, cloud-native architecture, full-stack product development, and practical AI/GenAI adoption . The successful candidate will lead multiple engineering teams, drive architectural and technical decisions, improve engineering velocity and quality, and help embed AI across the software development lifecycle and product engineering ecosystem. Key Responsibilities Lead and mentor multiple engineering teams responsible for building and delivering enterprise SaaS products. Define and execute the engineering strategy, technical roadmap, architecture and development standards . Drive development of highly scalable, secure, reliable and cloud-native SaaS applications. Provide technical leadership across backend, frontend, APIs, microservices and distributed systems . Drive adoption of AI/GenAI tools and capabilities across the engineering lifecycle, including development, testing, code quality, productivity and automation. Partner closely with Product, Architecture, DevOps, Security and other cross-functional teams to deliver business-critical capabilities. Establish engineering best practices around design, coding, testing, CI/CD, observability, performance and reliability . Own engineering delivery, quality, scalability and operational excellence across multiple product areas. Identify and resolve complex technical and architectural challenges. Build a high-performing engineering culture focused on innovation, ownership, collaboration and continuous improvement . Evaluate emerging technologies and identify opportunities to leverage AI and automation to improve engineering productivity and product capabilities. Drive technical modernization and evolution of existing
Job Details: Job Description: Intel's Design Quality and Reliability organization is seeking an AI Platform Engineer to architect and build an enterprise-grade AI platform for mission-critical engineering work. This platform will enable Intel engineers to analyze complex design, qualification, and reliability data; automate engineering workflows; access organizational knowledge; and make faster, evidence-based decisions throughout the product lifecycle. The successful candidate will combine strong software engineering fundamentals with expertise in AI-native and agentic development. They will be highly proficient with Agentic AI coding assistants and able to use these tools responsibly to accelerate architecture, implementation, testing, debugging, and documentation. This role requires close collaboration with Design, Quality and Reliability, Product Engineering, Manufacturing, IT, Information Security, and other Intel stakeholders. Responsibilities 1. Architect and develop Intel's reusable AI platform for Design Quality and Reliability. 2. Build AI agents and workflows for engineering data analysis, qualification planning, risk assessment, knowledge retrieval, reporting, and process automation. 3. Apply Agentic AI coding assistants to accelerate software development while maintaining rigorous engineering review and validation. 4. Integrate AI capabilities with Intel engineering databases, quality-management systems, internal APIs, spreadsheets, documentation repositories, and workflow tools. 5. Develop production-grade backend services, APIs, data pipelines, model gateways, and agent-orchestration components. 6. Establish shared platform capabilities for identity, access control, tool authorization, memory, observability, evaluation, and auditability. 7. Implement human approval, deterministic validation, and rollback controls for consequential engineering actions. 8.
Position Overview We are looking for a Software Engineer II to build and deliver scalable software solutions across our products. You will work on modern web applications and cloud-based services using Node.js, React, TypeScript, AWS, PostgreSQL, MSSQL, and Docker, while contributing to AI-enabled features and integrations. You will collaborate closely with other engineers, product managers, and cross-functional teams to develop reliable, maintainable, and production-ready solutions. This role provides an opportunity to work with modern AI technologies including Python, AWS Bedrock, MCP, RAG, and agentic AI workflows while developing strong expertise in cloud-native software engineering. What You'll Do Develop and maintain scalable backend services and APIs using Node.js, TypeScript, and JavaScript. Build responsive and maintainable frontend applications using React. Design and implement integrations with AWS services and contribute to cloud-native application development. Develop and maintain applications using PostgreSQL and MSSQL, including writing efficient queries and working with database schemas. Build, test, and deploy applications using Docker and modern CI/CD practices. Contribute to AI-enabled product features using Python, AWS Bedrock, RAG, MCP, and AI integration patterns. Work with the team to integrate LLM capabilities, APIs, tools, and data sources into production applications. Write clean, maintainable, and well-tested code following established engineering practices. Participate in code reviews, technical discussions, debugging, and production issue resolution. Develop unit and integration tests and contribute to improving application quality and reliability. Monitor application performance and troubleshoot issues across development and production environments. Collaborate with senior engineers and architects to implement technical solutions aligned with product and engineering requirements. Stay current with emerging technologies, particularly in
The Lead EMS/SCADA will be responsible for the design, development, integration, testing, and deployment of Energy Management Systems (EMS) and SCADA solutions for utility-scale Battery Energy Storage System (BESS) projects. The role will drive system architecture, control strategies, monitoring solutions, and communication interfaces to ensure optimal performance, reliability, and grid compliance. Source: Adani Group | Job ID: 58677
Who are we? FalconX is a pioneering team of operators, investors, and builders committed to revolutionizing institutional access to the crypto markets. Operating at the intersection of traditional finance and cutting-edge technology, FalconX addresses the industry's foremost challenges: Navigating the digital asset market can be complex and fragmented, with limited products and services that support trading strategies, structures, and liquidity found in conventional financial markets. As a comprehensive solution for all digital asset strategies from start to scale, FalconX operates as the connective tissue empowering clients with seamless navigation through the ever- evolving cryptocurrency landscape. Role Overview As an Engineering Manager for the Credit team, you will lead the core engine driving our credit systems, including the margin notification engine and loan booking system. Your primary focus will be on maximizing data accuracy, system reliability, and architectural integrity. You will lead a high-performing group of senior engineers, stream-lining technical processes, preventing over-engineering, and maintaining a high standard of delivery alongside key stakeholders. Role Split & Focus Areas Technical Leadership & Architecture (40%): Drive long-term system maintenance, architectural refactoring, and code quality. Ensure technical debt is addressed without blocking impactful business features. Process & Stakeholder Management (20–30%): Partner with product and business stakeholders to prioritize deliverables. Make decisive trade-offs and cut unnecessary overhead to maintain lean execution. People Management & Mentorship (20–30%): Manage, mentor, and guide senior technical talent across performance cycles, career growth, and talent acquisition. Key Responsibilities Oversee and maintain the reliability, precision, and efficiency of core credit applications (Margin Notification Engine, Loan Booking Systems). Drive architectural roadmap decision
About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the role: We are looking for a Senior Data Engineer to help build and evolve the data platform that powers critical business decisions and customer-facing experiences. You will own significant parts of our data architecture that is main powerhouse of DevRev Computer’s memory for accurate and efficient Answers. As a part of data team, you will design and operate scalable data systems, and work closely with Software Engineering, AI Agent teams, Data Science, and Product teams to turn complex data requirements into reliable, high-quality data products. This role is ideal for an experienced engineer who enjoys solving challenging problems involving large-scale data, distributed systems, database architecture, and performance optimization . You will have significant technical ownership and the opportunity to influence the direction of our agentic data platform while helping raise the engineering bar across the team. What you'll do: Own data architecture for large-scale, high-impact projects, making thoughtful tradeoffs across scalability, reliability, performance, maintainability, and operational cost. Design, build, and operate scalable data pipelines and data systems that reliably ingest, transform, store, and se
About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the role We are looking for a Quality Architect/Lead with hands-on experience building quality systems and has deep expertise in building and scaling test automation frameworks.The role requires an individual who applies systems thinking to solving complex problems. They should be able to understand the product from various perspectives and be able to effectively create testing programs that validate not just functionality but performance, reliability and user experience.DevRev is building a next generation AI native product that requires us to build novel test systems for the Agent AI platform. The role is mult-faceted and is going to continuously evolve with time. What you'll do Test Case Design and Documentation Actively use AI and intelligent agents to accelerate test generation, test maintenance, and coverage expansion. Leverage LLMs to convert requirements, user stories, and production incidents into high-quality automated test cases. Reduce reliance on manual test case creation by introducing AI-assisted automation workflows, with human review and ownership. Apply AI to optimize test selection, prioritization, and execution based on risk, code changes, and historical failures. Use AI to assist in identi
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Exa team and lead the charge in redefining enterprise storage by unifying block, file, and object protocols across hybrid-cloud environments. You will combine deep technical expertise in distributed systems with hands-on people leadership to guide architectural decisions and mentor high-impact engineers. This is a unique opportunity to build new engineering teams from the ground up and drive industry-leading innovation alongside Product and Architecture partners. Your work will directly impact how customers consume, scale, and operate mission-critical storage infrastructure. WHAT YOU'LL DO Establish and scale a new engineering organization focused on critical Exa platform services, ensuring a foundation of long-term success, technical excellence, and a high-performing culture. Own the successful delivery of complex, high-scale engineering features for the Exa platform, ensuring world-class security, reliability, and availability across multi-array and hybrid-cloud deployments. Define the technical vision and execution roadmap in close partnership with Product Management and Architecture, translating customer needs into a measurable business impact for Pure Storage. Drive a culture of operational rigor, owning the refinement of engineering processes around observability, CI/CD, and incident response, while actively mentoring the next generation of technical leads. WHAT YOU BRING Leadership and Scaling: Pro
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As the Engineering Manager for the FlashArray (FA) Foundation Quality Engineering team, you will build and lead a high-performing team of Quality Engineers and Software Test Developers in Bangalore. You will combine people leadership, technical judgment, and execution discipline to drive the quality, test automation, and release readiness of foundational FlashArray capabilities. Working at the intersection of software, hardware, firmware, and cloud infrastructure, you will partner closely with Development, Product Management, Release, and Support teams. Your focus will be translating complex roadmap requirements into modern test strategies, establishing strong shift-left CI/CD signals, and ensuring our products consistently deliver industry-leading reliability to our customers. WHAT YOU’LL DO Build & Lead a High-Performing Team: Hire, mentor, and coach a newly forming team of quality and automation engineers in Bangalore, fostering a culture of technical excellence, continuous learning, and shared ownership. Drive Quality Strategy & Shift-Left Automation: Define and execute the FA Foundation quality roadmap—including risk-based coverage, automated CI/CD pipeline integration, fault injection, and release-readiness criteria across software, firmware, and hardware. Partner Cross-Functionally for Delivery: Collaborate early in the design cycle with Development, Architecture, Product, and Release teams to impro
We are fueled by a moral imperative to advance mankind, and it all begins with our people, our product, and our purpose. Passion isn’t something we turn on and off; it’s woven into everything we do. If you thrive in high-challenge environments, are inspired by exceptional teammates, and are driven to grow beyond what you thought possible, MX is where you belong. Come build the future with us. Join an award-winning company that isn’t just shaping the financial industry, but transforming it in ways that create meaningful, lasting impact for millions of people. At MX, reliability is a product. Our infrastructure powers financial applications used by millions of people and processes billions of transactions for major financial institutions, and customers feel every second of downtime. We're building a new observability function that runs the way we run incident response: the system does the heavy lifting, and people handle judgment, customers, and the exceptions. As a Senior Observability Engineer, you build and operate an observability control plane. You scaffold baselines, score coverage, and turn every real incident into the detection the platform should have caught. This is a multiplier role: you raise the bar for every team through standards and automation instead of building each team's dashboards by hand. We call it the shepherd model. You shepherd Datadog and partner with our product engineering teams so they observe the right signals for their products. Service owners get real signal instead of noise, and leadership gets coverage and health as a program metric. This role shares the team pager. Observability and incident response run one on-call roster. You take shifts with the rest of the team and act as Incident Commander when an incident needs one. It is core to the role, not an afterthought. Engineering at MX runs hybrid infrastructure (AWS and bare metal) with services in Ruby, Go, and Java, messaging over NATS and RabbitMQ, and data on PostgreSQL an
Role Purpose: The Machine Learning Engineer IV will play a critical role in advancing Jumio's Fraud team's mission to develop and enhance state-of-the-art solutions for fraud detection for ID verification purposes. This role is essential for ensuring the highest standards of security and user verification through the application of advanced machine learning and deep learning techniques, ultimately contributing to Jumio's leadership in the online identity verification, eKYC, and AML solutions market Role Value: As a Machine Learning Engineer IV at Jumio, you will have the opportunity to significantly impact the security and user experience of our ID verification solutions. Your expertise in deep learning and computer vision will drive the development of innovative algorithms that keep Jumio at the forefront of the industry. By deploying and maintaining these models in production, you will ensure the robustness and reliability of our solutions, supporting our clients across diverse industries such as Financial Services, Travel, Sharing Economy, Fintech, and Gaming. Your contributions will be pivotal in maintaining Jumio's reputation as the leading provider of online identity verification solutions, helping to meet the growing demand for secure and seamless user verification globally. Example Responsibilities . Develop, maintain, and own key fraudulent CV models of Jumio, which shapes the whole fraud product offering of Jumio. Design and implement machine learning, deep learning, classical CV focused on fraud detection. Research to support the deployment of the advanced algorithms. Deploy models as AWS SageMaker endpoints or directly onto devices. Stay updated with the latest advancements in machine learning, deep learning, and computer vision by engaging with academic papers and attending industry conferences. Work collaboratively with other engineers and product managers in an Agile development environment. Experience and Qualifications Bach
Other cities to consider
More places hiring for this role
Get new reliability engineer jobs in India by email
Daily job updates · Unsubscribe anytime