Role: Application Reliability Engineer Location: Gurgaon Who we are Graviton Research Capital is a privately funded quantitative trading firm striving for excellence in financial markets research. We trade across a multitude of asset classes and trading venues using a diverse range of concepts, from time series analysis and stochastic models to machine learning and statistical inference. We analyse terabytes of data to identify pricing anomalies and drive innovation in financial markets. Key Responsibilities and Deliverables The ideal candidate will possess a strong background in technical support, with a passion for problem-solving and a commitment to excellence. As an Application Reliability Engineer, you will be responsible for: Monitor production services and respond quickly to alerts, incidents, and outages to ensure smooth operation and minimal downtime. Monitor trading systems and infrastructure., Triage issues across trading support services, databases, and infra; escalate and coordinate with the right owners, and drive root-cause analysis and ensure fixes are implemented for long-term stability. Serve as the first line of defense for trading operations. Proactively identify, address recurring issues, and build automation to reduce manual intervention. Improve observability by enhancing monitoring, logging, and alerting systems. Develop and maintain operational runbooks and SLO/SLA metrics. Eligibility and Required Skills Possess a degree in a highly analytical field, such as Engineering, or Computer Science 2-5 years of experience in Python, Shell/Bash scripting. Experience with Linux and shell/bash online tools. Hands-on experience with databases (SQL, NoSQL) Strong problem-solving and analytical skills Excellent communication skills Ability to remain calm and analytical under production pressure Good to have: Familiarity with monitoring/alerting stacks (Prometheus, Grafana, ELK, etc.) Familiarity with distributed messaging (Kafka) and caching systems (Red
Jobs in India
Application Reliability Engineer in India
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current application reliability engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
Position Overview We are looking for a Software Engineer II to build and deliver scalable software solutions across our products. You will work on modern web applications and cloud-based services using Node.js, React, TypeScript, AWS, PostgreSQL, MSSQL, and Docker, while contributing to AI-enabled features and integrations. You will collaborate closely with other engineers, product managers, and cross-functional teams to develop reliable, maintainable, and production-ready solutions. This role provides an opportunity to work with modern AI technologies including Python, AWS Bedrock, MCP, RAG, and agentic AI workflows while developing strong expertise in cloud-native software engineering. What You'll Do Develop and maintain scalable backend services and APIs using Node.js, TypeScript, and JavaScript. Build responsive and maintainable frontend applications using React. Design and implement integrations with AWS services and contribute to cloud-native application development. Develop and maintain applications using PostgreSQL and MSSQL, including writing efficient queries and working with database schemas. Build, test, and deploy applications using Docker and modern CI/CD practices. Contribute to AI-enabled product features using Python, AWS Bedrock, RAG, MCP, and AI integration patterns. Work with the team to integrate LLM capabilities, APIs, tools, and data sources into production applications. Write clean, maintainable, and well-tested code following established engineering practices. Participate in code reviews, technical discussions, debugging, and production issue resolution. Develop unit and integration tests and contribute to improving application quality and reliability. Monitor application performance and troubleshoot issues across development and production environments. Collaborate with senior engineers and architects to implement technical solutions aligned with product and engineering requirements. Stay current with emerging technologies, particularly in
About the job Senior Backend Engineer |100% Remote We are searching for a seasoned Sr Backend Engineer who will be responsible for developing and maintaining our backend systems, ensuring their efficiency, scalability, and reliability. The ideal candidate has a strong background in PHP and Java, with optional experience in Golang and Node.js. You should be well-versed in working with databases such as MongoDB, Redis, and Postgres, and have a solid understanding of cloud technologies, specifically AWS. Responsibilities Design, develop, and maintain backend systems and APIs to support our application's functionality. Collaborate with cross-functional teams, including front-end developers, product managers, and designers, to deliver high-quality solutions. Write clean, scalable, and well-documented code that adheres to industry best practices and coding standards. Perform code reviews and provide constructive feedback to peers to ensure code quality and consistency. Optimize and improve the performance of existing backend systems. Troubleshoot and debug production issues, providing timely resolutions. Stay up-to-date with emerging technologies and industry trends, identifying opportunities for innovation and improvement. Collaborate with DevOps teams to ensure smooth deployment and operation of backend services in the AWS cloud environment. Requirements Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent work experience). 5-10 years of professional experience as a Backend Engineer. Strong proficiency in Java/Golang is a must. Experience with Node js is a plus. In-depth knowledge of database technologies, including MongoDB, Redis, and Postgres. Solid understanding of cloud computing platforms, particularly AWS. Familiarity with containerization technologies such as Docker and orchestration tools like Kubernetes. Proficiency in writing efficient and optimized SQL queries. Experience with version control systems, such as Git. Excellent pr
Role Purpose: The Machine Learning Engineer IV will play a critical role in advancing Jumio's Fraud team's mission to develop and enhance state-of-the-art solutions for fraud detection for ID verification purposes. This role is essential for ensuring the highest standards of security and user verification through the application of advanced machine learning and deep learning techniques, ultimately contributing to Jumio's leadership in the online identity verification, eKYC, and AML solutions market Role Value: As a Machine Learning Engineer IV at Jumio, you will have the opportunity to significantly impact the security and user experience of our ID verification solutions. Your expertise in deep learning and computer vision will drive the development of innovative algorithms that keep Jumio at the forefront of the industry. By deploying and maintaining these models in production, you will ensure the robustness and reliability of our solutions, supporting our clients across diverse industries such as Financial Services, Travel, Sharing Economy, Fintech, and Gaming. Your contributions will be pivotal in maintaining Jumio's reputation as the leading provider of online identity verification solutions, helping to meet the growing demand for secure and seamless user verification globally. Example Responsibilities . Develop, maintain, and own key fraudulent CV models of Jumio, which shapes the whole fraud product offering of Jumio. Design and implement machine learning, deep learning, classical CV focused on fraud detection. Research to support the deployment of the advanced algorithms. Deploy models as AWS SageMaker endpoints or directly onto devices. Stay updated with the latest advancements in machine learning, deep learning, and computer vision by engaging with academic papers and attending industry conferences. Work collaboratively with other engineers and product managers in an Agile development environment. Experience and Qualifications Bach
You’ll shape the future of a business‑critical platform as the technical lead across both product engineering and cloud infrastructure. You’ll modernize a mature .NET application running on AWS today, while steering its evolution toward a cloud‑native, React/Node.js, AI‑enabled architecture. If you enjoy owning architecture end‑to‑end, from backend and frontend through CI/CD, DevOps, and AWS infrastructure, this role gives you real influence at Staff Engineer level and the opportunity to set engineering standards that others follow. You’ll spend your time leading complex .NET and React features, designing scalable AWS infrastructure with Infrastructure as Code, and building automation that makes releases fast, safe, and repeatable. You’ll work on performance, reliability, and modernization in equal measure—fixing what’s slowing the platform down today and designing what it will look like in the next generation. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead the architecture and development of enterprise .NET services and APIs that power a business‑critical platform. Design and operate AWS infrastructure (using AWS CDK in TypeScript) to support secure, scalable, multi‑environment deployments. Build and optimize CI/CD pipelines (AWS CodePipeline, CodeBuild, Windows build agents) to make shipping .NET and React changes fast and reliable. Drive modernization initiatives across the stack, including clean architecture, refactoring legacy components, and reducing technical debt. Design and tune PostgreSQL and MSSQL database solutions for performance, scalability, and reliability. Mentor engineers and influence engineering practices across teams, raising the bar on cloud, DevOps, and software design. These are the essentials you’ll need to get an interview Significant experience (typically 8+ years) delivering and operating scalable enterprise software, owning both application code and cloud infrastructure. Deep hands‑on expertise with C
Role Overview Build reliable software services that power products, platforms, and business decisions. As a Senior Software Developer, you’ll design and deliver scalable applications, backend services, and integrations that perform well in production and evolve with changing business needs. You’ll apply strong software engineering practices across APIs, data-intensive applications, cloud services, AI-enabled solutions, and deployment pipelines. You’ll help shape technical solutions, improve system reliability, and contribute to a high-quality engineering culture. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Design and develop scalable backend services and applications using Python or TypeScript. Lead the development of APIs, integrations, reusable software components, and AI-enabled features. Build reliable solutions for data ingestion, manipulation, service-to-service communication, and intelligent automation. Apply AI technologies and modern software engineering practices to improve product capabilities, developer productivity, and operational efficiency. Make sound technical decisions around architecture, performance, security, scalability, and maintainability. Deploy and operate applications using AWS services and CI/CD practices while improving testing, monitoring, documentation, and delivery standards. These are the essentials you’ll need to get an interview 5+ years of professional experience developing and delivering production software. Strong hands-on experience with Python; TypeScript or similar languages is also valuable. Proven experience building backend services, APIs, integrations, and service-oriented applications. Experience applying AI technologies, such as generative AI, machine learning services, intelligent automation, or AI-enabled application features. Strong understanding of software design principles, testing, debugging, performance optimization, and secure development. Experience working with cloud platfor
About the Role At Jumio, the Software Development Engineer IV - QA (SDE-IV, QA) is a senior technical role focused on ensuring the quality, performance, and reliability of highly scalable web portals and distributed backend systems. Our platform spans multiple Java Spring Boot microservices and customer-facing web portals , deployed across AWS ECS, EKS, and Lambda , and integrated through event-driven messaging using SNS/SQS . In this role you will design and drive the test automation strategy for both UI (Playwright) and API/service layers, set the quality bar for the team, and act as a force multiplier — mentoring other engineers and embedding quality earlier in the development lifecycle. You will work closely with development, product, and DevOps teams to ensure our products meet the highest standards of quality, scalability, and security. This is a hands-on senior IC role: you will write code, but you will also influence architecture, own cross-service test strategy, and make build-vs-buy decisions for testing tooling. T-Shaped Engineering Expectation As part of Jumio's engineering culture, you will adopt a T-shaped engineering approach. Beyond deep expertise in test automation and quality engineering, you will contribute across the development lifecycle — understanding software architecture, participating in design and API-contract discussions, reviewing application code, and ensuring our distributed systems are testable, observable, and resilient by design. Role Value This role is critical to ensuring the reliability, scalability, and security of Jumio's products. By architecting and maintaining automated testing frameworks across web, API, and event-driven layers, you will enable faster, higher-confidence releases and reduce production risk in a complex microservices environment. What You'll Do Test Architecture & Strategy Define and own the end-to-end automated test strategy across web portals and backend microservices, balancing UI, API, contract, integ
DeepIntent is the leading healthcare marketing platform, purpose-built to help marketers plan, activate, and optimize data-driven campaigns with speed and precision. Trusted by the world’s top healthcare brands and their agencies, DeepIntent uniquely unites media, identity, and real-world clinical data to power privacy-safe, omnichannel marketing across every screen. Backed by patented technology and proven outcomes, DeepIntent’s platform delivers measurable audience quality and script lift at scale. Learn more at www.deepintent.com . What You’ll Do: We are looking for a Senior Software Engineer – Platform Operations based in Pune, India, who will play a key role in ensuring the reliability, performance, and operational excellence of DeepIntent's platform and data ecosystem. This role requires a strong engineering mindset with the ability to troubleshoot complex technical issues, understand distributed data architectures, and collaborate across Engineering, Product, Analytics, and Customer-facing teams to deliver timely and effective solutions. As part of the Operations organization, you will work closely with Engineering to support production systems, improve operational processes, and drive platform stability. The ideal candidate is a self-motivated problem solver who is passionate about learning new technologies, improving system reliability, and delivering exceptional customer outcomes through engineering excellence. Serve as the engineering interface between Customer-facing teams, Analytics, Product, and Engineering organizations. Partner with Platform Support, Client Success, and other customer-facing teams to investigate and resolve complex platform-related issues. Analyze application, API, and data pipeline issues to identify root causes and drive timely resolution. Develop and standardize operational tools, and interfaces to support analytical and operational use cases. Monitor data pipeline executions, investigate failures, and implement corrective and pre
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO As Database Support Engineer, you’ll support various critical database platforms across Development, QA, UAT, and Production environments. The role partners closely with application teams, application support, and database engineers and operates within a Follow‑the‑Sun model to ensure availability, performance, and reliability of database services. Key responsibilities include: • Provide operational support for enterprise database platforms in both on-prem private cloud and public cloud • Monitor database health, capacity, performance, and availability, and respond to alerts, diagnose issues, and perform timely remediation • Perform routine maintenance activities (patching, upgrades, housekeeping etc) • Troubleshoot database‑related incidents and collaborate on root cause analysis • Work closely with application owners, application support teams, and DB Engineers • Provide guidance on database best practices and operational standards • Participate in cross‑team problem resolution and continuous improvement initiatives • Contribute to design, implementation and testing of automation and self service capabilities of DB platforms • Drive continuous improvement, identifying opportunities to reduce toil and increase platform efficiency. • Participate in a Follow‑the‑Sun operating model, including shift‑based coverage and handoffs WHAT’S REQUIRED • Bachelor’s degr
Staff Software Engineer Bengaluru, Karnataka, India Opportunity Get Well is seeking a visionary and technically adept Staff Software Engineer to architect, design, develop, and optimize our cloud-native healthcare platform while driving the adoption of AI-First and AI-Augmented Software Engineering practices. This role is pivotal in shaping the future of software development at Get Well by combining deep technical expertise with modern AI-assisted engineering workflows. As we evolve toward an AI-First engineering organization, this leader will champion the use of Generative AI, AI development assistants, and Agentic AI to improve developer productivity, software quality, and engineering velocity. The ideal candidate brings deep expertise in software architecture, cloud-native application development, AI-enabled engineering, and distributed systems. This role provides technical leadership across multiple engineering teams, ensuring high standards for architecture, code quality, reliability, security, and AI adoption. This is a hands-on leadership role where strategic thinking meets deep engineering execution within a complex healthcare environment. This position reports to the Director, Product Development and requires close collaboration with software engineers, AI engineers, product managers, DevOps, QA, and compliance specialists. Key Responsibilities Technical Leadership Define and drive the architecture of scalable, distributed healthcare platforms. Champion AI-First Software Engineering practices across the development lifecycle. Lead the adoption of AI-Augmented Development , including spec-driven development, AI-assisted coding, code reviews, testing, and documentation. Establish engineering standards, Cursor/AI coding guidelines, reusable patterns, and governance for responsible AI usage. Provide hands-on leadership in architecture, coding, design reviews, debugging, and performance optimization. Mentor enginee
Job Title: Senior QA Engineer - Performance Testing Paytm is India's leading mobile payments and financial services distribution company. A pioneer of the mobile QR payments revolution in India, Paytm builds technologies that empower small businesses with payments and commerce solutions. Paytm’s mission is to serve half a billion Indians and bring them into the mainstream economy through the power of technology. About the Role: We are seeking a skilled Performance Test Engineer to design, execute, and analyze performance tests to ensure application scalability, stability, and responsiveness under varying load conditions. The ideal candidate will have hands-on experience with industry-standard performance testing tools and a strong understanding of system architecture, monitoring, and troubleshooting. Expectations/ Requirements Develop comprehensive performance test strategies and plans aligned with system requirements, project timelines, and business goals. Understand application architecture and identify critical business transactions for performance validation. Design realistic workload models to simulate real-world usage scenarios. Create, maintain, and execute performance test scripts using tools such as JMeter, LoadRunner, Gatling, or similar. Conduct baseline, load, stress, and scalability testing to evaluate system behavior under different conditions. Monitor system performance using tools like Influx DB, Grafana, JVM monitoring tools, and MAT (Memory Analyzer Tool). Analyze test results to identify performance bottlenecks and system limitations. Collaborate with development and infrastructure teams to troubleshoot and resolve performance issues. Assess system scalability and recommend optimizations to improve performance and reliability. Generate detailed performance test reports, including metrics, findings, and actionable recommendations. Work with stakeholders to gather and validate Non-Functional Requirements (NFRs), SLAs, and KPIs. Perform API an
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Staff Software Engineer to lead technical direction across a major feature area or system domain at OneTrust. Staff Engineers here own outcomes, not just designs; they decide how ambiguous, cross-cutting problems get solved when no existing playbook applies, and their judgment carries weight across teams they don't formally manage. Your Mission Technical Leadership & Architecture Lead architecture and design for systems with significant scope and blast radius, ensuring decisions hold up under real growth, compliance, and reliability constraints; not just initial requirements. Paying attention to application performance Exercise judgment on where AI-assisted tooling accelerates delivery and where deeper human design thinking is required Cross-Team Collaboration Partner with Product, UX, and other engineering teams early, shaping problems before solutions are locked in. Build working relationships and technical credibility beyond your immediate team. Quality & Standards Set engineering practices for code revie
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge We're looking for a Staff Software Engineer to lead technical direction across a major feature area or system domain at OneTrust. Staff Engineers here own outcomes, not just designs; they decide how ambiguous, cross-cutting problems get solved when no existing playbook applies, and their judgment carries weight across teams they don't formally manage. Your Mission Technical Leadership & Architecture Lead architecture and design for systems with significant scope and blast radius, ensuring decisions hold up under real growth, compliance, and reliability constraints; not just initial requirements. Paying attention to application performance Exercise judgment on where AI-assisted tooling accelerates delivery and where deeper human design thinking is required Cross-Team Collaboration Partner with Product, UX, and other engineering teams early, shaping problems before solutions are locked in. Build working relationships and technical credibility beyond your immediate team. Quality & Standards Set engineering practices for code revie
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Staff SET Opportunity We are looking for an experienced Staff Software Engineer in Test to join our Identity Management Engineering (IDM) team serving the Privileged Access Team (PAM). This team is passionate about delivering large-scale, mission-critical software in a fast-paced Agile environment. In this role you'll be working with a team of highly-skilled and talented engineers, responsible for delivering sophisticated backend solutions that help Okta reliably operate at large scale and be highly available. As part of the team, you’ll be ensuring projects are completed with the highest quality and reliability using automation at every level for fast, robust and secure releases. What you’ll be doing Review requirements and design specs to develop relative test plans and test cases Automate API tests, end-to-end tests, reliability/scale tests Work with engineering management to scope and plan engineering efforts Communicate and document QE plans for scrum teams to review Review application code, identify bug and other areas of weakness, architect tools for future coverage Automate all critical features to maintain zero-debt cadence Release features with solid quality Respond to production issues/alerts and customer issues during on-call rotation Be a strong customer advocate with a strong quality DNA What you’ll bring to the role 5+ years of QE experience preferably in an enterprise SaaS company 3+ years experience in quality engineering for ente
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are You are an experienced Infrastructure Engineer Engineer who owns backend infrastructure end to end. You design multi-tenant, microservices-based systems that other engineering teams build on, and you make deliberate architectural tradeoffs around consistency, latency, scale, and cost. You are comfortable going deep — service mesh internals, database internals, distributed-systems failure modes — and equally comfortable defining the reliability and security contracts an enterprise AI platform depends on. Responsibilities Design, own, and evolve scalable microservices architectures on Kubernetes across GCP, Azure, and AWS, including multi-tenant isolation (namespaces, network policies, per-tenant resource quotas and RBAC). Build core platform and data-plane components in Golang and Python — data ingestion, knowledge-base indexing and vector/graph search, application connectivity, workflow automation, and ML operations — against explicit latency and throughput SLOs. Own service-to-service communication: gRPC/protobuf API contracts, service mesh (Istio/Linkerd), load balancing, retries, timeouts, and circuit breaking. Make and document architectural tradeoffs — partitioning
Other cities to consider
More places hiring for this role
Get new application reliability engineer jobs in India by email
Daily job updates · Unsubscribe anytime