We are fueled by a moral imperative to advance mankind, and it all begins with our people, our product, and our purpose. Passion isn’t something we turn on and off; it’s woven into everything we do. If you thrive in high-challenge environments, are inspired by exceptional teammates, and are driven to grow beyond what you thought possible, MX is where you belong. Come build the future with us. Join an award-winning company that isn’t just shaping the financial industry, but transforming it in ways that create meaningful, lasting impact for millions of people. At MX, reliability is a product. Our infrastructure powers financial applications used by millions of people and processes billions of transactions for major financial institutions, and customers feel every second of downtime. We're building a new observability function that runs the way we run incident response: the system does the heavy lifting, and people handle judgment, customers, and the exceptions. As a Senior Observability Engineer, you build and operate an observability control plane. You scaffold baselines, score coverage, and turn every real incident into the detection the platform should have caught. This is a multiplier role: you raise the bar for every team through standards and automation instead of building each team's dashboards by hand. We call it the shepherd model. You shepherd Datadog and partner with our product engineering teams so they observe the right signals for their products. Service owners get real signal instead of noise, and leadership gets coverage and health as a program metric. This role shares the team pager. Observability and incident response run one on-call roster. You take shifts with the rest of the team and act as Incident Commander when an incident needs one. It is core to the role, not an afterthought. Engineering at MX runs hybrid infrastructure (AWS and bare metal) with services in Ruby, Go, and Java, messaging over NATS and RabbitMQ, and data on PostgreSQL an
Jobs in India
Senior Observability Engineer in India
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current senior observability engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity Are you passionate about building foundational technology that fuels the world’s digital innovation? At New Relic, we provide the leading unified data platform for all things observability – helping engineers, developers, and operators make decisions using data at every stage of the software lifecycle. Our New Relic Control team is at the heart of that mission, creating groundbreaking capabilities that orchestrate, manage, and optimize observability agents and telemetry pipelines at scale. We’re looking for a Senior Product Manager to lead a new strategic initiative within our Pipeline Control team. In this role, you’ll be responsible for defining vision, strategy, and roadmap for an enterprise-grade solution that orchestrates observability pipelines to streamline telemetry data in flight. You'll collaborate closely with engineering and product design to solve complex challenges around instrumentation, configuration, and data management with new systems and experiences. If you love turning ambitious ideas into impactful enterprise products that delight customers, let's talk! What You'll Do Define and drive the product strategy for our next-generation observability pipelines solution and align your vision with broader company goals Lead cross-functional collaboration with Engineering, Product Design, Sales, and Marketing to translate customer insights and technical opportunities into impactful product capabilities Evangelize the product vision internally
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! As a Senior Salesforce Developer at New Relic, you will be responsible for building and maintaining our Salesforce org. You will work closely with our Sales Revenue & Finance teams to automate processes and build custom applications on the Salesforce platform. In this role, you will be responsible for the full development life cycle, from Design and Solution to deployment and maintenance. The ideal candidate will have 5+ years of experience as a Salesforce developer and be proficient in Apex, Lightning Web Components and Visualforce. They will also have experience with integrations and be able to work with other team members or independently. Duties & Responsibilities Design, develop, test, deploy, and maintain high-quality Salesforce.com solutions Configure Salesforce.com to meet business requirements Write Apex classes, triggers, Visualforce pages, and Lightning Web Components Integrate Salesforce.com with other systems (as needed) Perform data migrations Stay up to date on the latest Salesforce.com features Handle multiple projects simultaneously Work closely with business analysts, project managers, and other developers Adhere to coding standards and best practices Required Skills and Qualifications 5+ years of experience in Salesforce development, configuration, and customization Advanced Apex programming skills, including batch jobs, triggers, web services, and unit testing Experience with Lightning Components, Visualforce, and Force.com integration techno
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity As a Senior Software Engineer within the Container Fabric (CF) organization, you will be a key driver in evolving New Relic’s global internal platform. We are looking for an operations-heavy engineer with 5–8 years of relevant experience who can leverage open-source and custom tooling to orchestrate and maintain large-scale Kubernetes environments. You will play a "Captain" role—leading critical deliverables and mentoring junior engineers while maintaining the reliability of our global fleet. What You'll Do Architectural Leadership: Drive the design and implementation of internal tools, specifically focusing on Kubernetes Operators and Controllers to automate resource management. Platform Orchestration: Lead complex, large-scale infrastructure shifts. Operational Excellence: Take ownership of incident response, author comprehensive retrospectives, and implement systemic hardening to prevent recurrence using advanced overcommit strategies. This Role Requires Experience: 5–8 years in a DevOps, Site Reliability, or Infrastructure Engineering role. Kubernetes Mastery: Deep internals knowledge of Kubernetes and hands-on experience writing custom operators. Tooling Proficiency: Strong experience building production-grade tools and services, specifically for infrastructure automation. Operations-Heavy Mindset: A proven track record of Day 1/Day 2 operations for a large-scale Kubernetes fleet, handling high-severity incidents, and improving SLA compliance through auto
SmartBear delivers application integrity for modern tech stacks, ensuring continuous, measurable assurance that software just works as intended with governance to operate at AI speed and scale. SmartBear offers deep test automation, API lifecycle management, and observability capabilities. With integrations across the SDLC, it sets a new quality standard for application delivery teams. SmartBear is trusted by developers, testers, and software engineers across 32,000 organizations, including 75% of the largest financial institutions and industry leaders such as Adobe, JetBlue, and Microsoft. SmartBear’s open source tools are downloaded more than 100 million times a month and have earned over 30,000 GitHub stars from the developer community. With its best-loved brands, including Swagger, TestComplete, Reflect, QMetry, Zephyr, and more, SmartBear meets customers where they are to make our technology-driven world a better place. Learn more at www.smartbear.com , or follow us on LinkedIn , X , and Reddit . At SmartBear, you will be part of a dynamic team solving one of the most critical challenges facing modern businesses: ensuring the integrity of software in an AI-driven world. Whether you are working directly with customers, driving go to market strategies, supporting operations, building products, or enabling teams, your contributions help shape the future of software quality for organizations worldwide. Join us in our mission. Senior Full Stack Engineer – NodeJS/ReactJS All businesses today interact with their customers via an app, and the wide range of use cases and configurations available in the market today makes it extremely hard to deliver high-quality apps at speed. And end-users today do not have the time nor patience to interact with buggy apps. This amplifies the need for tooling that promotes rapid test automation, and LoadNinja is well-suited to accomplish that. Go to our product page if you want to know more about LoadNinja . You can even ha
SmartBear delivers application integrity for modern tech stacks, ensuring continuous, measurable assurance that software just works as intended with governance to operate at AI speed and scale. SmartBear offers deep test automation, API lifecycle management, and observability capabilities. With integrations across the SDLC, it sets a new quality standard for application delivery teams. SmartBear is trusted by developers, testers, and software engineers across 32,000 organizations, including 75% of the largest financial institutions and industry leaders such as Adobe, JetBlue, and Microsoft. SmartBear’s open source tools are downloaded more than 100 million times a month and have earned over 30,000 GitHub stars from the developer community. With its best-loved brands, including Swagger, TestComplete, Reflect, QMetry, Zephyr, and more, SmartBear meets customers where they are to make our technology-driven world a better place. Learn more at www.smartbear.com , or follow us on LinkedIn , X , and Reddit . At SmartBear, you will be part of a dynamic team solving one of the most critical challenges facing modern businesses: ensuring the integrity of software in an AI-driven world. Whether you are working directly with customers, driving go to market strategies, supporting operations, building products, or enabling teams, your contributions help shape the future of software quality for organizations worldwide. Join us in our mission. Senior Software Engineer – Java Solve challenging business problems and build highly scalable applications. Design, document and implement solutions in Java 8 and 17. Build and deploy services based on the AWS platform. Work with a high-performance team, working with continuous delivery and heavy test automation Product intro Zephyr provides test management capabilities to Atlassian products, working as a plugin/extension for Jira and Confluence, for managing both manual and automated testing processes. Using AI, users can transform thei
SmartBear delivers application integrity for modern tech stacks, ensuring continuous, measurable assurance that software just works as intended with governance to operate at AI speed and scale. SmartBear offers deep test automation, API lifecycle management, and observability capabilities. With integrations across the SDLC, it sets a new quality standard for application delivery teams. SmartBear is trusted by developers, testers, and software engineers across 32,000 organizations, including 75% of the largest financial institutions and industry leaders such as Adobe, JetBlue, and Microsoft. SmartBear’s open source tools are downloaded more than 100 million times a month and have earned over 30,000 GitHub stars from the developer community. With its best-loved brands, including Swagger, TestComplete, Reflect, QMetry, Zephyr, and more, SmartBear meets customers where they are to make our technology-driven world a better place. Learn more at www.smartbear.com , or follow us on LinkedIn , X , and Reddit . At SmartBear, you will be part of a dynamic team solving one of the most critical challenges facing modern businesses: ensuring the integrity of software in an AI-driven world. Whether you are working directly with customers, driving go to market strategies, supporting operations, building products, or enabling teams, your contributions help shape the future of software quality for organizations worldwide. Join us in our mission. Senior Software Engineer - C++ TestComplete Developing and maintaining the testing tool using C++. Collaborating with the team to implement and enhance backend functionalities. Tech stack: C++, C#, Delphi, Avalonia, Windows programming (COM, WinApi), GUI (DevExpress) Product intro All businesses today interact with their customers via an app, and the wide range of use cases and configurations available in the market today makes it extremely hard to deliver high-quality apps at speed. And end-users today do not have the time
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Location- Chennai Team: Engineering Enablement Group As a Senior Software Engineer in our Engineering Enablement Group, you will lead the re-design and evolution of our Mobile Branding framework — the system that enables customers to create custom-branded versions of the Appian mobile application for both iOS and Android. You will drive the architectural modernization of the end-to-end branding pipeline, from the customer-facing Forum application and provisioning tools to the backend build service running on Mac EC2 runners in AWS. By leveraging modern microservices, CI/CD automation, and cloud-native infrastructure, you will transform the current system into a more reliable, scalable, and maintainable platform that reduces manual intervention and accelerates customer delivery. We are looking for a technical leader who can bridge the gap between complex Ruby/Bash-based tooling, Appian process models, and AWS infrastructure to deliver a seamless mobile branding experience. Primary Qualifications: 6-9 Strong working experience with Android and iOS frameworks and mobile application development workflows. Familiarity with mobile build systems (Fastlane, Xcode, Gradle) and code-signing workflows. Experience with proficiency in Python, with experience in Ruby, Bash, or Go being a plus. Advanced experience with AWS infrastructure (S3, Lambda, EC2) and CI/CD pipeline design. Strong end-to-end knowledge of pipeline creation, deployment automation, and infrastructure-as-code (Terraform). Familiarity with monitoring, observability, and performanc
Senior Software Engineer (Backend) About Team Cloud Platform Engineering group designs, builds and manages platforms that allow Myntra’s tech product to be secured, reliable, deployed and run at scale. These horizontal platforms leverage cloud hosting platforms and ensure all Myntra’s hostings are agnostic to Cloud vendor. We also build a number of production automation like provisioning of infrastructure at scale and manage complex access management on servers. We have developed numerous in-house tools and platform for load/stress testing, zero trust, security and compliance, CI/CD, Observability at scale. And, we aggressively adopt from open sources and try to contribute back to the community. Tools and Platform Engineering This team builds and maintains centralized and high-scale platforms for Observability (centralized log collection, metric systems, monitoring systems), Security & Compliance (access management, secret management, database access, change management systems, Authentication and Authorization of services etc). The platforms developed by these teams are centralized tools used by all engineering teams for database access, and changes, infrastructure provisioning, on-call scheduling, onboarding new monitoring etc. The vault system and APIs, developed, deployed and managed by this team, is being used by all Myntra’s production services. This team consists of full-stack developers who are skilled in Python, Golang, ReactJs. Roles and Responsibilities Design, build and maintain central platform products to improve the security posture of Myntra Write maintainable, scalable, and efficient code. Design and architect technical solutions for the developer community at Myntra Work in a cross-functional team, collaborating with peers during the entire SDLC. Follow coding standards, code reviews, etc. Follow scrum sprint cycles and commitment to deadlines. Identify security gaps in or for software platforms and incorporate them into requirements
DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,000+ members, DataHub's metadata graph provides in-depth context of AI and data assets with best-in-class scalability and extensibility. The company's enterprise SaaS offering, DataHub Cloud, delivers a fully managed solution with AI-powered discovery, observability, and governance capabilities. Organizations rely on DataHub solutions to accelerate time-to-value from their data investments, ensure AI system reliability, and implement unified governance, enabling AI & data to work together and bring order to data chaos. About the job DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,000+ members, DataHub's metadata graph provides an in-depth context of AI and data assets with best-in-class scalability and extensibility. The company's enterprise SaaS offering, DataHub Cloud, delivers a fully managed solution with AI-powered discovery, observability, and governance capabilities. Organizations rely on DataHub solutions to accelerate time-to-value from their data investments, ensure AI system reliability, and implement unified governance, enabling AI & data to work together and bring order to data chaos. In this role, you will Build core capabilities for our SaaS Platform across multiple clouds Drive development of functional enhancements for Data Discovery, Observability & Governance for both OSS and SaaS offering Lead efforts around non functional aspects like performance, scalability, reliability Lead and mentor junior engineers Work closely with PM, Customers and OSS community Requirements Over 8+ years of experience building and scaling backend systems, preferably in cloud-first or SaaS environments. Solve complex tech
Title: Senior Site Reliability Engineer - I, Product Area Focus Location: Noida (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering teams. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedit
Role: Senior AI Engineer Location: Hyderabad, India (Hybrid) Department: Product Development About the Role GHX is building a cutting-edge LLM-powered document understanding platform focused on classification, structured data extraction, and intelligent orchestration at scale. This is a high-impact AI engineering role where you will own the full lifecycle—from problem framing to production deployment . Initially, you will focus on prompt engineering and evaluation systems , building the quality foundation for AI performance. Over time, the role expands into agent orchestration, system architecture, and migration of rule-based systems to LLM-driven pipelines . A strong foundation in software engineering (5+ years) is essential. This role demands engineering rigor across both traditional system design and AI system behavior . Core Responsibilities 1. Prompt Engineering Design prompts for diverse document classification and extraction tasks Treat prompts as formal specifications (precise, structured, and edge-case-aware) Develop few-shot, chain-of-thought, and structured output templates Manage prompt lifecycle: versioning, testing, and rollback 2. LLM Output Evaluation Create and maintain ground truth datasets Build automated evaluation pipelines (precision, recall, field-level accuracy) Identify and resolve conceptually incorrect outputs despite surface correctness 3. AI Agent Orchestration Design multi-agent workflows for document processing Implement tool-use patterns and integrate MCP servers Optimize orchestration for scale and efficiency 4. Software Engineering Develop production-grade APIs and backend services Apply Clean Architecture / DDD principles Write maintainable, testable Python code Contribute to CI/CD, deployment, and observability systems 5. Stakeholder Collaboration Act as a bridge between business stakeholders and AI systems Translate product requirements into technical architectures Communicate system behavior, limitations, and quality
Here's a summary of the role: Do you love building scalable cloud platforms and solving complex engineering problems with modern technologies? As a Senior Software Engineer at Diligent, you'll design and deliver high-performing , serverless applications that power our global SaaS platform. You'll work extensively with TypeScript, Node.js, AWS, and event-driven microservices, owning services from design to deployment and production monitoring. This is an opportunity to influence technical decisions, mentor engineers, and explore how AI can transform software development and engineering productivity. If you're passionate about cloud-native architectures, distributed systems, and building software that scales to millions of users, we'd love to meet you. Here's a breakdown of what you'll do (not all of it, just the important stuff): Design and build scalable backend services and event-driven microservices using TypeScript and AWS. Develop secure APIs and integrations that power reporting, analytics, and dashboard experiences. Build and maintain serverless solutions using AWS services such as Lambda, EventBridge , SQS, and DynamoDB. Drive engineering excellence through testing, observability, automation, and production readiness practices. Contribute to infrastructure-as-code and CI/CD pipelines using AWS CDK and modern DevOps practices. Mentor engineers, participate in architecture discussions, and champion the use of AI tools to improve development efficiency. These are the essentials you'll need to get an interview: 6-8 years of professional software engineering experience. Strong experience with TypeScript, Node.js, and modern backend development patterns. Hands-on experience building cloud-native applications on AWS. Strong understanding of serverless architectures and event-driven microserv
About the Role At FourKites we have the opportunity to tackle complex challenges with real-world impacts. Whether it’s medical supplies from Cardinal Health or groceries for Walmart, the FourKites platform helps customers operate global supply chains that are efficient, agile and sustainable. Join a team of curious problem solvers that celebrates differences, leads with empathy and values inclusivity. As a Senior Customer Engineer, you own the technical customer relationship end to end. You run discovery independently, design integration and agentic workflow architectures for complex enterprise problems, and deploy solutions live with customers — often before they know exactly how to articulate what they need. You write optimized, production-grade code at speed, grounded in strong data structures and algorithms fundamentals, because compressing the time from customer problem to working solution is how FDE delivers its value. You bring genuine innovation to hard problems — your solutions are technically sound, elegant, and often non-obvious. You are the primary technical contact for 2–4 major enterprise accounts, you mentor FDEs on the team, and you are building the skills that will take you into Staff-level technical leadership. What You'll Do Own complex integration and AI agent implementations end-to-end — from technical discovery through go-live and post-launch enhancement — as the primary technical contact for 2–4 enterprise accounts Design integration architectures with explicit attention to error handling, retry logic, observability, failure recovery, and multi-system authentication Design and deploy agentic AI workflows that orchestrate supply chain operations — from requirements through production, including regression testing and validation before each customer deployment Deploy AI agent workflows live with customers present — configuring and troubleshooting in the room during customer calls, not gathering requirements to build later Run customer discovery
₹2K – ₹2K/yr
Opportunity Overview: We are seeking a Senior Data Engineer to contribute to the design and delivery of our cloud-native healthcare data platform. You will implement scalable data solutions built on AWS, Apache Iceberg, Lake Formation, Glue Catalog, Athena, dbt, and modern orchestration frameworks. This role combines strong hands-on engineering with collaboration across platform, analytics, and business teams. What You'll Do Data Engineering Delivery Deliver complex data engineering projects in collaboration with cross-functional teams Drive technical execution from design through production deployment Implement scalable data patterns and reusable frameworks Design and implement batch and near-real-time pipelines Build reusable ingestion, transformation, validation, and publishing frameworks Support modernization of legacy workloads Contribute to Apache Iceberg implementation and optimization Apply standards for schema evolution, partitioning, compaction, and metadata management Ensure efficient storage and query performance Implement data quality frameworks and validation layers Support observability and monitoring practices Contribute to operational excellence and reliability improvements Participate in architecture and design discussions Conduct and participate in code reviews Mentor junior engineers and share best practices ISMS roles and responsibilities Good knowledge of Information security Oversee specific business processes within the ISMS. Responsible to manage the ISMS documentation, conduct risk assessments, and implement risk treatment plans. Risk Owners are responsible for identifying, assessing, and managing risks within their areas of responsibility. They are also responsible for implementing risk treatment plans. Conduct the BCP and other test related to information security continuity along with CISO Responsible for monitoring and reporting on the performance of the ISMS. Responsible for implementation of security policies and procedures and report
Other cities to consider
More places hiring for this role
Get new senior observability engineer jobs in India by email
Daily job updates · Unsubscribe anytime