Jobs in India

Senior Member Of Technical Staff Ml Systems And Infrastructure in India

999 active opportunities · Updated October 2026

Explore current senior member of technical staff ml systems and infrastructure jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. What You’ll Do: Architect the Future of AI Infrastructure: You will design, build, and own the end-to-end platform that supports the entire lifecycle of our ML models—from massive-scale distributed training to ultra-low-latency, highly-available inference. Optimize and Serve Cutting-Edge Models: You'll implement and scale sophisticated inference stacks for LLMs using frameworks like vLLM, TensorRT-LLM, or SGLang . You’ll solve complex challenges in throughput, latency, token streaming, and automated scaling to deliver a seamless user experience. Empower AI Innovation: You will act as a strategic partner to our AI Research and Data Science teams. You’ll create a seamless developer experience that accelerates their ability to experiment, fine-tune, and deploy groundbreaking models with velocity and confidence. Automate Everything: You'll develop robust CI/CD/CT (Continuous Training) pipelines using tools like Argo Workflows, ArgoCD, and GitHub Actions to automate model validation, deployment, and lifecycle management, ensuring our systems are both agile and rock-solid. What are we looking for Experience: 5+ years in infrastructure or software engineering, with at least 2+ years laser-focused on MLOps or ML infrastructu

PythonKubernetesCI/CDGit
E
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE We are looking for a highly skilled Senior Frontend Engineer to join the Portworx UI team, responsible for building intuitive, scalable user experiences for our Kubernetes-based platform. You will build production-quality web applications using React, collaborate closely with UX, Product, backend teams, and engineering leadership, and help shape frontend architecture and engineering practices. This role requires strong ownership, sound technical judgment, and the ability to solve complex problems independently while helping other engineers grow. WHAT YOU'LL DO We are primarily an in-office environment and therefore, you will be expected to work from the {{OFFICE_LOCATION}} office in compliance with Everpure's policies, unless you are on PTO, or work travel, or other approved leave. Design, develop, and maintain scalable, high-performance single-page applications using React and TypeScript. Own features end to end—from requirements and design through implementation, testing, release, and production support. Create reusable, accessible, responsive UI components using modern HTML, CSS, React patterns, and design systems. Collaborate with UX, Product Management, and backend teams to turn customer and product needs into polished, production-ready features. Contribute to frontend architecture, API integration, performance, maintainability, and engineering best practices. Write and maintain unit, integration, and end-to-

JavaScriptTypeScriptJavaReact
E
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE We are hiring a senior quality engineer to own System Testing for Pure’s FlashArray products. You will validate stability, resiliency, and performance under sustained, production‑like workloads , far beyond basic functional testing. You will design and run multi‑week, customer‑like scenarios that combine workloads, failovers, upgrades, and fault injections, and use the insights to influence architecture, design, and release decisions. This is a hands‑on, high‑impact role at the intersection of architecture, systems, and large‑scale testing—acting as a key quality gate before releases reach Pure’s customers. WHAT YOU'LL DO Own System Testing strategy for releases Define System Testing strategy and test plans for major features and releases, focusing on stability, longevity, and end‑to‑end behavior, not just feature correctness. Design realistic, high‑value scenarios Build scenarios that mirror Pure customer environments: mixed workloads (block, file, object), long‑running IO, failovers, NDUs, hardware events, and background operations (replication, snapshots, quotas, etc.), combining automation with targeted “tortures”. Drive execution and triage on System Testing beds Own System Testing environments (arrays, initiators, OSes, accessories); keep them healthy, representative, and well‑instrumented. Monitor runs, triage failures quickly, separate infra issues from product bugs, and file high‑quality defe

PythonAWSKubernetesLinux
E
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As the Team Lead for Initiator & Protocol Engineering, you will spearhead the critical bridge between our industry-leading FlashArray and the Linux/VMWare ecosystems. You will drive the performance and reliability of our storage protocol stacks—spanning NVMe over Fabrics and Fibre Channel—ensuring Pure Storage remains the gold standard for enterprise connectivity. Collaborating closely with cross-functional hardware and software teams, you’ll mentor a high-caliber engineering squad to solve complex kernel-level challenges and influence the global Linux upstream community. WHAT YOU'LL DO Own the Protocol Lifecycle: Lead the development, maintenance, and optimization of Linux and VMWare initiator stacks (NVMeoF, FC-SCSI, iSCSI) and target drivers to ensure seamless, high-performance integration with Pure FlashArray. Drive System Resilience: Architect enhancements for Fibre Channel and NIC driver stacks that improve RAS (Reliability, Availability, and Serviceability), specifically focusing on multipathing logic and link health monitoring. Technical Leadership & Mentorship: Guide a team of senior and junior engineers through complex project deliveries, conducting deep-dive code reviews and setting the technical bar for C/C++ and Python development within the kernel space. Solve the Impossible: Act as the final escalation point for the most challenging system-level bugs found in the field or internal testing, u

PythonAWSLinuxRest
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. About the Role We are looking for Staff System Software Engineer in Test to join our team. In this role, you will be responsible for design, development, automation and reporting of Integration and system tests spanning across firmware and device drivers. This role requires you to have significant technical breadth and deep understanding of low-level system software specifically in server class systems. You will be part of a new team responsible for integration of different system software deliverables and development of system tests spanning all the components. You will contribute to shaping the test strategy , guide best practices and solve complex problems while maintaining a strong hands-on focus. You will partner with development and other QA teams to deliver high quality scalable and reliable solutions. About the Team Integration and system test team is responsible for verification and validation of integrated components across Board management controller (BMC), Firmware and Linux device driver. The team is also responsible for management and maintenance of common tools and pipel

PythonCI/CDGitLinux
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Staff -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debu

PythonLinuxAIC++
T
📍 Hyderabad, Telangana, India
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Aliases: Staff Engineer, Senior Staff Engineer, Lead Engineer, Senior Technical Lead, Senior Architect, Platform Architect, Data Architect, Solutions Architect (Engineering), Principal Engineer at a smaller company About Truveta Truveta is the world's first health provider-led data platform with a vision of Saving Lives with Data. Our mission is to enable researchers to find cures faster, empower every clinician to be an expert, and help families make the most informed decisions about their care. Achieving Truveta's ambitious vision requires an incredible team of talented and inspired people with a unique combination of health, software, and big data expertise who share our company values. This opportunity Join the founding team of the Truveta India Development Center and play a pivotal role in shaping its future. As one of the early members, you will help build and scale a high-impact organization while contributing to the products and platforms that advance healthcare through data and AI. This is an unusual chance to influence both the technical direction and the culture of a growing global engineering hub. The Data Platform group in India owns parts of the pipeline that ingests, processes, stores, and serves healthcare data at a scale very few organizations operate at. Patients, doctors, and medical researchers deserve the same engineering rigor that has transformed other industries, and that is the work: correctness under volume, reliability under failure, and speed without cutting the corners that regulated health data does not allow anyone to cut. Who we need We are seeking engineers who think in platforms. You will own a processing domain or a multi-team data capability, define its architecture, and make that architecture legible enough that the teams building inside it can move quickly without asking permission. The problems at this level are the recurring ones: the class of data-quality defect that keeps retu

G
📍 India· Full-time
✓ Quality checkedCompany trend -75.6%

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As a Staff Backend Engineer at GitLab, you will help shape a major investment in our Software Supply Chain Security offering. In this role, you'll serve as a senior technical leader for backend systems that help customers secure how software is built, verified, and delivered inside the GitLab platform. You'll work on foundational capabilities across package policy enforcement, build provenance, artifact signing, and malicious package detection, with a strong focus on enterprise-grade security and performance. You'll define architecture before systems are built, write clear technical proposals, and guide i

CI/CDGitRestAI
G
📍 India· Full-time
✓ Quality checkedCompany trend -75.6%

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As Director, Engineering, Platform Operations & Productivity, you'll own three functions that all require hands-on technical depth, not just people management. Platform Staff is a small, senior, AI-native team that moves to wherever the organization needs the most leverage, from standing up early scaffolding for an initiative, to taking on a high-impact customer request that doesn't fit any existing team's charter, to stepping directly into a production crisis until it's resolved. This role is for a technical engineering director who has personally built distributed systems, not only managed people wh

GitRestAIRust
S
📍 Gurugram, India
✓ Quality checkedCompany trend -83.3%

Work Flexibility: Hybrid What You Will Do / Roles & Responsibilities As a member of the Global Product Engineering team you will gain in depth knowledge about Stryker’s Trauma & Extremities portfolio. You will learn how Bone Plates & Screws, External Fixators, Surgical Instruments and many other products improve people’s life. Together we are responsible for over 50'000 articles. We count on you to maintain and improve our high standard in design specification and performance. You ensure compliance with applicable regulatory and quality requirements, including FDA, EU MDR, ISO 13485, and internal standards, by driving design change management activities, support regulatory submissions, and serving as a subject matter expert during internal and external audits. In your field of responsibility, you will make sure that design requirements are met as well as apply the design change process and identify technical solutions in this context. You oversee the continuous improvement projects for existing products in our group, create and manage project plans, assign and follow up on work packages. You participate in and at times lead cross functional teams to perform design related Non-Conformities(NC), Corrective- and Preventive-Actions(CAPA). Your role will include frequent contact with other teams such as the international Regulatory Affairs, Tech. Publications, Design Quality or Marketing to address design related questions for international registrations and creations of technical files required for the implementation of EU-MDR requirements. Drive value improvement, design optimization, and continuous improvement initiatives to enhance product performance, quality, compliance, and cost effectiveness. What You Need <

D
📍 Mumbai, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the Role: As a Forward Deployment Architect, you will serve as a hands-on senior technical architect on the Applied AI Engineering team, owning the end-to-end design and delivery of AI-driven business transformation projects. You'll work closely with pre-sales teams to scope technical integration and implementation strategies, translating business requirements into architectural solutions. Once opportunities move to post-sales, you'll own the detailed technical designs from original scoping documents and drive execution including hands-on coding to build proofs-of-concept, custom integrations, and solution prototypes that validate technical feasibility. As the technical owner of the customer relationship, you'll partner cross-collaboratively to ensure successful delivery. Your role spans from understanding domain-specific customer needs to architecting scalable, agent-based AI solutions using DevRev's platform**, with direct involvement in implementing key technical components and debugging complex integration challenges. This position requires a unique blend of enterprise architecture expertise, AI solution design, hands-on development skills, customer empathy, and cross-functional collaboration. You'll act as

JavaScriptTypeScriptPythonJava
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary The Workplace Coordinator is responsible for the day-to-day operation and maintenance of workplace infrastructure, ensuring that all building systems operate efficiently, safely, and reliably. This role oversees HVAC systems, including chillers and AHUs, Building Management Systems (BMS), electrical panels, and coordinates preventive and corrective maintenance activities to support uninterrupted business operations. Key Responsibilities Technical Operations · Monitor and maintain HVAC systems including Chillers, AHUs, FCUs, and ventilation systems. · Operate and monitor Building Management System (BMS) for alarms, trends, and equipment performance. · Inspect electrical LT panels, UPS systems, DG synchronization (if applicable), and power distribution systems. · Monitor critical utilities including temperature, humidity, pressure, and energy consumption. · Ensure uninterrupted operation of critical infrastructure and respond promptly to system failures. Preventive & Corrective Maintenance · Plan and execute preventive maintenance schedules for HVAC and electrical systems. · Coordinate breakdown maintenance with

AILeanProcurementHR
W
📍 India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: As a member of the Global Technical Operations (TechOps), you will be a part of a team that focuses on operational reliability within a cloud-based infrastructure. You have hands-on cloud experience in architecting, building, deploying, managing databases, compute instances, and storage buckets. You have a passion for providing solutions through automation. You know that success is through collaboration and communication. What you'll be doing: Work in cross-functional teams to develop solutions and identify opportunities to bring efficiency and effectiveness. Research, evaluate, and incorporate new technologies/concepts into existing frameworks. Proactively identify areas to improve efficiency and effectiveness, recommend and implement solutions towards them. Develop and innovate operational practices, procedures for workflows, and documentation. Implement and contribute to IT security best practices. Automate tasks to ensure consistency and speed of deployment. Identify, analyze, and troubleshoot issues and work towards resolution. Explain technical solutions to bo

PythonGCPKubernetesCI/CD
O
📍 India· Full-time
✓ Quality checkedCompany trend -68.5%

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta is seeking a technical writer who is passionate about creating and delivering developer-centered technical content for developer.okta.com. We’re looking for a collaborative, creative team member who understands how to write clear, concise, and relevant content. Reporting to the Manager of Technical Documentation, you will work closely with a global network of engineers, product managers, and developer success professionals to understand real-world use cases for identity and access management. You will create practical technical content that guides Okta's broad technical community, including application developers and integration partners to adopt, integrate, and deploy our solutions successfully. If you have a passion for learning new technologies and want to help us drive the adoption and reach of our products and platform, this is the place for you! Location : Bengaluru, Karnataka, India Work Mode : Hybrid (2 days Onsite per week) Note : "This role requires in-person onboarding and travel to our Bengaluru, IN office during the first week of employment." What you’ll be doing Design, develop, edit, and deliver accurate and effective developer documentation for Okta products. Work closely with Engineering, Product Management, QA, Support, and members of the Information Development team to ensure quality and accuracy of content. Follow our documentation guidelines to produce high quality documentation. Prioritize projects when working

AWSGitRestAgile
G
📍 Gurugram, Haryana, India· Full-time
✓ High-confidence listingCompany trend +33.3%
Quick readStrong listing-quality and freshness signals

Location Details: Remote, India At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team... The Securities Analytics and Products Group is responsible for developing and maintaining sophisticated software solutions to safeguard GoDaddy's ecosystem. We are seeking a dedicated Senior Software Engineer with a strong background in software development and a keen focus on security. The ideal candidate will have proven hands-on experience in software development, with a consistent track record of working with the latest software technologies while prioritising security best practices. As a key member of our engineering team, you will play a crucial role in designing, developing, and implementing secure software solutions to protect our organisation from cyber threats. You will get to work with some of the brightest minds to build secure, highly available, fault-tolerant, and globally performant microservices-based platform deployed on the AWS cloud, using the newest technology stack. While the role is primarily backend-focused, you'll also get opportunities to contribute to frontend features as needed, giving you exposure across the full stack. What you'll get to do... Design, develop, and maintain secure, highly available, fault-tolerant, and globally performant code deployed on AWS cloud. Ensure code quality through extensive unit and integration testing Own frontend features and UI components across projects on an as-required cadence, from design through delivery Investigate and resolve production issues, ensuring your team's DevOps on-call responsibilities Contribute to the technical documentation, code reviews,

TypeScriptPythonReactAngular
🔔

Get new senior member of technical staff ml systems and infrastructure jobs in India by email

Daily job updates · Unsubscribe anytime