About the Role Adobe is seeking a Machine Learning Engineer to join the Adobe Genuine Engineering team. This group protects Adobe's ecosystem from fraud, abuse, and misuse using intelligent systems worldwide. In this position, you will build and develop machine learning models from scratch, including custom transformer-based frameworks, to identify fraudulent actions, stop account sharing, and protect the experience of hundreds of millions of users. You will manage the entire model lifecycle: raw behavioral data and feature engineering, architecture development, large-scale GPU training, deployment, and monitoring. The team is actively building in-house behavioral foundation models that learn identity-preserving representations from long sequences of user activity. This is a role for an engineer who wants to own deep learning systems end-to-end — not consume pre-built ones. Key Responsibilities Build and train deep learning models from scratch, including custom transformer and attention-based architectures for long behavioral event sequences. Own the full training stack: event tokenization, temporal and positional embeddings, self-supervised pretraining (e.g., masked modeling, contrastive learning), and downstream fine-tuning. Train large models efficiently on GPU infrastructure using mixed-precision training, gradient accumulation/checkpointing, efficient attention, and distributed strategies (DDP, FSDP, or equivalent). Build and optimize feature pipelines on Databricks and Spark, transforming raw behavioral events into high-quality model inputs. Translate prototypes into production ML systems — scalable, reliable, and observable — and drive inference performance through architectural and serving-side optimization. Contribute to MLOps practices: experiment tracking, model versioning, CI/CD, automated retraining, and
Jobs in India
Infrastructure Team Manager in India
669 active opportunities · Updated October 2026
Showing
15 jobs
Explore current infrastructure team manager jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
NVIDIA is seeking a Senior Staff SRE to build and operate reliable, scalable compute platforms that support global engineering workloads. This role spans Kubernetes, KubeVirt, bare-metal infrastructure, automation, observability, and AI-enabled operations. Join a team that solves complex infrastructure challenges, builds durable automation, and improves the reliability and operational experience of critical compute services. What you’ll be doing: Build, operate, and improve large-scale Kubernetes, KubeVirt, Linux, container, and bare-metal compute platforms, with a focus on performance, capacity, reliability, and operational scale. Lead bare-metal provisioning and lifecycle management in data centers, including PXE boot, DHCP, DNS, OS provisioning, hardware validation, and fleet automation. Develop automation, self-service capabilities, and observability solutions using APIs, Python or Go, Infrastructure as Code, configuration management, metrics, logs, traces, and service-health data. Define and operate SLOs, SLIs, error budgets, alerting, and incident-response practices; lead complex incident investigations, corrective actions, and blameless postmortems. Partner with infrastructure, security, hardware, data-center, and application teams to deliver global platform initiatives, and participate in an on-call rotation. What we need to see: BS in Computer Science, Engineering, a related technical field, or equivalent experience, plus 10+ years operating production infrastructure or platform services. Strong expertise in Kubernetes administration, KubeVirt, Docker, containerization, microservices, Linux systems, and resolving distributed-system challenges. <l
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Staff Software Engineer (Full Stack - Backend) Staff Software Engineer The Product Auth0 is a developer-friendly identity platform designed by developers, for developers, simplifying authentication and authorization for more than 100 million daily logins worldwide. Within this global platform, our team drives two of Auth0’s most critical, high-impact surfaces: The Manage Dashboard: The definitive front door for developers and enterprise administrators worldwide. It is the control center where our customers configure critical identity policies, manage complex organizations, and monitor active threat intelligence. The Auth0 Marketplace: Our rapidly expanding extensibility ecosystem. The Marketplace allows customers to seamlessly discover and install third-party integrations - such as security monitoring, consent management, and database connectors - to dynamically customize their identity pipeline. The Role As a Staff Engineer on this team, you will not just write code; you will own the technical strategy, architect highly resilient frameworks, design developer-friendly APIs, and mentor a high-performing team of engineers. Your architectural ownership will span the entire stack - from the React-based frontend down through our Node.js/Go services, robust operational pipelines, and PostgreSQL databases. What You Will Do: Product Ownership: Lead the technical lifecycle of your team’s products, ensuring they are robust, secure, and performant from the frontend to
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Staff SET Opportunity We are looking for an experienced Staff Software Engineer in Test to join our Identity Management Engineering (IDM) team serving the Privileged Access Team (PAM). This team is passionate about delivering large-scale, mission-critical software in a fast-paced Agile environment. In this role you'll be working with a team of highly-skilled and talented engineers, responsible for delivering sophisticated backend solutions that help Okta reliably operate at large scale and be highly available. As part of the team, you’ll be ensuring projects are completed with the highest quality and reliability using automation at every level for fast, robust and secure releases. What you’ll be doing Review requirements and design specs to develop relative test plans and test cases Automate API tests, end-to-end tests, reliability/scale tests Work with engineering management to scope and plan engineering efforts Communicate and document QE plans for scrum teams to review Review application code, identify bug and other areas of weakness, architect tools for future coverage Automate all critical features to maintain zero-debt cadence Release features with solid quality Respond to production issues/alerts and customer issues during on-call rotation Be a strong customer advocate with a strong quality DNA What you’ll bring to the role 5+ years of QE experience preferably in an enterprise SaaS company 3+ years experience in quality engineering for ente
About Target As a Fortune 50 company with more than 400,000 team members worldwide, Target is an iconic brand and one of America's leading retailers. Joining Target means promoting a culture of mutual care and respect and striving to make the most meaningful and positive impact. At Target, we have a timeless purpose and a proven strategy. Some of the best minds from different backgrounds come together to redefine retail in an inclusive learning environment that values people and delivers world-class outcomes. Target in India operates as a fully integrated part of Target's global team and supports the company's global strategy and operations. About the team The IT Data Platform (ITDP) team enables data-driven management of Target's technology ecosystem by bringing together trusted data and insights across technology assets, software delivery, infrastructure, reliability, security, engineering effectiveness, and technology operations. The Analytics team within ITDP transforms this data into metrics, analytical products, dashboards, predictive insights, and decision-support capabilities that help technology teams understand what is happening, why it is happening, where risk may be emerging, and where action is needed. The team is building toward an analytics capability that progresses from descriptive and diagnostic analytics to predictive and prescriptive insights, using statistical methods, applied data science, and GenAI where each approach is appropriate. As a Senior Data Analyst for Target's IT Data Platform Analytics team you'll: <p style="co
Software Engineer - III About the Role GHX is building a next-generation Intelligent Process Automation (IPA) platform powered by LLMs and AI-native document understanding. We extract structured data from complex healthcare procurement documents — Purchase Orders, invoices, contracts — at scale across cloud environments. As a member of the IPA engineering team, you will bridge strong software engineering with applied AI. You will design and ship Python services, integrate LLM APIs, build AI agents, create skills and validate the output of LLMs to build document extraction pipelines, and own evaluation infrastructure that ensures production quality. This is not a research role — it is a hands-on engineering role where AI fluency amplifies solid software craft. Core Responsibilities Python Development & Automation Build and maintain Python-based automation services and IPA workflows Develop platform-agnostic solutions supporting future migration across automation tooling Build reusable libraries, frameworks, and components for automation projects Integrate automation solutions with REST APIs and enterprise applications Deploy and monitor automation bots on AWS (Lambda, ECS, SQS, S3) AI & LLM Integration Integrate LLM APIs (like Anthropic Claude, OpenAI, Azure AI) into production pipelines Design classification and extraction prompts for diverse document types (POs, invoices, contracts) Write prompts that function as formal specifications — unambiguous, edge-case-aware Build and iterate few-shot, chain-of-thought, and structured output templates Own prompt library versioning, rollback strategy, and prompt lifecycle management AI & LLM Integration Integrate LLM APIs (like Anthropic Claude, OpenAI, Azure AI) into production pipelines Design classification and extraction prompts f
Description: Graviton is a privately funded quantitative trading firm striving for excellence in financial markets research. We are seeking a Software Engineer for our team in Gurgaon. Our Core Technology team has some of the best programmers in India working on cutting edge technologies to build a super fast and robust trading infrastructure handling millions of dollars worth of trading transactions every day. Requirements Contribute to all layers of backend systems including databases, APIs and applications Design, build and maintain applications for business requirements Architect scalable and reliable applications Write clean and modular code, following good coding standards and practices Troubleshoot and debug applications Be involved in the entire application lifecycle Collaborate with multidisciplinary team of front-end developers, engineers and system administrators Devise innovative solutions to address new and complicated challenges Build reusable code and libraries for future use Take lead on projects, as needed Work in a high paced competitive environment Qualifications The ideal candidate will have: Engineering degree in Computer Science (preferred) or any other discipline from a Tier 1 college. 3 to 4 years of relevant experience in designing and improving information systems Proficiency in Python programming Extensive experience working with relational databases and handling large datasets Good Understanding of object oriented and asynchronous programming Familiarity with front-end languages such as HTML, JavaScript and CSS Knowledge of Linux systems and bash scripting Good communication and interpersonal skills Willingness to continuously learn and improve Ability to work in fast paced environment under pressure and manage multiple high priority projects Good to have: Experience in Financial services space/domain Experience in and understanding of system design decisions Experience in leading small teams or projects Benefits: Our open and casua
JOB TITLE Cloud Compute Engineer A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. WHAT YOU'LL DO Design, build, and operate Kubernetes clusters on Amazon EKS, including cluster lifecycle management, networking, autoscaling, and workload scheduling. Manage and optimize EC2-based compute infrastructure, including instance selection, placement strategies, capacity planning, and utilization analysis. Operate and improve ECS-based services where applicable, ensuring consistency across our container runtime environments. Develop and maintain Infrastructure as Code (IaC) using Terraform to provision and manage compute resources at scale. Collaborate with development and platform teams to define compute patterns, containerization standards, and deployment best practices. Monitor compute environments for availability, performance, and cost, driving continuous optimization across the fleet. Contribute to architectural decisions around workload placement, multi-tenancy, OS image and container li
JOB TITLE SOFTWARE ENGINEER, TECHNOLOGY A CAREER WITH POINT72'S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU'LL DO We are looking for an experienced professional to work as part of the Finance Technology team. In addition to tactical development, you will be responsible for delivering and creating programs to modernize and scale the platform through technology upgrades, cloud technology adoption, and re-architecting business processes. You will work alongside world-class engineers, partnering directly with Finance stakeholders to understand existing workflows and deliver scalable replacements. This position offers deep domain exposure across FP&A, investor reporting, and compensation, and carries a high degree of ownership over the modernization roadmap. Specifically, you will: Build software applications and deliver software enhancements and projects supporting finance and investor processing. Work closely with business stakeholders to develop software solutions using test-driven and agile software development methodologies. Be responsible for system upgrades and features supporting resiliency and capacity improvements, automation and controls, and integration with internal and external vendors and services. Driving architecture of core platforms and accelerating modernizing leveraging AI tools Work with DevOps teams to manage and resolve operational issues and leverage CI/CD platforms while following DevOps practices within the team and projects. Con
JOB TITLE SOFTWARE ENGINEER, TECHNOLOGY A CAREER WITH POINT72'S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU'LL DO We are looking for an experienced professional to work as part of the Finance Technology team. In addition to tactical development, you will be responsible for delivering and creating programs to modernize and scale the platform through technology upgrades, cloud technology adoption, and re-architecting business processes. You will work alongside world-class engineers, partnering directly with Finance stakeholders to understand existing workflows and deliver scalable replacements. This position offers deep domain exposure across FP&A, investor reporting, and compensation, and contributes to the modernization roadmap. Specifically, you will: Build software applications and deliver software enhancements and projects supporting finance and investor processing. Work closely with business stakeholders to develop software solutions using test-driven and agile software development methodologies. Be responsible for system upgrades and features supporting resiliency and capacity improvements, automation and controls, and integration with internal and external vendors and services. Contributing to architecture of core platforms and accelerating modernizing leveraging AI tools Work with DevOps teams to manage and resolve operational issues and leverage CI/CD platforms while following DevOps practices within the team and projects. Continuously improve the platforms using th
Forward Deployed Senior Software Engineer (Migration Tooling – RunMyJobs) OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full-stack automation fabric solutions for mission-critical business processes. Our flagship SaaS platform, RunMyJobs (RMJ) , is the first composable automation platform specifically built for ERP environments. We enable organizations to orchestrate, manage, and monitor workflows across applications, services, and infrastructure — in the cloud or on premises. Our global team of automation experts, engineers, and customer success professionals work together to deliver seamless automation transformations. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are seeking a Forward Deployed Software Engineer focused on migration tooling and customer onboarding to RunMyJobs (RMJ) . This role is part of an exciting new Forward Deployed Engineering team within the Global Professional Services team, at the intersection of Product Engineering and Go To Market teams. In this role, you will operate at the intersection of engineering and delivery. Your primary focus will be designing, building, and enhancing migration frameworks, tooling, and automation accelerators that enable customers to smoothly transition from legacy schedulers and automation platforms into RMJ. You will work closely with: Migration Architects to design scalable and reusable migration patterns Professional Services to enable efficient customer onboarding Engineering & Product to improve platform capabilities based on field learnings Customers (occasionally) to validate requirements, troubleshoot edge cases, and ensure successful implementations This is a forward-deployed engineering role — highly technical, impact-driv
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. About the Role We are looking for Staff System Software Engineer in Test to join our team. In this role, you will be responsible for design, development, automation and reporting of Integration and system tests spanning across firmware and device drivers. This role requires you to have significant technical breadth and deep understanding of low-level system software specifically in server class systems. You will be part of a new team responsible for integration of different system software deliverables and development of system tests spanning all the components. You will contribute to shaping the test strategy , guide best practices and solve complex problems while maintaining a strong hands-on focus. You will partner with development and other QA teams to deliver high quality scalable and reliable solutions. About the Team Integration and system test team is responsible for verification and validation of integrated components across Board management controller (BMC), Firmware and Linux device driver. The team is also responsible for management and maintenance of common tools and pipel
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. We are looking for a high-impact Senior Software Engineer in Test (Sr. SET) to join our QA Core Team. In this role, you will partner closely with Infrastructure, Release Engineering, and Development teams to drive high availability, build scalable test automation, and optimize CI/CD release pipelines. You will take ownership of release-based testing, system observability, and automation services to ensure that weekly software releases are delivered seamlessly and with high quality across mission-critical systems. Key ResponsibilitiesTest Automation & System Quality Design, implement, and maintain scalable test automation frameworks for both server- and client-side systems using Java, TestNG, JUnit, and Selenium. Develop and execute comprehensive end-to-end and system test plans to validate new software changes against existing customer deployment patterns. Set up and manage automated test configurations, lab environments, and containerized test setups using Docker. Perform release-based validation, ensuring full compatibility, high availability, and uptime across weekly deployment cycles. CI/CD Pipeline Maintenance & Optimization Troubleshoot, debug, and optimize CI/CD pipelines (Jenkins, GitHub Actions, CircleCI, Buildkite) to reduce test execution times and eliminate flaky builds. Manage and optimize Maven-based build structures, dependency management, and automated release workflows. Write operational automation scripts using Python, Bash, or Gro
NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by phenomenal technology – and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are seeking a Senior Site Reliability Engineer – Storage, you will own the reliability, performance, and scalability of our global NAS, SAN, and Object Storage platforms that power critical internal and external services. You will combine deep storage expertise with strong automation and SRE practices to design, build, and operate highly available storage systems at scale. What you will be doing: Lead design, deployment, and operations of production NAS, SAN, and Object Storage platforms, ensuring reliability, performance, and security. Capture requirements from partner teams, architect storage solutions, and drive end‑to‑end implementation for new and existing services. Develop, maintain, and improve automation for provisioning, configuration, monitoring, incident response, and lifecycle management of storage infrastructure. Participate in on‑call and incident response, lead troubleshooting of complex storage and performance issues, and drive root cause analysis and preventive actions. Define and track SLOs/SLIs and error budgets for storage services, using observability and analytics to continuous
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. About the Team The Professional Services R&D team is a new, dynamic group at the forefront of innovation within Okta. Our mission is to design and build reusable, scalable assets and tools that empower our delivery teams and partners. By making customer engagements more efficient, streamlined, and cost-effective, we directly contribute to our customers' success and accelerate their time-to-value with Okta. This is a unique opportunity to join a strategic team from the ground up and shape the future of Okta's professional services. Position Summary As the DevOps Engineer for the R&D team, you will build and own the infrastructure and automation that enables us to develop and release software with speed and confidence. You will be responsible for creating and managing our CI/CD pipelines, defining our infrastructure as code, and ensuring our deployed assets are scalable, secure, and observable. You will be a key enabler of the team's agility, implementing the tools and processes that allow us to innovate and iterate quickly while maintaining a high bar for quality and reliability. Responsibilities Design, build, and maintain the team's CI/CD pipelines to automate the build, test, and deployment of our software assets. Manage and provision cloud infrastructure using Infrastructure as Code (IaC) principles and tools (e.g., Terraform, CloudFormation). Implement and manage monitoring, logging, and alerting solutions to ensure the health and performance of
Other cities to consider
More places hiring for this role
Get new infrastructure team manager jobs in India by email
Daily job updates · Unsubscribe anytime