Jobs in India

Lead Infrastructure Software Engineer in India

1,055 active opportunities · Updated October 2026

Explore current lead infrastructure software engineer jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

62/100

steady · 180 related jobs

Hiring trend

+72.7%

Job postings compared with the previous 30 days

Remote options

6.1%

Share of matching jobs listed as remote

RS
📍 Hyderabad, Telangana, India
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full stack automation fabric solutions for mission-critical business processes. With the first SaaS-based composable automation platform specifically built for ERP, we believe in the transformative power of automation. Our unparalleled solutions empower you to orchestrate, manage and monitor your workflows across any application, service or server — in the cloud or on premises — with confidence and control. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for an Engineering Manager, Products & Platforms to join our Product engineering team to lead a high-performing software engineering team while driving technical strategy across Redwood’s Workload Automation Platform. You will be instrumental in mentoring engineers, managing team deliverables, and guiding the design, development, and enhancement of scalable, secure software that powers enterprise data exchange for more than 1,000 customers worldwide. As an Engineering Manager, you will: People Leadership & Mentorship: Manage, coach, and grow a team of talented software engineers, supporting career development, conducting performance reviews, and fostering an inclusive, collaborative team culture. Technical Strategy & Architecture: Provide hands-on technical guidance, participate in design reviews, and define technical roadmaps for Java/Spring Boot microservices while ensuring high standards for architecture, security, and observability. Delivery & Platform Ownership: Oversee team execution, Sprint planning, and delivery timelines to ensure resilient, scalable core platform features and infrastructure. AI Integration: Research and apply AI/ML concepts and their usage to innovate and enhance the MFT (Managed

JavaAWSKubernetesAI
B
📍 India· Full-time
✓ Quality checked

About Us Blueshift is the Intelligent Customer Engagement Platform (CEP), headquartered in San Francisco, that empowers leading B2C brands to drive truly personalized, 1:1 marketing across every channel. Founded by repeat entrepreneurs who previously built Mertado (acquired by Groupon) and were part of the early team at Kosmix (acquired by Walmart), Blueshift leverages AI, including Predictive, Generative, and Agentic AI, to automate customer engagement for clients like ClearScore, LendingTree, Udacity, and U.S. News. Backed by top-tier VCs including Nexus Venture Partners, Storm Ventures, and SoftBank Venture Asia, the company has raised a total of $65 million in venture funding and is consistently recognized as a market leader and a Deloitte Technology Fast 500 award recipient. Blueshift is actively scaling its development center in Pune, India. As part of our team, you will drive innovation in cutting-edge technologies including machine learning, artificial intelligence, big data, and large-scale distributed data systems. This is an exciting career path for motivated individuals looking to build complex, impactful solutions that define the future of customer engagement. AI Solutions Engineer II As a Software Engineer in the AI Solutions team , you occupy a unique techno-functional position. You are not a researcher; you are an implementation specialist and problem-solver . You bridge the gap between our core AI infrastructure and real-world customer impact. You aren't just writing code; you are applying data engineering, analysis, and AI knowledge to help global brands realize the full potential of AI-driven marketing. Responsibilities End-to-End Solution Delivery: Lead the full lifecycle of custom AI projects—from initial customer design and technical architecture to testing and production implementation. Production Stewardship: Take ownership of the "last mile" of delivery. This includes triaging technical tickets, analyzing logs (Datadog/Kibana), and deb

PythonSQLAWSDocker
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug compl

PythonLinuxAIC++
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Debug Validation Lead will drive post-silicon debug and validation activities for next-generation AI compute silicon and systems. The role is responsible for leading teams focused on identifying, reproducing, analysing and resolving complex silicon, firmware and system-level issues during bring-up, characterization and product readiness. This position combines deep technical debugging expertise with strong cross-functional collaboration across multiple engineering disciplines. The role will work closely with architecture, RTL, firmware, software and systems teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The role will work closely with architecture, RTL, firmware, software, systems and platform teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The Team The Post-Silicon Debug and Validation team sits within the Architecture and Validation organisation and is responsible for bring-up, debug and validation of Graphcore silicon and systems. The

PythonGitLinuxAI
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role combines deep technical expertise with people leadership responsibilities, including team development, prioritisation, mentoring and delivery coordination across multiple projects and stakeholders. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debug complex issues, optimize workloads and continuously imp

PythonLinuxAIC++
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Senior -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to deb

PythonLinuxAIC++
P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Create and execute automated and manual test cases to ensure optimal system performance according to specifications. • Work closely with software development and support teams to deliver high-quality applications in a timely manner. • Develop and maintain comprehensive test plans, including manual and automated tests for functional, regression, and integration testing. • Build consensus between stakeholders and developers to define clear and testable acceptance criteria. • Document quality assurance and process flows, both existing and proposed. • Advocate for best quality assurance practices and testing techniques. • Ensure any new software changes meet business, legal, compliance, and technical requirements. • Develop and maintain test automation frameworks for various applications in trading domains. • Maintain and track quality assurance capacity, velocity, statuses, and deliverables with the team lead. WHAT’S REQUIRED • 5+ years of experience in quality assurance with test automation, software engineering, or business analysis roles. • Comprehensive expertise in processing trades, managing trading and lifecycle events, handling corporate actions, and managing cash flows. • Experience supporting fixed income products, Equities and Treasuries. • Solid understanding of position management for listed and OTC products. • Understanding of SDLC, Test lifecycle, and testing methodologies. • Experience creating and writing SQL que

JavaScriptPythonJavaSQL
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Staff -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to debu

PythonLinuxAIC++
O
📍 Bengaluru, India· Full-time
✓ High-confidence listingCompany trend -68.5%
Quick readStrong listing-quality and freshness signals

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services that our customers can trust. We embrace an automation-first mindset and continuously invest in platform engineering, observability, and operational excellence to enable our engineering teams to move quickly and safely. This role is ideal for an experienced Site Reliability Engineer who enjoys solving complex technical challenges at scale, building automation, and improving the reliability of production systems. You will serve as a key contributor within the EPG SRE organization, partnering closely with software engineers, architects, and product teams to design, build, and operate world-class cloud services. What You'll Be Doing Reliability & Operations Design, build, and operate large-scale cloud infrastructure and production services. Participate in an on-call rotation supporting highly available customer-facing systems. Lead incident response efforts and drive post-incident reviews focused on systemic improvements. Define, measure, and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets. Partner with engineering teams to improve service availability, scalability, performance, and resilience. Continuously improve observability through metrics, logging, tracing, dashboards, and alerting. Eng

PythonSQLPostgreSQLMySQL
AC
📍 Chennai, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. This role is open to people who are currently in a product leadership role or someone in an engineering leadership role that is interested to make the switch into product. This role combines experienced product management leadership with the pivotal responsibilities of a Group Product Lead, offering an opportunity to shape the product strategy and the missions of several teams. You will oversee the product management execution working closely with the product managers for several cross-functional software engineering teams around the world. This position reports to the Head of Product for the Platform Engineering Strategic Business Unit, which is responsible for Appian Cloud offering. An excellent candidate will come with a wide range of experience in a product leadership role in the cloud infrastructure domain at a software-as-a-service firm. The product portfolio owned by this position will cover a wide range of cloud native infrastructure and services – from the highly available, resilient, scalable, and efficient infrastructure components on which all Appian Cloud services operate to the the scalable, enterprise-grade managed services that power the backend data persistence and event streams for customer sites in Appian Cloud. You will provide product management leadership for initiatives that deliver efficient and scalable infrastructure, core compute capacity, and managed data planes as multi-tenant services to accelerate the innovations of Engineering teams throughout the department. Your portfolio will include aspects that intersect with AWS accou

AWSKubernetesRestAI
CH
📍 Hyderabad, TELANGANA, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Opportunity Overview: We’re looking for a Manager, Platform Engineering that can lead and grow a high-performing engineering team focused on Developer Experience, DevOps, SRE, and Quality. You will own the systems and processes that enable teams to build, test, release, and operate software with high velocity and reliability, driving engineering efficiency and operational excellence across the organization. What you’ll do: Lead a fast-paced, autonomous team of engineers focused on platform engineering, developer experience, DevOps, SRE, and quality engineering Own and drive the internal developer platform strategy and roadmap, improving how engineering teams build, test, deploy, and operate services Create transparency into engineering efficiency and system health through meaningful metrics across delivery, reliability, and quality Enable teams to move faster by improving CI CD pipelines, environments, tooling, and overall developer workflows Provide technical leadership across platform, infrastructure, and reliability, helping teams build scalable and resilient systems Ensure strong engineering practices across release processes, testing, quality, reliability, and security Define and enforce release guardrails, validation standards, and rollback mechanisms to improve production safety Improve environment stability and consistency across development, QA, and pre production environments Drive test strategy and automation maturity to improve overall product quality and confidence in releases Define and implement observability, monitoring, and alerting standards across systems Improve incident detection, response, and RCA practices, ensuring learnings translate into platform and system improvements Drive cloud infrastructure best practices across AWS, containers, and infrastructure as code Foster a culture of ownership, reliability, and continuous improvement within the team Provide innovative solutions for attracting, developing, and retaining top engineering talent I

AWSAIGoDevOps
O
📍 India· Full-time
✓ Quality checkedCompany trend -68.5%

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. We are looking for a manager to lead the Documentation (Doc) Tools team supporting the Builder Experience department. Reporting to the Information Development Director, you will be leading a team of engineers who support the help.okta.com and developer.okta.com sites as well as the teams who contribute and create content for those sites (Information Development, Builder Advocacy, and Content Strategy). What you’ll be doing Lead a team of engineers responsible for building and maintaining the help.okta.com and developer.okta.com content sites. You’ll lead the adoption of AI across your team’s day-to-day deliverables, integrating it thoughtfully to improve your internal stakeholder’s productivity, quality, and speed. Work closely with the Doc Tools leads to develop and implement our roadmap. Recruit great engineers to accelerate our work. Mentor, retain, and develop engineers as they advance in their own careers. Establishing clear priorities, expectations, and accountability for individuals and team Collaborate with the Doc Tools leads to assist in delivering projects on the Builder Experience roadmap. What you’ll bring to the role Have managed software engineering teams with end-to-end ownership over their technical stack and production performance. Hands-on experience with AI-assisted development tools such as GitHub Copilot, Codex and Claude, with the ability to integrate them effectively into day-to-day engineering workflows. Are excited to b

AWSAzureGCPCI/CD
P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

AI/ML – Investment Services A Career with Point72's AI/ML – Investment Services Team The AI/ML – Investment Services team at Point72 spearheads the development of cutting-edge AI solutions that seek to transform our business processes and enhance enterprise intelligence. The team aims to bridge the gap between business challenges and technological innovation, collaborating with stakeholders across the firm and leveraging expertise in generative AI, data engineering, and machine learning. WHAT YOU'LL DO Build and scale core backend services and platforms that power generative AI applications and data infrastructure used across the firm’s investment workflows Design and implement high-throughput, low-latency data pipelines to ingest, normalize, and serve both structured and unstructured data Develop robust APIs and microservices to support model inference, feature serving, and downstream applications Integrate generative AI tools and model-serving workflows into production, including embedding stores, retrieval components, and fine-tuning pipelines Optimize system performance, cost, and reliability through profiling, capacity planning, and architectural improvements Implement automated testing, continuous delivery pipelines, monitoring, and incident response practices to maintain production health Partner with data scientists, AI engineers, product owners, and operations to translate models and prototypes into scalable, production-grade solutions Mentor engineers, lead code reviews, and establish engineering best practices for maintainability, security, and observability Own end-to-end delivery, operational runbooks, and metrics-driven measurement of feature impact and system reliability WHAT'S REQUIRED Bachelor’s degree in computer science, software engineering, or a related technical field Minimum 5+ years of professional experience building backend systems and production services Demonstrated experience designing and operating large-scale data engineering pipelines

PythonJavaKubernetesCI/CD
J
📍 India· Full-time· Remote
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Backend Engineer (Senior Level) - SDE IV We're looking for a Senior Backend Engineer to lead the architecture and evolution of backend services that deploy and serve machine learning models in production. You'll work closely with ML Engineers, Platform, and Product teams to build scalable, reliable systems and drive technical direction across multiple teams. What You’ll Do Design and drive the long-term architecture of backend services for biometrics and ML model serving. Collaborate with core platform and backend teams on organization-wide architectural initiatives. Partner with business and engineering teams to design and deliver cross-cutting platform capabilities. Lead architectural reviews, mentor engineers, and promote engineering best practices. Build and maintain backend services for deploying and serving ML models Monitor service reliability, performance, and scalability in production Deploy and operate services on AWS using ECS + Fargate, SageMaker, or EC2 + Kubernetes Support real-time and batch inference workflows Contribute to CI/CD pipelines and deployment automation What We’re Looking For Strong expertise in backend development using Java and working knowledge of Python. Experience mentoring engineers and driving architectural decisions. Working knowledge of Python, especially for ML-related workflows Hands-on experience with AWS (e.g., DynamoDB, ECS, EC2, Redis, S3, SageMaker) Familiarity with Terraform or other infrastructure-as-code tools, and experience with CI/CD and production monitoring Experience with observability tools (Datadog, New Relic, etc.) Experience with containers and orchestration (Docker, ECS, etc.) Understanding of how ML models are deployed and served in production Experience with Kubernetes Nice to Have Experience with MLOps or ML platform engineering. Experience with asynchronous programming and event-driven systems. Jumio Values: IDEAL: Integrity, Diversity, Empowerment, Accountability, Leading Innovation Equal Opportunities :

PythonJavaRedisAWS
S
📍 Pune, Maharashtra, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

SonicWall is a cybersecurity forerunner with more than 30 years of expertise and is recognized as a leading partner-first company, ensuring our partners and their customers are never alone in the fight against cybercrime. With the ability to build, scale and manage security across the cloud, hybrid and traditional environments in real-time, SonicWall provides relentless security against the most evasive cyberattacks across endless exposure points for increasingly remote, mobile and cloud-enabled users. With its own threat research center, SonicWall can quickly and economically provide purpose-built security solutions to enable any organization—enterprise, government agencies and SMBs—around the world. For more information, visit www.sonicwall.com or follow us on Twitter , LinkedIn , Facebook and Instagram . As a Software Dev Senior Engineer , you will own the reliability, scalability, and operational excellence of our Cloud-based services. You will define and enforce reliability standards, drive the adoption of SRE practices across engineering teams, and build the systems and tooling that keep our production infrastructure healthy. We follow a DevOps model: Development and Operations teams are integrated, and the SRE function acts as the reliability layer — setting Service Level Objectives, managing error budgets, and continuously reducing toil through engineering. Key Responsibilities: Define, publish, and continuously refine Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs ) for all critical services, partnering with product and engineering leadership. Own the error budget framework: track consumption, enforce error budget policies, and drive reliability investments when budgets are at risk. Lead the design and implementation of comprehensive observability platforms — metrics, structured logging, and distributed tracing — to ensure full visibility into pro

PythonSQLPostgreSQLMongoDB
🔔

Get new lead infrastructure software engineer jobs in India by email

Daily job updates · Unsubscribe anytime