Jobs in India

Senior Infrastructure Architect in India

999 active opportunities · Updated October 2026

Explore current senior infrastructure architect jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.

G
📍 Pune, Maharashtra, India
✓ High-confidence listingCompany trend +33.3%
Quick readStrong listing-quality and freshness signals

Location Details: Pune, India At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a hybrid position. You’ll divide your time between working remotely from your home and an office, so you should live within commuting distance. Hybrid teams may work in-office as much as a few times a week or as little as once a month or quarter, as decided by leadership. The hiring manager can share more about what hybrid work might look like for this team. Join our Team Our team builds and operates the foundational infrastructure platforms that power GoDaddy's engineering organization. We own critical services including secrets management, software distribution, host security controls, and live patching for thousands of Linux systems running on OpenStack. This role sits at the intersection of Linux engineering, platform engineering, reliability engineering, and security. You will help define how core infrastructure services are designed, operated, automated, and scaled across the enterprise! What you'll get to do... Design, build, and operate highly available, scalable, and secure infrastructure platforms supporting large-scale Linux environments, with a focus on reliability, resiliency, and operational efficiency Lead the architecture, implementation, and operation of infrastructure services, including OpenStack, enterprise secrets management, package management, software promotion pipelines, and platform lifecycle management Develop and maintain automation solutions using infrastructure-as-code, Ansible, Python, Go, and self-service capabilities to improve efficiency and reduce operational overhead Build and improve observability and reliability practices through monitoring, logging, alerting, dashboards, managing incidents, analyzing underlying causes, disaster recovery, and service health reporting

PythonGitLinuxAI
E(
📍 India· Full-time
✓ Quality checkedCompany trend -93.7%

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are You are an experienced Infrastructure Engineer Engineer who owns backend infrastructure end to end. You design multi-tenant, microservices-based systems that other engineering teams build on, and you make deliberate architectural tradeoffs around consistency, latency, scale, and cost. You are comfortable going deep — service mesh internals, database internals, distributed-systems failure modes — and equally comfortable defining the reliability and security contracts an enterprise AI platform depends on. Responsibilities Design, own, and evolve scalable microservices architectures on Kubernetes across GCP, Azure, and AWS, including multi-tenant isolation (namespaces, network policies, per-tenant resource quotas and RBAC). Build core platform and data-plane components in Golang and Python — data ingestion, knowledge-base indexing and vector/graph search, application connectivity, workflow automation, and ML operations — against explicit latency and throughput SLOs. Own service-to-service communication: gRPC/protobuf API contracts, service mesh (Istio/Linkerd), load balancing, retries, timeouts, and circuit breaking. Make and document architectural tradeoffs — partitioning

PythonSQLAWSAzure
N
📍 Mumbai, India
✓ Quality checkedCompany trend -100%

NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect to join its NVIDIA Infrastructure Specialist Team. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer! What you'll be doing: Primary responsibilities will include building AI/HPC infrastructure for new and existing customers. Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, real-time monitoring, logging, and alerting. Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation, and refinement. Maintain services once they are live by measuring and monitoring availability, latency, and overall system health. Provide feedback to internal teams such as opening bugs, documenting workarounds, and suggesting improvements. What we need to see: BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields. At least 5+ years of professional experience in networking fundamentals, Ethernet or InfiniBand World. Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS, etc. Possess solid working knowl

PythonLinuxAIAnsible
E
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Senior Technical Architect on our Advanced Services team, you will bridge pre-sales strategy and post-sales delivery to architect enterprise data migration, disaster recovery, and hybrid-cloud storage solutions. In this dual consultative and hands-on role, you will collaborate closely with Sales, Engineering, and enterprise customers to translate complex business needs into resilient, automated IT infrastructure. You will drive revenue growth and project success by serving as the trusted technical authority across high-impact data center transformations. WHAT YOU'LL DO Architect & Scope High-Impact Solutions: Partner with Sales, Product Marketing, and customer leadership to evaluate complex enterprise environments, design high-availability storage architectures (replication, clustering, cloud integration), and scope Professional Services engagements that drive sales conversion and solution adoption. Lead Enterprise Data Migrations & Implementations: Own end-to-end technical execution of complex data center consolidations and block/file migrations using specialized appliances (e.g., Cirrus Data) and native toolsets, ensuring seamless cutovers with zero data loss and minimal operational disruption. Automate & Modernize Storage Operations: Develop custom scripts (Python, PowerShell, Shell) and integration workflows to automate replication, disaster recovery, and cloud-bursting processes within client

PythonAWSAzureLinux
W
📍 India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: As a member of the Global Technical Operations (TechOps), you will be a part of a team that focuses on operational reliability within a cloud-based infrastructure. You have hands-on cloud experience in architecting, building, deploying, managing databases, compute instances, and storage buckets. You have a passion for providing solutions through automation. You know that success is through collaboration and communication. What you'll be doing: Work in cross-functional teams to develop solutions and identify opportunities to bring efficiency and effectiveness. Research, evaluate, and incorporate new technologies/concepts into existing frameworks. Proactively identify areas to improve efficiency and effectiveness, recommend and implement solutions towards them. Develop and innovate operational practices, procedures for workflows, and documentation. Implement and contribute to IT security best practices. Automate tasks to ensure consistency and speed of deployment. Identify, analyze, and troubleshoot issues and work towards resolution. Explain technical solutions to bo

PythonGCPKubernetesCI/CD
N
📍 Pune, India
✓ Quality checkedCompany trend -100%

NVIDIA pioneers computer graphics, gaming, AI, and accelerated computing. We are looking for a Senior Solution Architect with full-stack software engineering experience to join our team and play an important role in developing Sales AI applications. This position offers the opportunity to design, build, and evolve solutions that bring generative AI and intelligent workflows into everyday sales experiences. You will work across the application stack and collaborate with Product, AI and machine learning, Data, Security, Solution Architecture, and Engineering teams to deliver secure, reliable, and scalable solutions used globally. What you’ll be doing: Collaborate with application teams to design, develop, and maintain scalable full-stack solutions for enterprise sales workflows. Guide technical solutions across front-end, back-end, APIs, data services, integrations, and cloud infrastructure. Translate product requirements and business needs into secure, maintainable solutions and intuitive user experiences. Integrate generative AI models, AI services, APIs, retrieval systems, and agentic workflows into production applications. Design application architectures that support performance, availability, observability, security, scalability, and long-term maintainability. Lead technical design discussions, compare implementation approaches, make informed architecture decisions, and evaluate emerging technologies. Improve engineering practices for testing, code quality, continuous integration and delivery, monitoring, documentation, and production readiness. Investigate complex issues and develop solutions that improve reliability and user experience. Mentor engineers, share technical knowledge, and contribute to engineering standards and collaborative team practices. What we need to see: <

JavaScriptPythonJavaReact
D
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. What You’ll Do: Architect the Future of AI Infrastructure: You will design, build, and own the end-to-end platform that supports the entire lifecycle of our ML models—from massive-scale distributed training to ultra-low-latency, highly-available inference. Optimize and Serve Cutting-Edge Models: You'll implement and scale sophisticated inference stacks for LLMs using frameworks like vLLM, TensorRT-LLM, or SGLang . You’ll solve complex challenges in throughput, latency, token streaming, and automated scaling to deliver a seamless user experience. Empower AI Innovation: You will act as a strategic partner to our AI Research and Data Science teams. You’ll create a seamless developer experience that accelerates their ability to experiment, fine-tune, and deploy groundbreaking models with velocity and confidence. Automate Everything: You'll develop robust CI/CD/CT (Continuous Training) pipelines using tools like Argo Workflows, ArgoCD, and GitHub Actions to automate model validation, deployment, and lifecycle management, ensuring our systems are both agile and rock-solid. What are we looking for Experience: 5+ years in infrastructure or software engineering, with at least 2+ years laser-focused on MLOps or ML infrastructu

PythonKubernetesCI/CDGit
G
📍 Hyderabad, TELANGANA, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Role: Senior Cloud Engineer Location: Hyderabad, India (Hybrid) Department: Product Development About GHX: GHX (Global Healthcare Exchange) is a leading healthcare technology company on a mission to simplify the business of healthcare and improve patient outcomes. Founded in 2000, GHX has built the GHX Global Network — the world’s largest cloud-based supply chain community connecting healthcare providers, suppliers, distributors, and partners to automate key processes, reduce costs, and increase operational efficiency. Its solutions span electronic trading, procurement automation, inventory and contract management, business intelligence, and data synchronization, helping healthcare organizations improve productivity and focus more on patient care. Over the years, GHX has enabled significant cost savings for the industry and continues to innovate with intelligent automation and AI-driven capabilities. Website: https://www.ghx.com/ LinkedIn: https://www.linkedin.com/company/ghx/ About Role: The Senior Cloud Engineer leads the design, implementation, and operations of the organization’s cloud infrastructure. This role is responsible for complex projects, high-level architectural planning, and making strategic technology decisions that support scalability, performance, security, and cost optimization. Acting as a technical leader, the Senior Cloud Engineer provides mentorship to junior and mid-level engineers while driving innovation, resiliency, and compliance in cloud environments. Key Responsibilities: Design and Implementation Lead the design and implementation of cloud-native architectures and hybrid cloud solutions. Oversee cloud engineering projects, ensuring solutions are secure, resilient, and cost-effective. Contribute to high-level architectural planning and strategic decision-making around technology adoption. Review and approve design proposals, Infrastructure as Code (IaC) tem

AWSAzureGCPKubernetes
N
📍 Bengaluru, India
✓ Quality checkedCompany trend -100%

We are seeking a Senior Software Engineer with strong infrastructure expertise to design, build, and operate the next generation of our enterprise Observability, Automation, and AI-driven Reliability Platform. This role will build highly scalable distributed systems and platform services spanning Storage, Compute, Network, VMware, OpenShift, and bare-metal infrastructure. The engineer will help transform infrastructure operations from reactive monitoring and manual remediation to proactive, predictive, and AI-driven autonomous operations. What You Will Be Doing: Design, build, and operate distributed software platforms for enterprise observability, telemetry, automation, and infrastructure reliability at large scale. Develop reusable platform services, APIs, automation frameworks, and control planes that enable self-service, reduce operational toil, and automate infrastructure operations across multiple engineering teams. Build scalable telemetry and event-processing systems spanning metrics, logs, traces, events, topology, and alerts, with the performance and efficiency to process billions of infrastructure signals. Build intelligent and AI-native reliability capabilities, including agentic workflows for anomaly detection, forecasting, root-cause analysis, automated debugging, and closed-loop remediation. Drive technical architecture and engineering direction across Storage, Compute, Network, and Platform domains, solving complex and ambiguous problems that span multiple teams. Engineer for production at scale, with strong focus on software quality, scalability, security, performance, observability, maintainability, and operational readiness. Provide technical leadership and mentorship, influence engineerin

PythonKubernetesAITerraform
DC
📍 Bengaluru, KARNATAKA, India
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Position Overview: We are looking for a Senior Software Engineer to drive technical excellence, architect complex systems, and elevate our engineering team. You will own critical technical decisions, lead major initiatives from conception to delivery, and set the standard for engineering quality across our products. As a senior engineer, you will architect and lead the development of sophisticated AI-enabled features and infrastructure. This includes designing MCP server architectures, building advanced RAG systems, implementing agentic AI workflows, and establishing patterns that scale across our product portfolio. You will combine deep technical expertise in both traditional software engineering and AI/ML to deliver production-grade solutions. What You'll Do Lead the design and implementation of AI integration infrastructure (MCP servers, orchestration layers, API gateways) Build sophisticated AI features including advanced RAG systems, agentic workflows, and multi-step reasoning Establish AI engineering best practices, security patterns, and quality standards Lead technical initiatives from requirements through production deployment Make critical architectural decisions balancing performance, scalability, cost, and maintainability Design AI evaluation frameworks and implement quality benchmarks Debug and resolve complex production issues across traditional and AI systems Required Qualifications Tech Stack Core: Node.js, React, TypeScript, AWS, PostgreSQL, MSSQL, Docker AI & Integration: Python, MCP, AWS Bedrock, LangGraph/Semantic Kernel, Vector Databases, RAG Core Technical Skills 5+ years professional development with proven track record of delivering complex systems Strong Node.js and JavaScript/TypeScript expertise Advanced React and frontend architecture skills Extensive AWS architecture experience Expert PostgreSQL database design, optimization, and performance tuning Deep understanding of microservices, distributed systems, and

JavaScriptTypeScriptPythonReact
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. About the Role We are looking for Staff System Software Engineer in Test to join our team. In this role, you will be responsible for design, development, automation and reporting of Integration and system tests spanning across firmware and device drivers. This role requires you to have significant technical breadth and deep understanding of low-level system software specifically in server class systems. You will be part of a new team responsible for integration of different system software deliverables and development of system tests spanning all the components. You will contribute to shaping the test strategy , guide best practices and solve complex problems while maintaining a strong hands-on focus. You will partner with development and other QA teams to deliver high quality scalable and reliable solutions. About the Team Integration and system test team is responsible for verification and validation of integrated components across Board management controller (BMC), Firmware and Linux device driver. The team is also responsible for management and maintenance of common tools and pipel

PythonCI/CDGitLinux
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary The Workplace Coordinator is responsible for the day-to-day operation and maintenance of workplace infrastructure, ensuring that all building systems operate efficiently, safely, and reliably. This role oversees HVAC systems, including chillers and AHUs, Building Management Systems (BMS), electrical panels, and coordinates preventive and corrective maintenance activities to support uninterrupted business operations. Key Responsibilities Technical Operations · Monitor and maintain HVAC systems including Chillers, AHUs, FCUs, and ventilation systems. · Operate and monitor Building Management System (BMS) for alarms, trends, and equipment performance. · Inspect electrical LT panels, UPS systems, DG synchronization (if applicable), and power distribution systems. · Monitor critical utilities including temperature, humidity, pressure, and energy consumption. · Ensure uninterrupted operation of critical infrastructure and respond promptly to system failures. Preventive & Corrective Maintenance · Plan and execute preventive maintenance schedules for HVAC and electrical systems. · Coordinate breakdown maintenance with

AILeanProcurementHR
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Debug Validation Engineer will drive post-silicon debug and validation activities for next-generation AI compute silicon and systems. The role is responsible for leading teams focused on identifying, reproducing, analysing and resolving complex silicon, firmware and system-level issues during bring-up, characterization and product readiness. This position combines deep technical debugging expertise with strong cross-functional collaboration across multiple engineering disciplines. The role will work closely with architecture, RTL, firmware, software and systems teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The role will work closely with architecture, RTL, firmware, software, systems and platform teams to improve debug methodologies, accelerate issue resolution and strengthen validation coverage. The Team The Post-Silicon Debug and Validation team sits within the Architecture and Validation organisation and is responsible for bring-up, debug and validation of Graphcore silicon and systems.

PythonGitLinuxAI
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Senior -Power and Performance Validation Engineer About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to senior leadership within Architecture and Validation, the Power and Performance Validation Lead will drive validation strategy and execution for advanced AI compute silicon and systems. The role is responsible for leading power, thermal and performance validation activities across pre-silicon and post-silicon environments to ensure products meet efficiency, reliability and scalability expectations. This role requires strong technical expertise and collaboration across multiple engineering disciplines to deliver robust validation methodologies, scalable automation frameworks and actionable performance insights. The Team The Power and Performance Validation team sits within the Architecture and Validation organisation and is responsible for validating the performance, efficiency and thermal behaviour of Graphcore silicon and systems. The team supports the full product lifecycle, from early architectural modelling through to first silicon bring-up, characterization and production readiness. Engineers work closely with cross-functional teams globally to deb

PythonLinuxAIC++
G
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software, and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers, and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to the Validation leadership team, the Senior Execution and Quality Validation Engineer will be responsible for executing validation plans, developing automation solutions, and improving product quality across Graphcore silicon and platform technologies. Working closely with architecture, design, verification, firmware, software, systems engineering, and validation teams, the successful candidate will contribute to scalable validation methodologies, improve test coverage and execution efficiency, and help ensure products meet high standards of functionality, reliability, and performance before customer deployment. The role requires strong technical skills, attention to detail, and a passion for improving validation quality through effective execution, automation, and continuous improvement. The Team The Validation Execution and Quality team sits within the Validation organisation and is responsible for improving validation effectiveness, test execution efficiency, product quality, and release readiness across Graphcore silicon and platform products. The team develops validation methodologies, automation

PythonCI/CDAIC++
🔔

Get new senior infrastructure architect jobs in India by email

Daily job updates · Unsubscribe anytime