Jobs in India

Infrastructure Team Manager in India

669 active opportunities · Updated October 2026

Explore current infrastructure team manager jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.

P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

JOB TITLE Observability Engineer A CAREER WITH POINT72'S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. WHAT YOU’LL DO Design observability capabilities that give engineering teams clear insight into application health, platform performance, and issues affecting users Build scalable collection pipelines for metrics, logs, and traces across cloud-based and on-premises environments Develop actionable alerting standards that reduce noise, shorten incident response, and highlight the most important signals Partner with application and infrastructure teams to define service health indicators and improve operational readiness before production launches Automate monitoring configuration, dashboard deployment, and reliability checks to support consistent observability across the technology environment Analyze production incidents to identify telemetry gaps and improve detection, diagnosis, and recovery Create dashboards and reporting views that help teams understand trends, capacity risks, and reliability outcomes Establish practical observability

PythonAgileAIHR
P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

JOB TITLE Site Reliability Engineer A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO You will play a highly critical operational role where you will apply a combination of software and systems engineering skills to develop and maintain a complex set of distributed, real-time systems that serve critical stakeholders in Point72’s Global Macro business. You will focus on optimizing the operations of existing systems and infrastructure in an efficient manner, through a strict adherence to automation and tooling Specifically, you will: Build out foundational technical components of an extensive SRE program across multiple complex systems, both new and existing • Collaborate with our development and quant teams to ensure that ongoing change is consistent with a pre-determined, measurable set of SLOs spanning multiple complex user interactions with our systems • Monitor system capacity and performance, identifying and addressing potential future bottlenecks and sources of instability before they become impactful to our stakeholders • Review and provide feedback on automation code developed by peers to maintain high standards of code quality and efficiency • Troubleshoot and resolve system issues, analyzing their impact on infrastructure and service operations • Participate in or lead design reviews with peers and stakeholders, evaluating and selecting the best technologies and automation strategies for our needs WHAT’S REQUIRED We are looking for highly motivated, proactive engineers

PythonAWSDockerKubernetes
P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

JOB TITLE User Access Certification Analyst A CAREER WITH POINT72'S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting with discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU'LL DO Develop and scale the firm’s user access certification program, driving measurable improvements in access hygiene and enforcement of least privilege Support execution of periodic access certifications across enterprise systems, assisting with multiple reviews at different stages simultaneously Analyze and validate large access-control datasets to identify excessive, inappropriate, privileged, or high-risk access Perform data quality assurance and reconciliation to ensure completeness, accuracy, and consistency of access and entitlement data before and during certifications Track certification progress and remediation status using metrics and dashboards, and deliver data-driven insights to accelerate closure Investigate data discrepancies and anomalies, coordinate with certifiers, application owners, technical subject matter experts, and development teams to drive timely resolution Validate remediation and deprovisioning outcomes through targeted data verification and reconciliation activities Produce clear, audit-ready reporting and artifacts on certification results, access risk, and remediation outcomes for internal and external audits Partner with system and application owners to interpret access data and support decisions that reduce access risk and uphold least privilege principles Innovate certification lifecycle processes by applying automation, reporting i

SQLAgileAIExcel
BA
📍 India· Full-time
✓ Quality checkedCompany trend -70%

About Bolna Bolna is Voice AI infrastructure built for India - and now for the world. We help businesses deploy intelligent voice agents that can call, converse, and convert in any language, at scale. From collections to customer support to sales, our agents handle millions of conversations so humans don’t have to. We’re a YC F25 company, backed by General Catalyst, with 1,050+ paying customers and growing fast. Our team of ~25 is based in Bengaluru. The Role Every voice AI agent Bolna deploys makes real-time judgment calls - when to speak, when to go silent, when a customer is done talking, when to hand off. We’re building automated systems to grade these calls at scale, using LLMs as judges of call quality. But before you trust a model’s judgment, you verify it against a human’s. That’s this role. You’ll listen to real calls, annotate what actually happened, and check whether our automated systems - LLM-as-judge evals and quantitative signal detection - got it right. It’s precise, high-attention work, and it sits right at the center of how we know our voice agents are actually working. This is an internship role for someone early in their career who wants hands-on exposure to how a voice AI company builds trust in its own AI. What You’ll Do Annotation Listen to and annotate real customer calls - transcription review, issue tagging, labeling - using tools like Label Studio Follow (and help sharpen) annotation guidelines for a multilingual environment (Hindi, English, Hinglish, ) Verifying LLM-as-Judge Evaluations For calls flagged by our automated eval pipeline, verify whether the model’s call was actually correct - for example, confirming whether a detected barge-in (agent/customer talking over each other) genuinely happened by listening to the audio Mark agreements and disagreements clearly, with reasoning, so we can measure and improve model accuracy over time All tools needed for this will be provided Verifying Quantitative Measures Check system-flagged quantit

SQLAIGoRust
BA
📍 India· Full-time
✓ Quality checkedCompany trend -70%

About Bolna Bolna is Voice AI infrastructure built for India. We help businesses deploy intelligent voice agents that can call, converse, and convert in any language, at scale. From collections to customer support to sales, our agents handle millions of conversations so humans don't have to. We're a YC F25 company, backed by General Catalyst, with 1,050+ paying customers and growing fast. Our team of ~25 is based in Bengaluru. The Role You're the person who makes our voice agents actually sound good and actually work for real customers. As an AI Solutions Engineer, you'll sit at the intersection of our product and our customers. Your primary job is to design, write, and iterate on the prompts and tools that power Bolna's voice agents, making them smarter, more natural, and more effective for each use case. You'll work closely with customers to understand their goals, build agent flows, test conversations, and fix what breaks. No heavy coding required-if you can vibe-code a basic script or write a solid system prompt, you're qualified. What You'll Do Write, test, and iterate on system prompts for voice agents across industries like D2C, fintech, healthcare, and logistics Listen to real call recordings, identify where agents fail, and fix them Build conversation flows and call pathways for new customer deployments Help onboard new customers-understand their use case, set up their agent, and get it live Maintain a growing library of prompts, templates, and best practices across Bolna's verticals Red-team agents-try to break them, find edge cases, and make them bulletproof Work with vernacular inputs: test agents in Hindi, Hinglish, and regional languages Feed insights back to product and engineering-you'll see what customers need before anyone else does What We're Looking For Must-have: You've spent serious time prompting ChatGPT, Claude, or similar LLMs-not just casually, but to actually build or solve something You're obsessive about language-you notice when a senten

S
📍 Bengaluru, KARNATAKA, India· Full-time
✓ High-confidence listingCompany trend -85.7%
Quick readStrong listing-quality and freshness signals

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. Bridge, a Stripe company, is a rapidly growing business and the number of developers integrating our APIs is growing quickly. We provide Slack-based support to our developers on a wide range of topics in close partnership with our internal engineering team. We’re a small but mighty team that’s expanding global coverage and is looking to make our first Product Support Specialists hires in Mexico City. About the team In this role, you’ll be working directly with developers integrating Bridge APIs and helping them resolve their issues. You will take ownership of complex, technical user issues and work across teams to resolve them. As part of the team, you’ll have a big impact to grow the Product Support operation and enhance various aspects such as capacity planning and forecasting, operational tools and systems, workflow optimization and automation, metrics and reporting, quality control, and more. What you’ll do Stripe is launching Stripe Delivery Centers - a brand new global team to design, implement and grow Stripe’s operations for the next decade. We are looking for dynamic and curious people that have a passion for solving global user issues, building operations, driving process improvements and want to play a front-line role in building this new operational capability for Stripe and accelerating Stripe’s growth. If you like challenging, scaled problems and are an amazing teammate, we want to hear from you! Responsibilities Analyze and troubl

SQLAIExcelHR
S
📍 Bengaluru, KARNATAKA, India· Full-time
✓ High-confidence listingCompany trend -85.7%
Quick readStrong listing-quality and freshness signals

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Stripe was built with simplicity in mind. We strive to deliver frictionless experiences for all of our users, whether they are an Independent Business, Startup, SMB, or Enterprise and our mission is to provide all Stripe Users with the best support experience possible. Today, Stripe handles over a million support cases per year and processes millions of internal transactions. We’re going to achieve excellence by thinking of support in a novel, solution-oriented way, and viewing operations as an integral enabler of all of Stripe’s growth. Stripe is launching Stripe Delivery Centers - a brand new global team to design, implement and grow Stripe’s operations for the next decade. We are looking for dynamic and curious people that have a passion for solving global user issues, building operations, drive process improvement and want to play a front-line role in building this new operational capability for Stripe and accelerating Stripe’s growth. If you like challenging, scaled problems and are an amazing teammate, we want to hear from you! What you’ll do The credit risk operations team plays a critical role in ensuring a healthy financial ecosystem for businesses around the world. This team will directly impact the company’s bottom line and growth capabilities by supporting new and emerging businesses. The Credit Operations team is responsible for conducting credit risk assessments, underwriting small and mid-size users, and proposing

GoExcelAccountingFinance
AG
📍 Ahmedabad, Gujarat, India· Full-time
✓ Quality checkedCompany trend -82.5%

We are seeking an experienced Business Development Professional with around 5-8 years of expertise in the Infrastructure Industry to join our team in Ahmedabad, Gujarat. The ideal candidate will be responsible for driving business growth, fostering strategic partnerships, and expanding our market presence in alignment with Adani's core values and culture. Source: Adani Group | Job ID: 50943

S
📍 India· Full-time
✓ Quality checkedCompany trend -85.7%

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Our mission is to enable effective financial decisions through reliable data, increased efficiency, and automation. We support Marketing, Sales, Seller, Accounting, Tax, Finance and Strategy (F&S), Finance Operations (FinOps), and Treasury functions across automation, data insights, and process improvements. You may work on a wide variety of critical business areas including Seller Systems — Responsible for building the systems and tooling that make sellers and internal revenue teams at Stripe dramatically more productive and effective. We partner with Sales, Finance, Legal, and Product to deliver a single "plane of glass" selling experience that spans deal creation and modeling, negotiation and approvals, contracting, onboarding, and activation. The Seller Systems team composes first‑party, custom Stripe components with best‑in‑class third‑party business systems to deliver configurable, auditable, and globally scalable workflows. Engineers on Seller Systems build services, APIs, integrations, data pipelines, and internal UIs that power seller productivity, reduce time‑to‑activation, improve deal velocity, and enable AI‑driven assistive workflows. Use cases include deal modeling and pricing engines, approval and orchestration platforms, CLM/CPQ integrations, onboarding automation, seller analytics, and AI‑assisted seller tooling. Finance Engineering — Responsible for building the robust and scalable infrastructure that powers

SQLMySQLAWSDocker
P
📍 India· Full-time
✓ Quality checkedCompany trend -90%

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is now at one of the most consequential inflection points in its history. Enterprises are rapidly shifting from human-driven API workflows to multi-agent systems — where AI agents autonomously discover, call, and collaborate with APIs, tools, CLIs, and each other. Postman already owns the API design, governance, and lifecycle layer for the enterprise; the next frontier is owning the runtime layer for how agents actually interact with all of it . Where Postman today is the API platform for human-first integration , we are building Postman into the agent interface fabric for AI-first integration . That transformation starts with Fabric Gateway — a brand new product, being built from scratch, that will serve as the policy-driven control plane governing how agents connect to APIs, services, and other agents at scale. This is a rare 0-to-1 opportunity inside a company with 40+ million developers already in the ecosystem. About the Team The Fabric Gateway team is building the API gateway for the AI era — one of Postman's most ambitious greenfield infrastructure products. We're a small, high-ownership team wo

KubernetesRestAIGo
GR
📍 Gurugram, Haryana, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Description: Graviton Research Capital LLP, Gurgaon is looking to hire Software Engineers for our Core Technology team which has some of the best programmers in India working on cutting edge technologies to build a super fast and robust trading infrastructure handling millions of dollars worth of trading transactions every day. As a Senior Software Engineer with Graviton your responsibilities will include: Designing and implementing a high-frequency automated trading system, that trades on multiple exchanges Building live reporting and administration tools for the trading system Performance optimization and improving the overall latency of systems, through algorithm research and using cutting edge tools and techniques End-to-end ownership of modules, including designing, development, deployment and support Growing the team through involvement in the regular hiring process and occasional campus recruitments Requirements : The ideal requirements for our candidates are: A degree in Computer Science 3-5 yrs Experience with C/C++ and object-oriented programming Experience in HFT industry Expertise in algorithms and data structures Excellent problem solving skills Strong communication skills A working knowledge of Linux systems Any of the following is a plus: A good understanding of TCP/IP and Ethernet Knowledge of any other programming language e.g. Java, Scala, Python, bash, Lisp, etc. Familiarity with parallel programming models and parallel algorithms Experience with big data environments e.g. Hadoop, Spark etc. Benefits: Our open and collaborative work culture gives you the freedom to innovate and experiment. Our cubicle free offices, non-hierarchical work culture and insistence to hire the very best creates a melting pot for great ideas and technological innovations. Everyone on the team is approachable, there is nothing better than working with friends! Our perks have you covered. Competitive compensation Annual international team outing Fully covered commuti

PythonJavaLinuxC++
P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

AI/ML – Investment Services A Career with Point72's AI/ML – Investment Services Team The AI/ML – Investment Services team at Point72 spearheads the development of cutting-edge AI solutions that seek to transform our business processes and enhance enterprise intelligence. The team aims to bridge the gap between business challenges and technological innovation, collaborating with stakeholders across the firm and leveraging expertise in generative AI, data engineering, and machine learning. WHAT YOU'LL DO Build and scale core backend services and platforms that power generative AI applications and data infrastructure used across the firm’s investment workflows Design and implement high-throughput, low-latency data pipelines to ingest, normalize, and serve both structured and unstructured data Develop robust APIs and microservices to support model inference, feature serving, and downstream applications Integrate generative AI tools and model-serving workflows into production, including embedding stores, retrieval components, and fine-tuning pipelines Optimize system performance, cost, and reliability through profiling, capacity planning, and architectural improvements Implement automated testing, continuous delivery pipelines, monitoring, and incident response practices to maintain production health Partner with data scientists, AI engineers, product owners, and operations to translate models and prototypes into scalable, production-grade solutions Mentor engineers, lead code reviews, and establish engineering best practices for maintainability, security, and observability Own end-to-end delivery, operational runbooks, and metrics-driven measurement of feature impact and system reliability WHAT'S REQUIRED Bachelor’s degree in computer science, software engineering, or a related technical field Minimum 5+ years of professional experience building backend systems and production services Demonstrated experience designing and operating large-scale data engineering pipelines

PythonJavaKubernetesCI/CD
N
📍 Bengaluru, India
✓ Quality checkedCompany trend -100%

We are seeking a highly skilled and experienced Staff Network Site Reliability Engineer (SRE) to join our Enterprise Network Operations and SRE team. In this role, you will be pivotal in implementing our vision for a reliable and efficient network infrastructure. The ideal candidate is passionate about network operations and committed to enhancing the user experience. You'll have the opportunity to solve complex network challenges using hands-on debugging and by focusing on network automation, observability, documentation, and operational excellence. This is a critical position focused on ensuring user satisfaction and brilliance in network operations. What you'll be doing: Owning the operational aspect of the network infrastructure, ensuring its high availability and reliability, actively working on network incidents and service requests. Partnering with architecture and deployment teams to guarantee that new implementations are supportable and align with production standards. Advocating for and implementing automation to reduce toil and improve operational efficiency. Minimizing manual operational tasks to achieve and maintain Service Level Objectives (SLOs). Monitoring network performance, identifying areas for improvement, and collaborating with relevant teams to implement refinements. Proactively identifying and mitigating network risks to promote continuous improvement. Collaborating with domain experts across functions to resolve production issues swiftly and effectively, ensuring customer happiness. Conducting blameless postmortems and following through on Root Cause Analyses (RCAs). Discovering opportunities for operational improvements and teaming up with colleagues to devise solutions that enhance excellence and sustainability in network operations. Developing knowledge base articles for automa

PythonLinuxAnsible
TA
📍 India· Full-time· Remote
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the Role Together AI runs one of the largest GPU fleets in the world. The Infra Agent Systems team builds the software systems that power and automate that infrastructure. We develop production AI agents that diagnose hardware failures, investigate incidents, correlate signals across the fleet, and automate operational workflows. Alongside these agents, we build the platform they run on, including knowledge graphs, retrieval systems, orchestration frameworks, and developer tooling. You’ll work across two areas: Infrastructure Agent Systems — Build production AI agents that help operate our GPU fleet by diagnosing failures, investigating incidents, gathering evidence from live systems, and assisting with remediation. These agents are used every day by our infrastructure and datacenter teams through APIs, CLI, dashboards, and Slack. Core Agent Platform — Build the platform that powers these agents, including knowledge graphs, search and retrieval, orchestration, evaluation, and the tooling that enables agents to reason, act, and continuously improve. We’re working on something that hasn’t really been done before: building knowledge graphs and self-improving AI agents that understand, operate, and continuously improve large-scale AI infrastructure. This is an opportunity to work at the intersection of AI agents, distributed systems, infrastructure, and automation , solving challenging engineering problems with real production impact. There’s an enormous amount to build, learn, and shape as we define the future of autonomous infrastructure. responsible for delivering the software but also for operating and supporting it in production. Why this Role You’ll work on two hard problems at the same time: making AI agents trustworthy enough to operate production infrastructure, and building the knowledge, retrieval, and distributed systems that make those agents effective. You’ll have the opportunity to build foundational systems from the ground up, work on infrastructur

TypeScriptPythonKubernetesGit
P
📍 Bengaluru, India· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

JOB TITLE IT Operations Engineer, Application Support A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Provide technical support for software applications and investigate, diagnose, and resolve application issues • Automate start-of-day and end-of-day checks for key applications. • Log and track incidents across applications in the production environment. • Implement monitoring and automation initiatives and develop custom solutions using Python, Shell, and/or Powershell scripts. • Create tactical support tools and scripts to improve the incident investigation process and enable • transparency into potential business impacts. • Prioritize and categorize incidents based on severity and impact. • Collaborate with the development team to improve applications based on user feedback. • Create and maintain documentation for responding to common errors and application incidents. • Assist with software applications deployment and configuration . • Provide training and assistance to users to ensure effective use of applications and systems. • Develop knowledge base resources to empower users to independently resolve common problems. WHAT’S REQUIRED • Bachelor's degree in computer science, information technology, or a related field. • Experience supporting middle- and back-office applications created in .Net, Java, C# etc. Ability to debug apps using of code, logs, alerts etc. • Literacy in complex SQL procedures/queries. • Ability to diagnose and troubleshoot technical issues.

PythonJavaSQLAgile
🔔

Get new infrastructure team manager jobs in India by email

Daily job updates · Unsubscribe anytime