Jobiba hiring network

Engineering Manager Platform Reliability Salary India Jobs

8,358 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current engineering manager platform reliability salary india jobs. Use filters to narrow by work mode, employment type, experience and date posted.

S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team Stripe's Financial Connections group is building technology that expands the scope of problems we tackle beyond card payments. We aspire to enable individuals and businesses to securely share their financial data to improve their product experiences. With more high-quality customer financial data, both Stripe and our users can offer better products to their customers. We have ambitious goals to enable all global businesses to run their entire financial life cycle on Stripe, and we want you to be part of our story. Technical Operations roles in Financial Connections are a dynamic and critical component of the organization's success. As Financial Connections TechOps Manager, you will lead a team of Integration Reliability Engineers who sit at the intersection of product and platform engineers, financial partners, and AI automation. Your role is to ensure that nothing is lost in translation: setting the standard for how the team operates, where it invests, and how the operational backbone of Stripe's open banking ecosystem stays reliable at scale. What you'll do Lead integration reliability and partner operations Own the health of Financial Connections' banking integrations at the team level—define what "reliable" looks like across the partner ecosystem, set the operational bar, and hold the team to it Serve as the management escalation point for complex partner issues, coordinating across financial partners, Stripe engineering, and product lea

aigorust
View job →

Pre-sales Manager, Inference and Agentic AI - Paytm Location: Noida Company: Paytm About Paytm: Paytm is a pioneer of digital payments in India, serving over 450 million consumers and 45 million merchants across payments, financial services, and commerce. Over the years, Paytm has built deep in-house capabilities across technology, data, and operations to operate at scale with high reliability. Paytm is building a full stack AI platform focussed on Inference and Agents, enabling large enterprises to deploy AI driven automation across sales, service, operations, and analytics. The Inference and Agentic AI team operates as a cross functional unit spanning engineering, product, data science, business management, and sales, and owns the full lifecycle of AI solutions from opportunity discovery to deployment and scale. Pre-Sales Manager | Paytm AI Key Responsibilities 🔹 Partner with Enterprise Sales teams to understand client requirements and map them to Paytm AI solutions. 🔹 Conduct discovery sessions, identify business challenges, and create solution recommendations. 🔹 Deliver customized product demos, solution walkthroughs, and Proof of Concepts (POCs) for enterprise clients. 🔹 Act as a technical advisor, addressing product, integration, architecture, and implementation-related queries. 🔹 Create solution documents, proposal inputs, RFP responses, technical FAQs, and sales enablement assets. 🔹 Collaborate with Product, Engineering, Onboarding, and Business teams to ensure seamless solution delivery. 🔹 Capture client feedback and provide insights to improve product capabilities and customer experience. 🔹 Support enterprise deal closures through strong solutioning, stakeholder management, and technical consulting. Ideal Candidate ✔ 3–6 years of experience in Pre-Sales, Solution Consulting, Solutions Engineering, Product Consulting, or Enterprise Technology roles. ✔ Strong understanding of APIs, integrations, SaaS platforms, enterprise applications, and solut

Client Onboarding Manager – Inference & Agentic AI | Paytm (Noida) About the Role Paytm is a pioneer of digital payments in India, serving over 450 million consumers and 45 million merchants across payments, financial services, and commerce. Over the years, Paytm has built deep in-house capabilities across technology, data, and operations to operate at scale with high reliability. Paytm is building a full stack AI platform focussed on Inference and Agents, enabling large enterprises to deploy AI driven automation across sales, service, operations, and analytics. The Inference and Agentic AI team operates as a cross functional unit spanning engineering, product, data science, business management, and sales, and owns the full lifecycle of AI solutions from opportunity discovery to deployment and scale. Key Responsibilities Own client onboarding from sales handover to go-live. Understand client workflows, systems, and integration requirements. Coordinate with Product, Engineering, and Client teams for seamless deployment. Manage onboarding timelines, milestones, and stakeholder communication. Conduct client training sessions and drive product adoption. Track onboarding KPIs, client satisfaction, and implementation success. Gather client feedback and support continuous product improvements. Ideal Candidate 2–5 years of experience in Client Onboarding, Implementation, Customer Success, or Solutions Engineering. Experience in SaaS, Fintech, Enterprise Technology, or AI products preferred. Good understanding of APIs, integrations, CRM systems, and enterprise workflows. Strong project management, problem-solving, and stakeholder management skills. Excellent communication and client-facing abilities. Bachelor’s degree in Engineering, Business, or related field. Location: Noida Why Join? Be part of Paytm’s fast-growing AI business and work closely with enterprise clients to deliver cutting-edge AI-driven automation solutions at scale.

Location Details: Pune, India At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a hybrid position. You’ll divide your time between working remotely from your home and an office, so you should live within commuting distance. Hybrid teams may work in-office as much as a few times a week or as little as once a month or quarter, as decided by leadership. The hiring manager can share more about what hybrid work might look like for this team. Join our Team Our team builds and operates the foundational infrastructure platforms that power GoDaddy's engineering organization. We own critical services including secrets management, software distribution, host security controls, and live patching for thousands of Linux systems running on OpenStack. This role sits at the intersection of Linux engineering, platform engineering, reliability engineering, and security. You will help define how core infrastructure services are designed, operated, automated, and scaled across the enterprise! What you'll get to do... Design, build, and operate highly available, scalable, and secure infrastructure platforms supporting large-scale Linux environments, with a focus on reliability, resiliency, and operational efficiency Lead the architecture, implementation, and operation of infrastructure services, including OpenStack, enterprise secrets management, package management, software promotion pipelines, and platform lifecycle management Develop and maintain automation solutions using infrastructure-as-code, Ansible, Python, Go, and self-service capabilities to improve efficiency and reduce operational overhead Build and improve observability and reliability practices through monitoring, logging, alerting, dashboards, managing incidents, analyzing underlying causes, disaster recovery, and service health reporting

pythongitlinux
View job →

Senior Product Manager, Robotics & Autonomy What we're doing isn't easy, but nothing worth doing ever is. At Diligent Robotics, we envision a future powered by robots that work seamlessly with human teams. We build artificial intelligence that enables service robots to collaborate with people and adapt to dynamic human environments. Our robots operate every day in hospitals, helping healthcare staff spend less time on routine work and more time caring for patients. Operating a real-world fleet gives us something few robotics companies have: continuous customer feedback and operational data that directly shapes the next generation of Physical AI. We're looking for a Senior Product Manager, Robotics & Autonomy to define and execute the product strategy for some of the most critical capabilities in our robotics platform. You'll work at the intersection of robotics, autonomy, AI, and software engineering to translate business priorities, customer needs, and technical opportunities into a clear product roadmap that drives measurable outcomes. This role is ideal for someone who understands complex autonomous systems and enjoys working alongside world-class engineers to bring ambitious technology from concept into production. Responsibilities Own the product strategy and roadmap for key Robotics and Autonomy initiatives, balancing customer impact, technical feasibility, and long-term platform investments. Define product requirements for autonomy, navigation, perception, fleet intelligence, simulation, and robotics platform capabilities. Partner closely with Engineering, AI, Robotics, Customer Success, Operations, and Leadership to align priorities across the organization. Translate customer feedback, fleet telemetry, and operational insights into product decisions that improve robot performance, reliability, and user experience. Prioritize investments using data, customer value, technical complexity, and business impact. Drive cross-functional execution from concep

agilemachine learningai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Role As a Sales Manager, Energy, you will build and lead a team of Account Directors focused on strategic growth across utilities, oil and gas, renewables, power generation, and energy services. The team will partner with complex organizations modernizing operations, improving reliability, accelerating the energy transition, and adopting enterprise AI responsibly at scale. You’ll help the team navigate regulated enterprise sales cycles, deepen relationships with business, technology, operations, engineering, security, and risk leaders, and drive adoption of OpenAI’s platform across safety-conscious, asset-intensive organizations. Key Responsibilities Recruit, develop, and lead a high-performing team of Energy Account Directors. Create a strong coaching culture through deal reviews, account strategy sessions, ride-alongs, and structured 1:1s. Define the Energy GTM strategy, including subsector segmentation, account prioritization, partner strategy, executive engagement, and territory planning. Drive disciplined pipeline generation, forecast accuracy, and operational rigor. Guide multi-stakeholder opportunities involving operations, engineering, digital, data, security, legal, risk, procurement, and executive leadership. Help customers translate AI and API capabilities into measurable outcomes across asset and field operations, grid and generation planning, engineering knowledge, customer service, commercial workflows, and enterprise productivity. Partner with Product, Solutions Architecture, Technical Success, Legal, Security, Finance, and policy experts to support responsible deployment. Provide structured feedback on customer requirements, integration blockers, reliability and governance needs, and emerging industry trends. What We’re Looking For 15+ years of enterprise sales, GTM, or sales leadership experience. Proven experience building and scaling enterprise sales teams responsible for complex strategic accounts and large revenue targets. Deep underst

awsgitrest
View job →
O
1mo ago

About the Team The compute infrastructure team runs the GPU fleet and large-scale compute clusters that serve the models backing ChatGPT and the API, while also supporting training workloads for our next generation models. We operate a large, modern GPU fleet and provide a unified platform for other OpenAI teams to seamlessly run production Applied AI and Research training workloads. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role You will be part of an engineer-first TPM team as a Technical Program Manager for Compute Infrastructure who owns the end-to-end delivery of large-scale GPU clusters, partnering with engineers to bring clusters online across external providers and partners. You’ll run a broad, parallel portfolio spanning hardware, networking, power, and cooling—driving execution, risk management, and crisp alignment from working teams through leadership to deliver production-ready capacity at scale. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead end-to-end delivery of both New Compute SKUs and large-scale GPU clusters across an external partner ecosystem while supporting capacity planning for training and inference. Ability to contextually drive multi-threaded bring-up programs spanning hardware, networking, power, and cooling—owning plans, dependencies, and critical paths. Interface with chip providers to derisk long-term onboarding to new hardware platforms by working across kernels, comms, hardware, and scheduling engineering teams. Build and operationalize program mechanisms (roadmaps, milestones, risk registers, runbooks) that make delivery predictable at massive scale. Partner with engineering to improve cluster turn-up reliability, repeatability, and automation

awsrestai
View job →

About Paytm Paytm is a pioneer of digital payments in India, serving over 450 million consumers and 45 million merchants across payments, financial services, and commerce. Over the years, Paytm has built deep in-house capabilities across technology, data, and operations to operate at scale with high reliability. Paytm is building a full stack AI platform focussed on Inference and Agents, enabling large enterprises to deploy AI driven automation across sales, service, operations, and analytics. The Inference and Agentic AI team operates as a cross functional unit spanning engineering, product, data science, business management, and sales, and owns the full lifecycle of AI. Role Overview Paytm is looking to hire Sales Operations Managers to drive financial and operational rigor across its AI Inference and Agentic AI business. This role sits at the core of sales operations, working closely with sales, business management, and central finance teams to ensure accurate billing, collections, and revenue recognition. The role involves owning the full order to cash lifecycle, strengthening revenue assurance, and managing procurement and vendor operations for the AI charter. The candidate will play a key role in building scalable, audit ready systems that improve financial control, reduce leakage, and enable efficient business growth. Key Responsibilities Order to Cash Operations Own end to end invoicing for enterprise AI deals from contract trigger to invoice generation, dispatch, and acknowledgement. Maintain a central invoicing tracker covering deal terms, billing milestones, invoice status, and collections. Coordinate with central finance and accounts receivable teams to ensure GST compliant, PO aligned, and accurately booked invoices. Drive collections follow ups with enterprise clients in partnership with business teams and escalate overdue receivables. Manage billing adjustments including credit notes, disputes, and corrections. Revenue Assurance and Financial Co

Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler’s high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You’ll Do (Role Expectations) Maintain h

pythonkuberneteslinux
View job →
O
26 days ago

About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.

awsazurerest
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
29 days ago

About the Team The Statsig team is responsible for the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Teams across ChatGPT, Codex, model measurement, monetization, business subscriptions, developer products, and shared infrastructure rely on Statsig to introduce capabilities safely, measure their impact, and make high-confidence product decisions. About the Role As a Product Lead on the Statsig team, you will define how experimentation, rollout, configuration, and analytics become a simple, reliable, and trusted part of how every OpenAI product team ships. You will set strategy across multiple product and platform workstreams, translate company-wide needs into durable capabilities, and help Statsig become a core part of OpenAI’s product development system. We’re looking for a product leader who combines strong product judgment, technical fluency, and deep analytical thinking. You should be comfortable navigating ambiguous customer needs, influencing teams across the company, and balancing rapid adoption with reliability, usability, and measurement quality. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. Travel requirements should be confirmed with the recruiter before publishing. In this role, you will: Define the product vision, strategy, and roadmap for experimentation, feature management, dynamic configuration, rollout safety, and analytics. Partner with product, engineering, research, data, design, and infrastructure leaders to turn recurring launch and measurement needs into reusable platform capabilities. Develop a deep understanding of workflows across ChatGPT, Codex, model measurement, monetization, subscriptions, and developer products, then establish clear priorities across competing needs. Drive adoption by making sophisticated experimentation and analytics c

REMOTEawsrestai
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

We're building a new product team in India to help shape how customers build, deploy, and operate AI applications on MongoDB. This is an early hire on a growing team—you'll join while the team is still taking shape, with the opportunity to have real impact from the start. You'll report to a Director of Product Management based in Europe, working async-first with partners across India, EMEA, and North America. The engineering team in India is being built in parallel, so you'll be working closely with engineers from the early days of the team. We are looking to speak to candidates based in Gurugram for our hybrid working model. What you'll own You'll be the end-to-end owner for a core area of our AI application platform—from discovery and vision through roadmap, delivery, and iteration. This is a senior individual contributor role where you lead through ownership, expertise, and cross-functional influence. Day to day, that means Defining product strategy and roadmap for your area, aligned with MongoDB's broader AI strategy and business goals Talking to customers — from AI-native startups to large enterprises — to understand how they build and deploy AI applications, where they struggle, and what MongoDB should do about it Writing clear product narratives, memos, and specs that align stakeholders, clarify trade-offs, to guide conversations with engineering, design, product marketing, etc Partnering with cross-functional stakeholders—including Engineering, Design, Product Marketing, Partners, Sales, and Developer Relations—to scope, prioritize, and ship, balancing experimentation with reliability and scale, and driving positioning, launches, and enablement Identifying new opportunities autonomously—spotting gaps, shaping problem spaces, researching industry trends and running discovery sessions with customers Define and own clear success metrics for your area and use them to prioritize, make trade-offs and communicate impact You'll develop deep expertise in how companie

mongodbawsazure
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

We're building a new product team in India to help shape how customers build, deploy, and operate AI applications on MongoDB. This is an early hire on a growing team—you'll join while the team is still taking shape, with the opportunity to have real impact from the start. You'll report to a Director of Product Management based in Europe, working async-first with partners across India, EMEA, and North America. The engineering team in India is being built in parallel, so you'll be working closely with engineers from the early days of the team. We are looking to speak to candidates based in Gurugram for our hybrid working model. What you'll own You'll be the end-to-end owner for a core area of our AI application platform—from discovery and vision through roadmap, delivery, and iteration. This is a senior individual contributor role where you lead through ownership, expertise, and cross-functional influence. Day to day, that means: Defining product strategy and roadmap for your area, aligned with MongoDB's broader AI strategy and business goals Talking to customers — from AI-native startups to large enterprises — to understand how they build and deploy AI applications, where they struggle, and what MongoDB should do about it Writing clear product narratives, memos, and specs that align stakeholders, clarify trade-offs, to guide conversations with engineering, design, product marketing, etc Partnering with cross-functional stakeholders—including Engineering, Design, Product Marketing, Partners, Sales, and Developer Relations—to scope, prioritize, and ship, balancing experimentation with reliability and scale, and driving positioning, launches, and enablement Identifying new opportunities autonomously—spotting gaps, shaping problem spaces, researching industry trends and running discovery sessions with customers Define and own clear success metrics for your area and use them to prioritize, make trade-offs and communicate impact You'll develop deep expertise in how compani

mongodbawsazure
View job →

About the Team Training Runtime builds the distributed systems that power OpenAI's largest model training runs - most recently GPT-5.5! The Data Movement area owns the infrastructure that keeps training jobs supplied with the right data at the right time, and keeps model state moving safely and efficiently across large clusters. Our work spans machine learning systems, distributed storage, high-throughput data loading, reliability engineering, and developer experience. Success means researchers can move quickly while training runs remain fast, reproducible, debuggable, and resilient at scale. About the Role We are looking for a deeply hands-on Technical Lead Manager to own datasets throughout our training infrastructure. This person will set the direction for how training jobs read data: the APIs, storage contracts, versioning model, benchmarks, debugging tools, and reliability guarantees that make data access consistent across current and future training frameworks. You will begin as the primary technical owner for dataset reads, working directly in the code while aligning researchers, training framework owners, storage teams, and infrastructure partners around a durable platform. The problem is deceptively hard at frontier scale: make enormous, heterogeneous datasets easy to consume, correct across distributed workers, observable when something goes wrong, and flexible enough to support pretraining, reinforcement learning, and multimodal training. In this role, you will Design and build a unified dataset read platform for multiple current and future training frameworks. Define dataset APIs, storage-format expectations, registration/versioning, and migration paths that make data access reproducible and maintainable. Build reliability into the read path, including stateful iteration, caching, fast restart, recovery, and clear operational contracts. Build terminal and web-based visualizers that let teams inspect text, multimodal, and reinforcement learning data late

pythonawsrest
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
15 days ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. REQUIREMENTS: Product Positioning & Technical Narrative Own the positioning and messaging for Baseten’s dedicated inference and platform capabilities, including autoscaling, routing, failover, release safety, observability, and cost/performance capabilities. Translate infrastructure-heavy product work into buyer narratives for ML engineering, platform engineering, security, compliance, procurement, and executive audiences. Partner with Product and Engineering to understand the technical architecture, customer value, roadmap tradeoffs, and proof points behind each capability. Define when a capability should be positioned as a platform differentiator, a dedicated inference requirement, a reliability story, a compliance story, or sales enablement. Build messaging that is technically credible without being overly implementation-focused or generic. Launch Strategy & GTM Execution Build and execute launch plans for major dedicated inference and serving platform capabilities, from early internal enablement through external announcement. Decide what deserves a full launch versus what should ship through docs, sales enablement, customer-specific materials, or targeted enterprise outreach. Create launch assets including messaging briefs, landing pages, blog posts, sales decks, one-pagers, FAQs, demo storylines, competitive talk tracks, and customer-facing proof points. Sequence launches and supporting assets based on custo

REMOTEkubernetesmachine learningai
View job →
🔔

Get new engineering manager platform reliability salary india jobs by email

Daily job updates · Unsubscribe anytime