Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. With the Okta's Auth0 organization’s increased dedication to ensuring customer availability expectations are exceeded in every way, you will play a key role as we evolve our system architecture to meet the demands of enormous growth and support the hundreds of millions of users who rely on us to provide uninterrupted access to business-critical Reporting to the Manager of Engineering, in this role as a SRE Operations Engineer, you will ensure smooth operations of our Customer Identity Cloud at Okta. Working closely with the SRE team, your primary focus will be on ensuring production systems remain operational at all times, while continually setting and achieving long-term operational success for the platform with potential career growth into Site Reliability Engineering. What you’ll be doing Executes operational work including updating/patching and maintaining the Engineering Service Desk queue Responsible for ensuring team requests are triaged and/or actioned in a timely manner Monitors Platform health and take steps to alleviate issues related to deployment and operations Assist with capacity, performance and scalability testing where required Escalation point for Platform issues from customer support teams Execute runbooks and update processes as required Interface with the SRE team to report core issues, required improvements and new feature requests What you’ll bring to the role General platform infrastructure knowledge, including high availability / l
Jobs in India
Platform Operations Specialist in India
1,228 active opportunities · Updated October 2026
Showing
15 jobs
Explore current platform operations specialist jobs across India. Filter by work mode, employment type, experience, department, date posted and distance.
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are You are an experienced Platform Engineer who owns backend infrastructure end to end. You design multi-tenant, microservices-based systems that other engineering teams build on, and you make deliberate architectural tradeoffs around consistency, latency, scale, and cost. You are comfortable going deep — service mesh internals, database internals, distributed-systems failure modes — and equally comfortable defining the reliability and security contracts an enterprise AI platform depends on. Responsibilities Design, own, and evolve scalable microservices architectures on Kubernetes across GCP, Azure, and AWS, including multi-tenant isolation (namespaces, network policies, per-tenant resource quotas and RBAC). Build core platform and data-plane components in Golang and Python — data ingestion, knowledge-base indexing and vector/graph search, application connectivity, workflow automation, and ML operations — against explicit latency and throughput SLOs. Own service-to-service communication: gRPC/protobuf API contracts, service mesh (Istio/Linkerd), load balancing, retries, timeouts, and circuit breaking. Make and document architectural tradeoffs — partitioning/sharding strat
NK Securities Research is a leading financial firm that leverages cutting edge technology and sophisticated algorithms to trade the financial markets. Founded in 2011, we have gained invaluable experience in the field of High Frequency Trading across different asset classes. Key Responsibilities: As a Software Developer - Platform, you will play a vital role in building and enhancing tools that empower our Quant, Infrastructure, Compliance, and Operations teams. In addition to web application development, you will be responsible for automating critical infrastructure processes, optimizing tasks, and delivering scalable solutions for our trading ecosystem. Application Development: Develop and maintain in-house software tailored for business and trading requirements. Enhance internal tools and applications to improve the user experience for different teams. Ensure robust and reliable trade monitoring systems through continuous innovation. Automation Development: Design and develop frameworks to automate infrastructure provisioning, configuration, and deployment using latest industry best practices Automate exchange-specific tasks such as connectivity management, order book monitoring, and trade execution workflows. Task Optimization: Identify and optimize repetitive tasks through scripting and configuration management. Ensure scalable and adaptable solutions to support multiple exchanges and regions. Infrastructure as Code (IaC): Use Ansible to codify infrastructure configurations, ensuring consistency and repeatability. Manage playbooks for server setups, network configurations, and middleware deployment. System Monitoring and Maintenance: Develop tools for system health checks, performance monitoring, and logging. Automate response mechanisms for critical alerts and incidents. Collaboration: Work closely with infrastructure, network, and trading teams to gather requirements and deliver robust solutions. Coordinate with exchange connectivity teams to ensure complianc
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Key responsibilities - • Lead Procure to Pay and Billing functions, oversaw vendor master management, invoice processing, order management, contract generation, billing, and inventory processes from the WPP SSC India. Supervised P & L management, billing operations, and proactively sought opportunities for business expansion. • Lead a team of 100+ FTEs. Managed team operations, fostered employee development, engaged stakeholders, and ensured service delivery excellence. Nurtured a culture of excellence and accountability within the team, provided mentorship, coaching, and professional development opportunities to enhance skill sets and drive career progression. Streamlined business processes and documentation, including the establishment and monitoring of KPIs and metrics. Continuously monitored SLAs, conducted team presentations, and spearheaded process quality enhancement initiatives. • Identify and implement automation solutions to optimize processes, reduce turnaround time, and enhance efficiency. Solicited feedback from internal and external stakeh
Who We Are Addepar is a global data and AI platform empowering investment professionals to turn complex financial information into actionable intelligence. Addepar unifies portfolio, market and client data in a total portfolio view and delivers AI-powered insights within investment and client workflows. More than 1,400 firms in nearly 60 countries use Addepar to manage and advise on nearly $9 trillion in assets. Its open platform integrates with nearly 650 software, data and consulting partners to power end-to-end investment operations across firms of all sizes and complexity. Addepar supports clients worldwide with offices in New York City, Salt Lake City, London, Edinburgh, Pune, Dubai, Geneva, Singapore and São Paulo. The Role We are currently seeking a Staff Software Engineer, Infrastructure to join the AI Platform team that powers seamless insights and interaction through natural language and data intelligence across our AI products. As a Staff Software Engineer, you’ll architect, build, and operate the backend and platform systems that power AI Platform. You’ll work across service design, distributed systems, cloud infrastructure, event-driven processing, observability, CI/CD, and production reliability, helping shape the technical direction of a platform that supports scalable, client-facing AI experiences. This role requires a strong software engineering foundation combined with deep infrastructure and systems thinking. We are looking for an engineer who can write high-quality production code, make sound architectural tradeoffs, and own platform capabilities end-to-end — not someone focused only on scripting, cloud configuration, or infrastructure tooling in isolation. You will collaborate closely with frontend, product, and AI/ML engineers to deliver reliable, secure, and scalable systems that align with Addepar’s standards of performance, resilience, and trust. Applicants must have legal authorization to work in the country where this role is based o
NVIDIA is seeking a Senior Staff SRE to build and operate reliable, scalable compute platforms that support global engineering workloads. This role spans Kubernetes, KubeVirt, bare-metal infrastructure, automation, observability, and AI-enabled operations. Join a team that solves complex infrastructure challenges, builds durable automation, and improves the reliability and operational experience of critical compute services. What you’ll be doing: Build, operate, and improve large-scale Kubernetes, KubeVirt, Linux, container, and bare-metal compute platforms, with a focus on performance, capacity, reliability, and operational scale. Lead bare-metal provisioning and lifecycle management in data centers, including PXE boot, DHCP, DNS, OS provisioning, hardware validation, and fleet automation. Develop automation, self-service capabilities, and observability solutions using APIs, Python or Go, Infrastructure as Code, configuration management, metrics, logs, traces, and service-health data. Define and operate SLOs, SLIs, error budgets, alerting, and incident-response practices; lead complex incident investigations, corrective actions, and blameless postmortems. Partner with infrastructure, security, hardware, data-center, and application teams to deliver global platform initiatives, and participate in an on-call rotation. What we need to see: BS in Computer Science, Engineering, a related technical field, or equivalent experience, plus 10+ years operating production infrastructure or platform services. Strong expertise in Kubernetes administration, KubeVirt, Docker, containerization, microservices, Linux systems, and resolving distributed-system challenges. <l
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: Responsible for leading the Cloud Automation Engineering function. Primary focus will be leading a team of other engineers in designing and implementing automation solutions to improve customer experience and increase productivity in our cloud estates. Responsible for maintaining and delivering automation solutions through infrastructure as code, ensuring security best practice, evangelising automation practice and tools, and supporting customer needs, both internal and external. What you'll be doing: Identify opportunities for improvement and automation of operations Design, build, test and implement use cases to drive automation adoption and improve operational efficiency Work closely with the IT Operations team to develop automated incident detection and response mechanisms. Implement proactive monitoring and alerting systems to quickly respond to and resolve critical issues, minimizing downtime and service disruptions Responsible for driving CSI initiatives to improve operations (processes/tools) working with various stakeholders Responsible for providing feedback at various leve
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: WPP is undertaking a major transformation of its enterprise Service Management capability, replacing its existing ServiceNow platform with a modern, AI-enabled service management platform that will establish a simpler, more standardised, and scalable global operating model. The current environment has evolved over many years and now requires a structured re-platforming programme to remove legacy complexity, reduce fragmentation, improve governance, and align more closely with industry best practice. This role will own the platform that underpins service management across Technology Operations and the wider Enterprise Technology team along with Finance, HR, and other corporate functions, making it a strategically important platform. Working with senior stakeholders and strategic implementation partners, you will define and deliver the future service management capability, ensuring the platform, operating model, governance, and service management processes support a modern, scalable, AI-enabled enterprise. The successful candidate will bring
Razorpay is one of India’s leading full-stack financial technology companies, powering the way businesses move, manage, and grow money. Founded in 2014 by Harshil Mathur and Shashank Kumar with a simple vision - to simplify payments for Indian businesses - we’ve since grown into a fintech powerhouse driving India’s digital payment revolution. Razorpay powers millions of businesses with a smarter, scalable stack that goes beyond transactions to help them truly build and grow. From building AI-native agentic payments, to AI-assisted fraud detection and real-time risk intelligence to automated reconciliation, smart payouts, and predictive financial insights, we are embedding intelligence across our stack to make money movement faster, safer, and more efficient. In close collaboration with ecosystem partners - including banks, networks, regulators - we are pioneering industry-first solutions that are shaping the next era of fintech Across India, Singapore and Malaysia, our products span everything from seamless checkouts to payroll automation - powering a fintech ecosystem that’s redefining how money moves across Asia. Today, that ecosystem supports everyone from early-stage startups to some of India’s largest enterprises, enabling them to accept, process, and disburse payments at scale while expanding into new ways of managing money more efficiently. Our scale speaks volumes: Razorpay processes $180+ billion in annualized transactions, powering leading businesses like Airbnb, Facebook, WhatsApp, Airtel, CRED, BookmyShow, Zomato, Swiggy, Lenskart, Mirae Asset Capital markets, Indian Oil, National Pension Scheme - and over 100 of India’s unicorns. With strong roots in India and growing operations in Southeast Asia, we are shaping the next chapter of financial technology across the region. We are backed by global investors including GIC, Peak XV Partners (formerly Sequoia Capital India & SEA), Tiger Global, Ribbit Capital, Matrix Partners, MasterCard, and Salesforce V
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Smartsheet is looking for an experienced Revenue Operations Analyst to support and accelerate the productivity of our Go-To Market organization. In this role, you will help define Smartsheet's growth strategy through insights produced with data analysis. You will work directly with senior leadership to inform strategic decision-making as part of a collaborative, motivated team. Your work will be instrumental in helping our Sales partners optimize their pipeline, increase retention, and close deals. Our ideal candidate is curious and displays an ability to translate business questions into analysis, reports, and recommendations. To be successful in this role, you are able to communicate to a diverse audience of internal stakeholders. You should have strong technical acumen with data analytics tools and languages, the ability to identify areas of opportunity within the business, and build innovative solutions. In 2005, Smartsheet was founded on the idea that teams and millions of people worldwide deserve a better way to deliver their very best work. Today, we deliver a leading cloud-based platform for work execution, empowering organizations to plan, capture, track, automate, and report on work at scale, resulting in more efficient processes and better business outcomes. You will report to our Manager, Revenue Operation Analytics. You Will: Act as a business partner to Sales leaders to understand and identify essential business questions and challenges Lead projects to develop answers and
About Paytm Paytm is India's leading mobile payments and financial services distribution company. Pioneer of the mobile QR payments revolution in India, Paytm builds technologies that help consumers and small businesses with payments, commerce, and credit. Paytm's mission is to serve half a billion Indians and bring them to the mainstream economy with the help of technology. Role Overview We are looking for an early-career Product Manager to join Paytm Lending. The role is about re-imagining how a lending business runs, from underwriting and risk to operations and customer journeys, using LLMs and AI agents. It is a hands-on role. You will dig into data, edge cases, and day-to-day operational problems to understand them from the ground up before designing solutions. We are looking for first-principles thinkers who enjoy building with LLMs and like going deep on real problems. Key Responsibilities - Identify workflows across lending (underwriting, risk, recon, operations, collections) where LLMs and agents can add value, and judge where they cannot. - Build and ship product solutions from prototype to production, working closely with engineering. - Dig into data, edge cases, and operational detail to understand problems at the root. - Partner with stakeholders across risk, operations, engineering, and business to align priorities and drive delivery. - Write clear product requirement documents and manage roadmaps and deliverables. - Use data to define hypotheses, run experiments (A/B), and track impact; use SQL or Excel to access and validate data independently. - Track key product and data metrics (OKRs and KPIs) and adjust strategy as needed. Requirements - Bachelor's degree in engineering, computer science and/or related field with MBA from top tier colleges. - 2 to 4 years of professional experience in product management, product analytics, or a similar role, with a strong customer-first mindset. - Genuine, hands-on experience building with LLMs.
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity As a Senior Software Engineer within the Container Fabric (CF) organization, you will be a key driver in evolving New Relic’s global internal platform. We are looking for an operations-heavy engineer with 5–8 years of relevant experience who can leverage open-source and custom tooling to orchestrate and maintain large-scale Kubernetes environments. You will play a "Captain" role—leading critical deliverables and mentoring junior engineers while maintaining the reliability of our global fleet. What You'll Do Architectural Leadership: Drive the design and implementation of internal tools, specifically focusing on Kubernetes Operators and Controllers to automate resource management. Platform Orchestration: Lead complex, large-scale infrastructure shifts. Operational Excellence: Take ownership of incident response, author comprehensive retrospectives, and implement systemic hardening to prevent recurrence using advanced overcommit strategies. This Role Requires Experience: 5–8 years in a DevOps, Site Reliability, or Infrastructure Engineering role. Kubernetes Mastery: Deep internals knowledge of Kubernetes and hands-on experience writing custom operators. Tooling Proficiency: Strong experience building production-grade tools and services, specifically for infrastructure automation. Operations-Heavy Mindset: A proven track record of Day 1/Day 2 operations for a large-scale Kubernetes fleet, handling high-severity incidents, and improving SLA compliance through auto
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Hardware Manufacturing Test Engineer at Everpure, you will spearhead manufacturing test enablement for our next-generation enterprise hardware platforms from New Product Introduction (NPI) through high-volume production ramp. Operating with high autonomy, you will design robust test strategies and collaborate cross-functionally with Firmware, Diagnostics, and global factory teams to deliver market-ready, high-yield hardware infrastructure. Your expertise will directly optimize manufacturing throughput, improve release quality, and accelerate the global deployment of the industry-defining Everpure Platform. WHAT YOU'LL DO Own NPI & Sustaining Test Readiness: Drive end-to-end manufacturing test strategy, test sequence definition, and coverage validation for multiple hardware infrastructure variants to ensure seamless factory execution and build readiness. Lead Debug & Yield Optimization: Independently lead technical triage and resolve complex system-level test fallout during critical build phases, analyzing failure paretos to isolate root causes across hardware, firmware, and test infrastructure. Scale Test Architecture & Frameworks: Enhance and standardize the Design for Manufacturing (DFM) hardware test architecture by deploying reusable automation frameworks and consolidating test execution into shared platforms to minimize factory cycle times. Drive Vendor & Cross-Functional Alignment: Colla
From $18K/yr
Bloomreach is building the world’s premier agentic platform for personalization .We’re revolutionizing how businesses connect with their customers, building and deploying AI agents to personalize the entire customer journey. We're taking autonomous search mainstream, making product discovery more intuitive and conversational for customers, and more profitable for businesses. We’re making conversational shopping a reality, connecting every shopper with tailored guidance and product expertise — available on demand, at every touchpoint in their journey. We're designing the future of autonomous marketing , taking the work out of workflows, and reclaiming the creative, strategic, and customer-first work marketers were always meant to do. And we're building all of that on the intelligence of a single AI engine — Loomi — so that personalization isn't only autonomous…it's also consistent.From retail to financial services, hospitality to gaming, businesses use Bloomreach to drive higher growth and lasting loyalty. We power personalization for more than 1,400 global brands, including American Eagle, Sonepar, and Pandora. About the SDM Team The Service Delivery Management team at Bloomreach is the engine behind every successful customer implementation. We don't just manage projects — we orchestrate complex, multi-workstream implementations that bring Bloomreach's AI-powered platform to life for the world's leading brands. What makes our SDM team different in 2026: AI-native operations — We use Von (our AI copilot), automated processes, and intelligent workflows to eliminate manual overhead and focus on high-value delivery Agentic tooling — ClickUp automations, Slack-integrated utilization tracking, and automated retrospective analysis are standard operating procedure Commercial ownership — SDMs own the services metrics for their projects, identify upsell opportunities, author/ review SOWs, and drive revenue through delivery
$30K – $35K/yr
SPECIFIC JOB RESPONSIBILITIES Defensive Operations (SecOps): Design and automate the Security Incident Response (SIR) and Vulnerability Response (VR) lifecycles. Build playbooks in Flow Designer to automate threat containment and remediation. Offensive Operations: Develop custom scoped applications to track penetration testing results, manage red-team engagement lifecycles, and automate the ingestion of reconnaissance data. Compliance & GRC: Configure and customize Integrated Risk Management (IRM) modules to map technical controls to frameworks like SOC2, ISO 27001, HIPAA, and FedRAMP. Integrations & Orchestration: Build robust, secure integrations (REST/SOAP, IntegrationHub , MID Servers) with our XDR, SIEM (Splunk/Sentinel), and cloud-native services (AWS/Azure/GCP). Multi-Tenancy & MSSP Architecture: Architect a scalable, multi-tenant environment that ensures strict data isolation between clients while allowing for unified " ClickOps " and Terraform-driven automation. AI & Innovation: Explore and implement Now Assist (GenAI) and AI-heavy workflows to automate security reporting and incident summarization. Defensive Operations (SecOps): Design and automate the Security Incident Response (SIR) and Vulnerability Response (VR) lifecycles. Build playbooks in Flow Designer to automate threat containment and remediation. Offensive Operations: Develop custom scoped applications to track penetration testing results, manage red-team engagement lifecycles, and automate the ingestion of reconnaissance data. Compliance & GRC: Configure and customize Integrated Risk Management (IRM) modules to map technical controls to frameworks like SOC2, ISO 27001, HIPAA, and FedRAMP. Integrations & Orchestration: Build robust, secure integrations (REST/SOAP, IntegrationHub , MID Servers) with our XDR, SIEM (Splunk/Sentinel), and cloud-native services (AWS/Azure/GCP)
Other cities to consider
More places hiring for this role
Get new platform operations specialist jobs in India by email
Daily job updates · Unsubscribe anytime