Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. With the Okta's Auth0 organization’s increased dedication to ensuring customer availability expectations are exceeded in every way, you will play a key role as we evolve our system architecture to meet the demands of enormous growth and support the hundreds of millions of users who rely on us to provide uninterrupted access to business-critical Reporting to the Manager of Engineering, in this role as a SRE Operations Engineer, you will ensure smooth operations of our Customer Identity Cloud at Okta. Working closely with the SRE team, your primary focus will be on ensuring production systems remain operational at all times, while continually setting and achieving long-term operational success for the platform with potential career growth into Site Reliability Engineering. What you’ll be doing Executes operational work including updating/patching and maintaining the Engineering Service Desk queue Responsible for ensuring team requests are triaged and/or actioned in a timely manner Monitors Platform health and take steps to alleviate issues related to deployment and operations Assist with capacity, performance and scalability testing where required Escalation point for Platform issues from customer support teams Execute runbooks and update processes as required Interface with the SRE team to report core issues, required improvements and new feature requests What you’ll bring to the role General platform infrastructure knowledge, including high availability / l
Jobs in India
Production Support Sre Analyst in Bengaluru
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current production support sre analyst jobs in Bengaluru. Filter by work mode, employment type, experience, department, date posted and distance.
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Build software used to support global trading across time zones • Work closely with business users to establish and refine requirements • Participate in identifying new technologies to continuously improve software systems • Implement DevOps practices within the team (GitHub/GitLab, Jenkins) • Provide production support to diagnose and resolve elevated application software production incidents and implement proactive remediation measures WHAT’S REQUIRED • Bachelor’s degree in computer science or another technical/scientific field • Minimum 8 years object-oriented programming experience with C#/.NET • Significant experience working with / understanding databases - primarily MS SQL Server • Must have experience in writing automated tests, unit tests, Test-Driven Development • Knowledge of design/architecture patterns, distributed systems, microservices, observability and monitoring, and containerization (Docker) • Willingness to work as part of a distributed Dev Team - 3 time zones (USA, Poland, India) • Strong problem solving and analytical skills • Exceptional verbal and written communication skills • Commitment to the highest ethical standards WE TAKE CARE OF OUR PEOPLE We invest in our people, their careers, their health, and their well-being. When you work here, we provide: • Health care benefits • Maternity, Adoption & related leave policies • Generous paternity and family care leave policies • Employee Assistance Prog
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE We are looking for a highly skilled Senior Frontend Engineer to join the Portworx UI team, responsible for building intuitive, scalable user experiences for our Kubernetes-based platform. You will build production-quality web applications using React, collaborate closely with UX, Product, backend teams, and engineering leadership, and help shape frontend architecture and engineering practices. This role requires strong ownership, sound technical judgment, and the ability to solve complex problems independently while helping other engineers grow. WHAT YOU'LL DO We are primarily an in-office environment and therefore, you will be expected to work from the {{OFFICE_LOCATION}} office in compliance with Everpure's policies, unless you are on PTO, or work travel, or other approved leave. Design, develop, and maintain scalable, high-performance single-page applications using React and TypeScript. Own features end to end—from requirements and design through implementation, testing, release, and production support. Create reusable, accessible, responsive UI components using modern HTML, CSS, React patterns, and design systems. Collaborate with UX, Product Management, and backend teams to turn customer and product needs into polished, production-ready features. Contribute to frontend architecture, API integration, performance, maintainability, and engineering best practices. Write and maintain unit, integration, and end-to-
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. Why is this a critical role at OneTrust? What is the challenge / type of challenges someone in this role will have the opportunity to take on? What impact does someone in this role have the opportunity to make within the company, for our customers and on a larger scale in the evolution of privacy and trust? An awareness of current issues affecting the industry and its technologies Create innovative, scalable, fault-tolerant software solutions for our clients and customer base Expand existing software to meet the changing needs of our key demographics What does this person do each day/each week? Describe a true to life day in the life for someone in this role. What do they do and how do they do it? Goal is to paint a real, genuine picture so candidates can see themselves in the role. Support production customers by monitoring and maintaining our cloud application & cloud infrastructure hosting it Build scripts for operational automation and incident response Handle processes surrounding cloud application deployment for our agile release Work with the monitoring, tuning, maintenanc
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO As Database Support Engineer, you’ll support various critical database platforms across Development, QA, UAT, and Production environments. The role partners closely with application teams, application support, and database engineers and operates within a Follow‑the‑Sun model to ensure availability, performance, and reliability of database services. Key responsibilities include: • Provide operational support for enterprise database platforms in both on-prem private cloud and public cloud • Monitor database health, capacity, performance, and availability, and respond to alerts, diagnose issues, and perform timely remediation • Perform routine maintenance activities (patching, upgrades, housekeeping etc) • Troubleshoot database‑related incidents and collaborate on root cause analysis • Work closely with application owners, application support teams, and DB Engineers • Provide guidance on database best practices and operational standards • Participate in cross‑team problem resolution and continuous improvement initiatives • Contribute to design, implementation and testing of automation and self service capabilities of DB platforms • Drive continuous improvement, identifying opportunities to reduce toil and increase platform efficiency. • Participate in a Follow‑the‑Sun operating model, including shift‑based coverage and handoffs WHAT’S REQUIRED • Bachelor’s degr
Who are we? FalconX is a pioneering team of operators, investors, and builders committed to revolutionizing institutional access to the crypto markets. Operating at the intersection of traditional finance and cutting-edge technology, FalconX addresses the industry's foremost challenges: Navigating the digital asset market can be complex and fragmented, with limited products and services that support trading strategies, structures, and liquidity found in conventional financial markets. As a comprehensive solution for all digital asset strategies from start to scale, FalconX operates as the connective tissue empowering clients with seamless navigation through the ever- evolving cryptocurrency landscape. Responsibilities Be part of a trading systems engineering team, dedicated to building out the core trading platforms. Work closely with cross functional teams to improve the system reliability, scalability and security. Engage in and improve the quality supporting the platform. Build and manage systems, infrastructure and applications through automation. Provide operational support to internal teams working on the platform. Work on improvements to bring in high efficiency, reduce latency, deploy systems faster. Practice sustainable incident response and blameless postmortems. Together with your engineering team, you will share an on-call rotation and be an escalation contact for service incidents. Implement and maintain rigorous security best practices across all infrastructure, with a focus on minimizing attack surface and ensuring data integrity. Monitor system health and performance with a keen eye for identifying and resolving issues before they affect trading activity. Manage user queries and service requests (often requiring in depth analysis of the technical and/or business logic of our systems). Proactive approach to problem analysis and resolution of production incidents. Manage Issue tracking and prioritisation of day to day production incidents. Manage platf
JOB TITLE IT Operations Engineer, Application Support A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Provide technical support for software applications and investigate, diagnose, and resolve application issues • Automate start-of-day and end-of-day checks for key applications. • Log and track incidents across applications in the production environment. • Implement monitoring and automation initiatives and develop custom solutions using Python, Shell, and/or Powershell scripts. • Create tactical support tools and scripts to improve the incident investigation process and enable • transparency into potential business impacts. • Prioritize and categorize incidents based on severity and impact. • Collaborate with the development team to improve applications based on user feedback. • Create and maintain documentation for responding to common errors and application incidents. • Assist with software applications deployment and configuration . • Provide training and assistance to users to ensure effective use of applications and systems. • Develop knowledge base resources to empower users to independently resolve common problems. WHAT’S REQUIRED • Bachelor's degree in computer science, information technology, or a related field. • Experience supporting middle- and back-office applications created in .Net, Java, C# etc. Ability to debug apps using of code, logs, alerts etc. • Literacy in complex SQL procedures/queries. • Ability to diagnose and troubleshoot technical issues.
JOB TITLE IT Operations Engineer, EQUITY TRADING technology A Career with point72’s technology TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source and AI solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. What you’ll do Provide operational and technical support for the firm’s trading platforms to ensure optimal performance Coordinate and execute software upgrades and releases across trading platforms Manage all production and UAT trading platforms Support trading systems during incidents, including both remediating the issue and ensuring ongoing communication with end-users and stakeholders Liaise with brokers, service providers, and other internal technology groups and stakeholders Design and implement tools and reports to enhance department efficiency Assist platform users during onboarding processes What’s REQUIRED 7+ years of application support experience within the financial services industry Experience working with Linux and other languages Ability to work effectively within a global team, adapting to varying time zones and flexible shift schedules Strong understanding of order management workflows Commitment to the highest ethical standards About point72 Point72 is a leading global alternative investment firm led by Steven A. Cohen. Building on more than 30 years of investing experience, Point72 seeks to deliver superior returns for its investors through fundamental and systematic investing strategies across asset classes and geographies. We aim to attract and retain the industry’s brightest talent by cultivating an investor-led culture and co
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Technical Services Engineer for Portworx® by Everpure™, you will empower innovators and enterprise customers to seamlessly deploy and scale multi-cloud data services for Kubernetes and containerized applications. In this high-impact role, you will redefine the cloud-native storage experience by bridging deep technical expertise with direct customer engagement. You will serve as a technical expert across all deployment phases, partnering closely with customer engineers, account teams, and core R&D to solve complex cloud-native architecture challenges. WHAT YOU'LL DO Drive Enterprise Customer Success: Deliver end-to-end technical expertise across all stages of Portworx® deployment and production environments, ensuring rapid issue resolution, minimal downtime, and maximum platform reliability for key enterprise accounts. Perform Complex Problem Diagnosis & Triage: Analyze application workflows, system logs, and containerized workloads across public/private cloud stacks to isolate, reproduce, and resolve multi-layer technical issues alongside engineering teams. Enable Pre-Sales and Onboarding: Partner with field systems engineers to guide technical installation, integration, and proof-of-concept (POC) execution for prospective and onboarding enterprise clients. Scale Organizational Knowledge: Author and maintain high-impact knowledge base articles, technical documentation, and FAQ guides to build self-ser
Who We Are Simpplr is the AI-powered intranet for unifying the digital workplace. It brings people, trusted knowledge, apps, and agents into a coherent digital experience. Powered by a proprietary EX Knowledge Graph, Simpplr synthesizes signals and context across connected systems to deliver personalized information and actions. The platform serves as a digital hub supporting communications, engagement, employee services, and work. With low-code extensibility and enterprise-grade security and governance, Simpplr enables confident operation at scale. More than 1,000 organizations — including AAA, the NHS, Penske, and Moderna — trust Simpplr to keep their workforce informed, aligned, and productive. Learn more at simpplr.com . About the role We are looking for a Lead Voice AI Engineer to build production-grade Voice Agents for frontline heavy verticals like healthcare, manufacturing, warehousing, retail, hospitality focusing on employee support, procurement, collections, logistics, ordering etc. You will lead the design of low-latency, real-time voice systems combining ASR, TTS, LLMs, conversational AI, enterprise workflows, knowledge retrieval, compliance, and human handoff. This is a hands-on technical leadership role for someone who can take Voice AI from architecture to production. Responsibilities Design and build the real-time voice runtime for live conversations. Build and optimize streaming ASR, TTS, VAD, endpointing, turn-taking, and barge-in. Build adaptive voice pipelines for high-noise frontline environments (60-112 dB), hospitals, factory floors, warehouses, including server-side noise cancellation, echo suppression, and dynamic ASR/TTS optimization for PSTN and mobile phone audio quality. Architect multi-provider speech routing across a broad multilingual matrix, including code-switching (e.g., Spanglish, Hinglish), where no single ASR or TTS provider covers all languages, and language detection, provider selection, and fallback chains must operate
Roles and Responsibilities Installation and configuration of NoSQL instances on single or multiple ports. ? Hands on experience of production on medium to big sized NoSQL databases Setting up and maintaining users and privileges management systems and Troubleshooting relevant access issues. Understand the transaction flowsand ACID compliance. Performing on-call support and should be able to provide the first level support . Configure and setup NOSQL databases like mongodb and Cassandra. Automation of repetitive tasks. Qualifications & Experience 3-6 years of Hands-on experience of working with NoSQL DBA . Some exposure to external tools like Percona , ProxySQL , HAP etc. Understanding of networking concepts . verbal and written communication skills. Experience in tools like shell , python . perl etc for automation. fundamentals on the linux system side and monitoring tools like top , iostats , sar etc. Clear understanding of NoSQL Replication process flows , threads , setting up multi node clusters and basic troubleshooting. Understanding of at least one of the backup and recovery methods for MySQL, fundamentals of SQL. Understand and tune complex SQL queries when needed.
A Career with Point72's Technology Team As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. What you'll do Lead the design, development, and operation of scalable, enterprise-grade AI/ML architectures and systems with a strong emphasis on reliability, availability, and performance. Lead and mentor a team of engineers, driving technical direction, code quality, and iterative delivery of large-scale solutions. Partner closely with data scientists, engineers, product teams, and compliance to integrate AI/ML solutions into existing and new products. Own the end-to-end lifecycle of GenAI services, including LLM inference, model serving, and proxy/gateway layers that support multiple downstream applications. Define and uphold engineering best practices around observability, scalability, security, and cost efficiency for AI/ML platforms. Evaluate tools, technologies, and processes to ensure the highest quality and performance of AI/ML systems. Stay abreast of the latest advancements in AI/ML technologies and methodologies, and translate them into pragmatic solutions for the business. Ensure compliance with industry standards and best practices in AI/ML. What's required Bachelor's or Master's degree in Computer Science, Engineering, or a related field. 10+ years of experience in software/AI/ML engineering, with a proven track record of successful delivery of complex, production-grade systems. Demonstrated experience building large-scale enterprise-grade services with high reliability, availability, and observability (SLO/SLA-driven en
Graphcore Senior Principal AI SoC Validation (Bring-up lead) Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Bengaluru which will play a central role in Graphcore's work building the future of AI computing. We are developing the next generation of AI compute, a large-scale system-on-chip (SoC) designed to power future high-performance AI systems. As the SoC Validation Lead, you will be responsible for enabling pre-production software to run reliably on new silicon quickly and efficiently, before showing that the silicon meets the highest standards of quality, reliability and functionality, ready for production deployment. You will lead a team delivering post-silicon validation across the full AI SoC, working across silicon, firmware, and platform levels. The role requires a deep technical understanding, strong hands-on debug experience, and the ability to collaborate effectively with hardware, software, and systems engineering teams. Key responsibilities Define and lead post-silicon validation strategy Develop and refine the overall post-silicon validation approach for our AI SoCs, ensuring reliable and timely delivery of validated silicon, architectural correctness, feature robustness, and at-scale system reliability. Drive cross-domain debug and issue resolution Lead investigation and resolution of complex issues spanning silicon, firmware, operating systems, and platform interactions. Ensure that fixes are effective and sustainable. Promote collaboration and shared understanding Work closely with
Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before it touches production. That founding instinct still defines how we work. We're accurate and evidence-driven, relentless about clarity, and we'd rather be certain than comfortable, building a groundbreaking platform that transforms how teams run and secure networks across every major cloud and vendor environment. Global leaders like Goldman Sachs, PayPal, S&P Global, IBM, and Dell trust Forward, alongside fast-growing enterprises and government agencies, realizing an average of $14.2 million in annual benefits, according to IDC. Backed by top-tier investors, including A. Capital, Andreessen Horowitz, Goldman Sachs, MSD Partners, Omega Venture Partners, Section 32, and Threshold Ventures, and headquartered in Santa Clara, we're most proud of our team: curious people who'd rather build what doesn't exist than accept how things have always been done. Forward is currently seeking a Senior Backend Software Engineer to work as part of our Platforms team. You will play a critical role in designing, developing, and scaling the core backend services and infrastructure that support our SaaS and on-prem deployments. Your contributions will have a direct impact on the stability, performance, and scalability of our platform, helping to ensure an exceptional experience for our customers. Responsibilities: Platform development: Contribute to the design and development of storage systems, job scheduling systems, data ingestion frameworks, monitoring frameworks etc to ensure high system performance and availability. Feature development: Build and maintain backend frameworks that support essential platform features Scalability & Reliability: Develop scalable, high-performing
Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before it touches production. That founding instinct still defines how we work. We're accurate and evidence-driven, relentless about clarity, and we'd rather be certain than comfortable, building a groundbreaking platform that transforms how teams run and secure networks across every major cloud and vendor environment. Global leaders like Goldman Sachs, PayPal, S&P Global, IBM, and Dell trust Forward, alongside fast-growing enterprises and government agencies, realizing an average of $14.2 million in annual benefits, according to IDC. Backed by top-tier investors, including A. Capital, Andreessen Horowitz, Goldman Sachs, MSD Partners, Omega Venture Partners, Section 32, and Threshold Ventures, and headquartered in Santa Clara, we're most proud of our team: curious people who'd rather build what doesn't exist than accept how things have always been done. Forward is currently seeking a Java Backend Software Engineer to work as part of our Apps - Server team. The work will involve developing our web server, REST APIs, and product core by writing clean and solid code that interacts with our other services and components. Responsibilities include: Developing new product features that leverage the network model to help users: visualize their network, understand how it behaves, see how it has evolved, answer specific questions, and plan changes Designing the data model for new product features Proposing and implementing REST APIs to support the Forward web application and to publish to customers Constructively reviewing product designs, technical design documents, and code changes Requirements: At least 5+ years of full lifecycle software development experience Expertise in Java (versi
Other cities to consider
More places hiring for this role
Get new production support sre analyst jobs in Bengaluru, India by email
Daily job updates · Unsubscribe anytime