At Bolna, we’re building tools that change how businesses leverage voice AI. We’re looking for a Software Engineer to build reliable, scalable systems that power millions of production conversations across languages, industries, and telephony environments. This is a high-impact, high-ownership role where you’ll work on core platform problems across distributed systems, real-time communication, developer infrastructure, and customer-facing products. Our team includes IIT alumni with experience at Bain, Atlassian, Uber, Zomato, and LinkedIn, and is backed by leading investors. Responsibilities Build systems that operate at scale: Design and build backend services that support high-volume, real-time voice AI conversations with strong reliability, performance, and fault tolerance. Own features end to end: Take problems from product requirements and technical design through implementation, testing, deployment, monitoring, and iteration. Improve platform reliability: Build systems that are observable, resilient, and easy to debug. Identify bottlenecks, reduce failure rates, and improve system availability. Work on real-time infrastructure: Solve problems across telephony, streaming audio, webhooks, queues, scheduling, concurrency, and low-latency communication. Build for developers and customers: Improve APIs, SDKs, integrations, dashboards, and internal tools that make the Bolna platform easier to use and operate. Raise the engineering bar: Contribute to technical design reviews, code quality, testing standards, documentation, incident response, and engineering best practices. Required Skills Strong engineering fundamentals: Solid understanding of data structures, algorithms, databases, networking, operating systems, and distributed systems. Backend development experience: 2+ years of experience building and operating production backend systems using Python, Go, Java, Node.js, or a similar language. Production ownership: Experience shipping software to production and own
Jobs in India
Incident Commander in Bengaluru
45 active opportunities · Updated October 2026
Showing
15 jobs
Explore current incident commander jobs in Bengaluru. Filter by work mode, employment type, experience, department, date posted and distance.
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. About the Role As a Site Reliability Engineer at Ema, you will own the stability, availability, and operational health of our agentic AI platform across customer environments. You'll work closely with Engineering and DevOps to provision infrastructure, drive deployment excellence, and keep production running at the quality bar our enterprise customers expect — 99.9%+ uptime, proactive incident response, and continuous improvement. What You'll Do Infrastructure & Deployment Design and provision cloud infrastructure (GCP, Azure, AWS) tailored to customer environments, with security, scalability, and compliance built in Execute on-call SaaS deployments with minimal downtime; automate and optimize deployment workflows end-to-end Production Stability & Observability Monitor logs, alerts, and metrics to maintain SLA commitments and catch issues before they escalate Diagnose and resolve production incidents with speed and rigor; drive root cause analysis and permanent fixes Collaborate with DevOps to enhance monitoring dashboards and alerting frameworks; deliver clear system health reporting to internal and customer stakeholders Documentation & Knowledge Management Maintain de
Principal Product Manager - Agentic Investigation & Reliability Experiences Sumo Logic is hiring a Principal Product Manager to lead how engineers and operators investigate incidents, understand reliability risk, and act on their operational and security telemetry. The observability category was built around collecting telemetry and giving people tools to navigate it: dashboards, queries, monitors, traces, and alerts. Customer expectations are now shifting. Teams don't just want more dashboards; they want help getting from a signal to a resolution, understanding what's broken, why, what's impacted, and what to do next. As AI agents move into production operations, this role owns how Sumo Logic brings intelligent, agent-assisted investigation and reliability workflows to customers, grounded in evidence, context, and enterprise governance. This is a senior, high-ownership role. It requires genuine observability domain background. You should have lived in this space and understand how monitoring, troubleshooting, and reliability actually work, combined with the ambition to define a new category of experience on top of it. What You Will Own The current data experiences. Log Search, Live Tail, query and query optimization, Metrics Search, Tracing, Dashboards, and the data-experience UI. This is a live, revenue-generating product with real customers, and keeping it strong is part of the job. You own its health, roadmap, and competitiveness today while steering it toward an AI-native future, focusing new investment where it strengthens investigation, speed, and value for both new and power users. The reliability and alerting surface. Monitors, Alerts, SLOs, Scheduled Searches, and the reliability workflows around them. You will own alerting accuracy, noise reduction, and operational health signals both as capabilities customers depend on today and as the foundation for more automated, agent-assisted detection and investigation. The agentic investigation experience. You
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO • Build software used to support global trading across time zones • Work closely with business users to establish and refine requirements • Participate in identifying new technologies to continuously improve software systems • Implement DevOps practices within the team (GitHub/GitLab, Jenkins) • Provide production support to diagnose and resolve elevated application software production incidents and implement proactive remediation measures WHAT’S REQUIRED • Bachelor’s degree in computer science or another technical/scientific field • Minimum 8 years object-oriented programming experience with C#/.NET • Significant experience working with / understanding databases - primarily MS SQL Server • Must have experience in writing automated tests, unit tests, Test-Driven Development • Knowledge of design/architecture patterns, distributed systems, microservices, observability and monitoring, and containerization (Docker) • Willingness to work as part of a distributed Dev Team - 3 time zones (USA, Poland, India) • Strong problem solving and analytical skills • Exceptional verbal and written communication skills • Commitment to the highest ethical standards WE TAKE CARE OF OUR PEOPLE We invest in our people, their careers, their health, and their well-being. When you work here, we provide: • Health care benefits • Maternity, Adoption & related leave policies • Generous paternity and family care leave policies • Employee Assistance Prog
A Career with point72’s Technology team As Point72 reimagines the future of investing, our Technology team is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open-source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. What you’ll do Provide high‑quality technical support via phone, chat, email, and self‑service channels while delivering white‑glove service to global employees. Diagnose and resolve incidents across end‑user technologies (Windows/Mac), Microsoft 365, collaboration tools, financial/trading apps, mobile devices, and enterprise applications. Troubleshoot identity, networking, and endpoint issues including Active Directory, VPN, device configuration, endpoint management, and mobile device management. Deliver first‑contact resolution whenever possible while logging, categorizing, prioritizing, and managing tickets in the ITSM platform according to SLAs. Escalate complex issues with clear technical documentation, diagnostic details, and handoff notes to downstream support teams. Create, improve, and maintain knowledge base content, capturing troubleshooting steps, solutions, and best practices. Participate in structured shift handovers and support follow‑the‑sun operational coverage. Identify opportunities for automation, self‑service enhancements, process improvements, and overall service optimization. Communicate clearly and professionally with users, manage expectations, provide timely updates, and maintain composure under pressure. Support operational excellence by contributing to continuous improvement initiatives and maintaining strong documentation discipline. What’s REQUIRED Bachelor’s degree in computer science or related discipline,
JOB TITLE IT Operations Engineer, EQUITY TRADING technology A Career with point72’s technology TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source and AI solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. What you’ll do Provide operational and technical support for the firm’s trading platforms to ensure optimal performance Coordinate and execute software upgrades and releases across trading platforms Manage all production and UAT trading platforms Support trading systems during incidents, including both remediating the issue and ensuring ongoing communication with end-users and stakeholders Liaise with brokers, service providers, and other internal technology groups and stakeholders Design and implement tools and reports to enhance department efficiency Assist platform users during onboarding processes What’s REQUIRED 7+ years of application support experience within the financial services industry Experience working with Linux and other languages Ability to work effectively within a global team, adapting to varying time zones and flexible shift schedules Strong understanding of order management workflows Commitment to the highest ethical standards About point72 Point72 is a leading global alternative investment firm led by Steven A. Cohen. Building on more than 30 years of investing experience, Point72 seeks to deliver superior returns for its investors through fundamental and systematic investing strategies across asset classes and geographies. We aim to attract and retain the industry’s brightest talent by cultivating an investor-led culture and co
We are seeking a highly skilled and experienced Staff Network Site Reliability Engineer (SRE) to join our Enterprise Network Operations and SRE team. In this role, you will be pivotal in implementing our vision for a reliable and efficient network infrastructure. The ideal candidate is passionate about network operations and committed to enhancing the user experience. You'll have the opportunity to solve complex network challenges using hands-on debugging and by focusing on network automation, observability, documentation, and operational excellence. This is a critical position focused on ensuring user satisfaction and brilliance in network operations. What you'll be doing: Owning the operational aspect of the network infrastructure, ensuring its high availability and reliability, actively working on network incidents and service requests. Partnering with architecture and deployment teams to guarantee that new implementations are supportable and align with production standards. Advocating for and implementing automation to reduce toil and improve operational efficiency. Minimizing manual operational tasks to achieve and maintain Service Level Objectives (SLOs). Monitoring network performance, identifying areas for improvement, and collaborating with relevant teams to implement refinements. Proactively identifying and mitigating network risks to promote continuous improvement. Collaborating with domain experts across functions to resolve production issues swiftly and effectively, ensuring customer happiness. Conducting blameless postmortems and following through on Root Cause Analyses (RCAs). Discovering opportunities for operational improvements and teaming up with colleagues to devise solutions that enhance excellence and sustainability in network operations. Developing knowledge base articles for automa
At Rockstar Games, we create world-class entertainment experiences. Become part of a team working on some of the most rewarding, large-scale creative projects to be found in any entertainment medium - all within an inclusive, highly-motivated environment where you can learn and collaborate with some of the most talented people in the industry. Rockstar is on the lookout for a talented Security Analyst within Security GRC to support Security Compliance Operations through recurring control validation, evidence review, reporting preparation, and remediation tracking. This role will help determine whether key security safeguards are operating as intended by performing manual and semi-automated validation activities, analyzing evidence from multiple sources, documenting control gaps, and preparing clear reporting that helps interpret validation results. This is a full-time permanent position based out of Rockstar’s studio in Bangalore, India. WHAT WE DO The Rockstar Games Security team is responsible for advancing the state of information security across the company globally in collaboration with numerous partners and stakeholders by prioritizing and executing security initiatives that drive down risk. We strive to understand the threat landscape affecting our development studios, the gaming industry, and the world at large to define information security policies, standards, and procedures to safeguard our business and protect our players. We lead efforts to build enterprise security controls ranging from endpoint protection technologies to security incidents and event monitoring solutions. We have a passion for identifying threats and vulnerabilities and coming up with clever solutions to mitigate or remediate those risks. RESPONSIBILITIES Perform recurring manual and semi-automated validation of security controls to confirm safeguards are operating as intended. Collect, review, and organize evidence from systems, tools, reports, an
About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the role We are looking for a Quality Architect/Lead with hands-on experience building quality systems and has deep expertise in building and scaling test automation frameworks.The role requires an individual who applies systems thinking to solving complex problems. They should be able to understand the product from various perspectives and be able to effectively create testing programs that validate not just functionality but performance, reliability and user experience.DevRev is building a next generation AI native product that requires us to build novel test systems for the Agent AI platform. The role is mult-faceted and is going to continuously evolve with time. What you'll do Test Case Design and Documentation Actively use AI and intelligent agents to accelerate test generation, test maintenance, and coverage expansion. Leverage LLMs to convert requirements, user stories, and production incidents into high-quality automated test cases. Reduce reliance on manual test case creation by introducing AI-assisted automation workflows, with human review and ownership. Apply AI to optimize test selection, prioritization, and execution based on risk, code changes, and historical failures. Use AI to assist in identi
Toradex is a global company strongly focused on engineering & technology. We’re powered by a diverse & uniquely gifted workforce. We pursue the best people to propel our innovative vision of embedded computing and IoT. If you’re interested in being a driving force at an agile technology company, engineering clever computing solutions & helping other companies bring their products to life, we should talk. Description We are looking for a DevOps Engineer to strengthen our cloud operations and engineering practices, with a focus on reliable website delivery, secure AWS foundations, and fast but controlled delivery of new services. The position combines AWS operations, infrastructure as code, CI/CD, automation, and pragmatic software engineering. The person should be confident working with services for edge delivery, compute, storage, databases, DNS, security, and observability without relying on manual console changes as the default operating model. The role also supports on-premises to cloud migration, global service optimization, and practical responses to increasing AI-driven traffic. We value candidates who can use modern AI-assisted development effectively to spin up proof-of-concept projects quickly, while still applying disciplined Git, review, security, and deployment practices. About you You enjoy building stable, secure, and maintainable infrastructure that supports business-critical services. You can work independently and take ownership of cloud environments, deployments, and operational improvements. You are comfortable balancing speed, reliability, cost, and security when making technical decisions. You communicate clearly with technical and non-technical stakeholders and explain trade-offs in a practical way. You document your work well and create clear runbooks and support material for future maintenance. You are methodical when troubleshooting incidents and stay calm when systems are under pressure. You are curious about modern traffic patt
Linux Engineer A Career with point72’s Technology team As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. What you’ll do Support, maintain, and troubleshoot Linux systems in production and development environments Assist with system provisioning, configuration, and lifecycle management Monitor system performance, availability, and capacity; respond to incidents and outages Work closely with developers and quants to support research and trading workloads Automate operational tasks using scripting and configuration management tools Assist with patching, upgrades, and vulnerability remediation Maintain documentation and operational runbooks Participate in on-call rotation (as appropriate for level) What’s REQUIRED Bachelor’s degree in computer science, information technology, engineering, or a related field. 5+ years of professional experience administering Ubuntu Linux systems and ZFS Strong understanding of Linux fundamentals including processes, memory, CPU, filesystems, and networking as well as systemd, cron, logging, and package management (apt) Experie
A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology group is constantly improving our company’s IT infrastructure, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts experimenting, discovering new ways to harness the power of open source solutions, and embracing enterprise agile methodology. We encourage professional development to ensure you bring innovative ideas to our products while satisfying your own intellectual curiosity. WHAT YOU’LL DO As Database Support Engineer, you’ll support various critical database platforms across Development, QA, UAT, and Production environments. The role partners closely with application teams, application support, and database engineers and operates within a Follow‑the‑Sun model to ensure availability, performance, and reliability of database services. Key responsibilities include: • Provide operational support for enterprise database platforms in both on-prem private cloud and public cloud • Monitor database health, capacity, performance, and availability, and respond to alerts, diagnose issues, and perform timely remediation • Perform routine maintenance activities (patching, upgrades, housekeeping etc) • Troubleshoot database‑related incidents and collaborate on root cause analysis • Work closely with application owners, application support teams, and DB Engineers • Provide guidance on database best practices and operational standards • Participate in cross‑team problem resolution and continuous improvement initiatives • Contribute to design, implementation and testing of automation and self service capabilities of DB platforms • Drive continuous improvement, identifying opportunities to reduce toil and increase platform efficiency. • Participate in a Follow‑the‑Sun operating model, including shift‑based coverage and handoffs WHAT’S REQUIRED • Bachelor’s degr
JOB TITLE Data Reliability Engineer A CAREER WITH CUBIST Cubist Systematic Strategies, an affiliate of Point72, deploys systematic, computer-driven trading strategies across multiple liquid asset classes, including equities, futures, and foreign exchange. The core of our effort is rigorous research into a wide range of market anomalies, fueled by our unparalleled access to a wide range of publicly available data sources. What you’ll do Ensure smooth day-to-day implementation of a large research infrastructure and the timely delivery of comprehensive and error-free data to Cubist’s portfolio managers across the globe Serve as a frontline owner for mission-critical data ETL pipelines that power trading and investment decision-making, ensuring reliability, accuracy, and timeliness. Actively manage and resolve data incidents in a fast-paced trading environment, partnering closely with investment professionals, data scientists, and external data vendors. Design and build tooling, automation, and robust documentation to improve operational efficiency, scalability, and data quality across the platform. Play a hands-on role in daily data operations, including data validation, remediation, and enrichment, with opportunities to continuously improve and modernize workflows through engineering best practices. What’s REQUIRED Bachelor’s degree in computer science or a related field. Strong proficiency in SQL Server and Python programming, with experience in AWS and both Windows and Linux environments. Exceptional attention to detail with a strong appreciation for well-defined processes and systems. 3+ years of experience in a client-facing support or operations role. Excellent organizational, communication, and interpersonal skills. Commitment to the highest ethical standards About point72 Point72 is a leading global alternative investment firm led by Steven A. Cohen. Building on more than 30 years of investing experience, Poin
TEGNA Inc. helps people thrive in their local communities by providing the trusted local news and services that matter most. With 64 television stations in 51 U.S. markets, TEGNA reaches more than 100 million people monthly across web, mobile apps, streaming, and linear television, while also maintaining a strong global presence in India with offices in Bangalore and Chennai that support technology, product, and business operations initiatives. Together, we are building a sustainable future for local news. Operation Analyst Position Overview TEGNA is looking for Operation Analyst to join the dynamic engineering team. This role is responsible for the day-to-day monitoring, support, and operational stability of TEGNA’s technology platforms, including broadcast, streaming, and OTT systems, while providing advanced helpdesk support to end users and business-critical applications. The position works closely with distributed IT teams, streaming, and digital teams, as well as vendors and service partners, to resolve incidents, manage tickets, and maintain service levels in a 24x7 support environment. Clear documentation, effective communication, and ownership of issues from intake through resolution are critical to ensuring reliable linear and digital content delivery. What You’ll Do Handle and resolve ServiceNow support tickets from intake to closure within defined SLAs. Provide Tier 1 & Tier 2 technical support for PC, Mac, Android, and iOS devices. Support collaboration platforms, workstreams, and AI tools. Monitor critical systems and applications proactively to prevent service disruptions. Perform routine health checks, respond to alerts, and provide operational support. Troubleshoot incidents and perform root cause analysis for recurring issues. Document resolutions, workarounds, SOPs, and knowledge base articles in ServiceNow. Maintain operational documentation and improve knowl
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. At Okta, our motto is "Always On" and nowhere do we embrace that more than in Technical Operations. We strive to build the most reliable and performant systems on the planet through the skillful use of automation. If you like to be challenged and have a passion for solving large-scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate is someone who exemplifies the ethics of, “If you have to do something more than once, automate it” and who can rapidly self-educate on new concepts and tools. You will work on: ● Mentoring, managing, and leading a team of SRE’s with a broad range of expertise and experience. ● Being an evangelist and advocate for security best practices, leading initiatives and projects to strengthen our security posture for our most critical infrastructure. ● Responding to production incidents, driving us to remediation as quickly as possible and determining how we can prevent them in the future. ● Triaging and troubleshooting complex production issues to ensure reliability and performance. ● Working closely with our stakeholders across the organization to ensure our new capabilities are aligned to our competing constraints of reliability, security, and delivery velocity. ● Partnering directly with recruiting and people ops to hire and retain the best talent in the world. ● Keep sharp eyes on our metrics, including vulnerability scanning and security posture, 
Other cities to consider
More places hiring for this role
Get new incident commander jobs in Bengaluru, India by email
Daily job updates · Unsubscribe anytime