Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Exa team and lead the charge in redefining enterprise storage by unifying block, file, and object protocols across hybrid-cloud environments. You will combine deep technical expertise in distributed systems with hands-on people leadership to guide architectural decisions and mentor high-impact engineers. This is a unique opportunity to build new engineering teams from the ground up and drive industry-leading innovation alongside Product and Architecture partners. Your work will directly impact how customers consume, scale, and operate mission-critical storage infrastructure. WHAT YOU'LL DO Establish and scale a new engineering organization focused on critical Exa platform services, ensuring a foundation of long-term success, technical excellence, and a high-performing culture. Own the successful delivery of complex, high-scale engineering features for the Exa platform, ensuring world-class security, reliability, and availability across multi-array and hybrid-cloud deployments. Define the technical vision and execution roadmap in close partnership with Product Management and Architecture, translating customer needs into a measurable business impact for Pure Storage. Drive a culture of operational rigor, owning the refinement of engineering processes around observability, CI/CD, and incident response, while actively mentoring the next generation of technical leads. WHAT YOU BRING Leadership and Scaling: Pro
Jobiba hiring network
Incident Commander Jobs
589 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current incident commander jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a ServiceNow Platform Support Specialist, you will play a critical role in supporting the ServiceNow platform. Your primary responsibility will be managing and resolving L2 incidents and the technical development and delivery of run-the-business (RTB) enhancements and keep-the-lights-on (KTLO) items. This role requires a hands-on, customer-focused individual with strong problem-solving skills who can manage platform support, troubleshooting, and system administration in a fast-paced, collaborative environment. You will also have opportunities to expand your skillset in many different directions within the ServiceNow ecosystem. WHAT YOU'LL DO Incident Monitoring & User Support: Provide second tier support for ServiceNow Actively monitor incident queue for escalations from the L1 teamTroubleshoot and resolve L2 issues related to ServiceNow (CMDB, CSM, WSD, ITSM, HRSD, LSD, among others) and L2 issues regarding ServiceNow integrations (SAP, Workday, NICE InContact, among others) Escalate complex or unresolved issues and requests for enhancement to the appropriate L3 teams, ensuring prompt follow-up and, when possible, a resolution.Maintain strong communication with users, providing clear guidance, solutions, and updates on issue resolution or feature requests.Participate in L2 on-call rotation to respond to critical incidents raised from the primary L1 on-call team in off-hours Run the Busine
We are fueled by a moral imperative to advance mankind, and it all begins with our people, our product, and our purpose. Passion isn’t something we turn on and off; it’s woven into everything we do. If you thrive in high-challenge environments, are inspired by exceptional teammates, and are driven to grow beyond what you thought possible, MX is where you belong. Come build the future with us. Join an award-winning company that isn’t just shaping the financial industry, but transforming it in ways that create meaningful, lasting impact for millions of people. Senior Cloud Platform Engineer At MX, we’re on a mission to empower the world to be financially strong. We power the financial experiences behind thousands of banks, credit unions, and fintechs, helping millions of people better understand, manage, and improve their financial lives. Our platform sits at the intersection of financial data, cloud-scale infrastructure, and trust, and it has to work, every time, at massive scale. We’re entering a pivotal phase of our technology evolution: moving from legacy on-prem infrastructure to a modern, cloud-native platform designed for resilience, security, and developer velocity. It’s a chance to build the next generation of MX’s platform thoughtfully, reliably, and with long-term impact in mind. If you enjoy working on complex systems, care deeply about uptime and reliability, and want your work to directly power the financial well-being of millions, this is where you’ll do the most meaningful work of your career. What You’ll Do Drive Cloud Migration: Drive the end-to-end migration of production workloads from on-premise data centers to GCP. Architect for Reliability: Design and implement production-grade Kubernetes environments and GCP architectures that prioritize 99.99+% availability. Operational Excellence: Improve incident triage and improved recovery timeby implementing cloud-aware diagnostics and automated recovery patterns. Empower Developers: Reduce friction in th
About the Role We are seeking an experienced Azure DevOps Engineer to design, implement, and maintain CI/CD pipelines, cloud infrastructure, and automation solutions on Microsoft Azure. This role bridges development and operations, ensuring reliable, secure, and scalable delivery of applications and infrastructure. Location: Hyderabad-India-Onsite Duration: Fulltime Responsibilities Design, build, and maintain CI/CD pipelines using Azure DevOps (Pipelines, Repos, Artifacts, Boards) Architect and manage Azure cloud infrastructure using Infrastructure as Code (ARM templates, Bicep, or Terraform) Automate build, test, and deployment processes across multiple environments Implement and manage containerization and orchestration (Docker, Azure Kubernetes Service) Monitor system performance, availability, and security using Azure Monitor, Log Analytics, and Application Insights Collaborate with development, QA, and security teams to streamline release management Implement Azure security best practices, identity management (Azure AD/Entra ID), and network architecture Manage cost optimization and governance across Azure subscriptions Troubleshoot production issues and support incident response Document infrastructure, pipelines, and operational procedures Required Qualifications Microsoft Certified: Azure Solutions Architect Expert, Azure Administrator Associate (required) 3+ years of hands-on experience with Azure DevOps and Azure cloud services Strong experience with Infrastructure as Code (Bicep, ARM templates, or Terraform) Proficiency scripting in PowerShell, Bash, or Python Experience with Git version control and branching strategies Solid understanding of networking, security, and identity concepts in Azure Experience with containerization (Docker) and orchestration (Kubernetes/AKS) Familiarity with monitoring and logging tools (Azure Monitor, Application Insights) Preferred Qualifications Additional certifications: Azure DevOps Engineer Expert Experience with multi-
Are you looking for an opportunity to help solve one of today's biggest business challenges? AI is changing the pace of business, and organizations everywhere are struggling to help their workforce, partners, and customers keep up. At Litmos, we're building the Learning Acceleration Platform that helps organizations build human capability faster—and we're looking for people who are passionate about making a meaningful impact for customers while growing alongside a collaborative, people-first team. Litmos is the Learning Acceleration Platform that helps organizations build capability faster, adapt at the speed business changes, and scale learning to anyone, anywhere. Combining an intuitive platform, AI-powered capabilities, trusted content, expert services, and a broad ecosystem of integrations, Litmos helps organizations accelerate workforce productivity, improve customer adoption and retention, enable high-performing partners, and reduce organizational risk through continuous learning. Organizations such as Hewlett Packard Enterprise, Graco, Sabre, and Russell Mineral Equipment trust Litmos to accelerate learning across their workforce, partners, and customers. Today, more than 11K customers with 30 million learners across 150 countries and 37 languages use Litmos to build the capabilities their organizations need to succeed. Backed by Francisco Partners, one of the world's leading technology investment firms, we're investing in the future of learning—and the people who are building it. Learn more at www.litmos.com . About the Role We are looking for a senior, self-directed SOC Analyst to join our lean global security operations team based in Pune. As the sole security operations resource during non-US business hours, you will own the full spectrum of alert triage, incident investigation, and response for your shift. This role requires someone who can function as both an L2 investigator and an L3 responder depending on the situation, without rel
Bloomreach is building the world’s premier agentic platform for personalization .We’re revolutionizing how businesses connect with their customers, building and deploying AI agents to personalize the entire customer journey. We're taking autonomous search mainstream, making product discovery more intuitive and conversational for customers, and more profitable for businesses. We’re making conversational shopping a reality, connecting every shopper with tailored guidance and product expertise — available on demand, at every touchpoint in their journey. We're designing the future of autonomous marketing , taking the work out of workflows, and reclaiming the creative, strategic, and customer-first work marketers were always meant to do. And we're building all of that on the intelligence of a single AI engine — Loomi — so that personalization isn't only autonomous…it's also consistent.From retail to financial services, hospitality to gaming, businesses use Bloomreach to drive higher growth and lasting loyalty. We power personalization for more than 1,400 global brands, including American Eagle, Sonepar, and Pandora. Senior Staff Security Engineer The Senior Staff Security Engineer owns current and target-state data architectures and reporting while also designing, implementing, and monitoring cloud (AWS/GCP) infrastructure security controls; deploying, securing, configuring, and operating SIEM and other security resources; identifying, triaging, and remediating infrastructure and web vulnerabilities; leading incident triage and external-researcher engagement; mentoring junior staff; and helping shape secure, scalable approaches for AI-enabled tooling, automation, and emerging product capabilities. Role summary and core responsibilities 6+ years of relevant experience Candidates must demonstrate proficiency in cloud security, network security, URL filtering, common security frameworks, and CVE lifecycle management Prac
TextNow is on a mission to make communications affordable and accessible for everyone. As a full MVNO operating our own mobile core network over LTE and 5G NSA, we have the unique advantage of controlling our network infrastructure end-to-end. We operate the HSS, PGW, and other critical network functions, giving us the flexibility to innovate and deliver exceptional service to millions of users. About the Role Join us in our mission to break down barriers to communication and free the flow of conversation for people everywhere. T extNow is looking for a new SecOps team member to secure, monitor , and enable automated response within our infrastructure. What You’ll Do Ensure Secure & Reliable Systems: Design, implement, and maintain security-focused infrastructure to protect TextNow’s services while ensuring reliability and scalability. Security Automation & Infrastructure as Code: Develop and enforce best practices using Terraform, Ansible, Crowdstrike , and AWS security tools , ensuring secure configurations, automated compliance checks, and infrastructure as code. Threat Detection & Incident Response: Participate in an on-call rotation to respond to security incidents, investigate vulnerabilities, and implement proactive measures to prevent future threats. Work closely with engineering teams to remediate security risks. Monitoring & Logging for Security: Improve observability by implementing security monitoring solutions, logging best practices, and alerting mechanisms to detect anomalies and suspicious activity. Access Control & Identity Management: Manage IAM roles, permissions, and policies to ensure least privilege access and enforce security controls across cloud and internal systems. Collaboration & Security Advocacy: Wo
We believe communication belongs to everyone. We exist to democratize phone service. TextNow is evolving the way the world connects and that's because we're made up of people with curious minds who bring an optimistic, yet critical lens into the work we do. We're the largest provider of free phone service in the nation. And we're just getting started. Join us in our mission to break down barriers to communication and free the flow of conversation for people everywhere. TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between! This role is about impact at scale. You’ll shape how TextNow builds and operates its systems in an AI-first environment where intelligent tooling is embedded into everyday engineering practice. Using AI is not optional, it’s expected. From design and architecture to implementation, testing, debugging, documentation, and operational analysis, you will actively leverage AI tools to increase velocity, improve code quality, and make better technical decisions. We provide a robust suite of AI-powered development tools and workflows to support you, and we expect you to continuously evolve how you use them to raise the bar for efficiency, clarity, and product excellence across the organization. What You'll Do Ensure System Reliability: Design, build, and maintain scalable, resilient, and highly available systems to support TextNow’s infrastructure and services. Automation & Infrastructure as Code: Develop and maintain automation using Terraform, Ansible, and other tools to enable efficient deployment, scaling, and operations of cloud-based systems (AWS preferred). Incident Response & On-Call Support: Participate in an on-call rotation, troubleshoot issues, and drive incident resolution to minimize downtime and improve syste
About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. Database Platform Manager Company: THG Ingenuity Location: Manchester (Head Office) Reports to: Director of Data Role Overview We are looking for a strong technical leader to run our database platform. Our estate runs on Google Cloud Platform following a recent migration to self-managed services, and there is a real opportunity here to shape the next phase. How we consolidate and modernise the estate, where managed and cloud-native services earn their place, which new technologies are worth adopting, and how the platform scales with the business are all live questions. You will lead the thinking on them and work with stakeholders across the business to agree and deliver the roadmap. You will lead our DBA team, who own every database across the group regardless of the application running on it and who run the platform as a 24/7 service. You will work hand in hand with our Data Reliability Engineering team, whose Principal Engineer is your peer and whose focus is automation, fleet reliability and SLOs. This role is weighted toward hands-on technical depth. You should be as credible in a design review or an incident bridge as you are in a planning session with senior stakeholders. Key Responsibilities Technical direction The technical roadmap for the estate. We want someone genuinely interested in emerging database technologies who evaluates them on merit and can articulate the pros and c
About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. Database Engineer (DBA) Company: THG Ingenuity Location: Manchester (Head Office) Reports to: Database Platform Manager Role Overview We are looking for a Database Engineer to help run and improve our database platform. Our estate runs on Google Cloud Platform following a recent migration to self-managed services. There is plenty still to do in the next phase of modernising the estate, and you will be hands-on in that work as well as in the day-to-day running of the platform. You will work alongside a small team of database engineers and closely with engineering and infrastructure teams, keeping our database environments secure, reliable and performant. There is real scope to grow here, and you will be supported to deepen your technical expertise and take on more as you do. The role participates in a rotating on-call rota and occasionally requires out-of-hours work to support deployments, maintenance or incident response. Key Responsibilities Database administration Install, configure, maintain and upgrade PostgreSQL and SQL Server environments across development, test and production Ensure database servers are securely configured, patched and compliant with operational standards Own backup, restore and maintenance strategies, and test recovery procedures regularly Maintain database security, access control and auditing practices Cloud and infrastructure &n
About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. Role summary This role is responsible for making sure customers can pay successfully and that the business is not losing revenue through avoidable payment failures. This role will identify issues, bring the right teams together and ensure that changes are implemented and measured. Role purpose Will be responsible for ensuring that all payment methods, payment providers and checkout journeys operate to an optimal standard across the business. The role will monitor payment acceptance, identify performance issues and opportunities, and work with payment service providers, acquiring banks, technology teams, product teams and commercial stakeholders to improve successful payment completion and reduce avoidable customer friction. The successful candidate will take ownership of payment-performance reporting, incident escalation and optimisation activity, ensuring that issues are identified quickly, investigated thoroughly and converted into measurable improvements in revenue and customer experience. Key responsibilities Payment performance monitoring Monitor payment performance across cards, digital wallets, PayPal, BNPL and other alternative payment methods. Track authorisation, acceptance and completion rates by brand, market, payment method, provider, issuer, device and customer journey. Identify deterioration, unusual trends and performance gaps requiring investigation. Develop appropriate alerts and thresholds so material payment issues
Who Are We HALA is a leading fintech player in the MENAP region that aims to redefine financial services and build the future bank of SMEs. HALA aims at empowering SMEs to start, run, and grow their businesses by providing them with cutting-edge financial and technological tools. HALA currently holds multiple entities in UAE, Saudi Arabia and Egypt (including HALA Payments and HALA Logistics) and offers solutions that enable merchants to digitize their payments as well as manage their sales and operations. Founded in 2017, HALA is currently licensed by the Saudi Arabian Central Bank. Job Description Operations Risk & Governance Manager Job Purpose The Operations Risk & Governance Manager ensures that operational processes, controls, and product launches are well governed, compliant, and managed within HALA’s risk appetite. The role works closely with Operations, Product, Technology, Risk, Compliance, Information Security, and Internal Audit to identify risks, strengthen controls, manage incidents, and support safe product launches. Key Responsibilities Operational Risk and Controls Maintain the Operations risk register and track mitigation actions. Conduct risk and control assessments across operational processes. Review the design and effectiveness of key operational controls. Identify control gaps and ensure corrective actions are completed. Monitor operational risk indicators, incidents, losses, and recurring issues. Escalate material risks and control failures to management. Product Launch Governance Manage the governance process for new product and service launches. Ensure all launches complete the required risk, compliance, operational, technology, security, and customer-impact assessments. Coordinate launch-readiness reviews with Product, Technology, Compliance, Risk, Operations, Finance, and Customer Experience. Confirm that operating procedures, controls, training, customer support, reporting, and incident processes are ready before launch. Maintain
DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,000+ members, DataHub's metadata graph provides in-depth context of AI and data assets with best-in-class scalability and extensibility. The company's enterprise SaaS offering, DataHub Cloud, delivers a fully managed solution with AI-powered discovery, observability, and governance capabilities. Organizations rely on DataHub solutions to accelerate time-to-value from their data investments, ensure AI system reliability, and implement unified governance, enabling AI & data to work together and bring order to data chaos. About the Role We're seeking an experienced DevOps/ Site Reliability Engineering (SRE) Engineer to join DataHub and drive the reliability, scalability, and operational excellence of our platform offerings. In this role, you'll work on technical initiatives across DataHub Cloud and our emerging enterprise deployment solution, which provides customers with enhanced control and flexibility for running DataHub in their preferred environments. Key Responsibilities Enterprise Platform Development: Partner with product and engineering teams to influence the development of advanced deployment capabilities. Collaborate with cross-functional teams to help build systems for seamless installation, upgrade, and rollback processes across various environments. Influence the design and help implement comprehensive monitoring and health check systems for distributed deployments. Partner with engineering teams to help develop self-healing and automated remediation capabilities. Platform Reliability and Operations: Establish and maintain SLAs/SLOs for both cloud and enterprise offerings. Lead incident response and post-mortem processes to drive continuous improvement. Optimise system performance, capacity planning, and cost efficiency. Work closely with product, engineerin
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Engineering at Brex Engineering at Brex is about building systems that scale with speed and intention. Our teams span Software, Data, Security, and IT, and operate with high autonomy and deep collaboration. We tackle hard technical problems, own our outcomes, and push for excellence at every level — from architecture to deployment. It’s an environment where engineering is a craft, and builders become leaders. What you’ll do As a Senior Software Engineer, Infrastructure (Release Engineering) at Brex, you will design, build, and operate the core systems that power Brex’s release, observability, and incident management processes. You will partner closely with product, platform, and operations teams to ensure releases are safe, fast, and reliable, and that our infrastructure scales securely as Brex grows. Where you’ll work This role will be based in our San Francisco office. We are a hybrid environment that combines the energy and connect
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Engineering at Brex Engineering at Brex is about building systems that scale with speed and intention. Our teams span Software, Data, Security, and IT, and operate with high autonomy and deep collaboration. We tackle hard technical problems, own our outcomes, and push for excellence at every level — from architecture to deployment. It’s an environment where engineering is a craft, and builders become leaders. What you’ll do As a Senior Software Engineer, Infrastructure (Release Engineering) at Brex, you will design, build, and operate the core systems that power Brex’s release, observability, and incident management processes. You will partner closely with product, platform, and operations teams to ensure releases are safe, fast, and reliable, and that our infrastructure scales securely as Brex grows. Where you’ll work This role will be based in our New York office. We are a hybrid environment that combines the energy and connections
Get new incident commander jobs by email
Daily job updates · Unsubscribe anytime