Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. What You’ll Be Doing Design, build, and operate highly scalable, reliable, and secure infrastructure powering our production systems across AWS and GCP. Lead major reliability and modernization initiatives, including container platform migrations (e.g., ECS to EKS/GKE) and microservice enablement across multi-cloud environments. Serve as a technical authority in Kubernetes (EKS and GKE), cloud infrastructure (AWS and GCP), and modern CI/CD practices (GitOps, automation pipelines). Partner with development teams to architect and enable microservice-based applications, ensuring production readiness, scalability, and observability. Implement and manage infrastructure as code (Terraform, Ansible) to automate provisioning, scaling, and configuration management across multiple cloud providers. Drive improvements in observability, performance, and cost efficiency through robust monitoring, logging, and alerting systems that span AWS and GCP. Champion SRE best practices — defining SLOs/SLIs, conducting blameless postmortems, and continuously improving incident response. Lead complex technical projects from conception to completion, managing timelines, and technical dependencies across teams. Mentor engineers across teams, fostering a culture of reliability, automation, and continuous learning. Collaborate with security and compliance partners to ensure infrastructure adheres to best practices and standards (e.g., IAM Federation, Workload Identity). Participate in t
Jobiba hiring network
Incident Commander Jobs
589 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current incident commander jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. What You’ll Be Doing Design, build, and operate highly scalable, reliable, and secure infrastructure powering our production systems across AWS and GCP. Lead major reliability and modernization initiatives, including container platform migrations (e.g., ECS to EKS/GKE) and microservice enablement across multi-cloud environments. Serve as a technical authority in Kubernetes (EKS and GKE), cloud infrastructure (AWS and GCP), and modern CI/CD practices (GitOps, automation pipelines). Partner with development teams to architect and enable microservice-based applications, ensuring production readiness, scalability, and observability. Implement and manage infrastructure as code (Terraform, Ansible) to automate provisioning, scaling, and configuration management across multiple cloud providers. Drive improvements in observability, performance, and cost efficiency through robust monitoring, logging, and alerting systems that span AWS and GCP. Champion SRE best practices — defining SLOs/SLIs, conducting blameless postmortems, and continuously improving incident response. Lead complex technical projects from conception to completion, managing timelines, and technical dependencies across teams. Mentor engineers across teams, fostering a culture of reliability, automation, and continuous learning. Collaborate with security and compliance partners to ensure infrastructure adheres to best practices and standards (e.g., IAM Federation, Workload Identity). Participate in t
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Position Overview: We are seeking a highly technical Staff Observability Site Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to delivering a world class, comprehensive, scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure as code —utilizing Terraform and strong coding proficiency in Go, Python, or Ruby —to automate the deployment of agents and collectors across complex distributed systems. Key Responsibilities Automated Infrastructure: Design, build, and maintain scalable observability infrastructure using tools like Terraform. Splunk Engineering: Optimize the collection, processing, and storage of log data to ensure high reliability and low latency of our Splunk services Incident Response: Participate in on-call rotations and lead post-incident reviews to drive systemic improvements and "observability-driven development." Automation: Eliminate "toil" by automating the deployment and scaling of observability agents and collectors. Required Skills & Experience (The Essentials) Log Management: Minimum 5+ Experience scaling and managing Splunk Cloud at scale (1000+ SVCs), including Workload Management (WLM) and HEC optimization. Visualization: Expertise in creating intuitive, actionable Splunk dashboards that correlate data across multiple sources. SRE Mindset: Minimum 5+ years of experience in an SRE, Dev
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Principal Forward Deployed Engineer, Okta for AI Agents About Okta for AI Agents Okta secures access for 20,000 organizations and billions of users. Okta for AI Agents extends that work to the agentic shift. Deploying an AI agent is not like deploying traditional software. You are putting professional work output into production, and it needs deep integration, continuous tuning, and change management. Every agent needs an identity, a scope, an audit trail, and a way to be shut down when it goes wrong. Most enterprises have not built this yet. We are. We hire builders who see the cracks in enterprise agent identity that everyone else has learned to live with. The Role You embed inside four to five of Okta’s most strategic enterprise customers as their dedicated technical partner for agent identity. You sit alongside their identity, platform, and security engineering teams, develop sample and bespoke code, and own the technical outcome from prototype through production. You are a builder-consultant. You go past architecture diagrams to code, debug, and ship bespoke agent identity solutions inside the customer’s environment. You ship secure agents faster for the customer, and you feed real field insight back to Okta product engineering. Responsibilities Become the customer’s trusted technical voice on agent security. Sit in their standups, design reviews, and incident reviews. Earn a seat on their architecture review board and security council for agent risk d
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Technology, Data & Intelligence Team Okta’s Technology, Data & Intelligence (TDI) team delivers the systems, tools, and services that power internal operations across the company. From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Staff Site Reliability Engineer Opportunity Okta Federal, Inc. is looking for an experienced Staff TDI Site Reliability Engineer to help build, improve, and maintain our cloud platform services that help Okta support the most sensitive national security missions. The Site Reliability Engineering team delivers foundational infrastructure capabilities that enable corporate engineering teams to operate securely, reliably, and at scale. You’ll play a key role in designing and implementing complex cloud-based engineering enablement systems, while ensuring compliance with strict government requirements. What you’ll be doing Operate and maintain enterprise grade solutions within air-gapped environments. Build, run, and monitor development tools, pipelines, and infrastructure with a security-first mindset. Operate autonomously within secure facilities. Maintain SLOs/SLIs for workloads with no dependency on external monitoring or SaaS tooling. Own runbooks and incident response procedures tailored to limited external escalation paths. Participate in POA&M remediation and support annual/recurring Authority to Oper
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta’s Technology, Data & Intelligence (TDI) team delivers the systems, tools, and services that power internal operations across the company. From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Staff Site Reliability Engineer Opportunity Okta Federal, Inc. is looking for an experienced Staff TDI Site Reliability Engineer to help build, improve, and maintain our cloud platform services that help Okta support the most sensitive national security missions. The Site Reliability Engineering team delivers foundational infrastructure capabilities that enable corporate engineering teams to operate securely, reliably, and at scale. You’ll play a key role in designing and implementing complex cloud-based engineering enablement systems, while ensuring compliance with strict government requirements. What you’ll be doing Operate and maintain enterprise grade solutions within air-gapped environments. Build, run, and monitor development tools, pipelines, and infrastructure with a security-first mindset. Operate autonomously within secure facilities. Maintain SLOs/SLIs for workloads with no dependency on external monitoring or SaaS tooling. Own runbooks and incident response procedures tailored to limited external escalation paths. Participate in POA&M remediation and support annual/recurring Authority to Operate activities. Support and run mission criti
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! As a Security Engineer you will identify and drive impactful projects to improve the security of Figma’s product, platform, and IT systems. We are hiring for multiple teams within Security Engineering: AI Security, Platform Security, Product Security, and Anti-Abuse. This is a remote first role. You will partner closely with teams across the company and focus on systemic security improvements and risk reduction. You will also participate in operational security responsibilities like security reviews, consulting, vulnerability triage, and security incident response. Examples of what you may work on across teams: AI Security Perform technical security assessments, code audits, and design reviews for new AI infrastructure, platforms, and products. Design and develop technical solutions to secure AI models, tooling, debugging workflows, and data pipelines. Advocate for secure practices across Figma’s AI infrastructure, platforms, and data systems. Build the next generation of internal AI-powered access insights and security tooling. Help run penetration testing and offensive security exercises against Figma’s AI infrastructure, platforms, and products. Platform Security Perform technical security assessments, code audits, and design reviews for changes to Figma’s cloud and corporate infrastructure. Design and develop solutions to prevent or mitigate cloud and corporate security risks. Advocate for secure practices within Figma’s cloud and corporate infrastructure. Build platforms and tooling t
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! Figma seeks an AI-Native Performance TPM to own performance prevention, diagnostics, and safe rollout for both our flagship products and next-gen AI features. This platform-level role spans desktop, browser and mobile; requires deep experience with performance testing (load, stress, endurance, interference), observability, and incident response. You’ll partner across Product, Platform, Performance Program and Product Support. Join us to keep Figma snappy quick! This is a full time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: This is a specialist Platform TPM owning horizontal, high-visibility performance programs that span: Flagship product performance (load times, FPS, memory) and AI features (model inference latency, throughput, cost, and reliability) Desktop, browser, native mobile (React Native / WebView), and WASM contexts Cross-org programs: observability/telemetry, regression prevention (performance CI), xfn performance forum, SEV mitigation, safe rollout of AI capabilities, and CE Planning (customer engineering / enterprise readiness) We’d love to hear from you if you have: 5+ years in performance engineering, performance TPM, platform TPM, or SRE with hands-on experience shipping performance programs for SaaS products Demonstrated experience with load, stress, performance or scalability testing and new-build comparisons Deep familiarity with web performance (FCP, LCP), rendering/FPS, WASM memory, mobile profiling (Xcode I
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! Figma's Security team is growing, and we're looking for a Security Operations Manager to lead the strategy and execution of our security operations program. In this role, you'll build and scale the systems, processes, and tooling that help protect Figma and our community. You'll partner closely with Security Engineering, Platform Security, IT, GRC, and Legal to strengthen our detection and response capabilities, improve operational resilience, and help shape the future of our DART and SOC functions. This is a full time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Own Figma's security monitoring and incident response program, from detection engineering through post-incident review and continuous improvement Build and automate security operations workflows, including alert triage, enrichment, investigation, and response actions using SOAR and custom tooling Develop and maintain incident response run books, escalation procedures, and communication plans for security events of varying severity Lead incident response preparedness initiatives, including tabletop exercises, red team engagements, and response capability assessments Improve the effectiveness of our SIEM and SOAR platforms by reducing noise, increasing signal fidelity, and closing detection coverage gaps Build and operationalize threat intelligence capabilities to identify adversary behaviors, prioritize investments, and strengthen detection and response programs Partner wi
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As the Operations Manager for Coinbase Bermuda (CBBM and CBSL), you'll report to the Chief Operating Officer and help run all aspects of our Bermuda operations, from product and regulatory operations to service delivery, technology, and outsourcing oversight. The Operations team drives country-level growth initiatives that align Coinbase Bermuda's business plans to our global strategy. You'll own operational execution across critical business areas and work directly with regulators, banking partners, and cross-functional teams to ensure our Bermuda entities operate with resilience and in full regulatory compliance. What you'll do: Own the oversight and management of operational functions within CB Bermuda and outsourced to intragroup or third-party entities, including implementing policies, managing SLA frameworks, and ensuring delivery in line with Bermuda regulations and contractual obligations. Build and maintain supervision frameworks across critical business areas of Coinbase Bermuda, ensuring alignment with both regulatory requirements and operational best practices. Lead the establishment of operational resilience frameworks, including risk identification, contingency planning, incident response coordination, and communication with regulators and service providers to uphold business continuity. Partner with internal and external stakeholders, including banking
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As the Operations Manager for Coinbase Bermuda (CBBM and CBSL), you'll report to the Chief Operating Officer and help run all aspects of our Bermuda operations, from product and regulatory operations to service delivery, technology, and outsourcing oversight. The Operations team drives country-level growth initiatives that align Coinbase Bermuda's business plans to our global strategy. You'll own operational execution across critical business areas and work directly with regulators, banking partners, and cross-functional teams to ensure our Bermuda entities operate with resilience and in full regulatory compliance. What you'll do: Own the oversight and management of operational functions within CB Bermuda and outsourced to intragroup or third-party entities, including implementing policies, managing SLA frameworks, and ensuring delivery in line with Bermuda regulations and contractual obligations. Build and maintain supervision frameworks across critical business areas of Coinbase Bermuda, ensuring alignment with both regulatory requirements and operational best practices. Lead the establishment of operational resilience frameworks, including risk identification, contingency planning, incident response coordination, and communication with regulators and service providers to uphold business continuity. Partner with internal and external stakeholders, including banking
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . The Core Infrastructure team within Coinbase's Platform product group builds the foundational systems that keep Coinbase online, secure, and scalable, owning the compute and networking platforms that power every product and service across the company. As the Group Product Manager for Core Infrastructure & Reliability, you'll own the product vision and multi-year strategy for Coinbase's cloud infrastructure, driving the design, operation, and scaling of the systems that underpin hundreds of billions of dollars in annual transaction volume. You'll partner deeply with Engineering, SRE, Security, and Finance to ensure Coinbase's infrastructure is reliable, cost-efficient, and resilient across multiple cloud environments and regions. What you’ll do: Own the product strategy and roadmap for Core Infrastructure, spanning compute, networking, multi-region and multi-cloud architecture, and platform reliability. Strengthen infrastructure reliability and resilience programs, defining platform-level SLOs, capacity planning, failover capabilities, and incident reduction targets to meet the uptime demands of a global financial platform. Lead evaluation and adoption of cloud infrastructure technologies (Kubernetes, service mesh, distributed storage, observability, infrastructure-as-code), making build-vs-buy decisions that balance cost, speed, and long-term scalability. A
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . CB Payments, Ltd (CBPL) is Coinbase’s UK‑incorporated electronic money institution, authorised and supervised by the FCA and registered as a cryptoasset exchange and custodian wallet provider, enabling customers to move between fiat and crypto and access Coinbase’s UK platform. This dual‑hatted role serves both as CBPL’s 2LoD Country Risk Manager and as a Risk Manager within Global Operational Risk Management (ORM), running the local risk framework for the UK entity and contributing to Coinbase’s enterprise‑wide operational risk programs so that risks remain within appetite while the UK business grows safely. As a Risk Manager - Country & Operational Risk, you'll serve a dual-hatted role across CBPL (Coinbase's FCA-authorised UK entity) and Global Operational Risk Management. Our Enterprise Risk Management team builds and runs the frameworks that keep Coinbase growing safely within risk appetite across every entity and jurisdiction. You'll own CBPL's second-line-of-defence risk framework covering operational, financial, regulatory, outsourcing, and conduct risks while co-owning global ORM program areas that strengthen our enterprise-wide risk posture. What you'll do: Own CBPL's entity-level risk framework, running RCSAs, incident/issue management, KRI/KCI monitoring, and risk-appetite reporting for the CBPL Risk Committee and Board. Lead 2LoD risk challenge on ne
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Security Operations Specialist on the Security Operations team, you'll defend Coinbase and its customers against some of the most sophisticated attackers in the world. This team - spanning CDRE, Insider Threat, Blockops and Threat Intelligence - protects billions of dollars in digital assets while scaling security coverage for the next billion crypto users. You'll own frontline incident response, build detection and automation capabilities, and partner across the organization to keep Coinbase safe as it launches new Web3 products globally. What you'll do: Own second-line triage and response for security alerts, leading incident management through resolution and driving post-incident improvements Build and maintain runbooks for repeatable response patterns, then define and implement automation to eliminate manual toil Partner with teams across Security Operations to develop monitoring strategies informed by attacker investigation findings Drive Security monitoring and incident response for emerging Web3 product launches Strengthen team capabilities by mentoring peers, sharing knowledge, and participating in 24/7 rotational coverage across time zones Required Skills and Experience: 3+ years of hands-on security operations experience including incident response, alert triage, and network/host forensics across cloud, SaaS, and container environments Demonstrated abi
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Senior Associate, FCM The FCM Operations team supports Coinbase Financial Markets, Inc. across US Futures, Prediction Markets, and Options by managing back-office operations including trade flow management, position management, margin monitoring, regulatory reporting, and business analytics. As a Senior Associate, FCM, you'll own exception management and operational readiness for new product launches, ensuring our derivatives infrastructure scales reliably while meeting regulatory expectations. What you'll do: Own exception management across cash, positions, trades, and margin, including triage, investigation, resolution, and implementation of durable fixes Lead operational readiness for new product launches, covering UAT/test case design, playbooks, cutover plans, and post-launch stabilization Partner with Product, Risk, Compliance, Accounting, and Engineering to design processes, controls, and reporting that scale and meet regulatory expectations Establish monitoring and exception reporting frameworks (SLAs, KPIs, dashboards) and run incident triage with clear escalation paths Drive root cause analysis and post-mortems, translating learnings into product and process improvements Identify and implement automation and tooling to reduce operational friction and manual work Required Skills and Experience: 5+ years in financial services or fintech operations with FCM, b
Get new incident commander jobs by email
Daily job updates · Unsubscribe anytime