Jobiba hiring network

Staff Infrastructure Engineer Jobs

3,518 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current staff infrastructure engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

D
Datadog
📍 Portugal• Full-time• Remote
1mo ago

This role is part of Datadog’s Security Agent team, which powers critical security capabilities across Workload Protection, Vulnerability Management, Cloud Security products, and other emerging security offerings. As a Staff Software Engineer, you will lead the design and development of low-level Linux instrumentation and runtime security technologies that help customers detect threats, monitor system activity, and protect cloud-native workloads at scale. You will work on complex technical challenges involving eBPF, Linux kernel internals, performance-sensitive systems, and large-scale data collection while influencing technical direction across multiple product teams. This role offers significant ownership, broad organizational impact, and the opportunity to shape the future of Datadog’s security platform. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the architecture and development of security agent capabilities that power runtime threat detection and workload protection across Datadog Security products. Design and build reusable eBPF-based monitoring functionality for process, file, and network visibility within Linux environments. Drive end-to-end delivery of new features, from technical strategy and design through implementation, testing, and rollout. Establish and evolve testing methodologies that improve platform coverage, detection quality, reliability, and performance. Partner with product, security, infrastructure, and engineering teams to deliver shared platform capabilities used across multiple Datadog products. Provide technical leadership by influencing engineering direction, mentoring peers, and helping resolve complex cross-functional challenges. Who You Are: You have significant experience building software in Linux environments,

REMOTElinuxaigo
View job →
D
Datadog
📍 Remote; Spain, Remote• Full-time• Remote
1mo ago

This role is part of Datadog’s Security Agent team, which powers critical security capabilities across Workload Protection, Vulnerability Management, Cloud Security products, and other emerging security offerings. As a Staff Software Engineer, you will lead the design and development of low-level Linux instrumentation and runtime security technologies that help customers detect threats, monitor system activity, and protect cloud-native workloads at scale. You will work on complex technical challenges involving eBPF, Linux kernel internals, performance-sensitive systems, and large-scale data collection while influencing technical direction across multiple product teams. This role offers significant ownership, broad organizational impact, and the opportunity to shape the future of Datadog’s security platform. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the architecture and development of security agent capabilities that power runtime threat detection and workload protection across Datadog Security products. Design and build reusable eBPF-based monitoring functionality for process, file, and network visibility within Linux environments. Drive end-to-end delivery of new features, from technical strategy and design through implementation, testing, and rollout. Establish and evolve testing methodologies that improve platform coverage, detection quality, reliability, and performance. Partner with product, security, infrastructure, and engineering teams to deliver shared platform capabilities used across multiple Datadog products. Provide technical leadership by influencing engineering direction, mentoring peers, and helping resolve complex cross-functional challenges. Who You Are: You have significant experience building software in Linux environments,

REMOTElinuxaigo
View job →
D
1mo ago

This role is part of Datadog’s Security Agent team, which powers critical security capabilities across Workload Protection, Vulnerability Management, Cloud Security products, and other emerging security offerings. As a Staff Software Engineer, you will lead the design and development of low-level Linux instrumentation and runtime security technologies that help customers detect threats, monitor system activity, and protect cloud-native workloads at scale. You will work on complex technical challenges involving eBPF, Linux kernel internals, performance-sensitive systems, and large-scale data collection while influencing technical direction across multiple product teams. This role offers significant ownership, broad organizational impact, and the opportunity to shape the future of Datadog’s security platform. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the architecture and development of security agent capabilities that power runtime threat detection and workload protection across Datadog Security products. Design and build reusable eBPF-based monitoring functionality for process, file, and network visibility within Linux environments. Drive end-to-end delivery of new features, from technical strategy and design through implementation, testing, and rollout. Establish and evolve testing methodologies that improve platform coverage, detection quality, reliability, and performance. Partner with product, security, infrastructure, and engineering teams to deliver shared platform capabilities used across multiple Datadog products. Provide technical leadership by influencing engineering direction, mentoring peers, and helping resolve complex cross-functional challenges. Who You Are: You have significant experience building software in Linux environments,

linuxaigo
View job →
M
Mongodb
📍 Ireland• Full-time
1mo ago

We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate actionable alerts to safeguard critical workloads. The Collections team is a newly formed team within MongoDB's Observability & Adoption Org focused on making telemetry onboarding and collection significantly easier across MongoDB. We own key parts of the observability collection stack, including onboarding experience, telemetry collection agents across the data and control planes, and ingestion services for metrics, logs, and traces that support both internal and customer observability in MongoDB, driving insights, recommendations, and alerting. Our mission is to reduce friction for teams implementing and iterating on Observability while partnering closely with development teams to instrument their services using shared best practices, helping define the conventions our telemetry should follow, and building collection and ingestion systems that are stable, performant, secure, well-documented, and self-service. We also work closely with the Data Pipeline and Storage & Query teams to help ensure MongoDB has a stable and performant observability stack end to end. This is an opportunity to join a team shaping how observability works across MongoDB and to have outsized impact on both the developer experience and the reliability of the platform underneath it. As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality observability data for internal and external use cases means we need to continually innovate and scale our systems to the next level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Grafana, Sp

javamongodbaws
View job →
M
Mongodb
📍 Ireland• Full-time
1mo ago

We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate actionable alerts to safeguard critical workloads. The Collections team is a newly formed team within MongoDB's Observability & Adoption Org focused on making telemetry onboarding and collection significantly easier across MongoDB. We own key parts of the observability collection stack, including onboarding experience, telemetry collection agents across the data and control planes, and ingestion services for metrics, logs, and traces that support both internal and customer observability in MongoDB, driving insights, recommendations, and alerting. Our mission is to reduce friction for teams implementing and iterating on Observability while partnering closely with development teams to instrument their services using shared best practices, helping define the conventions our telemetry should follow, and building collection and ingestion systems that are stable, performant, secure, well-documented, and self-service. We also work closely with the Data Pipeline and Storage & Query teams to help ensure MongoDB has a stable and performant observability stack end to end. This is an opportunity to join a team shaping how observability works across MongoDB and to have outsized impact on both the developer experience and the reliability of the platform underneath it. As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality observability data for internal and external use cases means we need to continually innovate and scale our systems to the next level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Grafana, Sp

javamongodbaws
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

We are seeking a Staff Engineer to join our growing Gurugram Products & Technology team to provide technical direction, direct architecture, and implement core parts of a new platform we are building to make it easier for customers to build AI applications using MongoDB. As a Staff Engineer on this new team, you will be responsible for providing technical leadership to teams developing cutting edge technologies related to enabling deployment at scale of AI applications. You will take on challenging, high-visibility projects that improve and enhance the performance, scalability, and reliability of the distributed systems infrastructure for this new product. MongoDB engineering teams pride themselves on building high-quality software and living MongoDB cultural values every day – we value intellectual curiosity and honesty, and building together in an environment that prioritizes collaboration over competition. We are looking to speak to candidates who are based in Gurugram for our hybrid working model. Position Expectations Work closely with research, product management, product engineering, product design, peers as well as other teams within the company to define the first version and future evolution of the service Design, build and deliver well-tested core pieces of the platform in collaboration with other vested parties Contribute to shaping architecture, code reviews and development practices, developer experience as the teams and product grow Mentor fellow engineers and assume ownership and accountability of projects Qualifications Strong background in building core components for high scale compute and data distributed systems 8+ years experience of building distributed systems, and/or foundational cloud services at scale and an interest in working with Python, Go and Java Proven success in designing, writing, testing, debugging, performance tuning, possessing a strong grip on the foundational materials of computer science and maint

pythonjavamongodb
View job →
M
Mongodb
📍 California• Full-time• From $211K/yr
1mo ago

The Application Modernization Platform (AMP) team is tackling one of the industry's most critical challenges: leveraging Generative AI to transform rigid, legacy applications into modern, microservices-based architectures powered by MongoDB. We are building a comprehensive, SaaS-like platform, encompassing both the "brain" (multi-agent reasoning and orchestration) and the "hands" (the deployment platform and modernization toolset). This solution requires a robust platform foundation and infrastructure designed for a "build once, run anywhere" model, ensuring seamless operation regardless of a client's security or network constraints. A key challenge is balancing the need to tune our tools for each customer's unique tech stack and restrictive environments with making them easily extensible and scalable for common application modernization challenges. We seek an engineering leader for this high-visibility initiative. This role requires defining the high-level strategy and technical direction across all AMP engineering pillars, leading the execution of solving uniquely complex application modernization puzzles, and delivering an enterprise-grade product. The leader will minimize deployment friction, meet customer compliance requirements, and help shape the future of how global enterprises leverage GenAI. The ideal candidate is a hands-on technical leader who excels at leveraging GenAI capabilities, architecting complex distributed systems, and designing the orchestration agents necessary to reliably and fluidly run the entire software development lifecycle. This role will be based in North America's West Coast (PST), and offers a hybrid working model. The ideal candidate for this role will have 10+ years of software development and operations experience, with a focus on building platforms and distributable software infrastructure Deep experience in building data warehouses and core components for data processing systems Have experience in using GenAI in building comple

mongodbawsazure
View job →
M
Mongodb
📍 Dublin• Full-time
1mo ago

The worldwide data management software market is massive (IDC forecasts it to be $138 billion by 2026). At MongoDB, we are transforming industries and empowering developers to build amazing apps that people use every day. We are the leading modern data platform and the first database provider to IPO in over 20 years. Join our team and be at the center of innovation and creativity. MongoDB is seeking a Sr. Staff Software Engineer to join the Atlas Core Data Services organization. The organization is responsible for building MongoDB Atlas, our database as a service offering and fastest growing product, along with the API Platform and Developer Tools. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. The Atlas Core Data Services organization builds the software that manages the Atlas cluster infrastructure hosted on the three major cloud providers (AWS, Azure, and GCP), as well as the software that manages the MongoDB database hosted on that infrastructure. We are constantly challenged to design features that ensure Atlas clusters are secure, available, durable, and performant while running large-scale, critical workloads. The Sr. Staff Engineer in this role will drive innovation across the organization and the company, setting technical standards and direction that enable future growth and velocity. We are looking for engineers with the experience and high standards needed to lead at that scale. Our organization champions a strong culture of inclusivity, diversity, and collaboration. If you want to be a deeply technical leader on a collaborative team that applies systems expertise to build the foundational infrastructure of a popular database, join us. Let's build a faster, more reliable, and highly scalable database platform together. We are looking to speak to candidates who are based in Dublin for our hybrid working model. Responsibilities Define standards and vision for the mission-critical Atlas SaaS data

javamongodbaws
View job →
M
Mongodb
📍 British Columbia• Full-time• From C$209K/yr
1mo ago

The Application Modernization Platform (AMP) team is tackling one of the industry's most critical challenges: leveraging Generative AI to transform rigid, legacy applications into modern, microservices-based architectures powered by MongoDB. We are building a comprehensive, SaaS-like platform, encompassing both the "brain" (multi-agent reasoning and orchestration) and the "hands" (the deployment platform and modernization toolset). This solution requires a robust platform foundation and infrastructure designed for a "build once, run anywhere" model, ensuring seamless operation regardless of a client's security or network constraints. A key challenge is balancing the need to tune our tools for each customer's unique tech stack and restrictive environments with making them easily extensible and scalable for common application modernization challenges. We seek an engineering leader for this high-visibility initiative. This role requires defining the high-level strategy and technical direction across all AMP engineering pillars, leading the execution of solving uniquely complex application modernization puzzles, and delivering an enterprise-grade product. The leader will minimize deployment friction, meet customer compliance requirements, and help shape the future of how global enterprises leverage GenAI. The ideal candidate is a hands-on technical leader who excels at leveraging GenAI capabilities, architecting complex distributed systems, and designing the orchestration agents necessary to reliably and fluidly run the entire software development lifecycle. This role will be based in North America's West Coast (PST), and offers a hybrid working model. The ideal candidate for this role will have 10+ years of software development and operations experience, with a focus on building platforms and distributable software infrastructure Deep experience in building data warehouses and core components for data processing systems Have experience in using GenAI in building comple

mongodbawsazure
View job →
O
Okta
📍 Bellevue, Washington; Chicago, Illinois; New York, New York; San Francisco, California; Washington, DC• Full-time• From $194K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Team The Site Reliability team is dedicated to architecting and owning the foundational infrastructure tooling and CI/CD platforms that support Okta’s SRE ecosystem. In this development-focused role, you will leverage a modern tech-stack to build durable, automated systems that maximize platform reliability and engineering velocity. The ideal candidate is someone who enjoys analyzing systems and identifying areas of opportunity to improve system performance, availability and capacity. They are part systems administrator, part network administrator, and part developer. What you’ll be doing Maintain a highly available cloud infrastructure edge for the Okta identity platform Automate AWS infrastructure with Terraform and/or Chef Evolve the system by introducing changes to improve efficiency, scalability, and velocity What you’ll bring to the role 8+ years of operations experience configuring, deploying, monitoring and troubleshooting applications and

pythonawsdocker
View job →
O
Okta
📍 Bengaluru, India• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. What You’ll Be Doing Design, build, and operate highly scalable, reliable, and secure infrastructure powering our production systems across AWS and GCP. Lead major reliability and modernization initiatives, including container platform migrations (e.g., ECS to EKS/GKE) and microservice enablement across multi-cloud environments. Serve as a technical authority in Kubernetes (EKS and GKE), cloud infrastructure (AWS and GCP), and modern CI/CD practices (GitOps, automation pipelines). Partner with development teams to architect and enable microservice-based applications, ensuring production readiness, scalability, and observability. Implement and manage infrastructure as code (Terraform, Ansible) to automate provisioning, scaling, and configuration management across multiple cloud providers. Drive improvements in observability, performance, and cost efficiency through robust monitoring, logging, and alerting systems that span AWS and GCP. Champion SRE best practices — defining SLOs/SLIs, conducting blameless postmortems, and continuously improving incident response. Lead complex technical projects from

pythonsqlpostgresql
View job →
O
Okta
📍 Bellevue, Washington; Chicago, Illinois; Toronto, Ontario, Canada• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Role We are looking for a highly curious, self-directed Full-stack engineer who thrives on solving complex technical challenges for our Digital Technology team. You are a versatile developer comfortable navigating everything from front-end interfaces (React/Next.js) and CMS environments (AEM) to cloud infrastructure (Vercel/Cloudflare). You possess an 'ownership' mindset: you don’t wait for assignments, you identify gaps and help drive plans, troubleshoot production incidents through to resolution, and drive technical alignment between frontend, backend, and design teams. You have a sharp eye for great UX/UI. You enjoy partnering closely with designers to refine interactive experiences, ensuring that technical implementations not only work flawlessly but also feel intuitive and polished for the end user. You are a proactive innovator who experiments with new technologies, including AI-driven tools, to optimize our systems and improve developer workflows. You will: Participate in cross-team initiatives from end to end, including code reviews, design reviews, operational robustness discuss

javascriptjavareact
View job →
O
Okta
📍 Bengaluru, India• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Job Overview: We are looking for a highly experienced Staff Engineer to join the FGA Developer Tooling team and lead the evolution of our end to end developer experience across both OSS and SaaS. This team owns the SDKs in Go, JavaScript, .NET, Python, Java and other languages, along with CLI workflows, IDE integrations, GitHub automation, developer documentation, and release strategy. All development is done in the open as open source, and we actively welcome and review community contributions. Our guiding principle is One developer experience, many deployment models. As a Staff Engineer, you will define technical direction, ensure cross language consistency, influence API design in partnership with FGA Core, and raise the quality bar across all developer facing tooling. <h2 id="Respo

javascripttypescriptpython
View job →
O
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Fullstack Staff Software Engineer - PAM Okta is the identity standard. The Okta Identity Cloud is an independent and neutral platform that securely connects the right people to the right technologies at the right time. We help organizations do two things - secure and manage their extended enterprise, and transform their customers' experiences. With over 14,000 customers, 7000+ app integrations, and well over 200 million registered users, we are only getting started. The Okta Privileged Access Management (PAM) is an identity-centric approach to a common and critical privileged access use case. Our elegant Zero Trust architecture is purpose-built for the modern cloud and helps customers solve challenging security and operations pain points at scale. We're looking for a staff-level fullstack engineer to join a team of highly skilled and talented team players who're proud of what they own and deliver. Our elite team is fast, creative, and flexible; with a weekly release cycle and individual ownership, we expect great things from our engineers and reward them with stimulating new projects, new technologies, and the chance to have significant equity in a company that is changing the cloud computing landscape forever. You Will Leverage cutting-edge AI pair-programmers and LLMs (such as Copilot and Claude) to accelerate the development of secure, enterprise-grade Privileged Access Management (PAM) products. Proven expertise leveraging AI

javascriptjavareact
View job →
O
Okta
📍 Washington• Full-time• From $174K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Technology, Data & Intelligence Team Okta’s Technology, Data & Intelligence (TDI) team delivers the systems, tools, and services that power internal operations across the company. From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology. The Staff Site Reliability Engineer Opportunity Okta Federal, Inc. is looking for an experienced Staff TDI Site Reliability Engineer to help build, improve, and maintain our cloud platform services that help Okta support the most sensitive national security missions. The Site Reliability Engineering team delivers foundational infrastructure capabilities that enable corporate engineering teams to operate securely, reliably, and at scale. You’ll play a key role in designing and implementing complex cloud-based engineering enablement systems, while ensuring compliance with strict government requirements. What you’ll be doing Operate and maintain enterprise grade solutions within air-gapped environments. Build, run, and monitor development tools, pipelines, and infrastructure with a security-first mindset. Operate autonomously within secure facilities. Maintain SLOs/SLIs for workloads with no dependency on external monitoring or SaaS tooling. Own runbooks and incident response procedures tailored to limited external escalation paths. Participate in POA&M remediation and support annual/recurring Authority to Oper

pythonawsci/cd
View job →
🔔

Get new staff infrastructure engineer jobs by email

Daily job updates · Unsubscribe anytime