Jobiba hiring network

Infrastructure Team Manager Jobs

4,730 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current infrastructure team manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

S
1mo ago

Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we’re changing the way people think about and interact with personal finance. We’re a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we’re at the forefront. We’re proud to come to work every day knowing that what we do has a direct impact on people’s lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The Role You will be the technical leader for Digital Identity at SoFi under the Sofi Technology Solution Group: the platform group that powers identity, authorization, and entitlements for every product and every member across the company. Digital Identity runs Tier-0 infrastructure: the highest criticality rating at SoFi. Every product line, banking, lending, investing, credit cards, crypto depends on these platforms to know who a member is, what they're entitled to, and what they're authorized to do. When these platforms are down, SoFi is down. You'll define the technical strategy for this group. You'll architect solutions for complex, ambiguous problems: multi-person access patterns, cross-organizational platform convergence, and data integrity at financial-services scale. You'll build the engineering processes and culture that let a lean team operate Tier-0 infrastructure with confidence. And you'll push the boundaries of how we build, leveraging AI to accelerate development, prototype faster, and experiment with approaches that would have been impractical two years ago. What You'll Own Platform Technical Strategy Digital Identity operates multiple Tier-0 platforms spanning identity resolution, entitlement management, and fine-grained authorization. You own the technical strategy across all of the

gitrestai
View job →
C
Cloudflare
📍 In Office• Full-time
1mo ago

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations Bengaluru, India About the Role Infrastructure Engineering is responsible for the world’s most reliable, observable, performant, and safe network ecosystem. Our customers rely on our products and systems to safely modify, troubleshoot, and release products without external impact. Our external customers rely on us to provide seamless and predictable incident, traffic, policy management, resulting in the fastest and safest netwo

pythonsqlmysql
View job →
A
Asana
📍 Warsaw• Full-time• $372K – $432K/yr
1mo ago

We're looking for a Senior Platform Reliability Engineer who brings strong software engineering skills and a deep understanding of system behavior under load and stress. This role is a good fit for someone who wants to own reliability as a first-class concern – building the foundational systems that protect Asana's platform, not just responding when things go wrong. You'll build core platform systems like load shedding, rate limiting, circuit breakers, and traffic controls that protect Asana under real-world load. This is deep, cross-cutting work that shapes stability and performance of our entire infrastructure – and you'll partner closely with other platform teams to make reliability something that's built in, not bolted on. Our tech stack includes: AWS, Kubernetes (EKS), CloudFront, Istio, Cilium, MySQL (RDS), OpenSearch, DynamoDB, Redis, Terraform, Datadog, TypeScript, Scala, Go, and Python. (Yeah, we know this sounds like buzzword bingo – but we want this post to actually show up in your searches.) Why this role? Reliability as a first-class feature : You won't be patching things up after the fact. You'll build the systems that make Asana resilient by design. Foundational work : Load shedding, traffic management, ingress/egress – these are the building blocks that protect everything else. You'll own them. Strong collaboration, reasonable hours : You'll work closely with infrastructure teams in Warsaw and Reykjavik, making deep collaboration practical without constant timezone gymnastics. Room to grow : This is a new team, and you'll help shape what Platform Reliability Engineering looks like at Asana – whether that means leading projects, mentoring others, or defining our technical direction. In this role, success means shipping systems that other teams rely on by default – because they make the platform safer, not because they're mandatory. We're especially interested in people who think like backend engineers but obsess over failure modes, capacity plan

typescriptpythonsql
View job →
M
Mongodb
📍 Palo Alto• Full-time• From $145K/yr
1mo ago

MongoDB serves as the data layer powering the most important AI applications in the world today. This role exists to ensure the market knows what MongoDB does and to secure the next generation of builders before they default to another platform. The VP of AI Marketing Strategy & Ecosystem will shape the training data, documentation, integrations, and ecosystems that determine what gets built and how, thereby positioning MongoDB as the generational data platform for agentic applications. Builders and agents alike need to reach for MongoDB by default. That requires presence in the systems and tools that influence how applications get built, not just the channels that reach the humans building them. This position will report to the CMO and be part of the Marketing Leadership Team. The VP of AI Marketing Strategy & Ecosystem will work closely with the Chief Product Officers, Chief Technology Officer, Chief Customer Officer, and regional marketing leaders. This role elevates AI visibility and representation, ecosystem partnerships, and developer content into a single VP seat, working closely with product and partner teams to align efforts that currently span the company. This role can be based out of our San Francisco or Palo Alto offices, or remotely in the region. What you will own AI market narrative: own the strategic case for MongoDB as the default data platform for agentic applications and AI-mediated technology selection broadly; ensure the story holds up under scrutiny from analysts and competitors; and ensure it resonates with builders. Partner with Product Management and Product Marketing to keep the AI story integrated into MongoDB's core narrative, not a separate one AI visibility and representation: own how MongoDB appears in AI-generated responses, agent framework recommendations, and developer tool suggestions; This goes beyond content; It includes documentation quality, technical accuracy, and the underlying infrastructure that determines how Mong

reactmongodbaws
View job →
O
Okta
📍 Toronto• Full-time• From C$110K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Team The Auth0 Platform Tools team owns the incident management tooling, Slack-based tooling, StatusPage, and local development environments that Auth0 engineers rely on every day. That includes incident.io and the services we have built around it, Statuspage, custom Slack bot applications that automate our incident response and engineering operations workflows, the customer-facing web application behind status.auth0.com, Vivaldi, and Tilt - the tools engineers use to run Auth0 locally. We are seeking an engineer to help build new features across all of these tools. Our stack is primarily TypeScript and Node.js, with a React and Next.js front end, backed by Postgres and Redis, and deployed on Kubernetes on AWS. A significant portion of our incident and engineering operations automation is built on Tines, a no-code automation platform. Prior no-code experience is welcome, but we expect you to learn Tines here and become effective with it. Current initiatives include extending our incident tooling to meet FedRAMP requirements, taking full ownership of the status page, and improving how we communicate incident status to customers. There is real room to improve along the way, from test coverage to resilience to inherited technical debt. We build for two audiences: Auth0 engineers, who depend on our tooling every day, and Auth0's customers, who rely on the status page during incidents. We are looking for an engineer who cares about both and enjoys working wi

typescriptpythonreact
View job →
R
1mo ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Storage Platform team builds and operates the platform that powers database access across Robinhood. We own relational (Postgres/Aurora), key-value (DynamoDB), and caching systems, along with the SDKs, control plane automation, and data plane services that enable safe and reliable access at scale. Our mission is to standardize and strengthen how services connect to storage, improve reliability and performance, and reduce operational overhead through automation. We manage thousands of databases and hundreds of caching clusters supporting millions of users and critical brokerage workloads. Availability is our highest priority — our systems are designed to meet strict uptime targets, including no downtime during market hours. As a Staff Software Engineer , you will design and evolve the core infrastructure that underpins Robinhood’s storage systems. You’ll lead complex distributed systems initiatives such as horizontal sharding, proxy-based query routing, connection pooling, and cross-shard transactions. You’ll work on improving database reliability, performance, and cost efficiency across multi-region deployments. This role has a direct impact on system availability, laten

vuesqlpostgresql
View job →
R
1mo ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Storage Platform team builds and operates the platform that powers database access across Robinhood. We own relational (Postgres/Aurora), key-value (DynamoDB), and caching systems, along with the SDKs, control plane automation, and data plane services that enable safe and reliable access at scale. Our mission is to standardize and strengthen how services connect to storage, improve reliability and performance, and reduce operational overhead through automation. We manage thousands of databases and hundreds of caching clusters supporting millions of users and critical brokerage workloads. Availability is our highest priority — our systems are designed to meet strict uptime targets, including no downtime during market hours. As a Senior Software Engineer , you will build and improve core infrastructure used by many engineering teams, with a focus on reliability, performance, and operational excellence. You’ll deliver key components of data plane and control plane systems (for example: connection pooling, query routing, automation workflows, and observability) and help evolve patterns for safe, consistent database access. You’ll work closely with peers to design pragmatic s

vuesqlpostgresql
View job →
R
1mo ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Cloud Networking team’s mission is to build scalable, secure, and reliable networking infrastructure that powers communication across all Robinhood services. We enable our engineering teams to build and operate microservices seamlessly by providing foundational networking capabilities that are resilient and transparent. We’re looking for a Staff Software Platform Engineer to design, build, and evolve our foundational platform for large-scale services—focusing on AWS, Kubernetes (K8s) on Amazon EKS, modern networking, and a robust Istio service mesh to deliver secure, reliable, and performant systems. This role is part of the Network Service Discovery and Communication (SDC) team, responsible for service discovery, traffic management, and resilient service-to-service communication across our platform. This role is based in our Bellevue office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams. What you’ll do Lead technical s

pythonvueaws
View job →
R
1mo ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. Robinhood has been tapped as the brokerage and initial trustee for Trump Accounts, and is working alongside BNY and The U.S. Department of Treasury to develop and operate the program, which aims to expand market access for the next generation of Americans. In this role, you will help build a world-class, intuitive platform, leveraging Robinhood’s industry-leading financial technology to deliver a standalone web and app experience for this historic national initiative. The Government Products team focuses on building and scaling systems that support long-term investing and account management for millions of customers. This team works closely with product, data, and infrastructure partners to deliver reliable, secure, and intuitive experiences. You will help shape systems that enable customers to plan for their financial future with confidence! As an iOS Engineer on the Government Products team, you’ll help build mobile experiences for a historic initiative designed to expand market access for the next generation of Americans. You’ll ship intuitive, secure, and scalable features that make long-term investing and account management feel simple and accessible, while partnering w

vueawsrest
View job →
R
1mo ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. Robinhood has been tapped as the brokerage and initial trustee for Trump Accounts, and is working alongside BNY and The U.S. Department of Treasury to develop and operate the program, which aims to expand market access for the next generation of Americans. In this role, you will help build a world-class, intuitive platform, leveraging Robinhood’s industry-leading financial technology to deliver a standalone web and app experience for this historic national initiative. The Government Products team focuses on building and scaling systems that support long-term investing and account management for millions of customers. This team works closely with product, data, and infrastructure partners to deliver reliable, secure, and intuitive experiences. You will help shape systems that enable customers to plan for their financial future with confidence! As an Android Engineer on the Government Products team, you’ll help build mobile experiences for a historic initiative designed to expand market access for the next generation of Americans. You’ll ship intuitive, secure, and scalable features that make long-term investing and account management feel simple and accessible, while partneri

javavueaws
View job →
C
Coinbase
📍 - Canada• Full-time• Remote• From C$191.1K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Software Engineer on the Compute Platform team within the Platform group, you'll own the primary compute orchestration infrastructure that every service at Coinbase runs on. Built largely on CNCF technologies including Kubernetes and Istio, this platform underpins the scalability, reliability, and efficiency of our entire product suite. You'll design and ship tooling, automation, and net-new capabilities that make it easy for hundreds of engineers to deploy and operate critical services, while partnering closely with Security, Reliability, and Observability teams to raise the bar across the stack. What you'll do: Own the design, build, and operation of Kubernetes cluster management tooling and automation that keeps our compute platform reliable and self-healing at scale. Build developer-facing tooling and workflows that improve how engineers across Coinbase interact with Kubernetes, with a heavy emphasis on integrating AI-driven processes and support. Deliver net-new compute capabilities for service owners, such as one-off jobs, cron scheduling, deployment strategies, EFS support, and automated right-sizing. Drive operational excellence by automating toil, reducing on-call burden, and continuously improving platform observability and incident response. Partner with Security, Reliability, and Observability teams to ensure the compute platform meets C

REMOTEawsgcpkubernetes
View job →
C
Coinbase
📍 - Canada• Full-time• Remote• From C$191.1K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Software Engineer, Core Reliability on the Infra Reliability team within Platform , you'll help Coinbase scale 50x by improving reliability, security, and deployment safety across our production environment. This team owns the systems that secure service configurations and secrets, reduce customer-facing incidents, and strengthen deployment infrastructure supporting thousands of services and hundreds of daily releases. You'll lead high-impact reliability projects that make our entire service environment more resilient and safer for customers. What you'll do: Own the design and delivery of reliability projects and features that improve resiliency across Coinbase's service environment in partnership with other engineering teams. Partner with critical T0/T1 services to understand architecture, improve scalability, and reduce operational toil. Build and enhance systems that securely manage service configurations and secrets at scale. Improve canary-based release systems and expand deployment capabilities to support thousands of services and hundreds of daily deployments with fewer incidents. Drive reliability best practices and strengthen reliability culture across engineering teams at Coinbase. Required Skills and Experience: 5+ years of software engineering experience designing, building, and maintaining production services in service-oriented architectures

REMOTEawsazuregcp
View job →
S
Stripe
📍 New York• Full-time
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Secure Devices is responsible for ensuring that every client endpoint at Stripe adheres to our rigorous security standards. Our services play a crucial role in detecting and preventing data loss, restricting software execution to only approved software, and providing attestation capabilities to securely manage device identities. We operate both on-device and backend services across multiple platform types. Our users-first approach ensures that we’re empowering Stripes to be as productive as possible while protecting user data. What you’ll do As a software engineer on Secure Devices, you will work at the intersection of software development, security, and client platform engineering. You will work with teams across Security, Infrastructure and Corporate Engineering to drive strategic projects to better secure Stripe endpoints, build infrastructure for supporting new platforms, and operate services critical to securing over 10,000 Stripe devices. Responsibilities Contribute to the secure design and implementation of Stripe’s mobile expansion initiative Act as the subject matter expert on iOS security by advising partner teams on iOS security best practices and secure-by-design architectures Design, build and maintain Stripe’s endpoint security software. This includes developing telemetry and prevention capabilities via macOS system extensions that run on all Stripe macOS devices Collaborate closely with partner teams to define and measure the

awslinuxrest
View job →
Z
Zscaler
📍 Netherlands• Full-time• Remote
20 days ago

Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer to join our Cloud Infrastructure & Operations team. This is a remote role based in the Netherlands, reporting to the Senior Director, Software Engineering. As a Staff SRE, you will leverage your expertise in Linux/UNIX System Administration to build scalable infrastructure and manage platforms like Kubernetes using automation and high security standards. You will troubleshoot complex Linux networking and security issues, manage firewall technologies, and ensure secure access across our global platforms and applications. What you’ll do (Role Expectations) Create and maintain highly scalable solutions based on KVM LINUX, Kubernetes, and Public Cloud Providers Analyze and troubleshoot systems performance and issues across the OS and Applications Maintain platform security and observability using nftables and robust monitoring tools Manage and deploy systems and s

REMOTEpythonawsdocker
View job →

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Developer Productivity group is responsible for making Stripe’s developers happy and productive. We work on tools, processes, and code refactoring to accelerate Stripe engineering as Stripe scales. We’re looking for people with an interest in building the tools to improve the day to day experience of Engineers in Stripe. The ideal candidate will have a passion for solving developer experience problems, and a pragmatic ability to ship results iteratively—powered by a mix of technical expertise across some or all of: language processing tools; version control systems; build systems; and distributed systems engineering. You’ll be working on a mix of engineer-facing systems, CI infrastructure platforms, and big-data engineering. What you’ll do We have a ton of important work to do, which is why we’re hiring! Our active projects change all the time, but here are a few examples of recent projects so you can get an idea of the types of work we do: Build and manage systems to handle CI at massive scale—including batching, speculative stacking, merge-race inhibition, and more. Build and manage systems to handle our enormous CI test suite–identifying and managing flaky tests, assessing and reproducing flakiness, and optimizing for test effectiveness. Enhance our CI systems for reliability, including adaptive response to available capacity, resilience and self-recovery from outages, and highly-leveraged observability. Detect and isolate code breaka

🔔

Get new infrastructure team manager jobs by email

Daily job updates · Unsubscribe anytime