Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Staff Engineer on the FinHub - Ledger team within the Platform group, you'll own the architectural direction of Coinbase's core ledgering systems that process billions of transactions. This team builds the foundational infrastructure that powers financial accuracy and reliability across the company. You'll make critical technical decisions, drive platform improvements, and lead high-impact cross-functional projects that define how we build scalable, highly available systems at Coinbase's growing scale. What you’ll do: Architect and build foundational backend systems with a focus on performance, scalability, and reliability across ledgering and transaction processing infrastructure Drive strategic technical direction for complex, cross-team initiatives spanning engineering, product, and finance Establish and evolve platform best practices, frameworks, and architectural standards that raise the quality bar for the broader engineering organization Provide deep technical mentorship, guide design decisions, and develop senior engineers across the team Shape the team's technical roadmap by collaborating with engineering leadership to identify high-leverage investments and anticipate scaling bottlenecks Embed AI into engineering and operational practices to drive system improvements and workflow efficiency Required skills and experience: 8+ years building and ope
Jobiba hiring network
Staff Infrastructure Engineer Jobs
3,518 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current staff infrastructure engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. The Global Support & Partnerships team is a critical and unique part of helping Lyft succeed in our purpose to Serve and Connect. We support Lyft's international growth by developing integration platforms that seamlessly link global customers to an exceptional customer care experience. We go beyond simply supporting customers; we build the foundational infrastructure that turns every support interaction into a moment of genuine connection. Whether supporting our established North American community or welcoming new users worldwide, we ensure a unified, reliable, and efficient experience for riders, drivers, applicants and support agents alike. We are looking for an experienced technical leader who can support our global ambition by designing, owning, and scaling the Global Support Platform. Responsibilities: Set the technical vision and strategy for a rapidly evolving product area, making high-judgment calls on architecture and technology direction Drive the adoption of AI and machine learning solutions to optimize customer support workflows and enhance operational efficiency across the platform. Lead a team of talented engineers who ship code and tackle hard engineering problems, maintaining a high bar for technical excellence Translate high-level business goals into actionable engineering projects. Own the technical roadmap from conception to delivery, managing cross-team dependencies and mitigating risks. Drive the responsible adoption of AI development tools across engineering teams - modeling effective use, establishing best practices, and mentoring engineers to improve productivity without compromising code quality or security. Champion improvements in system, observability, performance, and tech debt reduction, extending your influence beyond your immediate team. Establish best practices f
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The Developer Infrastructure organization’s mission is to enable Airbnb engineers and AI Coding Agents to ship high quality software with confidence and speed. We work on infrastructure that accelerates Airbnb’s engineering productivity across all stages of the build/test/deploy software development lifecycle. The Difference You Will Make: Build infrastructure has the potential to be one of the highest-leverage investments for our engineering organization as we attempt to accelerate the iteration loop of engineers and AI agents operating within our developer platform. To deliver these improvements, you will operate within the CI+Build team, which owns developer infrastructure that includes remote build systems, CI clusters, flaky test management, and merge queues. The team’s goal is to keep mainline green and PR cycle times fast. Your role will give you the opportunity to define and deliver world-class build systems at Airbnb’s scale, while driving technical direction for build infrastructure spanning multiple platform teams. Our internal customer teams deliver systems used by product engineers across backend, web, and mobile. Furthermore, to be successful in this role, you will need to be comfortable diving into the application architecture of some of our largest internal backend services as we manage a build graph that continues to grow in complexity. A Typical Day: Architecting improvements to our existing Bazel installation, and remote build/cache infrastructure to improve correctness, reproducibility, and speed. Setting the technical roadmap for software build
Discord has a highly engaged community of millions of daily active users who use the platform for many different reasons, but there’s one thing that nearly everyone does: play video games. Discord plays a uniquely important role in the future of gaming, and we are focused on making it easier and more fun for people to hang out before, during, and after playing games. We're looking for someone who gets excited about building data infrastructure at massive scale and cares deeply about the gaming communities we serve. Someone with passion for building lovable products for Discord users and Discord engineers. We're building the next generation Data Platform that powers decisions for one of the most vibrant platforms in the world. If you're the kind of Staff Engineer who sees ambiguity as an invitation — who builds systems that make whole teams faster, shapes technical strategy across the org, and wants your decisions to directly impact millions of gamers worldwide — we want to talk to you. To learn more about Discord's Data Platform, read our engineering blog, including how we built our modern data stack leveraging open-source tools! This position is based in our San Francisco office. What You'll Be Doing Own the technical direction of major Data Platform initiatives like GRC, scaling, or stream processing, from problem to production system and post-launch health Be a force multiplier: build reusable primitives, unblock teammates, and raise the quality bar across the team Stay a top code contributor. You write real production code while helping others reach the same bar Partner with XFN stakeholders to translate business problems into clear technical strategy Champion debate, decide, and commit and 80/20 thinking to drive robust technical solutions supporting multiple team What you should have 7+ years of software engineering experience with a track record of driving large, ambiguous technical initiatives end-to-end Expertise in distributed systems and data infrastructu
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Staff Software Engineer - External Observability Platform Location: Bellevue, WA (Hybrid: 3 days/week in-office) Team: Infrastructure & Observability Platform Engineering About the Role Snowflake’s Data Cloud processes exabytes of data across multi-cloud global environments every day. Delivering seamless reliability and real-time visibility to thousands of global enterprise customers requires an Observability Platform built on hyper-scalable backend distributed systems. We are seeking a Staff / Lead Software Engineer to architect, design, and scale our External Observability Platform . In this role, you will lead the technical strategy for customer-facing telemetry, system metrics, audit logs, distributed tracing, and actionable operational insights. You will build high-throughput, low-latency infrastructure capable of ingesting, processing, and serving petabytes of telemetry data with strict SLA guarantees. You will join a team of world-class engineers in our Bellevue, WA office. To be successful, you must be deeply technical, capable of leading complex cross-functional architecture initiatives, and skilled at mentoring senior engineers while holding your own with the brightest technical minds in the industry. Key Responsibilities Architect & Scale Distributed Infr
Here’s a summary of the role: As a Staff Ruby on Rails Engineer at Diligent, you will set technical direction for secure, scalable, and high-performing SaaS applications and services that power our governance platform. This role is ideal for you if you thrive on solving the hardest technical problems, shaping architecture across multiple teams, and driving how the organization builds software — including how we responsibly adopt AI into our engineering practices and products. You will own critical systems end-to-end, partner closely with Product, Security, DevOps, and other engineering leaders to shape technical roadmaps, and mentor engineers across the department. A key part of this role is helping define our AI strategy: embedding AI responsibly into our systems and workflows, advising on where AI tools and capabilities can meaningfully improve delivery, and raising the bar on how the team uses them safely and effectively. Here’s a breakdown of what you’ll do (not all of it, just the important stuff): Champion the design, delivery, and evolution of secure, scalable Ruby on Rails applications and services, driving architecture decisions across multiple teams and codebases, with responsibility for scalability, reliability, and the underlying infrastructure required to run them effectively. Set technical direction for major projects and platform initiatives, from solution design and prototyping through implementation and production ownership. Own and evolve the technical roadmap by identifying, shaping, and driving new technical initiatives and investments. Develop and evolve web applications and services with a strong focus on scalability, maintainability, reliability, and long-term platform health. Identify systemic pain points across services and propose pragmatic architectural improvements, including decomposition of monoliths and evolution toward service-oriented or microservices patterns where it adds valu
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role This is where security meets innovation at enterprise scale. As a staff security engineer, applications at WRITER, you'll be building the security foundations that protect the AI systems powering some of the world's most recognizable brands. You'll work at the intersection of application security, AI infrastructure, and developer enablement—partnering with engineering teams to embed security into every line of code while ensuring our platform remains both powerful and trustworthy. The opportunity is massive: you'll help define how enterprise AI applications are secured, from threat modeling our LLM architectures to building automated security controls that scale across our growing platform. This isn't about saying "no"—it's about finding creative ways to say "yes, and here's how we do it securely." You'll tackle challenges that most security engineers never encounter: securing AI agents, protecting training data pipelines, and designing controls for systems tha
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Responsibilities and Duties We are seeking a highly skilled System Tests & Diagnostics Engineer to develop, extend, and integrate specialized silicon validation and diagnostics tools for next-generation AI SoCs. Unlike traditional validation roles focused on executing test plans, this position is responsible for developing the diagnostic software and stress tools that expose hardware failures, characterize silicon behavior, and improve platform observability throughout bring-up and validation. You will work closely with Arm engineers to understand and extend existing diagnostics technologies while developing Graphcore-specific capabilities for future AI hardware. Role Summary You will work with existing Arm-developed diagnostics technologies and extend them to support Graphcore's next-generation AI silicon. You will be responsible for developing system-level diagnostics and stress tools that integrate with an existing framework to detect data integrity, computational correctness, performance, and reliability issues across CPUs, AI accelerators, memory, storage, PCIe, firmware, BMC, and other platform components. Examples include silent data corruption (SDC) tests, power transient stress tools, and platform diagnostics, with opportunities to develop new diagnostics as future hardware capabilities evolve. This role requires close collaboration with hardware architects, firmware enginee
Join Truecaller – The place where innovation meets impact! Truecaller's mission is to build trust in communication by making it safer, smarter, and more efficient. Born in Sweden, trusted by the world, and here’s why we stand out: We are trusted by over 500 million active users every month across 190+ countries We identify over 15 billion calls daily, helping users avoid spam and scams We are powered by a team of 400+ employees from 45+ nationalities We always look for people who take initiative, own their work, and keep raising the bar. An entrepreneurial mindset matters here, especially when it turns bold ideas into real actions. We stay collaborative and focused, always searching for smarter paths forward. If you want to make an impact and grow with a team that inspires millions, you’ll fit right in. The role: As a Staff Software Engineer, Backend, you will lead multi-team initiatives, leveraging your deep technical expertise to ensure alignment, quality, and long-term maintainability across the organization. You will balance hands-on technical guidance with strategic leadership, establishing engineering best practices that transcend your immediate team. By mentoring experienced engineers and building strong relationships with cross-functional stakeholders, you will drive project delivery and ensure our backend infrastructure remains scalable and robust enough to support millions of users worldwide What you will do Lead the team to ensure we build our product efficiently and effectively. Influence product priorities and scope, offering valuable input in decision-making. Communicate complex decisions clearly, weighing the pros and cons of different solutions. Mentor and guide experienced engineers, helping them develop both technically and culturally. Leverage your experience to establish best practices across the engineering team and beyond. Drive project delivery by aligning with cross-functional stakeholders and ensuring the team has everything needed. L
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About the Role & Team Every AI insight, every experiment, every cohort at Amplitude starts with a query. Our in-house OLAP engine, Nova , processes trillions of events in real time — turning raw behavioral data into fast, trustworthy answers that power decisions for thousands of product teams worldwide. We’re entering a world where AI agents don’t just assist product teams — they ship features, run experiments, and make prioritization calls autonomously. What makes that possible is agents’ ability to verify their work against real product data continuously. That makes Nova the critical infrastructure in the loop, and as non-stop agents become the main source of queries, the demand on Nova’s throughput, correctness, and operational rigor grows dramatically. We’re looking for a Staff Software Engineer who wants to go deep on both the engine internals and
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. About the Team DevX owns engineering velocity at Amplitude: build systems, CI/CD, developer environments, and internal tooling. We're building a software factory: automated workflows that remove manual bottlenecks from how engineers ship code. The Role We're looking for a Staff Software Engineer – DevX (Hybrid – San Francisco) who bridges infrastructure and application thinking and can accelerate how the whole team develops, tests, and ships in the cloud. You'll set architecture for our developer platform, lead our software factory work, and push our development model toward cloud-first workflows. This is a high-leverage, low-oversight role. You'll own initiatives end to end, from an ambiguous problem to production, and set technical direction for a foundational team. What You'll Do Cloud development platform: Unify and scale our existing loca
Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida/ Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering team. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedite
Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida / Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your product-area specific reliability roadmap, to optimize operations, increase efficiency in our use of cloud resources and our developer’s time, harden security posture, and increase feature velocity of our developers Work closely with multiple teams to optimize the operations of their microservices - and improve the lives of the engineers within your product area engineering team. Responsibilities Support the engineering teams within your product area by maintaining and executing a reliability roadmap of opportunities for improvement for reliability, maintainability, security, efficiency, and velocity - and help for realizing those opportunities. Collaborate with development infrastructure, Global SRE, and your product area engineering teams to establish and continually refine your reliability roadmap. Participate in defining, evolving, and managing SLOs for several teams within your product area. Participate in on-call rotations within your product area to understand operations workload so you can continually work to improve the on-call experience and reduce operational workload for running microservices and related components. Complete projects to optimize and tune on-call experience for your engineering teams. Continually improve the lifecycle of microservices and architectural components from inception and design, through deployment, operation, and refinement. Write code and automation to reduce operational workload, increase efficiency, improve security posture, eliminate toil, and enable Sumo’s developers to deliver features more rapidly. Work closely with the developer infrastructure teams to expedite
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange™️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Sr. Production Engineer to join our team. This role is available as a hybrid opportunity 3 days a week in San Jose, CA or Remote reporting to Production Engineering in the Cloud Infrastructure & Operations department. Join Zscaler to be a force multiplier for the reliability of a global platform processing 200+ billion transactions daily across tens of millions of enterprise users. In this role, you will provide the technical vision and hands-on execution to drive an "automation-first" culture across the company. By maturing our observability and architectural standards, you will directly reduce our Mean Time to Mitigate (MTTM) and shape the scalability of our globally distributed, multi-cloud infrastructure. What you’ll do (Role Expectations) Implement highly available, scalable infrastructure across AWS, GCP, and bare-metal environments Drive an "automation-first" culture by wr
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Engineering at Brex Engineering at Brex is about building systems that scale with speed and intention. Our teams span Software, Data, Security, and IT, and operate with high autonomy and deep collaboration. We tackle hard technical problems, own our outcomes, and push for excellence at every level — from architecture to deployment. It’s an environment where engineering is a craft, and builders become leaders. What you’ll do As a Staff Software Engineer in Banking, you will help shape the technical direction of one of Brex’s most strategic and complex product areas. The Banking org is both a product and platform org, it owns the Brex Business Account product and AP offerings like Bill Pay and Vendors, and the underlying money movement platform and partner integrations that power those experiences. In this role, you’ll work across customer-facing product surfaces and core financial infrastructure, driving architecture, reliability, and
Get new staff infrastructure engineer jobs by email
Daily job updates · Unsubscribe anytime