Jobiba hiring network

Lead Software Engineer Distributed Systems Jobs

6,753 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead software engineer distributed systems jobs. Use filters to narrow by work mode, employment type, experience and date posted.

M
Mongodb
📍 Toronto• Full-time• From C$158K/yr
1mo ago

MongoDB Search and Vector Search allows users to execute complex search queries and build RAG applications using the MongoDB Query Language. Our users are free to focus on relevance and data retrieval instead of the machinery needed to search data at scale. We are looking to speak to candidates who are based in Toronto for our hybrid working model. Candidate Profile: 5+ years of hands-on experience designing, building, testing, and maintaining industrial-strength backend software in a complex codebase Proficient in modern programming languages and techniques Experienced in developing distributed systems, cloud services, and SaaS products Excellent verbal and written technical communication skills; enthusiasm for collaborating closely with colleagues and mentoring other engineers A growth mindset and the desire to learn quickly through taking on challenges, reflecting on outcomes, and incorporating feedback A strong sense of ownership over their work, from initial design all the way through maintaining code in production You will: Build and design our integrated search platform, written in Java Work with a collaborative team that prioritizes sound technical decision-making and building systems that our customers love and that we are proud of as engineers Lead projects and own subsystems Help determine the team’s roadmap and the architecture of our system Success measures: In 3 months you’ll have contributed to the development of an existing project and completed several improvements or bug fixes In 6 months you’ll be reviewing code and project designs, and be an active participant in team meetings In 12 months you’ll have a thorough understanding of the systems the team owns and have led a project. You’ll have had a positive impact on our code, product, and team processes About MongoDB MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the data platform for the AI era, enabling builders to cr

javamongodbaws
View job →
O
Okta
📍 Toronto• Full-time• From C$160K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Streaming Foundations team builds services and operates data pipeline infrastructure to support event streaming, messaging, and analytics use cases. We are looking for a Software Engineer who is passionate about distributed systems, platform engineering, and solving data-intensive problems at scale. In this high-impact role, you will get to work with engineers throughout the organization to build foundational infrastructure that allows Auth0 to scale for years to come. What you’ll be doing Help set the technical direction for the team and influence the engineering roadmap for the Platform’s streaming capabilities Design and lead the implementation of our most complex and critical systems for data-intensive use cases. Research and champion new technologies and architectural patterns to solve strategic challenges and scale the platform. Lead and influence cross-functional initiatives, ensuring technical alignment and successful execution across multiple teams. Improve the operational posture of our systems by designing for observability, reliability, and scalability, and by mentoring others in operational best practices. Coach and mentor senior engineers and act as a technical leader across the engineering organization. Collaborate with different stakeholders like product teams whenever needed. What you’ll bring to our teams 7+ years of software development experience in a fast-paced, agile environment Experience working with Golang or Java is preferred H

typescriptjavareact
View job →
D
Datadog
📍 New York• Full-time• From $244K/yr
1mo ago

Role Summary: Datadog is seeking a Staff Software Engineer to help shape the future of our Bring Your Own Cloud (BYOC) Logs offering by unifying observability pipelines with log management software that customers deploy and manage in their own infrastructure. This role will focus on building and scaling systems that process, route, and store high-volume observability data within customer-managed infrastructure. You will operate as a hands-on technical leader, driving architecture, cross-team delivery, and product direction across a complex and evolving space. This is a high-impact opportunity to influence product strategy, mentor engineers, and solve deeply technical challenges at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Make customer-controlled deployments feel like a managed Datadog product: deployment, upgrades, configuration, observability, diagnostics, reliability, and secure operation across diverse customer cloud environments Build and scale high-throughput systems for log processing, routing, and transformation across distributed environments Lead cross-team initiatives, aligning engineers, product managers, and stakeholders to deliver complex, multi-team projects Design and implement software that runs reliably that customers deploy and operate within their own cloud infrastructure. Improve system performance, scalability, and cost efficiency through thoughtful trade-off analysis and capacity planning Contribute hands-on to critical code paths, debugging, and deployment challenges in customer environments Who You Are: You have significant experience building software that is installed, deployed, and operated in customer environments rather than only as a fully managed SaaS service. You have strong expertise in distributed systems,

awsazuregcp
View job →
M
Mongodb
📍 San Francisco• Full-time• From $126K/yr
1mo ago

Join the Atlas Search Query team to design and develop the next generation of Search query architecture, optimization, and execution. Atlas Search is a growing cloud service that allows users to execute complex search and vector search queries using the MongoDB Query Language. Our users can focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team is building a cloud-based distributed system responsible for the core components of search including data ingestion, performance, query language, query execution, for both relevance-based search and vector search. Our product is being adopted quickly and there are many interesting projects. This is a technical role where you will be responsible for the success of complex Search Query feature development. We are looking to speak to candidates who are based in San Francisco, CA for our hybrid working model. What You’ll Do Lead complex projects across the MongoDB ecosystem, for instance, development of a new Search aggregation framework within the MongoDB aggregation framework. Set project level strategy, architect features, and lead projects to successful execution Identify, design, and implement features enhancing our query language, performance, and operability Perform code reviews with peers and make recommendations on how to improve our software development processes Influence and grow team members through active mentoring and leading by example What We Look For 5+ years experience in data management/search systems, ideally with a strong query processing and optimization background Experienced in the development and maintenance of stateful distributed systems Eager to shape the technological direction of a complex system and have the ability to lead initiatives through collaboration with others Experienced in debugging and profiling multithreaded applications written in Java and Rust Bonus: experience with designing high-volume query engines, such as a datab

javamongodbaws
View job →
M
Mongodb
📍 San Francisco• Full-time• From $126K/yr
1mo ago

Join the Atlas Search team to design and develop the next generation of Semantic and Vector Search infrastructure. Atlas Search is a growing cloud service that allows users to execute complex search and vector search queries using the MongoDB Query Language. Our users can focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team is building a cloud-based distributed system responsible for the core components of search including data ingestion, performance, query language, query execution, for both relevance-based search and vector search. Our product is being adopted quickly and there are many interesting projects. This is a technical role where you will be responsible for the infrastructure and features enabling our at-scale cloud service powering vector and semantic search. We are looking to speak to candidates who are based in the San Francisco Bay Area for our hybrid working model. What You’ll Do Lead complex projects across the MongoDB ecosystem, for instance, development of a new Search deployment framework within the MongoDB managed cloud Set project level strategy, architect features, and lead projects to successful execution Identify, design, and implement features enhancing our reliability, performance, security and efficiency Perform code reviews with peers and make recommendations on how to improve our software development processes Influence and grow team members through active mentoring and leading by example What We Look For 5+ years experience in data management/search systems, ideally with a strong distributed systems and infrastructure background Experienced in the development and maintenance of concurrent, stateful services Eager to shape the technological direction of a complex system and have the ability to lead initiatives through collaboration with others Experienced in writing features, debugging and optimizing multithreaded applications written in Java Familiarity with LLM

javamongodbaws
View job →
L
Lyft
📍 Toronto• Full-time• $136K – $170K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Marketplace teams are at the heart of our products and decision-making, owning everything from rider pricing to driver earnings, incentives, and efficient matching. We’re looking for passionate, driven engineers to build systems that empower our riders and drivers to have the best transportation experience possible through prediction, adaptivity, and personalization. We’re looking for someone who is excited about working in a fast-paced, innovative, and impactful environment to create reliable solutions to distributed computing, ML, and data problems. The Pricing team is a centerpiece of Lyft’s Marketplace org, determining prices for all rideshare products and supporting new initiatives. We work with Product & Science to solve and implement complex pricing requirements, balancing the needs of riders, drivers, and the business goals. As an owner of one of the most critical flows in the company, you will work on a wide array of challenges such as latency-sensitive concurrency problems, large scale distributed systems, and experimentation. If you’re interested in playing a large part in demand / supply management and improving the Lyft customer experience, this could be a great fit for you. Responsibilities: Help define the roadmap and architecture based on technology and business needs Unblock, support, effectively communicate, and obtain buy-in across teams to achieve results Lead projects of multiple people from idea to positive execution Write clear, scalable and clear design documentation Write well-crafted, well-tested, readable, maintainable code Utilize your expertise in Python, Golang, AWS to deliver robust and scalable solutions Participate in code reviews to ensure code quality and distribute knowledge Proactively participate in resolving ongoing incidents Share your kno

pythonawsazure
View job →
C-
CLEAR - Corporate
📍 New York• Full-time• $225K – $300K/yr
19 days ago

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. Today, CLEAR is well-known as a leader in digital and biometric identification, reducing friction for our members wherever an ID check is needed. We’re looking for a Senior Software Engineer to establish our Observability framework and foundations. You will join us to accelerate building and scaling our innovative systems that support our growing identity platform. You will drive on Observability best practices to find and fix gaps in our observability and our overall systems. You will also lead practices such as load testing, capacity planning, game days, chaos testing, and incident post-mortems. What You Will Do: Embed within the Engineering pillar to deeply understand the product and implement observability across all key flows Facilitate and build load testing cases, ensuring we understand the limits and scaling factors of our services and systems Contribute to observability and support the design of new services and systems, ensuring highly reliable and scalable concepts are implemented Build and lead practices such as game days, chaos engineering, and failure analysis Build long-term capacity plans, with an eye toward reliability and cost-efficiency Who You Are: 6+ experience writing production-grade software in a modern language, such as Java and Python. Strong knowledge of distributed systems concepts (think CAP theorem), microservices architecture, and distributed tracing . Experience with modern observability systems such as Datadog. Experience with performance debugging tools and patterns. You should be able to read a f

pythonjavagit
View job →
G
Godaddy
📍 India• Full-time
21 days ago

Location Details: Remote, India At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team... This position is for a Staff Software Engineer within GoDaddy’s engineering organization, concentrating on developing scalable, fault-tolerant systems and promoting technical excellence across various teams and projects. As a Staff Software Engineer, you will serve as a technical leader at the division level. You will impact system design and architecture, as well as how engineering is carried out across teams. This position suits engineers who excel in uncertain environments, like tackling difficult technical problems, and can provide clarity, structure, and delivery for large projects. Our teams focus on greenfield and open-ended projects that require firm technical ownership, architectural vision, and collaboration across functions. You will assist in establishing engineering standards, support teams through mentoring, and create systems that boost reliability, scalability, and developer efficiency organization-wide. This role is important for encouraging innovation throughout the engineering organization. It includes promoting AI-assisted development workflows, modern cloud-native methods, and reusable platform features that speed up development across teams. What you'll get to do... Technical Execution & Delivery Lead delivery of complex engineering initiatives while driving predictable execution and long-term technical quality Architect and implement highly available, fault-tolerant distributed systems using AWS, PostgreSQL/Aurora, and MongoDB Architect and lead greenfield projects from concept to production, bringin

pythonreactnode.js
View job →
C
Coinbase
📍 United States• Full-time• Remote• From $218K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Staff Software Engineer on the Core Automation team within the Platform group, you'll architect and build the Agentic AI systems that are transforming how Coinbase operates. This team is reimagining customer support and compliance processes for a fully AI-driven world, designing intelligent agents, orchestration frameworks, and measurement systems that deliver delightful customer experiences at scale. You'll own the technical direction for production AI systems, working across cross-functional teams to bring this vision to reality while building primitives that scale automation across the company. What you'll do: Architect and build Agentic AI systems that power Coinbase's compliance automation and other Operations, from intelligent agents through orchestration and guardrails Design foundational APIs and measurement frameworks that ensure AI agents are grounded, relevant, and reliably deliver customer delight with minimal hallucination Lead technical direction for distributed systems underpinning AI automation, defining architecture patterns and strategic roadmaps in partnership with engineering leadership Build reusable primitives and orchestration solutions that enable AI-powered automation to scale across multiple domains beyond the initial customer support and compliance focus Mentor engineers on AI system design techniques, coding standards, and production-

REMOTEawsaigo
View job →

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity At New Relic, we provide our customers real-time insights, so they can innovate faster. Our software delivers insightful observability tools across different technologies and distributed systems, enabling software engineering teams to quickly identify, understand and tackle issues, analyze performance and get the most of their software and infrastructure. Database Observability is a critical pillar of New Relic's platform strategy. We are looking for an experienced Engineering Manager (M3) to lead a senior, high-performing team building next-generation Database Observability products. Your team will own the entire lifecycle of critical telemetry data flows from lightweight database agents and high-throughput ingestion pipelines to intelligent DB recommendation engines and autonomous DB AI Agents. You will lead a team that includes senior and Lead-level engineers with deep domain expertise in distributed systems and AI. Your primary value will come from setting strategic technical direction, enabling their best work, and fostering a high-accountability culture while partnering closely with Product and Design to deliver features that directly drive New Relic's Database Observability. What you'll do Manage a full-stack engineering team (6–8 engineers) spanning backend systems, database telemetry, agent engineering, and UI workflows. Own end-to-end delivery sprint planning, roadmap execution, system quality, and operational excellence for critical database ingesti

pythonjavasql
View job →
N
Nuro
📍 Mountain View• Full-time• From $160.4K/yr
1mo ago

Who We Are Nuro is a self-driving technology company on a mission to make autonomy accessible to all. Founded in 2016, Nuro is building the world’s most scalable driver, combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the Nuro Driver™, to support a wide range of applications, from robotaxis and commercial fleets to personally owned vehicles. With technology proven over years of self-driving deployments, Nuro gives the automakers and mobility platforms a clear path to AVs at commercial scale, empowering a safer, richer, and more connected future. About the Team A rider taps "request ride" and within seconds, an autonomous vehicle is matched, dispatched, and on its way. The On-Road Experience team owns the systems that make that moment happen. The On-road Experience team builds the real-time distributed system behind Nuro's ride-hailing product. This is foundational infrastructure where correctness and performance have consequences on a real road. About the Role This is the perfect role if you love architecting performant, reliable distributed systems, processing real-time vehicle data, and building the foundational APIs that enable seamless product experiences. The position demands technical excellence, a deep understanding of system design, and a knack for solving tough problems. Come join us in defining the future! What you'll own Backend services for ride planning, vehicle matching, and real-time trip execution Partner API integrations (Uber and others) routing ride requests to Nuro vehicles Telemetry pipelines surfacing live vehicle state to riders and operations About You 4+ years of backend engineering experience with strong Go proficiency Experience designing and operating distributed systems where reliability isn't optional End-to-end product ownership — from requirements through launch and iteration Solid hands-on experience with GCP (CloudSQL, BigQuery, Redis, or equivalent) A track record of technical lead

sqlredisgcp
View job →
W-
Wolt - English
📍 Helsinki• Full-time
18 days ago

About Wolt At Wolt, we create technology that brings joy, simplicity and earnings to the neighborhoods of the world. In 2014 we started with delivery of restaurant food. Now we’re building the delivery of (almost) everything and you’ll find us in over 500 cities in 30 countries around the world. In 2022 we joined forces with DoorDash and together we keep on dreaming big and expanding across the globe. Working at Wolt isn’t always easy, but it’s definitely exciting. Here you’ll learn more, build more, and ship more than in most other companies. You’ll be challenged a lot, but also have a lot of fun on the way. So, if you’re a self-starter with drive and entrepreneurial spirit, this could be the ride of your life. What you’ll do: Build and maintain high-throughput backend services using Go . Collaborate with product managers, designers, and frontend developers to ship features that support internal support agents across the globe. Design systems that are scalable , resilient , and easy to maintain. Lead and contribute to architectural discussions and technical decision-making. Write well-tested code and help the team maintain high code quality standards. Our humble expectations: 7+ years of professional software engineering experience, with a proven track record of building and scaling complex systems. 2+ years of production experience in Golang , with the ability to mentor others and drive best practices across the team. Strong hands-on experience with both SQL and NoSQL databases — especially Cassandra. Solid understanding of designing and operating low-latency, high-throughput distributed systems . Nice to have Background in Node.js or other backend languages. Familiarity with cloud infrastructure (AWS, GCP) and event-driven architectures. Previous on-call experience , with a pragmatic approach to reliability and incident management. What we value A product-oriented mindset — you think beyond the ticket, understand the “why” behind the work, and aim to create real

node.jssqlaws
View job →
GW
Get Well Network
📍 Bengaluru• Full-time
19 days ago

Staff Software Engineer Bengaluru, Karnataka, India Opportunity Get Well is seeking a visionary and technically adept Staff Software Engineer to architect, design, develop, and optimize our cloud-native healthcare platform while driving the adoption of AI-First and AI-Augmented Software Engineering practices. This role is pivotal in shaping the future of software development at Get Well by combining deep technical expertise with modern AI-assisted engineering workflows. As we evolve toward an AI-First engineering organization, this leader will champion the use of Generative AI, AI development assistants, and Agentic AI to improve developer productivity, software quality, and engineering velocity. The ideal candidate brings deep expertise in software architecture, cloud-native application development, AI-enabled engineering, and distributed systems. This role provides technical leadership across multiple engineering teams, ensuring high standards for architecture, code quality, reliability, security, and AI adoption. This is a hands-on leadership role where strategic thinking meets deep engineering execution within a complex healthcare environment. This position reports to the Director, Product Development and requires close collaboration with software engineers, AI engineers, product managers, DevOps, QA, and compliance specialists. Key Responsibilities Technical Leadership Define and drive the architecture of scalable, distributed healthcare platforms. Champion AI-First Software Engineering practices across the development lifecycle. Lead the adoption of AI-Augmented Development , including spec-driven development, AI-assisted coding, code reviews, testing, and documentation. Establish engineering standards, Cursor/AI coding guidelines, reusable patterns, and governance for responsible AI usage. Provide hands-on leadership in architecture, coding, design reviews, debugging, and performance optimization. Mentor enginee

pythonjavareact
View job →
R
27 days ago

Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Tokenization team's mission is to bring both public and private equity to be tradable 24/7 on decentralized exchanges (DEXs) across the Robinhood Chain. We work at the intersection of blockchain technology, financial market infrastructure, and modern distributed systems - with a focus on accessibility, security, and scalability. Equity markets have been closed nights and weekends for a century; we're building the infrastructure that changes that. As a Senior Software Engineer , you will design and build the core systems that issue, custody, and settle tokenized equities on-chain and keep them in sync with traditional brokerage and ledger systems. You'll lead technically complex projects end to end, make architectural decisions on a foundational Robinhood initiative, and work in a domain where correctness is non-negotiable - the systems you build move real customer assets, continuously, across multiple chains and jurisdictions. This role is based in our Toronto, ON office, with in-person attendance expected at least 3 days per week . At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office

pythonjavaaws
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Marketplace teams are at the heart of our products and decision-making, owning everything from rider pricing to driver earnings, incentives, and efficient matching. We’re looking for passionate, driven engineers to build systems that empower our riders and drivers to have the best transportation experience possible through prediction, adaptivity, and personalization. We’re looking for someone who is excited about working in a fast-paced, innovative, and impactful environment to create reliable solutions to distributed computing, ML, and data problems. The Pricing team is a centerpiece of Lyft’s Marketplace org, determining prices for all rideshare products and supporting new initiatives. Rider Engagement develops rider-facing engagement levers and optimizes user pricing experience to drive both short term and long term business outcomes. We work with Product & Science to solve and implement complex pricing requirements, balancing the needs of riders, drivers, and the business goals. As an owner of one of the most critical flows in the company, you will work on a wide array of challenges such as latency-sensitive concurrency problems, large scale distributed systems, and experimentation. If you’re interested in playing a large part in demand / supply management and improving the Lyft customer experience, this could be a great fit for you. Responsibilities: Drive high-impact projects and innovate new solutions to provide the best user experience. Work closely with cross-functional teams and partner teams to develop solutions based on technology and business needs, and advance team’s goals and priorities Independently lead features from idea to positive execution and launch Unblock, support and communicate with internal partners to achieve results Write well-crafted, well-tested, readable, maintaina

pythonawsrest
View job →
🔔

Get new lead software engineer distributed systems jobs by email

Daily job updates · Unsubscribe anytime