Jobiba hiring network

Distributed Systems Engineer Jobs

1,306 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current distributed systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

S
Supabase
📍 Remote• Full-time
1mo ago

Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. We're looking for an engineer to help build the future of distributed Postgres. You'll work on Multigres, our open-source distributed database system that brings horizontal scaling to Postgres. In this role, you'll architect and implement critical distributed database infrastructure including sharding, consensus protocols, and materialization systems. You'll collaborate closely with our Postgres, networking, and infrastructure teams to push the boundaries of what's possible with distributed databases. What You’ll Be Responsible for: Design and implement query routing logic for sharded databases Build consensus and replication systems to support distribute durability Develop materialization pipelines for migrations and change data capture Contribute to connection pooling infrastructure and intelligent workload isolation Collaborate with the open-source community on Multigres development You Might Be a Good Fit If You have: This role requires deep technical expertise in distributed databases and systems. For detailed qualifications, see our contributor qualifications document . Key areas of expertise: Database sharding, relational algebra, and Postgres internals Consensus protocols (Raft, Paxos, FlexPaxos) and distributed transactions Stream processing, materialization, and change data capture Building robust, performant distributed systems with strong observability Low-latency infrastructure and network protocol optimization. What We Offer Fully Remote We hire globally. We believe you can do your best work from anywhere. There are no Supabase offices, but we provide a WeWork membership or co-working allowance you can use anywhere in the world. ESOP Every team member receives ESOP (equity ownership) in the comp

aigorust
View job →
C
Cohere
📍 Toronto• Full-time• From £215K/yr
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About the role. We’re building the next generation of agentic AI infrastructure at Cohere. This team sits at the intersection of ML systems, distributed infrastructure, and developer experience, creating the platform that powers autonomous AI agents at scale. You’ll work on hard, forward-looking problems with few established patterns, including secure code execution, agent state management, model routing, identity and authentication, and resource management for long-running agent workflows. This role is a strong fit for someone who combines systems depth with ML intuition. You should be comfortable building reliable infrastructure, thinking through distributed systems tradeoffs, and understanding how emerging agentic capabilities shape platform design. What you’ll work on. Secure execution environments for agent-generated code Identity, authentication, and trust boundaries for agents Model routing and orchestration across different model types and environments Rate limiting, quotas, and resource management for agent workflows State management, memory, and filesystem abstractions for agents. In this role you will: Turn emerging M

kubernetesgitrest
View job →
C
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by building high-performance, scalable and reliable machine learning systems? Do you want to help define and build the next generation of AI platforms powering advanced NLP applications? We are looking for Members of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will work closely with many teams to deploy optimized NLP models to production in low latency, high throughput, and high availability environments. You will also get the opportunity to interface with customers and create customized deployments to meet their specific needs. You may be a good fit if you have: 5+ years of engineering experience running production infrastructure at a large scale Experience designing large, highly available distributed systems with Kubernetes, and GPU workloads on those clusters Experience with Kubernetes dev and production coding and support Experience with GCP, Azure, AWS, OCI, multi-cloud on-prem / hybrid serving Experienc

awsazuregcp
View job →
P
1mo ago

About Pinecone Pinecone is the knowledge infrastructure for AI at scale. Its leading vector database and knowledge engine, Pinecone Nexus, power accurate, performant AI applications for more than 9,000 customers and 800,000 developers worldwide. Pinecone's mission is to make AI knowledgeable. Pinecone is based in New York and raised $138M in funding from Andreessen Horowitz, ICONIQ, Menlo Ventures, and Wing Venture Capital. About the Team and Role: Join a team that builds robust, real-time distributed systems for a cutting-edge database. We care about performance, reliability, scalability, and most of all learning and having fun together. Whether you’re a seasoned coder or just getting started, if you’re passionate about technology and eager to learn, you’ll fit right in. Who we are: We show up to work, ready to collaborate and build technologies that make a difference, with people who genuinely care. We chase improvements such as tail latencies, bytes throughput, cache hit rate, and operational cost efficiency. We believe learning is ongoing and that even the most complex problems can have simple solutions. What You’ll Do: Collaborate with teammates to design and build database features that power AI applications. Learn how to tune performance and support reliability in distributed systems (don’t worry, we’ll guide you). Help Pinecone run smoothly on popular cloud providers. Take ownership of your work and grow your skills every day. Have fun. Who You Are: 5+ years of work experience - programming in Rust, Go, C++, or a comparable language. You’re genuinely curious about distributed systems and eager to dive deep into technical challenges. You approach problems with creativity and persistence, and you’re comfortable asking thoughtful questions or seeking feedback. You’re excited to learn, value constructive feedback, and appreciate mentorship. Bonus Points: You have hands-on experience with cloud platforms (AWS, GCP, Azure) or have demonstrated an ability to pick u

awsazuregcp
View job →
O
Okta
📍 Toronto• Full-time• From C$184K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Get to know the team The Developer Platform team at Auth0 (an Okta company) owns the platform that developers build identity on. Increasingly they build alongside AI agents, and that raises the bar on everything underneath: interfaces have to hold up whether a person or an agent is calling them, and the systems behind them have to stay reliable and coherent as usage grows. We move fast, we own problems end to end, and we care deeply about the platform we put in front of the developers and agents who depend on it. The opportunity We're hiring a Principal Engineer (P5) to serve as the technical leader and compass for the Developer Platform team. You'll work across the breadth of the platform, tackling the highly complex, vaguely specified problems that span it and turning them into clear technical direction the team can execute against, without day-to-day oversight. Above all, you'll own how the platform is architected to scale: the distributed systems behind it, the reliability and consistency guarantees developers depend on, and the coherence that keeps it easy to build on as usage grows. You'll champion the team's technical execution, raise the engineering bar, mentor the people around you, and partner with tech leads across teams to keep the wider platform aligned. You'll have real influence over how our platform holds up in a world where developers and agents are both first-class consumers. What you'll be doing Own the platform architecture: Set the tech

node.jsawsrest
View job →
G
Godaddy
📍 Canada• Full-time• From C$107K/yr
1mo ago

Location Details: Canada, Remote At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) , and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Contribute to the development of GoDaddy’s eCommerce and SSO infrastructure and Kubernetes systems on AWS. On a day-to-day basis you will be working on the team who designs, writes, tests and deploys the infrastructure and application management software for GoDaddy’s eCommerce applications. Expect to learn every day. What you'll get to do... Work as a polyglot engineer, writing and maintaining Infrastructure as code with frameworks/ ecosystems such as Java, Unix CLI, and NodeJS Build and operate infrastructure workflows and deployment pipelines using Kubernetes, Argo Workflows, Argo CD, and GitOps practices Design, build, and own services and APIs in Java, running on Kubernetes-based platforms across AWS and distributed systems Develop and support application and infrastructure delivery pipelines, enabling reliable releases of eComm, Auth and Infrastructure services Collaborate closely with other GoDaddy departments to help advance security and technical standards, maintain regulatory compliances while operating eComm & Auth platforms Your experience should include... 5+ years of strong backend software engineering experience in Java Hands-on experience with Kubernetes, including Helm, Kustomize, or equivalent tools to deploy and manage backend services Experience building and operating high-volume, mission-critical production systems on AWS with continuous deployment (CD) practices Strong experience with infrastructure as code, supporting backend applications and services Experience with observability a

javanodejssql
View job →
NR
New Relic
📍 Hyderabad, India• Full-time
1mo ago

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity At New Relic, we provide our customers real-time insights, so they can innovate faster. Our software delivers insightful observability tools across different technologies and distributed systems, enabling software engineering teams to quickly identify, understand and tackle issues, analyze performance and get the most of their software and infrastructure. We are seeking a Senior Engineer to help build next-generation Database Observability products. In this high-impact role, you will own the entire lifecycle of critical data flows from developing lightweight database agents and high-throughput ingestion pipelines to building intelligent DB Recommendation engines and autonomous DB AI Agents. You will build cutting-edge capability that moves observability from passive monitoring to automated, AI-driven database diagnostic and remediation workflows. If you are energized by hard distributed systems problems, extreme scalability, and harnessing agentic AI to solve complex developer challenges, this is a career-defining opportunity to shape the future of intelligent database observability at New Relic.We look forward to talking with you! What you'll do . Architect, develop, and scale high-throughput, low-late

typescriptpythonjava
View job →
N
Nuro
📍 Mountain View• Full-time• From $160.4K/yr
1mo ago

Who We Are Nuro is a self-driving technology company on a mission to make autonomy accessible to all. Founded in 2016, Nuro is building the world’s most scalable driver, combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the Nuro Driver™, to support a wide range of applications, from robotaxis and commercial fleets to personally owned vehicles. With technology proven over years of self-driving deployments, Nuro gives the automakers and mobility platforms a clear path to AVs at commercial scale, empowering a safer, richer, and more connected future. About the Team A rider taps "request ride" and within seconds, an autonomous vehicle is matched, dispatched, and on its way. The On-Road Experience team owns the systems that make that moment happen. The On-road Experience team builds the real-time distributed system behind Nuro's ride-hailing product. This is foundational infrastructure where correctness and performance have consequences on a real road. About the Role This is the perfect role if you love architecting performant, reliable distributed systems, processing real-time vehicle data, and building the foundational APIs that enable seamless product experiences. The position demands technical excellence, a deep understanding of system design, and a knack for solving tough problems. Come join us in defining the future! What you'll own Backend services for ride planning, vehicle matching, and real-time trip execution Partner API integrations (Uber and others) routing ride requests to Nuro vehicles Telemetry pipelines surfacing live vehicle state to riders and operations About You 4+ years of backend engineering experience with strong Go proficiency Experience designing and operating distributed systems where reliability isn't optional End-to-end product ownership — from requirements through launch and iteration Solid hands-on experience with GCP (CloudSQL, BigQuery, Redis, or equivalent) A track record of technical lead

sqlredisgcp
View job →
N
Nuro
📍 Mountain View• Full-time• From $193.9K/yr
1mo ago

Who We Are Nuro is a self-driving technology company on a mission to make autonomy accessible to all. Founded in 2016, Nuro is building the world’s most scalable driver, combining cutting-edge AI with automotive-grade hardware. Nuro licenses its core technology, the Nuro Driver™, to support a wide range of applications, from robotaxis and commercial fleets to personally owned vehicles. With technology proven over years of self-driving deployments, Nuro gives the automakers and mobility platforms a clear path to AVs at commercial scale, empowering a safer, richer, and more connected future. About the Role The Mapping team in Nuro takes a machine learning-first path to unblock geographic capability with lower costs. This team plays a crucial role in the advancement of autonomous driving systems by creating and improving different components in the full lifecycle of ML models for multiple teams in this organization. We are searching for an engineer with experience building reliable and scalable machine learning infrastructure and a strong desire to contribute to the future of robot navigation for logistics and transportation. About the Work Mainly focus on building and improving HD map generation and release pipelines. Develop scalable workflows to manage first party and third party map data. Develop and manage APIs for internal users to access data. Improve and refactor existing workflows and toolings to boost efficiency. Engage with other mapping teams to help identify issues and establish long-term relationships that include knowledge sharing. About You BS in Computer Science, Robotics or another quantitative area. You have experience in one or more of the following areas: large-scale distributed systems; data storage and processing systems; advanced algorithms using C++ and Python; multithreading; and software performance tuning and optimization. Ability to efficiently develop, debug, and support new technologies in a changing environment. Str

pythonmachine learningai
View job →
S
Smartsheet
📍 -REMOTE, USA-• Full-time• Remote• From $1.3M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. We are hiring a Software Engineer I to join our Engineering team. You will get your hands on full-stack services, dig into distributed systems problems, and work right at the intersection of traditional software engineering and AI-driven development. We are not just dabbling in AI here. We use it to write better code, catch bugs faster, and troubleshoot smarter, and we want engineers who want to get good at that too and help their teammates do the same. You will ship real features, sit in on architectural conversations that actually shape the product, and grow alongside a team that cares as much about doing good work as they do about doing meaningful work. You Will: Build scalable front-end and back-end services for the next generation of applications at Smartsheet (Kotlin, Java, Typescript, React) Solve challenging distributed systems problems and work with modern cloud infrastructure (AWS, Kubernetes) Take part in code reviews and architectural discussions as you work with other software engineers and product managers Forge a strong partnership with product management and other key areas of the business Enhance existing application code with new features and strike a balance when making technical decisions (build vs refactor vs simplify) Actively use AI tools to improve personal and team efficiency across coding, testing, design, and troubleshooting, exploring AI integration opportunities within team processes and features, and coaching others on effective AI use You Have: 1+ years software development experience build

REMOTEtypescriptpythonjava
View job →
S
Smartsheet
📍 -REMOTE, USA-• Full-time• Remote• From $1.5M/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. As a Software Engineer II at Smartsheet, you will help build the next generation of our platform while tackling meaningful distributed systems challenges alongside a talented engineering team. You will collaborate closely with product managers and fellow engineers through code reviews and architectural discussions, and you will have the opportunity to mentor junior engineers as you grow in your own career. This role is a strong fit for someone with a couple of years of hands-on development experience who is ready to take on more ownership, sharpen their skills in cloud infrastructure, and help shape how the team uses AI tools to work smarter and faster . You Will: Build scalable back-end services for the next generation of applications at Smartsheet (Kotlin, Java) Solve challenging distributed systems problems and work with modern cloud infrastructure (AWS, Kubernetes) Take part in code reviews and architectural discussions as you work with other software engineers and product managers Mentor junior engineers on code quality and other industry best practices Forge a strong partnership with product management and other key areas of the business Actively use AI tools to improve personal and team efficiency across coding, testing, design, and troubleshooting, exploring AI integration opportunities within team processes and features, and coaching others on effective AI use You Have: 2+ years software development experience building highly scalable, highly available applications 2+ years of programming experience with full st

REMOTEtypescriptjavaaws
View job →
T
Twilio
📍 - US• Full-time• Remote• $138.7K – $173.4K/yr
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as our next Software Engineer (L2) Who we are & why we’re hiring Twilio powers real-time business communications and data solutions that help companies and developers worldwide build better applications and customer experiences. Although we're headquartered in San Francisco, we're on a journey to becoming a globally antiracist company that supports diversity, equity & inclusion wherever we do business. We employ thousands of Twilions worldwide, and we're looking for more builders, creators, and visionaries to help fuel our growth momentum. About the job This position is needed to fulfill the critical role of a Software Engineer within Twilio Sendgrid, to be both hands-on in creating, deploying and managing highly available, very large scale distributed systems. Our systems processed 12 billion emails on Black Friday 2026 and we continue to scale! You will be a key contributor in one or more of these areas -&nbs

REMOTEjavasqlmysql
View job →
D
1mo ago

As a Staff Engineer on the Data Platform Experience team, you'll help shape how Datadog engineering teams build, operate, and evolve products on the Observability Data Platform. You'll lead the design and delivery of shared platform capabilities that reduce developer friction, improve operational visibility, and enable engineering teams to move faster with confidence. This role combines deep distributed systems expertise with technical leadership across multiple teams, influencing platform strategy while remaining hands-on in the code. You'll have the opportunity to solve company-wide challenges spanning cost intelligence, operational tooling, platform health, and developer experience. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Lead strategic engineering initiatives that improve how product teams build, operate, and evolve services on the Observability Data Platform. Design and build scalable platform capabilities for cost intelligence, including cloud cost allocation, trend analysis, and optimization recommendations. Develop operational intelligence and self-service tooling that helps engineering teams understand platform health, troubleshoot incidents, and improve operational efficiency. Drive reusable platform services and developer workflows that increase engineering autonomy while reducing operational complexity across multiple products. Provide technical leadership across teams by influencing architecture, mentoring engineers, and raising engineering standards through hands-on technical contributions. Participate in the team's on-call rotation and continuously improve platform reliability, observability, and operational excellence. Who You Are: You have experience designing and building large-scale SaaS or cloud platforms with deep expertise i

javakubernetesai
View job →
D
Datadog
📍 New York• Full-time• From $192K/yr
1mo ago

Senior Software Engineer - Streaming Platform Client Data streams are mission-critical at Datadog, powering near real-time communication across the vast majority of our services. Our Streaming Platform group builds the core infrastructure and abstractions that ensure Datadog remains a trusted partner for engineers worldwide. See our blog post . The Streaming Platform Client team sits at the heart of this ecosystem. We own the Rust client library (producers and consumers) with language bindings for Java, Go, and Python. We focus on building intuitive APIs and abstractions that make a powerful distributed system easy to adopt and operate for the hundreds of internal users of our library. Our library runs on critical data paths that handle hundreds of millions of messages per second making performance, observability, and reliability paramount. We also develop and operate the service that bridges the clients fleet with the platform's control plane, handling complex balancing, scaling, and static stability challenges. We are seeking a Senior Software Engineer to help us evolve these features. You will collaborate directly with our users, tackle performance-critical code, and solve complex distributed systems challenges across the control plane, client libraries, and data plane. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Work within a distributed, high-impact team spanning Europe and the US, building critical technologies that power data pipelines for dozens of internal teams and hundreds of services. Architect and implement resilient interactions between our client libraries and the control plane. Optimize our high-throughput, low-level streaming library to push the boundaries of performance and efficiency. Champion the developer experience by providing

pythonjavagit
View job →
D
Datadog
📍 New York• Full-time• From $192K/yr
1mo ago

Senior Software Engineer - Streaming Platform Client Data streams are mission-critical at Datadog, powering near real-time communication across the vast majority of our services. Our Streaming Platform group builds the core infrastructure and abstractions that ensure Datadog remains a trusted partner for engineers worldwide. See our blog post . The Streaming Platform Client team sits at the heart of this ecosystem. We own the Rust client library (producers and consumers) with language bindings for Java, Go, and Python. We focus on building intuitive APIs and abstractions that make a powerful distributed system easy to adopt and operate for the hundreds of internal users of our library. Our library runs on critical data paths that handle hundreds of millions of messages per second making performance, observability, and reliability paramount. We also develop and operate the service that bridges the clients fleet with the platform's control plane, handling complex balancing, scaling, and static stability challenges. We are seeking a Senior Software Engineer to help us evolve these features. You will collaborate directly with our users, tackle performance-critical code, and solve complex distributed systems challenges across the control plane, client libraries, and data plane. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Work within a distributed, high-impact team spanning Europe and the US, building critical technologies that power data pipelines for dozens of internal teams and hundreds of services. Architect and implement resilient interactions between our client libraries and the control plane. Optimize our high-throughput, low-level streaming library to push the boundaries of performance and efficiency. Champion the developer experience by pro

pythonjavagit
View job →
🔔

Get new distributed systems engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More distributed systems engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.