Jobiba hiring network

Distributed Systems Engineer Jobs

1,306 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current distributed systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

Software Engineer Build technology where every nanosecond matters. At Graviton, software isn't just a tool that supports trading. It is the infrastructure behind every research breakthrough, every trading decision and every competitive advantage. As a Software Engineer , you'll work on systems where performance, reliability and precision matter at an extraordinary scale. You'll partner closely with software engineers and quantitative researchers to build technology that processes enormous volumes of market data, powers quantitative research and supports live trading. You'll take on problems that don't have obvious answers — from designing high-performance systems and distributed infrastructure to eliminating bottlenecks measured in microseconds and building tools that make our researchers and engineers faster. Your work will go into production, be measured against real-world performance and have a direct impact on how our trading systems operate. If you enjoy solving hard engineering problems, understanding systems at a deep level and pushing technology to its limits, you'll feel right at home. What You'll Work On You'll work across the engineering stack that powers our quantitative research and trading platforms. Depending on your team, your work may include: Designing and building high-performance, low-latency systems in modern C++. Building distributed systems that process and analyze massive volumes of market data . Designing systems where latency, throughput and reliability directly influence trading performance . Working on Linux systems, networking, concurrency and multithreaded applications. Profiling systems, identifying bottlenecks and optimizing performance at the hardware and software level. Building robust infrastructure that supports quantitative research and live trading. Debugging complex production systems and solving problems where correctness and reliability are critical. Designing internal platforms and developer tools that accelerate research an

linuxrestai
View job →
C-
CLEAR - Corporate
📍 New York• Full-time• $225K – $300K/yr
16 days ago

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. Today, CLEAR is well-known as a leader in digital and biometric identification, reducing friction for our members wherever an ID check is needed. We’re looking for a Senior Software Engineer to establish our Observability framework and foundations. You will join us to accelerate building and scaling our innovative systems that support our growing identity platform. You will drive on Observability best practices to find and fix gaps in our observability and our overall systems. You will also lead practices such as load testing, capacity planning, game days, chaos testing, and incident post-mortems. What You Will Do: Embed within the Engineering pillar to deeply understand the product and implement observability across all key flows Facilitate and build load testing cases, ensuring we understand the limits and scaling factors of our services and systems Contribute to observability and support the design of new services and systems, ensuring highly reliable and scalable concepts are implemented Build and lead practices such as game days, chaos engineering, and failure analysis Build long-term capacity plans, with an eye toward reliability and cost-efficiency Who You Are: 6+ experience writing production-grade software in a modern language, such as Java and Python. Strong knowledge of distributed systems concepts (think CAP theorem), microservices architecture, and distributed tracing . Experience with modern observability systems such as Datadog. Experience with performance debugging tools and patterns. You should be able to read a f

pythonjavagit
View job →
G
Godaddy
📍 India• Full-time
18 days ago

Location Details: Remote, India At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team... This position is for a Staff Software Engineer within GoDaddy’s engineering organization, concentrating on developing scalable, fault-tolerant systems and promoting technical excellence across various teams and projects. As a Staff Software Engineer, you will serve as a technical leader at the division level. You will impact system design and architecture, as well as how engineering is carried out across teams. This position suits engineers who excel in uncertain environments, like tackling difficult technical problems, and can provide clarity, structure, and delivery for large projects. Our teams focus on greenfield and open-ended projects that require firm technical ownership, architectural vision, and collaboration across functions. You will assist in establishing engineering standards, support teams through mentoring, and create systems that boost reliability, scalability, and developer efficiency organization-wide. This role is important for encouraging innovation throughout the engineering organization. It includes promoting AI-assisted development workflows, modern cloud-native methods, and reusable platform features that speed up development across teams. What you'll get to do... Technical Execution & Delivery Lead delivery of complex engineering initiatives while driving predictable execution and long-term technical quality Architect and implement highly available, fault-tolerant distributed systems using AWS, PostgreSQL/Aurora, and MongoDB Architect and lead greenfield projects from concept to production, bringin

pythonreactnode.js
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
23 days ago

About the Team The ChatGPT organization at OpenAI supports our mission by bringing advanced AI capabilities to hundreds of millions of users worldwide. The Image Generation team is responsible for one of the fastest-growing experiences in ChatGPT, enabling users to create, edit, and transform images through natural language. Recent advances in our multimodal image models have dramatically improved image quality, instruction following, editing precision, consistency, and text rendering, unlocking entirely new creative and professional workflows. We work at the intersection of research, infrastructure, and product to build the systems that power image generation at global scale. Our team partners closely with researchers, product engineers, designers, and platform teams to bring state-of-the-art image capabilities to millions of users while continuously pushing the boundaries of what AI-powered creation can do. About the Role We are looking for an experienced Backend Engineer to join the Image Generation team and help build the systems that power image creation and editing across ChatGPT. You'll work on the core backend infrastructure that enables users to generate, edit, and iterate on visual content using cutting-edge multimodal AI models. This includes building highly scalable services, orchestration systems, APIs, storage platforms, and distributed infrastructure that support billions of image generations and editing workflows. You'll partner closely with product, research, and mobile teams to transform breakthrough AI capabilities into reliable, performant experiences used by millions around the world. In this role, you will: Design, build, and operate backend systems that power image generation and image editing experiences in ChatGPT. Develop scalable APIs, services, and infrastructure that support multimodal AI workflows. Optimize reliability, latency, throughput, and cost across large-scale distributed systems. Partner with researchers to productionize new im

REMOTEawsrestai
View job →
G
Godaddy
📍 India• Full-time
24 days ago

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Contribute to the development of GoDaddy’s eCommerce and SSO infrastructure and Kubernetes systems on AWS. On a day-to-day basis you will be working on the team who designs, writes, tests and deploys the infrastructure and application management software for GoDaddy’s eCommerce applications. Expect to learn every day. What you'll get to do... Work as a polyglot engineer, writing and maintaining Infrastructure as code with frameworks/ ecosystems such as Java, Unix CLI, and NodeJS Build and operate infrastructure workflows and deployment pipelines using Kubernetes, Argo Workflows, Argo CD, and GitOps practices Design, build, and own services and APIs in Java, running on Kubernetes-based platforms across AWS and distributed systems Develop and support application and infrastructure delivery pipelines, enabling reliable releases of eComm, Auth and Infrastructure services Collaborate closely with other GoDaddy departments to help advance security and technical standards, maintain regulatory compliances while operating eComm & Auth platforms Your experience should include... 5+ years of strong backend software engineering experience in Java Hands-on experience with Kubernetes, including Helm, Kustomize, or equivalent tools to deploy and manage backend services Experience building and operating high-volume, mission-critical production systems on AWS with continuous deployment (CD) practices Strong experience with infrastructure as code, supporting backend applications and services Experience with observability and l

javanodejssql
View job →
R
Replit
📍 Foster City• Full-time• Remote
25 days ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: As a New Grad Software Engineer, you'll join a team of exceptional builders working on products that are reshaping how the world creates software. You'll have the opportunity to work on everything from our AI-powered development platform to the distributed systems that enable real-time collaboration for millions of developers. This is a chance to define your career while defining the future of software development. You'll work on problems that matter, with the autonomy to drive solutions and the support to grow into a technical leader. What you will build: Product features that delight users and make it possible for anybody to create software AI coding agent that understands intent and generates production-ready applications Cloud infrastructure that provides instant, powerful development environments at global scale Platform features that enable one click deployments and scale to millions of users Required skills and experience: Recent graduate (2027) with a degree in Computer Science, Computer Engineering, or related field Strong programming skills in a modern language (JavaScript/TypeScript, Python, Go, Rust) Full-stack capabilities with experience in React, Node.js, and database technologies Growth orientation - eager to learn new technologies and take on increasing responsibility Collaborative spirit - you work well in cross-functional teams and value diverse perspectives What we value : Problem-solving mindset: Ability to approach complex operational challenges systematically and devise effective solutions Self-directed and autonomous: Capable of working independently while collaborating effectively with cross-functional teams Strong communication skills: Ability to explain complex technical conce

REMOTEjavascripttypescriptpython
View job →
O
Okta
📍 Bengaluru• Full-time
25 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. We are looking for an experienced Principal Software Engineer to work on our next-generation Imports Platform team. Imports Platform team is leading a strategic initiative to modernize Okta's identity lifecycle management capabilities by architecting and migrating from a legacy monolithic system to a highly scalable, distributed microservices platform. This critical service orchestrates the importing, syncing, and provisioning of identities and access policies—users, groups, roles, entitlements—from external directory services including Active Directory, Office 365, and LDAP-based systems. As a Principal Software Engineer on the Imports Platform team, you will be a cross-team technical leader who takes difficult, ambiguously defined problems and drives them from ideation through production impact without oversight. You will own projects from zero to landing—defining scope, planning execution, making architectural decisions, and articulating measurable impact across the group. You will generate novel solutions to complex distributed systems challenges, guide the team's technical direction, and get stakeholder buy-in on architectural strategy spanning multiple teams. Your sphere of influence extends beyond the Imports Platform team to adjacent teams within the group and cross-functional partners in Product, Design, and SRE. You will participate in group-level strategy, break down strategic initiatives into actionable technical milestones, and drive cross-team

javasqlmysql
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

The Application Modernization Platform (AMP) team is dedicated to solving one of the industry's biggest challenges: transforming rigid, legacy applications that suffer from poor scalability and high operating costs into modern, microservices-based architectures on MongoDB. To accelerate this transition, MongoDB is building a dedicated Platform and Infrastructure team to develop the Application Modernisation Platform and Infrastructure . As we help customers modernize their application and data ecosystems, we face the challenge of deploying complex tooling into highly restrictive client environments and architecting automated verification frameworks to ensure data equivalence. This new team will architect the platform foundation and infrastructure that enables a "build once, run anywhere" model, ensuring our AI-powered modernisation suite operates seamlessly regardless of a client's security or network constraints. We are looking for engineers to join this high-visibility initiative, where you will solve unique distributed systems puzzles and help shape the future of how global enterprises leverage data and AI. We are looking for an experienced Software Engineer who thrives on solving infrastructure constraints and building developer-centric modernisation platforms with a strong background in building software testkits/frameworks. The ideal candidate will be designing and building automated frameworks that validate functional equivalence, performance benchmarks, and data integrity. From leveraging LLMs for unit test generation to building contract testing frameworks, your work will be the safety net for the world’s largest enterprise migrations. This role will be based in our India office in Gurgaon and offers a hybrid working model. Position Expectations Contribute high-quality, well-tested code to the modernization and framework team and its surrounding services Collaborate effectively with Product Management, other engineers, and designers to build and deliver on

javasqlmysql
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

We are seeking a Software Engineer to join our growing Gurugram Products & Technology team to develop and expand core parts of a new platform we are building to make it easier for customers to build AI applications using MongoDB. As a Software Engineer on this new team, you will be responsible for developing cutting edge technologies related to enabling deployment at scale of AI applications. You will take on challenging, high-visibility projects that improve and enhance the performance, scalability, and reliability of the distributed systems infrastructure for this new product. MongoDB engineering teams pride themselves on building high-quality software and living MongoDB cultural values every day – we value intellectual curiosity and honesty, and building together in an environment that prioritizes collaboration over competition. We are looking to speak to candidates who are based in Gurugram for our hybrid working model. Position Expectations Work closely with research, product management, product engineering, product design, peers as well as other teams within the company to implement the first version and future evolution of the service Design, build and deliver well-tested core pieces of the platform in collaboration with other vested parties Contribute to shaping architecture, code reviews and development practices, developer experience as the teams and product grow Mentor fellow engineers and assume ownership and accountability of projects Qualifications Strong background in building core components for high scale compute and data distributed systems 3+ years experience of building distributed systems, and/or foundational cloud services at scale and an interest in working with Python, Go and Java Proven success in designing, writing, testing, debugging, performance tuning, possessing a strong grip on the foundational materials of computer science and maintaining distributed and/or highly concurrent software systems in large, long-

pythonjavamongodb
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
1mo ago

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will build the model runtime within the inference engine that executes complex, frontier models at scale on OpenAI’s custom silicon. The runtime will sit between models running on the hardware and the upper layers of the cluster serving software stack, translating demanding inference workloads into efficient execution while optimizing for throughput, latency, utilization, and reliability. You will work across model architecture, distributed systems, compilers, kernels, and silicon to design a production-grade runtime comparable in ambition to systems such as vLLM and SGLang, but customized and optimized for OpenAI’s AI accelerator. Your work will shape how new model capabilities map onto the platform and how quickly custom silicon can deliver meaningful performance in production. In this role, you will: Design and implement the LLM inference runtime for frontier models running on custom silicon. Build scheduling, continuous batching, memory management, KV-cache management, and execution orchestration for high-performance inference. Develop distributed execution strategies across chips, hosts, and racks, including model partitioning, communication, and synchronization. Optimize end-to-end latency, throughput, memory efficiency, and hardware utilization across diverse model architectures and serving workloads. Partner with kernel, compiler, architecture, and silicon teams to co-design interfaces and remove performance bottlenecks across the stack. Enable new

REMOTEpythonawsrest
View job →
C
Coinbase
📍 United States• Full-time• Remote• From $218K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Staff Software Engineer on the Core Automation team within the Platform group, you'll architect and build the Agentic AI systems that are transforming how Coinbase operates. This team is reimagining customer support and compliance processes for a fully AI-driven world, designing intelligent agents, orchestration frameworks, and measurement systems that deliver delightful customer experiences at scale. You'll own the technical direction for production AI systems, working across cross-functional teams to bring this vision to reality while building primitives that scale automation across the company. What you'll do: Architect and build Agentic AI systems that power Coinbase's compliance automation and other Operations, from intelligent agents through orchestration and guardrails Design foundational APIs and measurement frameworks that ensure AI agents are grounded, relevant, and reliably deliver customer delight with minimal hallucination Lead technical direction for distributed systems underpinning AI automation, defining architecture patterns and strategic roadmaps in partnership with engineering leadership Build reusable primitives and orchestration solutions that enable AI-powered automation to scale across multiple domains beyond the initial customer support and compliance focus Mentor engineers on AI system design techniques, coding standards, and production-

REMOTEawsaigo
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
1mo ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Marketplace teams are at the heart of our products and decision-making, owning everything from rider pricing to driver earnings, incentives, and efficient matching. We’re looking for passionate, driven engineers to build systems that empower our riders and drivers to have the best transportation experience possible through prediction, adaptivity, and personalization. We’re looking for someone who is excited about working in a fast-paced, innovative, and impactful environment to create reliable solutions to distributed computing, ML, and data problems. The Pricing team is a centerpiece of Lyft’s Marketplace org, determining prices for all rideshare products and supporting new initiatives. Rider Engagement develops rider-facing engagement levers and optimizes user pricing experience to drive both short term and long term business outcomes. We work with Product & Science to solve and implement complex pricing requirements, balancing the needs of riders, drivers, and the business goals. As an owner of one of the most critical flows in the company, you will work on a wide array of challenges such as latency-sensitive concurrency problems, large scale distributed systems, and experimentation. If you’re interested in playing a large part in demand / supply management and improving the Lyft customer experience, this could be a great fit for you. Responsibilities: Drive high-impact projects and innovate new solutions to provide the best user experience. Work closely with cross-functional teams and partner teams to develop solutions based on technology and business needs, and advance team’s goals and priorities Independently lead features from idea to positive execution and launch Unblock, support and communicate with internal partners to achieve results Write well-crafted, well-tested, readable, maintaina

pythonawsrest
View job →
A
Amplitude
📍 Remote• Full-time• From $24K/yr
1mo ago

About the Role & Team We’re looking for an Engineering Manager to lead the Delivery and SDK team within Statsig at Amplitude. This team builds the systems that connect developers to Statsig and safely deliver product experiences to end users around the world. The team owns two foundational areas: Delivery: The globally distributed systems that evaluate Experiments and Feature Gates and deliver the resulting experiences to end users. These systems sit directly in the critical path of our customers’ applications, making availability, scale, correctness, and latency essential. SDKs: More than 30 SDKs spanning web, mobile, server, edge, TV, gaming, and emerging platforms. Our goal is to meet developers wherever they build and give them an ergonomic, reliable way to integrate Statsig into any application or technology stack. This is a uniquely broad technical leadership role. You’ll work across distributed systems, mobile platforms, server runtimes, edge environments, and developer tooling—sometimes all in the same week. We’re looking for a hands-on, technically versatile leader who enjoys moving between technologies, learning unfamiliar systems, and solving problems across abstraction boundaries. You don’t need to be the world’s foremost expert in a single language or platform. You do need the curiosity and technical judgment to work effectively across many of them—and the ability to build a team that can make a complex, polyglot ecosystem feel simple and dependable to developers. What You’ll Do Lead and grow the team responsible for Statsig’s global delivery infrastructure and ecosystem of 30+ SDKs. Experience leading a team of 8-10 team members Define the technical strategy for highly available, low-latency evaluation and delivery systems operating at global scale. Make Statsig exceptionally easy to adopt by continuously improving SDK ergonomics, performance, reliability, consistency, and documentation. Expand our SDK coverage as new languages, frameworks, runtime

restaigo
View job →
A
Amplitude
📍 Remote• Full-time• From $24K/yr
1mo ago

About the Role & Team We’re looking for an Engineering Manager to lead the Delivery and SDK team within Statsig at Amplitude. This team builds the systems that connect developers to Statsig and safely deliver product experiences to end users around the world. The team owns two foundational areas: Delivery: The globally distributed systems that evaluate Experiments and Feature Gates and deliver the resulting experiences to end users. These systems sit directly in the critical path of our customers’ applications, making availability, scale, correctness, and latency essential. SDKs: More than 30 SDKs spanning web, mobile, server, edge, TV, gaming, and emerging platforms. Our goal is to meet developers wherever they build and give them an ergonomic, reliable way to integrate Statsig into any application or technology stack. This is a uniquely broad technical leadership role. You’ll work across distributed systems, mobile platforms, server runtimes, edge environments, and developer tooling—sometimes all in the same week. We’re looking for a hands-on, technically versatile leader who enjoys moving between technologies, learning unfamiliar systems, and solving problems across abstraction boundaries. You don’t need to be the world’s foremost expert in a single language or platform. You do need the curiosity and technical judgment to work effectively across many of them—and the ability to build a team that can make a complex, polyglot ecosystem feel simple and dependable to developers. What You’ll Do Lead and grow the team responsible for Statsig’s global delivery infrastructure and ecosystem of 30+ SDKs. Leading a team of 8-10 team members Define the technical strategy for highly available, low-latency evaluation and delivery systems operating at global scale. Make Statsig exceptionally easy to adopt by continuously improving SDK ergonomics, performance, reliability, consistency, and documentation. Expand our SDK coverage as new languages, frameworks, runtimes, and comp

restaigo
View job →
S
1mo ago

Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. We're looking for an engineer to own the deployment and operational infrastructure of Multigres, our distributed Postgres platform. You'll be responsible for building and maintaining the Multigres Operator, ensuring reliable cloud deployments, and creating the tooling that powers our Kubernetes-based infrastructure. What You’ll Be Responsible for: Build and maintain the Multigres Operator - Maintain our Go-based Kubernetes operator that orchestrates distributed Postgres deployments Architect cloud deployment infrastructure - Design and implement robust deployment patterns for EKS and other Kubernetes platforms Manage storage and networking layers - Work with CSI drivers, persistent volumes, and cross-cloud networking to ensure data reliability and connectivity Develop deployment tooling - Create internal tools and automation for provisioning, scaling, and managing Multigres clusters Ensure operational excellence - Build monitoring, alerting, and diagnostic capabilities into the deployment layer Collaborate across teams - Work with database engineers, SRE, and product teams to deliver seamless deployment experiences You Might Be a Good Fit If You have: Strong systems programming skills - Proficiency in Go and experience building production-grade operators or controllers Deep Kubernetes expertise - Hands-on experience with Kubernetes internals, custom resources, and cloud-managed Kubernetes services (EKS, GKE, AKS) Database operations knowledge - Understanding of database deployment patterns, backup/restore, replication, and high availability Distributed systems experience - Familiarity with consensus protocols, failure scenarios, and designing for resilience Cloud infrastructure background - Experience with cl

kubernetesrestai
View job →
🔔

Get new distributed systems engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More distributed systems engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.