We’re looking for an Intermediate Software Developer, Backend who can help us support the development organization to deliver value to customers in a reliable, efficient, and safe manner. You’ll be working in a focused team that owns one piece of the production application environment and the developer experience, you will execute on defined projects to achieve team-level goals. In line with Hootsuite's distributed workforce strategy, our flexible work arrangement allows for a hybrid model. This role is open to applicants located in Bucharest, Romania. WHAT YOU’LL DO: Write software - tools, libraries, automation, services Design and build our infrastructure platform Identify and implement new platform features Research and evaluate new technologies Refactor, rewrite or retire existing platform features Operate our developer experience and production application environments Diagnose and repair our distributed systems Perform maintenance, upgrades and migrations Control or eliminate repetitive tasks, alert noise, and business-as-usual work Enable development teams Provide executable interfaces to our infrastructure platform Provide tools and best practices to support the entire software development lifecycle Participate in a flexible on-call rotation Communicate by writing documentation, participating in meetings, and showing off your work at demos WHAT YOU’LL NEED: A degree in Computer Science or Engineering or equivalent experience working in a software engineering role An ability to write software and working knowledge of software engineering practice (Java programming language and strong working knowledge of object-oriented programming concepts) Proven experience creating stable, reliable, performing and maintainable code Familiarity with data modeling and schema design Knowledge of data structures and algorithms Open Communication: clearly conveys thoughts, both written and verbally, listening attentively and asking questions for clarification
Jobiba hiring network
Distributed Systems Engineer Jobs
1,301 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current distributed systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
We’re looking for an Senior Software Developer, Backend who can help us support the development organization to deliver value to customers in a reliable, efficient, and safe manner. You’ll be working in a focused team that owns one piece of the production application environment and the developer experience, you will execute on defined projects to achieve team-level goals. In line with Hootsuite's distributed workforce strategy, our flexible work arrangement allows for a hybrid model. This role is open to applicants within commutable distance to Luxembourg. WHAT YOU’LL DO: Write software - tools, libraries, automation, services Design and build our infrastructure platform Identify and implement new platform features Research and evaluate new technologies Refactor, rewrite or retire existing platform features Operate our developer experience and production application environments Diagnose and repair our distributed systems Perform maintenance, upgrades and migrations Control or eliminate repetitive tasks, alert noise, and business-as-usual work Enable development teams Provide executable interfaces to our infrastructure platform Provide tools and best practices to support the entire software development lifecycle Participate in a flexible on-call rotation Communicate by writing documentation, participating in meetings, and showing off your work at demos WHAT YOU’LL NEED: A degree in Computer Science or Engineering or equivalent experience working in a software engineering role An ability to write software and working knowledge of software engineering practice (Java programming language and strong working knowledge of object-oriented programming concepts) Proven experience creating stable, reliable, performing and maintainable code Familiarity with data modeling and schema design Knowledge of data structures and algorithms Open Communication: clearly conveys thoughts, both written and verbally, listening attentively and asking questions for clarific
As a Senior Product Manager for Serverless at Datadog, you will define and deliver products that help developers monitor and operate serverless applications at scale. You’ll own the strategy and execution for Datadog’s AWS Serverless observability offering, building experiences that provide visibility into distributed systems and simplify debugging and operations. This role sits at the intersection of cloud infrastructure, developer experience, and AI-powered workflows, and is ideal for a PM who thrives in highly technical product areas. You will work cross-functionally and with external partners to shape how customers build and run modern serverless applications. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own and drive the roadmap for Datadog’s AWS Serverless observability products, including Lambda, Fargate, and Step Functions Define how developers monitor, debug, and operate serverless systems across distributed environments Partner with engineering and design to deliver end-to-end product capabilities from concept through launch and iteration Collaborate with AWS product teams to align roadmaps and deliver joint solutions for shared customers Engage with customers to understand serverless adoption patterns and validate product direction Define and track success metrics such as adoption, usage, and impact on developer workflows Who You Are: 5+ years of product management experience building technical products in areas such as cloud infrastructure, developer platforms, or observability Strong understanding of distributed systems, cloud-native architectures, and modern application development practices Familiarity with serverless technologies, containers, Kubernetes, or microservices environments Comfortable working closely with engineers and discu
About the Role Anyscale is seeking a Senior / Staff Product Manager to lead Ray Data, our scalable data processing library for ML and AI workloads. This is a uniquely challenging role that requires balancing open source growth with commercial differentiation - driving rapid adoption in the open source Ray Data ecosystem while building compelling proprietary features for Anyscale RunTime, our high-performance commercial engine. You'll own the entire Ray Data product roadmap in a competitive landscape, working closely with the engineering team, the field team, enterprise customers, and the open source community. Success requires: Deeply ingraining yourself into the end-user experience to understand the nature of the product and its gaps and tradeoffs Working closely with customers and open source users to draw the subtle line between growth and commercialization Strategic thinking about which parts of the ML/Data lifecycle to focus on, identifying opportunities where our architectural strengths create the most value. Thinking deeply about and clearly articulating the product strategy to stakeholders Key Responsibilities Drive the Ray Data product roadmap - Balance open source Ray Data feature development with Anyscale Runtime commercial differentiation to ensure that both Ray Data becomes the open source standard for AI data processing and Anyscale Runtime remains sufficiently compelling. Drive open source Ray Data adoption - Focus on community growth, developer experience, and ecosystem integrations Market Positioning & Enablement - Work closely with Product Marketing on strategic market positioning, field enablement, and competitive analysis to maintain differentiation. Customer engagement - Drive key customer engagements assisting sales and field engineering teams. Required Qualifications 4+ years of product management experience with technical products Strong technical background in distributed systems, ML infrastructure, or data processing Experience working
MongoDB’s mission is to empower innovators to create, transform, and disrupt industries by unleashing the power of software and data. We enable organizations of all sizes to easily build, scale, and run modern applications by helping them modernize legacy workloads, embrace innovation, and unleash AI. Our industry-leading developer data platform, MongoDB Atlas, is the only globally distributed, multi-cloud database and is available in more than 115 regions across AWS, Google Cloud, and Microsoft Azure. Atlas allows customers to build and run applications anywhere—on premises, or across cloud providers. With offices worldwide and over 175,000 new developers signing up to use MongoDB every month, it’s no wonder that leading organizations, like Samsung and Toyota, trust MongoDB to build next-generation, AI-powered applications. The Escalation Manager is a critical role within Technical Services. As a member of our global Incident and Escalation Management team, they work internally with our Engineering, Services, Sales and Product Management teams, as well as externally with customers and partners, to coordinate and drive the resolution of critical technical issues and incidents. Transparency is key and is achieved by providing timely and accurate updates to senior management regarding active escalations, as well as important detail on the status of the customer account. Individuals in this role are highly organized, proactive and professional. You are one who excels in fast-paced environments and can assess business impact, mobilize cross-functional teams, and drive technical escalations with urgency and ownership. We are looking for someone who has a customer-focused mindset with excellent communication and expectation-setting abilities. You have a technical background in Support, Services, DevOps, Systems Engineering, or Database environments, and are experienced in incident response or crisis management. You will have strong negotiation and objection-handling skill
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! The mission of the Engineering TPM team is to drive Figma's most important cross-company engineering efforts, and we are looking for a Technical Program Manager (TPM) to partner with our Infrastructure team. The TPM provides oversight of the most important efforts that require coordinated technical execution across the Org to succeed. This is a role focused on enabling Figma's infrastructure teams to scale, improve performance, and deliver on critical projects. These large-scale efforts will involve collaboration across numerous backend, infrastructure, and security teams and cross-functional stakeholders, prioritization, decision-making, tracking execution, and driving operational excellence. We're looking for someone that can work in a TPM greenspace environment and is passionate about people, technology, and program management. Progress over process is our mantra. This is a full-time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Lead the execution, coordination, and risk management of Figma's infrastructure projects, ensuring seamless integration with minimal performance impact Drive key infrastructure initiatives, including reliability, storage, distributed systems, cloud-native perfo
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As a Member of Technical Staff on AI Infrastructure, you will build and maintain the foundational systems and distributed infrastructure that power AI model post training, inference, and data pipelines. You will collaborate with engineering and research teams to ensure performance, scalability, and reliability of critical AI systems. What You’ll Do Design and implement large-scale, distributed AI infrastructure and services Optimize performance for GPU/xPU accelerators and cloud environments Build tools for observability, reliability, and scaling of AI workloads Partner with cross-functional teams to define AI infrastructure requirements and roadmap Contribute to architectural design and system longevity About You Have experience with GenAI infrastructure systems, distributed systems, cloud computing, and high-performance infrastructure Are proficient in programming languages like Python, Go, or similar Understand scaling challenges specific to AI workloads and accelerators Thrive in fast-paced, collaborative engineering environments The reasonably estimated base salary for this role ranges from $256,000.00 to $276,
About the Team The Systems Integration team is responsible for building the infrastructure, tooling, and validation systems that ensure our device software our device software is reliable, testable, and ready to ship. We design and maintain build systems, CI pipelines, automated test frameworks, and hardware-in-the-loop labs to enable rapid, safe product launches. Our work spans build systems, developer tools, systems integration, and cross-team collaboration to ensure developers can build reliably and ship with confidence. About the Role We are looking for an engineer to help evolve OpenAI’s Consumer Products build and continuous integration systems for a fast-growing engineering organization. This role sits at the intersection of developer productivity, build systems, distributed infrastructure, software quality, and on-device software. You will work on the systems that determine how quickly and confident engineers can move: Bazel-bazed builds, Buildkite pipelines, test coverage, remote caching and execution, CI observability, and tooling that helps engineers understand and fix failures quickly. Our mission is to enable OpenAI to ship software running on consumer devices rapidly with a high bar for correctness, reliability, and safety. The best version of this work is invisible when it succeeds: builds are fast, tests are trusted, CI failures are understandable, and engineers can focus on shipping products instead of fighting infrastructure. This role is based in San Francisco, CA. We use a hybrid work model of four days in the office per week and offer relocation assistance to new employees. In This Role, You Will Own and evolve Bazel and yocto-based build and test workflows in a polyrepo environment Design and maintain Starlark rules, macros, toolchains, and integrations that make builds hermetic, reproducible, and easy for teams to adopt Improve CI performance and reliability across Buildkite pipelines, including queue time, build time, cache hit rates, retry b
About the Team The Systems Integration team is responsible for building the infrastructure, tooling, and validation systems that ensure our device software our device software is reliable, testable, and ready to ship. We design and maintain build systems, CI pipelines, automated test frameworks, and hardware-in-the-loop labs to enable rapid, safe product launches. Our work spans build systems, developer tools, systems integration, and cross-team collaboration to ensure developers can build reliably and ship with confidence. About the Role We are looking for an engineer to help evolve OpenAI’s Consumer Products build and continuous integration systems for a fast-growing engineering organization. This role sits at the intersection of developer productivity, build systems, distributed infrastructure, software quality, and on-device software. You will work on the systems that determine how quickly and confident engineers can move: Bazel-bazed builds, Buildkite pipelines, test coverage, remote caching and execution, CI observability, and tooling that helps engineers understand and fix failures quickly. Our mission is to enable OpenAI to ship software running on consumer devices rapidly with a high bar for correctness, reliability, and safety. The best version of this work is invisible when it succeeds: builds are fast, tests are trusted, CI failures are understandable, and engineers can focus on shipping products instead of fighting infrastructure. This role is based in San Francisco, CA. We use a hybrid work model of four days in the office per week and offer relocation assistance to new employees. In This Role, You Will Own and evolve Bazel and yocto-based build and test workflows in a polyrepo environment Design and maintain Starlark rules, macros, toolchains, and integrations that make builds hermetic, reproducible, and easy for teams to adopt Improve CI performance and reliability across Buildkite pipelines, including queue time, build time, cache hit rates, retry b
About the Role The Engineering Acceleration team builds and operates the foundational systems that engineers use to build, test, and ship ChatGPT, the API, and OpenAI's infrastructure. We are looking for an engineer to help evolve OpenAI's build and continuous integration systems for a fast-growing engineering organization. This role sits at the intersection of developer productivity, build systems, distributed infrastructure, and software quality. You will work on the systems that determine how quickly and confidently engineers can move: Bazel-based builds, Buildkite pipelines, test selection, remote caching and execution, CI observability, and tooling that helps engineers understand and fix failures quickly. Our mission is to make OpenAI one of the most productive engineering organizations in the world while preserving a high bar for correctness, reliability, and safety. The best version of this work is invisible when it succeeds: builds are fast, tests are trusted, CI failures are understandable, and engineers can focus on shipping useful systems instead of fighting infrastructure. In This Role, You Will Own and evolve Bazel-based build and test workflows across a large, polyglot monorepo. Design and maintain Starlark rules, macros, toolchains, and integrations that make builds reproducible, hermetic, and easy for product teams to adopt. Improve CI performance and reliability across Buildkite pipelines, including queue time, build time, cache hit rates, test sharding, retry behavior, and flake isolation. Build systems that reduce unnecessary CI work through affected-target detection, dependency graph analysis, test selection, caching, batching, and smarter scheduling. Improve local development workflows so engineers can reproduce CI behavior, debug build failures, and iterate quickly without learning every detail of the build stack. Operate and optimize build infrastructure across Docker/OCI images, Kubernetes-based runners, cloud resources, and remote cache/exec
Planning, managing, and executing the commissioning of various control systems, including Distributed Control Systems (DCS), Programmable Logic Controllers (PLC), Digital Electro-Hydraulic (DEH) systems and Vibration Monitoring Systems (VMS). Source: Adani Group | Job ID: 49442
About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Opportunity: Datadog’s Senior Staff Engineers are technical leaders operating at the forefront of large-scale systems design, building the infrastructure that will support our next five years of growth and beyond. They do this in three major ways: As individual contributors, they bring world-class technical depth to build industry-leading systems in areas such as observability data platforms, distributed query engines, and real-time event streaming at global scale. As technical leaders, they apply broad architectural perspective and deep systems thinking to align design decisions across teams and domains. They work across complex, multi-team problem spaces to define long-term technical direction, drive large-scale initiatives forward, and ensure consistent execution. As engineering stewards, they play a key role in evolving our systems and engineering culture. They actively participate in Datadog’s senior technical community, bringing external insights and internal experience to elevate engineering standards and mentor the next generation of technical leaders. Examples of projects a Senior Staff Engineer may lead include designing and launching a new distributed data storage engine capable of handling hundreds of millions of records per second, building the real-time infrastructure behind a new observability product, or re-architecting a core service to support exponential growth in throughput and complexity. What You’ll Do: Be the technical owner of multiple critical systems or architecture areas, often spanning several t
Principal Engineer - Backend About Us: Paytm is India’s leading digital payments and financial services company, which is focused on driving consumers and merchants to its platform by offering them a variety of payment use cases. To merchants, Paytm offers acquiring devices like Soundbox, EDC, QR and Payment Gateway where payment aggregation is done through PPI and also other banks’ financial instruments. To further enhance merchants’ business, Paytm offers merchants commerce services through advertising and Paytm Mini app store. Operating on this platform leverage, the company then offers credit services such as merchant loans, personal loans and BNPL, sourced by its financial partners. About the role: As a Principal Engineer, you will help define the technical design and implementation roadmap across multiple solutions and will work with engineering leadership to ensure we resource and equip our squads with the right expertise to deliver those solutions. Requirements: 8 to 12 years of strong software design/development experience in building massively large scale distributed internet systems and products Hands on experience in Advance Java, Spring boot, AWS, Node, Agentic AI, LLM, RAG, Cursor, Copilot Experience and knowledge of open source tools & frameworks, broader cutting edge technologies around server side development Should be an active contributor to developer communities like Stack overflow, Top coder, Git hub, Google Developer Groups (GDGs). Superior organization, communication, interpersonal and leadership skills. Must be a self-starter who can work well with minimal guidance and in fluid environment. Preferred Qualifications : Bachelor's/Master's Degree in Computer Science or equivalent Skills that will help you succeed in this role: Expertise in Java, DB: RDBMS, Messaging: Kafka/RabbitMQ, Caching: Redis/Aerospike, Micro services, AWS Strong experience in scaling, performance tuning & optimization at both API and storage layers Problem
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The AI team is building cutting-edge solutions that bring the power of AI directly to edge devices while seamlessly integrating with cloud infrastructure. We are looking for a Lead Software Engineer to design and develop high-performance, scalable services to support AI workloads across edge and cloud environments. What You Might Do Design, build, and maintain services that power AI-driven applications, ensuring scalability and performance. Develop APIs and microservices that facilitate seamless integration between cloud-based AI models and edge devices. Optimize data pipelines and storage solutions for real-time AI inference and processing. Implement security and privacy best practices for distributed AI systems. Work closely with AI researchers, infrastructure engineers, and frontend developers to deliver end-to-end AI-driven solutions. Build and optimize an agent orchestration runtime that enables tool use, memory management, and multi-step reasoning across LLMs, APIs, and edge-connected systems. Develop robust logging, monitoring, and alerting systems to ensure system reliability.
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The AI team is building cutting-edge solutions that bring the power of AI directly to edge devices while seamlessly integrating with cloud infrastructure. We are looking for a Senior Software Engineer to design and develop high-performance, scalable services to support AI workloads across edge and cloud environments. What You Might Do Design, build, and maintain services that power AI-driven applications, ensuring scalability and performance. Develop APIs and microservices that facilitate seamless integration between cloud-based AI models and edge devices. Optimize data pipelines and storage solutions for real-time AI inference and processing. Implement security and privacy best practices for distributed AI systems. Work closely with AI researchers, infrastructure engineers, and frontend developers to deliver end-to-end AI-driven solutions. Build and optimize an agent orchestration runtime that enables tool use, memory management, and multi-step reasoning across LLMs, APIs, and edge-connected systems. Develop robust logging, monitoring, and alerting systems to ensure system reliabilit
Get new distributed systems engineer jobs by email
Daily job updates · Unsubscribe anytime