Planning, managing, and executing the commissioning of various control systems, including Distributed Control Systems (DCS), Programmable Logic Controllers (PLC), Digital Electro-Hydraulic (DEH) systems and Vibration Monitoring Systems (VMS). Source: Adani Group | Job ID: 49442
Jobiba hiring network
Distributed Systems Engineer Jobs
1,306 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current distributed systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
About Datadog: We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—providing always-on alerting, metrics visualization, logs, and application tracing for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Opportunity: Datadog’s Senior Staff Engineers are technical leaders operating at the forefront of large-scale systems design, building the infrastructure that will support our next five years of growth and beyond. They do this in three major ways: As individual contributors, they bring world-class technical depth to build industry-leading systems in areas such as observability data platforms, distributed query engines, and real-time event streaming at global scale. As technical leaders, they apply broad architectural perspective and deep systems thinking to align design decisions across teams and domains. They work across complex, multi-team problem spaces to define long-term technical direction, drive large-scale initiatives forward, and ensure consistent execution. As engineering stewards, they play a key role in evolving our systems and engineering culture. They actively participate in Datadog’s senior technical community, bringing external insights and internal experience to elevate engineering standards and mentor the next generation of technical leaders. Examples of projects a Senior Staff Engineer may lead include designing and launching a new distributed data storage engine capable of handling hundreds of millions of records per second, building the real-time infrastructure behind a new observability product, or re-architecting a core service to support exponential growth in throughput and complexity. What You’ll Do: Be the technical owner of multiple critical systems or architecture areas, often spanning several t
Principal Engineer - Backend About Us: Paytm is India’s leading digital payments and financial services company, which is focused on driving consumers and merchants to its platform by offering them a variety of payment use cases. To merchants, Paytm offers acquiring devices like Soundbox, EDC, QR and Payment Gateway where payment aggregation is done through PPI and also other banks’ financial instruments. To further enhance merchants’ business, Paytm offers merchants commerce services through advertising and Paytm Mini app store. Operating on this platform leverage, the company then offers credit services such as merchant loans, personal loans and BNPL, sourced by its financial partners. About the role: As a Principal Engineer, you will help define the technical design and implementation roadmap across multiple solutions and will work with engineering leadership to ensure we resource and equip our squads with the right expertise to deliver those solutions. Requirements: 8 to 12 years of strong software design/development experience in building massively large scale distributed internet systems and products Hands on experience in Advance Java, Spring boot, AWS, Node, Agentic AI, LLM, RAG, Cursor, Copilot Experience and knowledge of open source tools & frameworks, broader cutting edge technologies around server side development Should be an active contributor to developer communities like Stack overflow, Top coder, Git hub, Google Developer Groups (GDGs). Superior organization, communication, interpersonal and leadership skills. Must be a self-starter who can work well with minimal guidance and in fluid environment. Preferred Qualifications : Bachelor's/Master's Degree in Computer Science or equivalent Skills that will help you succeed in this role: Expertise in Java, DB: RDBMS, Messaging: Kafka/RabbitMQ, Caching: Redis/Aerospike, Micro services, AWS Strong experience in scaling, performance tuning & optimization at both API and storage layers Problem
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The AI team is building cutting-edge solutions that bring the power of AI directly to edge devices while seamlessly integrating with cloud infrastructure. We are looking for a Lead Software Engineer to design and develop high-performance, scalable services to support AI workloads across edge and cloud environments. What You Might Do Design, build, and maintain services that power AI-driven applications, ensuring scalability and performance. Develop APIs and microservices that facilitate seamless integration between cloud-based AI models and edge devices. Optimize data pipelines and storage solutions for real-time AI inference and processing. Implement security and privacy best practices for distributed AI systems. Work closely with AI researchers, infrastructure engineers, and frontend developers to deliver end-to-end AI-driven solutions. Build and optimize an agent orchestration runtime that enables tool use, memory management, and multi-step reasoning across LLMs, APIs, and edge-connected systems. Develop robust logging, monitoring, and alerting systems to ensure system reliability.
Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role The AI team is building cutting-edge solutions that bring the power of AI directly to edge devices while seamlessly integrating with cloud infrastructure. We are looking for a Senior Software Engineer to design and develop high-performance, scalable services to support AI workloads across edge and cloud environments. What You Might Do Design, build, and maintain services that power AI-driven applications, ensuring scalability and performance. Develop APIs and microservices that facilitate seamless integration between cloud-based AI models and edge devices. Optimize data pipelines and storage solutions for real-time AI inference and processing. Implement security and privacy best practices for distributed AI systems. Work closely with AI researchers, infrastructure engineers, and frontend developers to deliver end-to-end AI-driven solutions. Build and optimize an agent orchestration runtime that enables tool use, memory management, and multi-step reasoning across LLMs, APIs, and edge-connected systems. Develop robust logging, monitoring, and alerting systems to ensure system reliabilit
Scale's LLM post-training platform team builds our internal distributed framework for large language model training. The platform powers MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs. It also serves as the underlying training framework for the data quality evaluation pipeline. Scale is uniquely positioned at the heart of the field of AI as an indispensable provider of training and evaluation data and end-to-end solutions for the ML lifecycle. You will work closely with Scale’s ML teams and researchers to build the foundation platform which supports all our ML research and development works. You will be building and optimizing the platform to enable our next generation LLM training, inference and data curation. If you are excited about shaping the future AI via fundamental innovations, we would love to hear from you! You will: Build, profile and optimize our training and inference framework. Collaborate with ML and research teams to accelerate their research and development, and enable them to develop the next generation of models and data curation. Research and integrate state-of-the-art technologies to optimize our ML system. Ideally you’d have: Passionate about system optimization Experience with multi-node LLM training and inference Experience with developing large-scale distributed ML systems Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc. Strong software engineering skills, proficient in frameworks and tools such as CUDA, Pytorch, transformers, flash attention, etc. Strong written and verbal communication skills to operate in a cross functional team environment. Nice to haves: Demonstrated expertise in post-training methods and/or next generation use cases for large language models including instruction tuning, RLHF, tool use, reasoning, agents, and multimodal, etc. Compensation packages at Scale for eligible roles include base salary, equity,
The worldwide data management software market is massive. At MongoDB we are transforming industries and empowering developers to build amazing apps that people use every day. We are the leading modern data platform and the first database provider to IPO in over 20 years. Join our team and be at the forefront of innovation and creativity. MongoDB is seeking a Software Engineer 3 to join the Atlas Clusters Platform team. The team is responsible for building MongoDB Atlas, our database as a service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. The Atlas Clusters Platform team develops the foundational orchestration platform behind MongoDB Atlas. Our systems drive cluster planning and execution, evolve the Atlas control plane toward service-oriented architecture, and provide critical infrastructure that help Atlas run safely and efficiently across cloud environments. We are looking to speak to candidates who are based in New York City for our hybrid working model. What you’ll do Build and design new features for MongoDB Atlas Contribute to and lead complex technical projects Work closely with product and design teams, considering the user’s perspective while building technical solutions Work with customers and support engineers to fix issues Collaborate with team members to develop our codebase, best practices, and design principles Learn from and mentor other team members We’re looking for someone who Has at least 3 years of professional software development experience Is skilled at writing large-scale, distributed backend systems in a compiled language (Java, C#, Go, etc.) Is comfortable working across the stack of a modern web application (e.g. React, TypeScript, Kubernetes) Has experience with at least one major cloud provider technology (AWS, Azure, GCP) Has led the launch of a new feature and maintained it in production Is eager to solve tough problems Has excellent
MongoDB is seeking a Senior Software Engineer to join the Atlas Clusters Organization. The organization is responsible for building MongoDB Atlas, our database as a service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. We are forming a new Atlas Clusters team in the Dublin area. We are looking to speak to candidates who are based in Dublin for our hybrid working model. What you’ll do Build and design new features for MongoDB Atlas Contribute to and lead complex technical projects Work with stakeholders throughout MongoDB to build our roadmap and product offerings Work with customers and support engineers to fix issues and become part of our on-call rotation Collaborate with team members to develop our codebase, best practices, and design principles Foster an inclusive and respectful work environment according to MongoDB's Core Values We’re looking for someone who Has at least 6 years of professional software development experience Is skilled at writing large-scale, distributed backend systems in a compiled language (Go, Java, C#, etc.). Has experience with at least one major cloud provider technology (AWS, Azure, GCP) Has led the launch of a new module and maintained it in production Is eager to solve tough problems Has excellent communication skills Is curious, collaborative, and motivated Success Measures In 3 months, you'll have shipped code into production and collaborated with the team to solve tough problems In 6 months, you'll have contributed to a large project and joined our on-call rotation In 12 months, you'll have designed new features, led development work, and become a go-to expert on parts of the system About MongoDB MongoDB is built for
MongoDB is seeking an Engineering Manager to join the Atlas Growth Team. The team is responsible for improving and creating new features for MongoDB Atlas, our developer data platform that accounts for 65% of the company’s revenue. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. The Atlas Growth team focuses on improving the customer experience via new product experiences and refinements to existing user flows. The team works closely with cross-functional partners, as well as other MongoDB engineering teams to bring new visions to life. We are constantly challenged to design features such as new onboarding flows, monetization improvements, and advanced cluster management tools for our large B2B customer base. We are looking to speak to candidates who are based in Dublin for our hybrid working model. What You’ll Do Manage a team of engineers to design, build and test new features for MongoDB Atlas Contribute to and lead complex technical projects Work with cross-functional stakeholders to design the team's roadmap, defining delivery dates that balance technical feasibility with the pace of the market Work closely with product, design and analytics teams, considering the user’s perspective while building technical solutions Collaborate with team members to develop our codebase, best practices, and design principles Learn from and mentor an impassioned array of team members We’re Looking for Someone Who Has at least 5 years of professional software development experience Has at least 2 years of people management experience Is skilled at writing large-scale, distributed backend systems in a compiled language (Java, C#, Go, etc.) Is comfortable working across the stack of a modern web application (e.g. React, TypeScript, React Testing Library) Has experience with at least one major cloud provider technology (AWS, Azure, GCP) Has a deep understanding of product analytics Has experience with A/B testing and
Engineering Manager - Backend About Us: Paytm is India’s leading digital payments and financial services company, which is focused on driving consumers and merchants to its platform by offering them a variety of payment use cases. To merchants, Paytm offers acquiring devices like Soundbox, EDC, QR and Payment Gateway where payment aggregation is done through PPI and also other banks’ financial instruments. To further enhance merchants’ business, Paytm offers merchants commerce services through advertising and Paytm Mini app store. Operating on this platform leverage, the company then offers credit services such as merchant loans, personal loans and BNPL, sourced by its financial partners. About the role: As a Engineering Manager, you will help define the technical design and implementation roadmap across multiple solutions and will work with engineering leadership to ensure we resource and equip our squads with the right expertise to deliver those solutions. Requirements: 8 to 12 years of strong software design/development experience in building massively large scale distributed internet systems and products Hands on experience in Advance Java, Spring boot, AWS, Node, Agentic AI, LLM, RAG, Cursor, Copilot Experience and knowledge of open source tools & frameworks, broader cutting edge technologies around server side development Should be an active contributor to developer communities like Stack overflow, Top coder, Git hub, Google Developer Groups (GDGs). Superior organization, communication, interpersonal and leadership skills. Must be a self-starter who can work well with minimal guidance and in fluid environment. Preferred Qualifications : Bachelor's/Master's Degree in Computer Science or equivalent Skills that will help you succeed in this role: Expertise in Java, DB: RDBMS, Messaging: Kafka/RabbitMQ, Caching: Redis/Aerospike, Micro services, AWS Strong experience in scaling, performance tuning & optimization at both API and storage layers Probl
Role Purpose: At Jumio, you will work for one of the market leaders in the global identity verification space that is helping to make the digital world a safer place for everyone. As a Software Development Engineer in the MLOpsTeam, you will develop the blueprint for highly scalable and performant ML model serving. Role Value: As a Software Engineer (SDE III), you will drive the continuous improvement of the infrastructure and applications to manage the lifecycle of ML assets (data, models) to better developer experience and strengthen governance capabilities. Secondly, you will design and implement robust ML infrastructure for model deployment, serving, and optimization. You will work on efficient CI/CD pipelines for ML models and leverage advanced compilers or hardware optimization to maximize inference performance while optimizing costs. We welcome you to challenge us to impact our software development processes and tools. Example Responsibilities: Upgrade ML assets (models, data) management systems for better developer experience and robust governance capabilities Build and optimize model serving infrastructure with a focus on inference latency and cost optimization Architect efficient inference pipelines that balance latency, throughput, and cost across various acceleration options Implement cost-efficient, enterprise-scale solutions Collaborate in a cross-functional, distributed team for continuous system improvement Work with MLEs, QA Engineers, and DevOps Engineers Evaluate and implement new technologies and tools Contribute to architectural decisions for distributed ML systems Experience and Qualifications : 5+ years of experience in software engineering with Python Experience with model lifecycle management (MLFlow, Weights & Biases or equivalent) Experience with data management ecosystem (quality, transformation, catalog) Experience with ML frameworks, particularly PyTorch Experience optimizing ML models with hardwar
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: We connect Airbnb’s community with the right information, in the right place, at the right time. We tailor Messaging & Notifications so hosts on Airbnb can streamline their operations, and travelers get just the information they need to enjoy their stay worry-free. Additionally, we are building new connections within our community to help enrich the experience of hosting & traveling on Airbnb: easing the process of hosting, and adding meaning to our guest’s trips. The data team utilizes industry-leading tools, builds scalable data systems and applies cutting-edge ML models to provide insights and empower all products in the Communication and Connectivity (CnC) organization. The Difference You Will Make: At CnC, data is foundational to our organization’s success.This role will lead key initiatives to design and build large-scale, distributed data systems - both batch and real-time processing. The data will power machine learning models and unlock new product features. You’ll be at the center of cross-functional collaboration, bridging backend, frontend/client, and machine learning engineering teams. CnC is applying GenAI and large language models (LLMs) to power products that enhance the Airbnb experience in various surfaces including highly used ones like Messaging. We're building a robust ML platform to power our product ambitions. A Typical Day: Shape the team’s long-term vision and roadmap in close collaboration with cross-functional partners across Airbnb Build strong relationships with partner engineering teams, including backend, client, data science, analytics,
About Us Blueshift is the Intelligent Customer Engagement Platform (CEP), headquartered in San Francisco, that empowers leading B2C brands to drive truly personalized, 1:1 marketing across every channel. Founded by repeat entrepreneurs who previously built Mertado (acquired by Groupon) and were part of the early team at Kosmix (acquired by Walmart), Blueshift leverages AI, including Predictive, Generative, and Agentic AI, to automate customer engagement for clients like ClearScore, LendingTree, Udacity, and U.S. News. Backed by top-tier VCs including Nexus Venture Partners, Storm Ventures, and SoftBank Venture Asia, the company has raised a total of $65 million in venture funding and is consistently recognized as a market leader and a Deloitte Technology Fast 500 award recipient. Blueshift is actively scaling its development center in Pune, India. As part of our team, you will drive innovation in cutting-edge technologies including machine learning, artificial intelligence, big data, and large-scale distributed data systems. This is an exciting career path for motivated individuals looking to build complex, impactful solutions that define the future of customer engagement. AI Solutions Engineer II As a Software Engineer in the AI Solutions team , you occupy a unique techno-functional position. You are not a researcher; you are an implementation specialist and problem-solver . You bridge the gap between our core AI infrastructure and real-world customer impact. You aren't just writing code; you are applying data engineering, analysis, and AI knowledge to help global brands realize the full potential of AI-driven marketing. Responsibilities End-to-End Solution Delivery: Lead the full lifecycle of custom AI projects—from initial customer design and technical architecture to testing and production implementation. Production Stewardship: Take ownership of the "last mile" of delivery. This includes triaging technical tickets, analyzing logs (Datadog/Kibana), and deb
About the Role & Team We’re looking for an Engineering Manager to lead the Data Infrastructure team within Statsig Experiment at Amplitude. You will lead a multidisciplinary team of software engineers, data engineers, and data scientists responsible for the systems that power experimentation at scale. The team owns three critical areas: Data ingestion: Collecting and importing experiment exposures, custom events, OpenTelemetry data, and real user monitoring data across SDKs, streaming systems, cloud storage, and customer data warehouses. Data computation: Building distributed computation systems that transform raw data into accurate, timely experiment results. Stats engine: Developing and productionizing the statistical methods that help customers make trustworthy decisions from their experiments. This is not a traditional data engineering management role. We are looking for a leader with a solid data science and statistical foundation who can connect advances in experimentation methodology with scalable production systems. You will help set our technical and scientific direction, translating new statistical methods and machine learning research into capabilities that customers can use reliably at scale. You’ll partner closely with data scientists, engineers, product managers, and customers to advance the state of experimentation. The ideal candidate is equally comfortable discussing causal inference and statistical power with data scientists, distributed computation architectures with engineers, and experimentation strategy with customers. What You’ll Do Lead and grow the team responsible for Statsig’s data ingestion, experiment computation, and stats engine. Define the technical and scientific strategy for advancing experimentation across both Statsig Cloud and warehouse-native deployments. Partner with data scientists and engineers to turn new statistical and causal inference methods into scalable, reliable product capabilities. Evolve our data and computatio
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Data Governance team makes sure Plaid handles consumer and customer data responsibly — and can prove it. Our mission is to enforce Plaid's privacy commitments and regulatory obligations in the systems themselves rather than in policy documents: we build the platform and controls that govern how data flows through Plaid — where it lives, who can use it, for what purpose, and for how long. That includes verifiable deletion of consumer data on request, enforcement of data-use restrictions so downstream systems can only use data in permitted ways, and the cataloging and classification that let Plaid know what data it holds and how sensitive it is. We operate at the scale of Plaid's entire data footprint, and correctness and auditability matter to us as much as throughput. As a Staff Software Engineer on Data Governance, you will set the technical direction for how Plaid enforces data governance at scale. You'll lead the design of distributed backend systems that reliably delete, restrict, and track data across dozens of services, making architectural decisions whose blast radius spans the whole company. You'll drive multi-quarter initiatives from ambiguous privacy and regulatory requirements through
Get new distributed systems engineer jobs by email
Daily job updates · Unsubscribe anytime