Jobiba hiring network

Distributed Systems Engineer Jobs

1,306 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current distributed systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

PE
Private Employer
📍 Bengaluru• Full-time• Hybrid
1mo ago

About Bazaarvoice At Bazaarvoice, we create smart shopping experiences. Through our expansive global network, product-passionate community & enterprise technology, we connect thousands of brands and retailers with billions of consumers. Our solutions enable brands to connect with consumers and collect valuable user-generated content, at an unprecedented scale. This content achieves global reach by leveraging our extensive and ever-expanding retail, social & search syndication network. And we make it easy for brands & retailers to gain valuable business insights from real-time consumer feedback with intuitive tools and dashboards. The result is smarter shopping: loyal customers, increased sales, and improved products. The problem we are trying to solve : Brands and retailers struggle to make real connections with consumers. It's a challenge to deliver trustworthy and inspiring content in the moments that matter most during the discovery and purchase cycle. The result? Time and money spent on content that doesn't attract new consumers, convert them, or earn their long-term loyalty. Our brand promise : closing the gap between brands and consumers. Founded in 2005, Bazaarvoice is headquartered in Austin, Texas with offices in North America, Europe, Asia and Australia. It’s official: Bazaarvoice is a Great Place to Work in the US , Australia, India, Lithuania, France, Germany and the UK! What you’ll be doing Lead, hire and grow a high-calibre team of frontend & backend engineers, and their line managers. Define and implement a roadmap based on critical business need, that delivers the valuable features, scale, and reliability our clients need, as well making systems more resilient and scalable. Drive business-significant and complex initiatives by collaborating across geographically distributed teams and partners Coach and mentor engineers globally in support of their growth and adherence to best practices. Who you are – Requirements for success in this rol

awsci/cdrest
View job →

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! The integration team is responsible for developing and scaling machine learning algorithms and infrastructure for LLM post-training, with a focus on large-scale, distributed RL methods. We strive for excellence in both engineering and science by meticulously designing experiments and design docs. While tasks are assigned according to everyone’s expertise, there is a global team effort to write production code and support the team research efforts, depending on individual interests and organizational needs. In particular, this role aims to enhance the global quality of the post-training codebase by implementing new tools to ease and support research, optimizing post-training algorithms, and scaling distributed RL to unprecedented levels. Please Note: We have offices in London, Paris, Toronto, San Francisco, New York but we are also remote-friendly! Applicants for this role may work anywhere between UTC−06:00 and UTC+01:00. As a Member of Technical Staff, you will: Design and write high-performing and scalable software for training models. Develop new tools to support and accelerate research and LLM training. Coordinate with other

pythonkubernetesgit
View job →

We’re looking for an Engineering Manager to lead our Sensitive Data Scanner (SDS) Telemetry team. The SDS group’s mission is to be the world’s easiest-to-use tool to discover, classify, manage, and report sensitive data risks across cloud, on-premise, and code environments. This team builds and scales the detection capabilities that scan all telemetry data flowing into Datadog — logs, APM spans, and RUM events — operating in streaming, at processing time, and at very large scale. You’ll lead a small, close-knit team based in Paris, with the opportunity to shape how the team grows as SDS Telemetry’s scope expands. It’s a chance to combine hands-on technical leadership with direct customer and product impact in the security and observability space. At Datadog, we place value in our office culture — the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a team of engineers building real-time sensitive data detection across Datadog’s Logs, APM, and RUM telemetry pipelines Partner closely with the Logs, APM, and RUM teams, plus Datadog’s Trust & Safety team, to align on roadmap and integration priorities Shape product direction by working closely with Product, grounding decisions in customer needs and business impact Stay hands-on: contribute to design decisions and participate in the team’s on-call rotation Recruit, mentor, and develop engineers as the team grows beyond its initial size Help build a strong engineering culture as part of Datadog’s broader Sensitive Data Scanner group Who You Are: You have experience building and shipping revenue-generating products, with strong product acumen and a customer-first mindset You have hands-on experience with Go and/or Java, and a track record building distributed, streaming systems at scale You have experience managing engineers — or are

javaaigo
View job →
TI
TEGNA India
📍 Chennai• Full-time
16 days ago

TEGNA Inc. helps people thrive in their local communities by providing the trusted local news and services that matter most. With 64 television stations in 51 U.S. markets, TEGNA reaches more than 100 million people monthly across web, mobile apps, streaming, and linear television, while also maintaining a strong global presence in India with offices in Bangalore and Chennai that support technology, product, and business operations initiatives. Together, we are building a sustainable future for local news. Senior DevOps Engineer About TEGNA TEGNA Inc. (NYSE: TGNA) helps people thrive in their local communities by providing trusted local news and services. With 64 television stations across 51 U.S. markets, TEGNA reaches more than 100 million people monthly across digital, mobile, streaming, and television platforms. We are focused on innovation, technology excellence, and building scalable solutions that create meaningful impact. Position Overview TEGNA is looking for a highly skilled Senior DevOps Engineer with strong expertise in AWS, Kubernetes, and Infrastructure as Code to design, automate, and manage scalable cloud infrastructure. The ideal candidate will have hands-on experience operating Kubernetes workloads in production, building CI/CD pipelines, and implementing monitoring and security best practices. This role requires deep technical expertise, strong troubleshooting skills, and the ability to work in fast-paced, distributed environments. You will play a key role in ensuring platform reliability, automation maturity, and production stability across cloud-native microservices systems. What You’ll Do Design and manage cloud infrastructure using Infrastructure as Code (AWS CDK, CloudFormation, Terraform). Build and maintain CI/CD pipelines using GitHub Actions and Jenkins to enable automated and reliable deployments. Deploy, manage, and scale Kubernetes clust

pythonsqlmongodb
View job →
TI
TEGNA India
📍 Bengaluru• Full-time
16 days ago

TEGNA Inc. helps people thrive in their local communities by providing the trusted local news and services that matter most. With 64 television stations in 51 U.S. markets, TEGNA reaches more than 100 million people monthly across web, mobile apps, streaming, and linear television, while also maintaining a strong global presence in India with offices in Bangalore and Chennai that support technology, product, and business operations initiatives. Together, we are building a sustainable future for local news. Operation Analyst Position Overview TEGNA is looking for Operation Analyst to join the dynamic engineering team. This role is responsible for the day-to-day monitoring, support, and operational stability of TEGNA’s technology platforms, including broadcast, streaming, and OTT systems, while providing advanced helpdesk support to end users and business-critical applications. The position works closely with distributed IT teams, streaming, and digital teams, as well as vendors and service partners, to resolve incidents, manage tickets, and maintain service levels in a 24x7 support environment. Clear documentation, effective communication, and ownership of issues from intake through resolution are critical to ensuring reliable linear and digital content delivery. What You’ll Do Handle and resolve ServiceNow support tickets from intake to closure within defined SLAs. Provide Tier 1 & Tier 2 technical support for PC, Mac, Android, and iOS devices. Support collaboration platforms, workstreams, and AI tools. Monitor critical systems and applications proactively to prevent service disruptions. Perform routine health checks, respond to alerts, and provide operational support. Troubleshoot incidents and perform root cause analysis for recurring issues. Document resolutions, workarounds, SOPs, and knowledge base articles in ServiceNow. Maintain operational documentation and improve knowl

awsazuregcp
View job →
R
Roblox
📍 San Mateo• Full-time• From $399.4K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Content Platform team at Roblox powers the infrastructure behind every asset used across the Roblox ecosystem—enabling creators and developers to bring their visions to life at global scale. From 3D models and images to videos and audio, our platform manages the complete lifecycle of all assets essential for immersive experiences, supporting one of the largest services in the world at over 100+ million requests per second. Our mission is to deliver a seamless, reliable, and innovative content system that empowers creators, supports record-breaking games, and ensures the highest standards of performance and safety for our community. As the Technical Director for Content Platform, you will lead multidisciplinary engineering teams responsible for the technical and product vision of Roblox’s asset infrastructure. You will own the lifecycle of every asset—from creation and upload to storage, indexing, delivery, and rendering in the game client. Your leadership will be critical in scaling our systems, optimizing distributed infrastructure, and enabling new possibilities for creators and players alike. You Will: Define and drive the long-term strategy, architecture, and priorities for the Cont

awsgitai
View job →
M
Modal
📍 San Francisco• Full-time
12 days ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Specifically, you'll be working on the distributed object storage system that underpins every container image, volume, and checkpoint on Modal: hundreds of petabytes of data, replicated across multiple cloud object stores and a CDN, cached on local NVMe across a large fleet of workers in many datacenters, and shared peer-to-peer within each datacenter. You'll make cold starts feel local when the data is hundreds of milliseconds away, designing the caching, preloading, and peer-to-peer layers that hide object-store latency and keep public ingress off saturated uplinks. You'll own durability and cost at petabyte scale, from streaming and batch replication between origins, to garbage collecti

M
12 days ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for a strong technical lead to guide the engineers designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. You'll lead the team responsible for Modal's machines layer: the fleet of bare metal and cloud hosts that every Function, Sandbox, and training job runs on, and the control plane that provisions, images, monitors, and repairs them. You'll own the full lifecycle of a machine, from accepting and benchmarking new hardware from a growing set of providers, to network bring-up, kernel and image management, GPU and disk health tracking, and automated remediation of unhealthy hosts. You'll manage a team of 3–8 engineers while staying hands-on across the stack which involves BMCs, firmware, PXE, bootloaders, Linux networking, drivers, and distributed control-plane services, and you'll shape our long-

M
Modal
📍 San Francisco• Full-time
12 days ago

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for a strong technical lead to guide the engineers designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. You'll lead the team responsible for the distributed object storage system that underpins every container image, volume, and checkpoint on Modal: hundreds of petabytes of data, replicated across multiple cloud object stores and a CDN, cached on local NVMe across a large fleet of workers in many datacenters, and shared peer-to-peer within each datacenter. You'll set technical direction for the primitives that other teams (filesystems, training, sandboxes) build on, balancing durability, latency, throughput, and cost. You'll own the roadmap from today's hardest problems (garbage collection at petabyte scale, active-active replication, rate limiting that protects the upstream without wasting ut

M
Mongodb
📍 Ireland• Full-time
1mo ago

MongoDB is seeking an Engineering Manager to join the Atlas Organization. The organization is responsible for building MongoDB Atlas, our database-as-a-service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. The Atlas Data Federation & Archiving team is an engineering team responsible for the Atlas capabilities that allow customers to move data from hot to cold storage and run federated queries over that data. The team builds Atlas Data Federation, a distributed query engine that lets users query data across Atlas Clusters and cloud object storage through a unified service. The team also builds Atlas Online Archive which allows customers to move data from Atlas Clusters into fully managed cloud object storage while preserving a seamless query experience across hot and cold datasets. We are forming a new Atlas Data Federation & Archiving team in the Dublin area. The Engineering Manager who fills this position will be pivotal in growing that team. We are looking to speak to candidates who are based in Cork and would like a hybrid or in-office working model. What you’ll do Lead a team of motivated individual contributors who are eager to learn and grow Contribute to the code, design, and architecture of the systems your team develops Work with stakeholders throughout MongoDB to build our roadmap and product offerings Work with customers and support engineers to fix issues and become part of our on-call rotation Collaborate with team members to develop our codebase, best practices, and design principles Foster an inclusive and respectful work environment according to MongoDB's Core Values We’re looking for someone who Has at least 6 years of professi

javamongodbaws
View job →
M
Mongodb
📍 Ireland• Full-time
1mo ago

MongoDB is seeking an Engineering Manager to join the Atlas Organization. The organization is responsible for building MongoDB Atlas, our database-as-a-service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. The Atlas Data Federation & Archiving team is an engineering team responsible for the Atlas capabilities that allow customers to move data from hot to cold storage and run federated queries over that data. The team builds Atlas Data Federation, a distributed query engine that lets users query data across Atlas Clusters and cloud object storage through a unified service. The team also builds Atlas Online Archive which allows customers to move data from Atlas Clusters into fully managed cloud object storage while preserving a seamless query experience across hot and cold datasets. We are forming a new Atlas Data Federation & Archiving team in the Dublin area. The Engineering Manager who fills this position will be pivotal in growing that team. We are looking to speak to candidates who are based in Dublin and would like a hybrid or in-office working model. What you’ll do Lead a team of motivated individual contributors who are eager to learn and grow Contribute to the code, design, and architecture of the systems your team develops Work with stakeholders throughout MongoDB to build our roadmap and product offerings Work with customers and support engineers to fix issues and become part of our on-call rotation Collaborate with team members to develop our codebase, best practices, and design principles Foster an inclusive and respectful work environment according to MongoDB's Core Values We’re looking for someone who Has at least 6 years of profes

javamongodbaws
View job →
D
Datadog
📍 New York• Full-time• From $280K/yr
1mo ago

Datadog is seeking a Director of Product Management to lead our AI Observability portfolio and shape how organizations build, monitor, and scale AI systems in production. This role leads LLM Observability and helps define the next wave of innovation across GPU Monitoring, Distributed AI Monitoring, and emerging research-oriented tooling such as Model Lab. You will set the vision and strategy for this rapidly growing area, expanding established products while incubating new capabilities that deliver deep visibility into AI infrastructure, model performance, and distributed AI environments. As AI becomes core to modern applications, this team plays a critical role in ensuring customers can deploy and scale AI with confidence. We’re looking for a builder-minded product leader with strong technical depth and hands-on curiosity - someone who has built or worked closely with AI-powered products and understands the realities of production AI. You will lead a team of product managers and partner closely with engineering and design to advance Datadog’s leadership in AI observability. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the vision and strategy for AI-driven products, ensuring alignment with overall company goals and customer needs. This will include managing our embed program to enhance the capabilities of existing products as well as developing dedicated and independent AI products. Lead and mentor a team of product managers, helping them grow and advance their careers while ensuring the delivery of high-quality, AI-powered features. Collaborate with cross-functional teams including engineering, data science, marketing, and sales to deliver AI product solutions that meet customer needs and business objectives. Identify new opportunities for

machine learningaigo
View job →
LA
16 days ago

We are looking for a Full Stack Developer to join our growing engineering team. In this role, you will design, build, and operate scalable software platforms that support analytics and AI solutions — owning the full journey from intuitive user interfaces to robust cloud-native backends. What This Involves: Front-End Development Develop and maintain user interfaces for web applications using React and Next.js. Translate wireframes and design mockups into functional, accessible UI components. Identify and resolve front-end performance bottlenecks to ensure smooth user experiences. Write and maintain unit and integration tests for front-end components (e.g. Jest, React Testing Library). Back-End Development Develop and maintain high-quality back-end services and APIs using Python. Deploy, operate, and monitor applications in cloud environments (AWS, Azure, or GCP). Manage containerized applications using Docker and Kubernetes. Contribute to the design and evolution of scalable, cloud-native software architectures. Contribute to and maintain CI/CD pipelines for web and back-end applications. Collaboration and Quality Work closely with data scientists, engineers and project managers to deliver integrated end-to-end solutions. Support the development and deployment of AI and analytics solutions. Write clean, well-documented, and maintainable code across the full stack. Participate in technical discussions, code reviews, and continuous improvement initiatives. Adhere to internal and client-mandated data protection and compliance policies, ensuring all handling, storage, and sharing of data meets required security and privacy standards. Requirements: Bachelor’s degree in Computer Science or related fields. 5+ years of software development experience, with meaningful time on both front-end and back-end systems. Experience designing systems in cloud-native or distributed environments is a plus. Excellent communication and collaboration skills — comfortable working acros

pythonreactaws
View job →

Are you ready to do your life’s work at the heart of the autonomous revolution? NVIDIA’s SWQA organization is seeking a world-class Software QA Test and Tool Developer to join our Automotive Platform team, where the code you validate ensures the safety of millions on the road. In this role, you won't just be testing software; you will be architecting the security and reliability of the next generation of intelligent vehicles. We are looking for engineers who are as comfortable navigating low-level product architecture as they are deep-diving into complex product use cases with passion for quality. This is a high-impact, hands-on position focused on our industry-leading automotive products, offering a rare opportunity to influence the core of our tech stack. You will also build the tools and frameworks that define performance standards for systems running on Linux and QNX. What you’ll be doing: Design, execute, and automate comprehensive test cases and test scenarios to validate our automotive platforms using various test methodologies to identify and track actionable defects and track them to closure. Participate in deep-dive reviews of product requirements and technical designs, providing critical feedback to ensure features are built for testability and security from day one. Partner closely with project management, hardware teams, and software developers to provide rigorous technical analysis of bugs and publish data-driven statistical reports for global team members. Architect and maintain a distributed test automation framework capable of managing high-concurrency workloads across an extensive automation farm of hundreds of concurrent systems. Develop sophisticated test libraries and automation solutions to accelerate development cycles and expand automated test coverage for re

pythonlinuxai
View job →
S
13 days ago

About the role (Remote) We're hiring a Senior Product Designer to join our global UX team. You'll own design within your product cluster — working directly with the engineers and PMs in those areas to define problems, shape solutions, and ship work that holds up. This is an end-to-end role. You lead discovery, define the interaction model, produce specs, and stay close through implementation. You're also expected to contribute to the systems and standards the whole team relies on. The Senior Product Designer will report to our Product Design Manager, and will work with our EU, UK and US based designers. You're part of a distributed team, so async clarity and proactive communication are as important as craft. What you’ll do Lead the design of complex AI-powered interactions in your product cluster, accounting for trust, transparency, fallback states, and human-in-the-loop considerations Embrace the use of modern AI-based tools like Claude, Figma Make, Cursor, Vercel, Loveable or equivalent not as theatre, but as a meaningful multiplier to your workflow. Partner with our UX research team to lead interviews, usability tests, and competitive analysis and translate findings into clear product direction and guidance Develop presentations, wireframes, mockups, and prototypes that effectively communicate interaction and design intent across web and mobile surfaces. Partner with PMs to shape requirements and with engineers to protect design intent through delivery Contribute to Showpad's design systems — adding components, documenting patterns, raising consistency across product areas Communicate design rationale through strategic storytelling: you can walk a PM, an engineer, and a senior stakeholder through the same decision and each one gets what they need Establish scalable AI-assisted workflows for your own work — research synthesis, rapid prototyping, concept generation — and help others on the team use them effectively Deliver UX solutions that measurably improve

aiaccounting
View job →
🔔

Get new distributed systems engineer jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More distributed systems engineer opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.