NVIDIA’s DGX Cloud organization is seeking a Senior Data Engineer to become part of its data team! We develop the reliable data foundation that supports fleet health, capacity, utilization, cost, reliability, and operational decision-making throughout DGX Cloud. Our platform supports engineering, operations, finance, and product teams managing and expanding large GPU fleets across cloud service providers and NVIDIA Cloud Partners. We are looking for a practical engineer and technical lead to take charge of a key part of the Navigator data platform. We develop the systems that transform distributed infrastructure telemetry and operational data into dependable, managed data products that support fleet health, capacity, utilization, cost, and operational decisions. We are seeking a hands-on, platform-minded engineer to build and evolve the systems that turn distributed infrastructure telemetry and operational data into reliable, governed data products. You will work across ingestion, transformation, data quality, platform architecture, security, observability, and self-service consumption to help make Navigator and the DGXC data platform a dependable source of truth. We do expect strong engineering fundamentals, experience operating production systems, and the ability to learn new platforms and domains quickly. What you'll be doing: Own systems end to end. For example, work from ambiguous customer and operational needs through architecture, implementation, deployment, observability, incident response, and ongoing support. Construct data pipelines and products. Such as designing and maintain batch and streaming ingestion, transformation, reconciliation, and serving paths for fleet, capacity, utilization, cost, scheduling, and operational telemetry. Build shared libraries, workflow and DAG or equivalent experience abstractions to evolve the data platform. Develop deployment tooling, data
Jobiba hiring network
Distributed Systems Engineer Data Platform Delivery Database Retrieval Jobs
1,301 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current distributed systems engineer data platform delivery database retrieval jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! The Data Platform team at Figma builds and operates the foundational systems that power analytics, AI/ML, and data-driven decision-making across the company. We serve a diverse set of stakeholders, including AI researchers, machine learning engineers, data scientists, product engineers, and business teams that rely on data for insights and strategy. Our team owns and scales critical platforms such as the Snowflake data warehouse, ML Datalake, orchestration and pipeline infrastructure, and large-scale data ingestion and processing systems, managing all data flowing into and out of these platforms. Despite being a small team, we take on high-scale, high-impact challenges. In the coming years, we're focused on building the data infrastructure layer for Figma's AI-powered products, driving cost and performance optimizations across our data stack, scaling our ingestion and reverse ETL capabilities for new product use cases, and strengthening data quality, reliability, and compliance at every layer. If you're passionate about building scalable, high-performance data platforms that empower teams across Figma, we'd love to hear from you! This is a full-time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Design and build large-scale distributed data systems that power analytics, AI/ML, and business intelligence across Figma. Develop batch and streaming solutions to ensure data is reliable, efficient, and scalable across the company. Manage and evo
About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. About the role We are looking for a Senior Data Engineer to help build and evolve the data platform that powers critical business decisions and customer-facing experiences. You will own significant parts of our data architecture that is main powerhouse of DevRev Computer’s memory for accurate and efficient Answers. As a part of data team, you will design and operate scalable data systems, and work closely with Software Engineering, AI Agent teams, Data Science, and Product teams to turn complex data requirements into reliable, high-quality data products. This role is ideal for an experienced engineer who enjoys solving challenging problems involving large-scale data, distributed systems, database architecture, and performance optimization. You will have significant technical ownership and the opportunity to influence the direction of our agentic data platform while helping raise the engineering bar across the team. Responsibilities Own data architecture for large-scale, high-impact projects, making thoughtful tradeoffs across scalability, reliability, performance, maintainability, and operational cost. Design, build, and operate scalable data pipelines and data systems that reliably ingest, transform, store, and serv
We're looking for an ML Data & Platform Engineer to own the infrastructure that powers our speech AI models: the pipelines that source and prepare training data, and the platform that trains, evaluates, and serves them in production. Speech AI has a data problem most ML teams don't, and you'll be at the centre of solving it, working as part of our ML team to remove friction across the entire lifecycle and get better models into production faster. This is a broad, cross-functional role suited to someone who enjoys working across the full stack: data infrastructure, distributed systems, and production ML, and who takes ownership of problems end to end rather than waiting to be told what to fix. What you'll do Designing, building, and maintaining scalable data pipelines for ingesting, transforming, validating, and storing large datasets used to train our models Developing and maintaining web scraping and data acquisition solutions to keep training datasets fresh, high-quality, and available at scale Building and operating the infrastructure that lets the ML team deploy and evaluate new models quickly, and that serves models efficiently and reliably in production Optimising infrastructure for both iteration speed and production reliability, including GPU utilisation, job scheduling, and training efficiency Implementing observability (monitoring, logging, alerting) across data pipelines and ML systems to catch issues early and keep things running smoothly Troubleshooting complex issues across distributed systems, spanning data infrastructure, training, and inference Continuously improving our data and MLOps practices, and helping shape the roadmap for how our platform evolves as we scale What you'll need Strong proficiency in Python and SQL, with a solid backend or data engineering foundation Hands-on experience with containerisation and orchestration (Docker, Kubernetes), and working with a major cloud provider Experience building data pipelines and ETL/ELT processe
Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . As a Senior Staff Software Engineer on the Data Platform team within Platform , you'll define and lead the technical strategy for Coinbase's data infrastructure, spanning ingestion, transformation, warehousing, streaming, and serving systems. This is a foundational role at the intersection of distributed systems, data engineering, and AI-readiness, reporting to the Senior Director of Engineering. You'll set architectural direction, drive multi-quarter roadmaps, and transition the organization from managed-service dependency toward engineering-built, platform-grade infrastructure that powers everything from fraud detection to modern multi-agent AI architectures. What you'll do: Own the technical strategy and architecture for Data Platform, setting direction across data ingestion, transformation, warehousing, streaming, and serving systems while driving engineering-led cost reduction at the infrastructure layer. Architect data infrastructure to natively support AI and ML workloads, ensuring pipelines, data lake systems, and compute can power ML training, feature stores, real-time inference, and multi-agent AI architectures at scale. Drive the evolution to near-real-time data availability, enabling downstream teams across Coinbase to act on fresher data for fraud detection, financial reporting, and analytics. Build alignment and secure commitment from senior leadership
About the Team The OpenAI Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role As a Research Engineer, Distributed Data Systems, you will design and scale the infrastructure that powers large-scale multimodal training and evaluation at OpenAI. You’ll manage distributed data pipelines, collaborate closely with researchers to translate requirements into robust systems, and harden pipelines that serve as the backbone for OpenAI's rapid iteration cycles. We’re looking for engineers who are detail-oriented, have strong experience with distributed systems, and excel at building reliable infrastructure in high-stakes environments. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and maintain data infrastructure systems such as distributed compute, data orchestration, distributed storage, streaming infrastructure, machine learning infrastructure while ensuring scalability, reliability, and security. Ensure our data platform can scale by orders of magnitude while remaining reliable and efficient. Partner with researchers to deeply understand requirements and translate them into production-ready systems. Harden, optimize, and maintain critical data infrastructure systems that power multimodal training and evaluation. You might thrive in this role if you: Have strong experience with distributed systems and large-scale infrastructure with a strong interest in data. Are detail-oriented and bring rigor to building and maintaining reliable systems. Demonstrate excellent software enginee
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity The Telemetry Data Platform group at New Relic builds the foundation for all of our products: data ingest, storage, and query. As an engineer working on NRDB, you’ll be contributing directly to the proprietary telemetry database technology at the core of our business. We own our software from top to bottom and are directly responsible for its quality and reliability. Each member of the team shares our pager rotation and will occasionally be on-call to respond to system failures; so we prioritize work that keeps the lights on and the pager quiet, in addition to the work that powers all of our new products and streams of data. If the idea of working on systems that process millions of messages per second and handle exabytes of data excites you, then you may be an excellent fit! What you'll do Develop new features with a focus on optimizing performance and efficiency Collaborate with the team to implement scalable solutions and enhance application performance Identifying and acting on opportunities to improve the reliability of our services This role requires 2+ years of professional experience in distributed SaaS software development. Proficiency in Java programming, expertise with algorithms and data structures, and building high-throughput software following best-practices. Deeper understanding of distributed systems and their core challenges. Experience using the command line to manage, investigate, and fix things when they’re broken. Expe
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Staff Software Reliability Engineer - Data Platform About the Team The Data Platform team is responsible for the foundational data services, systems, and data products for Okta that benefit our users. Today, the Data Platform team solves challenges and enables: Streaming analytics Interactive end-user reporting Data and ML platform for Okta to scale Telemetry of our products and data Our elite team is fast, creative and flexible. We encourage ownership. We expect great things from our engineers and reward them with stimulating new projects, new technologies and the chance to have significant equity in a company. Okta is about to change the cloud computing landscape forever. About the Position This is an opportunity for experienced Software Reliability Engineers to join our fast growing Data Platform organization that is passionate about scaling high volume, low-latency, distributed data-platform services & data products. In this role, you will get to work with engineers throughout the organization to build foundational infrastructure that allows Okta to scale for years to come. As a member of the Data Platform team, you will be responsible for designing, building, and deploying the systems that power our data analytics and ML. Our analytics infrastructure stack sits on top of many modern technologies, including Kinesis, Flink, ElasticSearch, and Snowflake. We are looking for experienced Software Engineers who can help desi
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Data Platform organization builds and operates the systems that power how data is stored, moved, and consumed across Robinhood. This organization spans three core pillars: Storage (Postgres, DynamoDB, and caching systems), Streaming (real-time event infrastructure), and Data Lake (ingestion and compute systems built on Delta Lake). Together, these platforms support transactional workloads, real-time data processing, and large-scale analytics that are critical to Robinhood’s products and operations. The team owns the full lifecycle of data—from low-latency order path systems to near real-time and batch analytics—serving millions of users and internal teams across the company. As a Senior Staff Software Engineer , you will serve as the technical lead across the Data Platform organization, shaping architecture and guiding execution across multiple teams. You’ll work on complex distributed systems challenges such as database sharding and proxy architectures, real-time streaming and CDC systems, and large-scale data ingestion and compute platforms. You’ll define and drive key technical bets, partner with engineering leaders to align platform capabilities with business needs,
The Infrastructure Engineering team is responsible for building and maintaining a self-service internal development platform that enables MongoDB engineering teams to reliably deploy and operate their own production services and products. We work with numerous engineering teams across the company to understand their infrastructure requirements and development workflows, develop broadly applicable self-service platform services and tooling, continuously monitor how platform services are being utilized, and look for ways to improve developer productivity through automation and education. We are big open source enthusiasts and use a number of open source tools in our stack (contributing upstream whenever possible). Some of the tools we use regularly include Go, AWS, Kubernetes, Crossplane, Terraform, Helm, Drone, Prometheus, and Grafana. However, technology is nothing without a stellar team of engineers that are focused on doing high quality work and working as a team to solve complex distributed computing and platform engineering problems. This is where you come in! We are looking to speak to candidates who are based in Gurugram for our hybrid working model. Our ideal candidate Has built and operated large-scale distributed systems in cloud providers (AWS strongly preferred) Has a strong backend programming background. Fluency in Go is strongly preferred; deep experience with another compiled or strongly-typed backend language is acceptable Has experience working with AI coding agents and can demonstrate building high quality context to yield high quality outputs Has experience designing and implementing medium-to-large software projects, including driving design reviews and mentoring less-senior engineers Pragmatic, detail-oriented, self-motivated, and understands the benefits of collaboration Strong experience operating production Kubernetes clusters, not just deployed to it Has practical experience defining and operating against SLI/SLOs for services they owned Str
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE: Are you passionate about driving innovation in the data storage industry? Are you skilled in crafting technical solutions that exceed customer expectations? If so, we have an exciting opportunity for you! Everpure (formerly Pure Storage), a leader in the data storage and flash technology space, is seeking a talented and motivated Senior Pre-Sales Systems Engineer to join our dynamic team. As a Senior Pre-Sales Systems Engineer, you will play a crucial role in understanding our customers' unique challenges and tailoring Everpure solutions to meet their specific needs. Collaborating closely with the sales team, you will act as a technical expert during the sales process, helping to showcase the value of our products and services. WHAT YOU’LL DO: Develop an exhaustive understanding of what drives a customer’s business and what motivates their decision making Connect the dots from technology solutions, inclusive of the Everpure portfolio and others from the ecosystem, to measurable customer business outcomes Partner closely with account managers, specialists and channel partners to create a seamless and holistic customer experience and strategy to drive revenue growth and net new business Delight customers and teammates with your technical leadership and domain expertise on storage products, distributed storage architectures, file systems, and competitive storage offerings in the DAS, NAS and SAN product spaces
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Everpure Cloud Azure Native is a generally available service that brings enterprise-grade block storage natively to Azure. With the first version already live, we are expanding the service into new use cases and markets while improving its reliability, operability, and customer experience. Our Foundation team owns a Go-based control-plane service that coordinates how customers provision and use storage in Azure. Around it, we work with a modern cloud stack including Temporal and other platform services for workflows, automation, and observability. This is a production cloud service: the code you write directly shapes how customers deploy, scale, and operate storage in their Azure environments. You’ll work on a core storage service in a major public cloud , as part of a joint effort between Everpure and Microsoft. You’ll design and evolve APIs and service behavior in the critical path of real customer workloads, collaborating closely with engineers across both companies. Clear API contracts, long-lived interfaces, test automation, and CI/CD are fundamental to how we build. You’ll have the opportunity to own services end to end and solve complex distributed-systems problems in the public cloud. WHAT YOU'LL DO Own and evolve a production cloud service that powers Everpure Cloud Azure Native, taking features from idea and design through deployment and operation in Azure for real customers. Build new capabilities and i
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Data Access team within Infra / Storage builds EaaS (Entities-as-a-Service), Roblox’s large-scale managed OLTP access and management platform powering tens of millions of QPS across thousands of services. EaaS abstracts the complexity of distributed databases and caching systems behind a consistent and productive developer experience, enabling teams across Roblox to safely build and operate stateful systems at massive scale without requiring deep database expertise. The team operates at the intersection of distributed systems engineering, reliability engineering, and platform architecture — solving hard infrastructure problems that directly impact Roblox-wide scale and stability. As a Senior Software Engineer you will work on some of Roblox’s hardest backend infrastructure challenges around scalability, reliability, workload governance, adaptive flow control, and distributed systems. You Will Build and evolve EaaS, Roblox’s managed OLTP access platform powering tens of millions of QPS across hundreds of services. Design infrastructure that abstracts distributed databases and caching systems behind a consistent, safe, and highly productive developer experience. Drive reliability and scal
MongoDB’s mission is to empower innovators to create, transform, and disrupt industries by unleashing the power of software and data. We enable organizations of all sizes to easily build, scale, and run modern applications by helping them modernize legacy workloads, embrace innovation, and unleash AI. Our industry-leading developer data platform, MongoDB Atlas, is the only globally distributed, multi-cloud database and is available in more than 115 regions across AWS, Google Cloud, and Microsoft Azure. Atlas allows customers to build and run applications anywhere—on premises, or across cloud providers. With offices worldwide and over 175,000 new developers signing up to use MongoDB every month, it’s no wonder that leading organizations, like Samsung and Toyota, trust MongoDB to build next-generation, AI-powered applications. Atlas Search is a multi-cloud service that allows users to execute complex full text and vector search queries using the MongoDB Query Language . Our users are free to focus on relevance and data retrieval instead of the machinery needed to search data at scale. Our team builds and maintains the instances and supporting infrastructure powering Atlas Search. This platform deploys and monitors search deployments, providing a highly scalable yet observable system for customers and engineers. The Atlas Search product is quickly gaining traction with customers and we are shipping core infrastructure components that enable this growth. This role is based in San Francisco, CA with an in-office or hybrid work model. Successful candidates will have the following qualities: 2+ years of hands-on experience designing, building, testing, and maintaining industrial-strength backend software and automation in complex codebases Experience developing distributed systems and multithreaded applications Familiarity with public cloud platforms, distributed infrastructure, and metric-based development Experience with at least one modern statically typed program
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity The Streaming Services team is responsible for routing and replicating New Relic's platform data through globally distributed cloud environments in a high-throughput, low-latency, and cost-aware streaming pipeline. Our streaming services process and route over a billion messages per minute across different continents, regions, and cloud providers, serving New Relic’s customer experience by making it easy for development teams to prioritize reliability. You will collaborate with this globally distributed team to develop expertise and best practices for New Relic development teams to operate highly reliable streaming services. What you'll do Own, build, maintain, and scale our streaming services and their support tools. Participate in an on-call rotation and bake stability into everything, continually seeking automation opportunities for built-in reliability. Participate in architectural definitions with a high degree of innovation and creativity. Own and improve your team processes. Develop automation and tooling to make our services more scalable and reliable. Use available innovation time to bring your creative ideas to life. This role requires Experience developing back-end services that use Flink, Kafka, or other streaming platforms. Large scale is a plus. Experience in writing software in Java, and you are not afraid of adapting, learning, and working with different languages and frameworks. Experience with distributed systems, concurrency, and scaling in
Get new distributed systems engineer data platform delivery database retrieval jobs by email
Daily job updates · Unsubscribe anytime