Jobiba hiring network

Lead Data Engineer Jobs

6,753 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead data engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox's data infrastructure processes petabytes of data daily, powering analytics, ML, and product decisions for a platform serving 200M+ daily active users. As a Principal Software Engineer in our Data Infra org, you will be the primary technical leader driving the strategic vision, long-term architecture, and massive scalability of our distributed data platforms that power Roblox. You will own and drive the next-generation architecture of our core platforms, which span Kafka, Flink, Spark, Trino, Druid, Airflow and Data Catalog. This role operates under high ambiguity, demanding unparalleled ownership to redefine the limits of infrastructure handling exabyte-scale workloads, and providing a unique opportunity to lead the future evolution of our global data ecosystem. You Will: Define Multi-Year Technical Strategy: Own and drive the end-to-end architectural vision for Roblox's core data platforms spanning Kafka, Flink, Spark, Trino, Druid, Airflow, and Data Catalog systems. Turn multi-year company strategies into concrete, production-grade infrastructure blueprints. Lead Cross-Functional Alignment: Partner closely with executive leadership, platform governance, data science, and product e

javaawsgcp
View job →
C
Coinbase
📍 - USA• Full-time• Remote• From $186.1K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Senior Software Engineer, Data Layer As a Senior Software Engineer on the Data Layer team within the Platform group, you'll shape the API (GraphQL) platform that connects every client application to Coinbase's backend services, handling the majority of user traffic across Consumer, Base, and Institutional products. You'll own the critical systems that route and serve API requests at scale, improving reliability, performance, and developer experience for hundreds of engineers building on this foundation layer. What you'll do: Own and deliver projects end to end, from scoping and system design through implementation, rollout, and production validation, driving measurable progress on the team's highest-priority initiatives. Design and build high-reliability, low-latency systems serving millions of users, tackling challenges like caching, upstream service optimization, and efficient connection management at scale. Build and improve the API framework and tooling that hundreds of engineers depend on, making it fast and easy for teams across the company to build, test, and ship independently. Drive operational excellence across T0 services: own SLOs, improve observability, lead incident response, and reduce operational toil so the team can invest in high-leverage work. Partner cross-functionally with product engineering teams to design API schemas, support service launches,

REMOTEjavaawsgraphql
View job →

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The Unified Data Store (UDS) team is the architect of Airbnb’s global system-of-record. We design, build, and operate the mission-critical storage platform that powers every user profile, listing, reservation, and financial transaction on the platform. Supporting over 150 million users worldwide, our work is the bedrock of Airbnb’s reliability and efficiency. As a Staff Engineer in our Brazil Engineering Hub , you will join a high-impact group of technical leaders who value craft and operational excellence. You won't just be managing data; you will be building a modern distributed infrastructure service that enables hundreds of product teams to ship features with total confidence. The Difference You Will Make: We are looking for a hands-on technical leader who thrives on solving deep architectural challenges and leading through ambiguity. As a Staff Engineer, you will serve as the Technical North Star for the UDS Client Stack, ensuring our data access layer is seamless, high-performance, and future-proof. A Typical Day: Define Technical Strategy: Lead the multi-year roadmap and long-term architecture for the UDS client stack, balancing immediate execution with systemic platform evolution. Architect for Scale: Design and operate a high-performance data access layer that abstracts complexities like indexing, replication, and global consistency models. Drive Engineering Excellence: Lead deep-dive design reviews and establish best practices for building fault-tolerant distributed systems across the organization. Empower Developers: Act as a bridge between infrastructure and

About the Team Training Runtime builds the distributed systems that power OpenAI's largest model training runs - most recently GPT-5.5! The Data Movement area owns the infrastructure that keeps training jobs supplied with the right data at the right time, and keeps model state moving safely and efficiently across large clusters. Our work spans machine learning systems, distributed storage, high-throughput data loading, reliability engineering, and developer experience. Success means researchers can move quickly while training runs remain fast, reproducible, debuggable, and resilient at scale. About the Role We are looking for a deeply hands-on Technical Lead Manager to own datasets throughout our training infrastructure. This person will set the direction for how training jobs read data: the APIs, storage contracts, versioning model, benchmarks, debugging tools, and reliability guarantees that make data access consistent across current and future training frameworks. You will begin as the primary technical owner for dataset reads, working directly in the code while aligning researchers, training framework owners, storage teams, and infrastructure partners around a durable platform. The problem is deceptively hard at frontier scale: make enormous, heterogeneous datasets easy to consume, correct across distributed workers, observable when something goes wrong, and flexible enough to support pretraining, reinforcement learning, and multimodal training. In this role, you will Design and build a unified dataset read platform for multiple current and future training frameworks. Define dataset APIs, storage-format expectations, registration/versioning, and migration paths that make data access reproducible and maintainable. Build reliability into the read path, including stateful iteration, caching, fast restart, recovery, and clear operational contracts. Build terminal and web-based visualizers that let teams inspect text, multimodal, and reinforcement learning data late

pythonawsrest
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. As a Principal Engineer on the Data Platform team, you will be a technical anchor — setting direction, solving the hardest engineering problems, and elevating the entire team. You will work at the intersection of systems design, long-term architecture, and real-world engineering execution. In this specific role, you will lead the evolution of our declarative data pipelines, focusing on streaming ingestion and transformations. What You'll Do Drive the technical strategy and architecture for key Data Platform initiatives, focusing heavily on real-time data movement and incremental processing. Lead design reviews and set engineering standards across teams. Identify and resolve high-impact technical challenges and systemic bottlenecks to reduce latency and improve the efficiency of dynamic tables at cloud scale. Mentor and grow senior engineers; raise the engineering bar. Partner closely with product and infrastructure leadership on roadmap and direction. Continuous hands-on technical deliverables in the most critical areas. What We're Looking For 14+ years of software engineering experience with deep expertise in distributed systems. Demonstrated track record of defining and delivering platform-scale technical initiatives. Expert-level knowledge of streaming and/or batch data

aigorust
View job →

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of a best-in-class family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from a diverse group of backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a senior validation lead engineer to lead at-scale rack validation efforts for next-generation AI hyperscale systems. This role focuses on post-silicon system validation across the full lifecycle, ensuring functional, electrical, and thermal performance meets product objectives. You will own end-to-end blade and rack validation including planning, development, execution, and debug while collaborating across firmware, systems, and hardware teams. The Team The Rack Validation team is responsible for ensuring system readiness and quality at scale. The team works cross-functionally with firmware, silicon, and system engineering teams to validate complex AI compute platforms. Responsibilities and Duties Lead post-silicon validation of AI compute blades and racks including test planning, development, and automation. Drive provisioning and integration of system components (SoC FW, BMC, RMC, OS) for rack-level readiness. Own execution against program achievements and report validation progress and risks. Triage test failures, collect debug data, and collaborate on root cause analysis. Track

pythonci/cdlinux
View job →
C
Coinbase
📍 Singapore• Full-time• Remote• From S$143.7K/yr
1mo ago

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Software Engineer, Data Layer As a Software Engineer on the Data Layer team within the Platform group, you'll shape the API (GraphQL) platform that connects every client application to Coinbase's backend services, handling the majority of user traffic across Consumer, Base, and Institutional products. You'll own the critical systems that route and serve API requests at scale, improving reliability, performance, and developer experience for hundreds of engineers building on this foundation layer. What you'll do: Own and deliver projects end to end, from scoping and system design through implementation, rollout, and production validation, driving measurable progress on the team's highest-priority initiatives. Design and build high-reliability, low-latency systems serving millions of users, tackling challenges like caching, upstream service optimization, and efficient connection management at scale. Build and improve the API framework and tooling that hundreds of engineers depend on, making it fast and easy for teams across the company to build, test, and ship independently. Drive operational excellence across T0 services: own SLOs, improve observability, lead incident response, and reduce operational toil so the team can invest in high-leverage work. Partner cross-functionally with product engineering teams to design API schemas, support service launches, and ens

REMOTEjavaawsgraphql
View job →
J
Justworks
📍 Phoenix• Full-time• $113K – $141K/yr
1mo ago

Who We Are At Justworks, you’ll enjoy a welcoming and casual environment, great benefits, wellness program offerings, company retreats, and the ability to interact with and learn from leaders in the startup community. We work hard and care about our most prized asset - our people. We’re helping businesses get off the ground by enabling them to focus on running their business. We solve HR issues. We’re data-driven and never stop iterating. If you’d like to work in a supportive, entrepreneurial environment, are interested in building something meaningful and having fun while doing it, we’d love to hear from you. We're united by shared goals and shared motivations at Justworks. These are best summed up in our company values, which are reflected in our product and in our team. Our Values If this sounds like you, you’ll fit right in. Who You Are You're a seasoned IT support professional who pairs white-glove customer support with the technical depth to solve complex problems independently in a fast-paced environment. You bring sound judgment, professionalism, and the ability to navigate ambiguity without close oversight. As the first IT hire in our new Phoenix office, you'll be the local face of IT while also serving as a key technical resource for Justworkers across all our locations. In this hybrid role (3 days in office, 2 days remote), you'll report to the IT Support Manager and partner with our global IT, Network, and AV teams to keep operations running smoothly across the organization. Beyond core support, you'll bring strong network and AV capabilities to keep the Phoenix office running smoothly — and act as a trusted technical partner to leadership during high-visibility moments. Your Success Profile What You Will Work On IT Support (Primary) Day-to-Day Support: Serve as the primary onsite contact in Phoenix for technical issues across macOS, Windows, mobile devices, and peripherals. Actively support Justworkers across all Justworks offices — providing remote sup

pythonawsgit
View job →
S
Stripe
📍 New York• Full-time
1mo ago

About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the Team Every business running on Stripe eventually asks the same questions: What's happening with my business? What does it mean for my company? How do I reconcile my books? And how do I get my Stripe data into my own systems and apps? This is all becoming even more important in the era of AI, where merchant’s agents need the right access to, and understanding of, Stripe’s data. The Data Products team exists to answer those questions — at any scale, for any business. We build the products that turn Stripe's rich transaction data into understanding and action: Sigma for deep SQL-powered analytics, Stripe Data Pipeline for delivering Stripe data directly into merchants' own warehouses and applications, and databases for surfacing the right insight, in context, right inside the Dashboard. Our users range from a founder running their first revenue report to an enterprise finance team reconciling millions of transactions across dozens of markets, to any agent working on a merchant’s behalf. What they all share is a need to trust their data, understand their business, and move fast. We make that possible. About the Role As the PM leading Data Products, you will own the product strategy and execution for Stripe's merchant-facing data portfolio — defining what it means for Stripe to be the best data platform for our merchants and agents. You'll set the direction across Sigma, SDP, and other data products, lead a team of PMs, and work closely with engineering, d

sqlairust
View job →
E
18 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE 1touch.io is a technology company focused on automated, real-time discovery, mapping, and tracking of sensitive personal data. Its AI-powered platform helps enterprises improve data privacy, security, and governance across complex on-premises and cloud environments. Together, 1touch.io and Everpure transform enterprise data from passive storage into an intelligent, context-aware, and governed foundation - making it AI-ready at the source so organizations can securely understand, trust, and activate their data at scale. As a Java Engineer, you will design and deliver core platform capabilities that help organizations discover, classify, and protect sensitive data across complex environments. You will work on highly scalable backend systems, collaborating with engineering, product, and customer-facing teams to solve challenging technical problems where performance, resilience, and precision matter. WHAT YOU'LL DO Design and build software that discovers, classifies, and safeguards sensitive data across diverse environments and platforms. Lead the design, implementation, testing, deployment, and continuous improvement of platform capabilities. Develop reliable, scalable, and high-performing backend systems designed for complex, real-world operating environments. Extend platform capabilities by integrating new data sources, storage systems, and services. Solve engineering challenges related to scale, concurrency, perf

javaawsrest
View job →
A
1mo ago

About The Role & Team Amplitude is only as useful as the data inside it. The Data Connections team owns how that data gets in and out — importing behavioral and customer data from cloud data warehouses like Snowflake, Databricks, and BigQuery, from cloud object storage like S3 and Azure Blob Storage, and pushing enriched event data back out to warehouses, object storage, streaming destinations, and downstream advertising and marketing platforms. That means batch and streaming pipelines moving billions of events a day, connections that have to keep working across dozens of customer-controlled systems, credentials and configuration that have to stay correct and secure, and latency and reliability targets that customers build their own pipelines on top of. Recent work includes launching new warehouse export destinations, migrating our import pipelines onto a durable workflow engine, building low-latency streaming export, and supporting cross-region and cross-cloud customer storage. You'll lead a team of around 10 engineers building this, reporting to our Head of Data Platform, and you'll be the person accountable for the roadmap, the reliability, and the growth of the people on it. The stack is Java, Temporal, Kafka, Kubernetes, Terraform, DynamoDB, S3, and Snowflake, running on AWS and GCP. Responsibilities Set the direction and roadmap for data import and export, with clear priorities and tradeoffs you can explain to engineers, PMs, and customers Lead, coach, and grow a team of around 10 engineers — hiring, career development, feedback, and performance Partner with your tech leads on architecture across ingestion, transformation, and delivery, without becoming the bottleneck for every decision Own reliability, SLOs, and cost efficiency for pipelines customers depend on daily Expand the set of destinations and sources we support, and make each new integration cheaper to build than the last Work directly with Product, Design, and other Data Platform teams to ship e

javaawsazure
View job →
M
Mongodb
📍 Cork• Full-time
1mo ago

MongoDB is seeking an Engineering Manager to join the Atlas Organization. The organization is responsible for building MongoDB Atlas, our database-as-a-service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. The Atlas Data Federation & Archiving team is an engineering team responsible for the Atlas capabilities that allow customers to move data from hot to cold storage and run federated queries over that data. The team builds Atlas Data Federation, a distributed query engine that lets users query data across Atlas Clusters and cloud object storage through a unified service. The team also builds Atlas Online Archive which allows customers to move data from Atlas Clusters into fully managed cloud object storage while preserving a seamless query experience across hot and cold datasets. We are forming a new Atlas Data Federation & Archiving team in the Dublin area. The Engineering Manager who fills this position will be pivotal in growing that team. We are looking to speak to candidates who are based in Cork and would like a hybrid or in-office working model. What you’ll do Lead a team of motivated individual contributors who are eager to learn and grow Contribute to the code, design, and architecture of the systems your team develops Work with stakeholders throughout MongoDB to build our roadmap and product offerings Work with customers and support engineers to fix issues and become part of our on-call rotation Collaborate with team members to develop our codebase, best practices, and design principles Foster an inclusive and respectful work environment according to MongoDB's Core Values We’re looking for someone who Has at least 6 years of professi

javamongodbaws
View job →
M
Mongodb
📍 Dublin• Full-time
1mo ago

MongoDB is seeking an Engineering Manager to join the Atlas Organization. The organization is responsible for building MongoDB Atlas, our database-as-a-service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. The Atlas Data Federation & Archiving team is an engineering team responsible for the Atlas capabilities that allow customers to move data from hot to cold storage and run federated queries over that data. The team builds Atlas Data Federation, a distributed query engine that lets users query data across Atlas Clusters and cloud object storage through a unified service. The team also builds Atlas Online Archive which allows customers to move data from Atlas Clusters into fully managed cloud object storage while preserving a seamless query experience across hot and cold datasets. We are forming a new Atlas Data Federation & Archiving team in the Dublin area. The Engineering Manager who fills this position will be pivotal in growing that team. We are looking to speak to candidates who are based in Dublin and would like a hybrid or in-office working model. What you’ll do Lead a team of motivated individual contributors who are eager to learn and grow Contribute to the code, design, and architecture of the systems your team develops Work with stakeholders throughout MongoDB to build our roadmap and product offerings Work with customers and support engineers to fix issues and become part of our on-call rotation Collaborate with team members to develop our codebase, best practices, and design principles Foster an inclusive and respectful work environment according to MongoDB's Core Values We’re looking for someone who Has at least 6 years of profes

javamongodbaws
View job →
E
Everpure
📍 Bengaluru• Full-time
18 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. Portworx by Everpure(Formerly Pure Storage) Everpure Acquired Portworx in October 2020, Creating the Industry's Most Complete Kubernetes Data Services Platform for Cloud Native Applications. This acquisition represents Pure’s largest to date and our deeper expansion into the fast-growing market for multi-cloud data services to support Kubernetes and containers. WHAT YOU'LL DO Lead the design, development, and deployment of Portworx's Kubernetes-native backup and restore solution, serving as a technical anchor for the team. Design and build cloud-native services that operate reliably at scale, with a strong emphasis on performance, quality, and efficiency in both design and implementation. Deeply leverage Kubernetes constructs (operators, CRDs, controllers, CSI) to build robust, production-grade data protection capabilities. Ensure security, resilience, stability, and high availability are first-class considerations in every design and implementation decision. Design and develop robust, well-defined APIs that integrate seamlessly with the frontend and Web UI. Drive architectural decisions and technical direction for the product, mentor engineers, and set the bar for engineering quality through design reviews and code reviews. Own end-to-end delivery of complex features — from design documents through implementation, testing at scale, and production readiness. We are primarily an in-office environment and therefore, you will

pythonsqlmysql
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
13 days ago

About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the Role As a Deployment Lead Life Sciences, you will define how OpenAI delivers complex systems to customers. You will own how they are built, shipped, and adopted. You’ll translate business outcomes into a technical plan, run day-to-day execution across FDEs, Researchers, and Customer Engineers, and partner with customer teams to ensure delivery supports their goals. You will focus on the Life Sciences vertical, partnering with pharmaceutical companies, clinical research organizations, and other data and services providers to deploy next-generation AI capabilities across their drug discovery, development, and operations. You will own delivery end-to-end: embedding with Life Sciences customers to map workflows and success criteria, ensuring components ship on time, and leading readiness and change management for adoption. You’ll track progress, manage dependencies, make sequencing decisions, and drive 0→1 prototypes through MVP and scale. You will also share field insights with Product and Research to guide roadmap and priorities. Success will be measured first and foremost by impact - deployments that deliver measurable value against customer goals, drive adoption, and become critical to their workflows. Additional measures of success include delivery reliability (milestones hit, low reopen/churn), operating leverage (patterns reused across deployments), judgment under pressure, and product impact (field signal that shifts roadmaps/architectures). This is a high-trust, high-autonomy role. Success requires deep technical project management expertise, extreme ownership of outcomes, and an ability to immerse in customer workflows and partner with customer teams to solve complex engineering problems at pace. This role is based in New York. We us

REMOTEartificial intelligenceaiExcel
View job →
🔔

Get new lead data engineer jobs by email

Daily job updates · Unsubscribe anytime