Jobiba hiring network

Lead Software Engineer Data Infrastructure Jobs

6,753 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead software engineer data infrastructure jobs. Use filters to narrow by work mode, employment type, experience and date posted.

SA
Scale AI
📍 San Francisco• Full-time• From $302.4K/yr
19 days ago

Director of Engineering, Physical AI Role Overview The Director of Engineering will report to the General Manager of Physical AI, and will be responsible for leading a multi-disciplinary engineering organization. In this senior leadership role, you will own the execution of the Physical AI Data Engine — the platform powering the next generation of Physical AI/Embodied AI. You will collaborate closely with Operations and GTM to guide product direction and help solve the data bottleneck that stands between today's robotics research and real-world deployment. This role requires significant ownership in a fast-paced environment and you will motivate internal teams to set the pace for business growth. Travel will come into play. Key Responsibilities: Set and drive the technical vision across data collection infrastructure, teleoperation systems, ML training pipelines, model evaluation frameworks, annotation tooling, and research Lead a multidisciplinary engineering organization—spanning engineering managers, software engineers, ML engineers, and ML research scientists—while designing the organizational structure, talent strategy, and culture required to scale rapidly without compromising on quality or strategic alignment Maintain exceptional technical and operational excellence by deeply understanding team deliverables, asking incisive questions, identifying slipping standards early, and knowing precisely when to step in Drive cross-functional alignment across Engineering, Operations, and GTM on platform architecture, release processes, and shared priorities Collaborate with researchers and clients to architect and deliver scalable, production-grade data infrastructure tailored for complex robotics workloads Required Qualifications: Bachelor's degree in Engineering, Robotics, Computer Science, or a related technical field 8+ years of engineering experience in fast-paced environments, including 4+ years direct people management demonstrated history of recruiting, mentorin

typescriptpythonaws
View job →
P
Pinterest
📍 US; Remote, United States• Full-time• Remote• From $189.3K/yr
1mo ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . The Conversion Visibility Modeling team enables a performant ads marketplace and helps prove value to advertisers by connecting Pinterest onsite activity with conversions that happen offsite (both digital and physical) in a privacy-preserving way. As a Machine Learning Engineering Manager on this team, you will lead a hybrid team of ML engineers and backend software engineers to build end-to-end identity and conversion visibility solutions across modeling, serving, and data infrastructure, so advertisers retain accurate, privacy-aware performance visibility as signals fragment and degrade. You will set the technical direction for high-impact ML systems that feed ranking, bidding, measurement, and reporting across Pinterest’s ads stack. What you’ll do: Attract, hire, develop, and lead a hybrid team of ML engineers and backend softwar

REMOTEawsgitrest
View job →
R
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Engineering Manager, Home Infrastructure The Home Infrastructure team builds the mission-critical backend and data systems that power Roblox’s Homepage and Experience Details Page, two of the highest-traffic surfaces on Roblox. These surfaces reach the vast majority of Roblox’s daily active users and are core drivers of discovery, engagement, retention, and platform growth. We are a full-stack product infrastructure team responsible for content distribution across Roblox. Our systems support multiple modes of user interaction, including exploratory browsing, directed discovery, and personalized content recommendations across the many types of content that make up the Roblox ecosystem. This team sits at the intersection of large-scale distributed systems, machine learning-powered personalization, data infrastructure, and product experimentation. We partner closely with Machine Learning, Data Science, Product, Design, Frontend, Ads, Marketplace, Virtual Economy, and other teams across Roblox to build the platforms that help users find the most relevant and engaging content. As Engineering Manager for Home Infrastructure, you will lead a team of Backend and Data Engineers responsible for the e

awsgitmachine learning
View job →
R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Engineering Manager, Home Infrastructure The Home Infrastructure team builds the mission-critical backend and data systems that power Roblox’s Homepage and Experience Details Page, two of the highest-traffic surfaces on Roblox. These surfaces reach the vast majority of Roblox’s daily active users and are core drivers of discovery, engagement, retention, and platform growth. We are a full-stack product infrastructure team responsible for content distribution across Roblox. Our systems support multiple modes of user interaction, including exploratory browsing, directed discovery, and personalized content recommendations across the many types of content that make up the Roblox ecosystem. This team sits at the intersection of large-scale distributed systems, machine learning-powered personalization, data infrastructure, and product experimentation. We partner closely with Machine Learning, Data Science, Product, Design, Frontend, Ads, Marketplace, Virtual Economy, and other teams across Roblox to build the platforms that help users find the most relevant and engaging content. As Engineering Manager for Home Infrastructure, you will lead a team of Backend and Data Engineers responsible for the e

awsgitmachine learning
View job →
SA
Scale AI
📍 San Francisco• Full-time• From $134.4K/yr
19 days ago

At Scale, we believe that the next frontier of artificial intelligence is embodied. The Physical AI team is focused on building general AI that can reason and act in the physical world. By leveraging Scale’s massive, industry-leading data infrastructure, we are partnering with frontier labs to build Foundation Models for Physical AI that will redefine the future of automation. To support our rapid hardware-software iteration cycles and ensure a world-class R&D environment, we are looking for a Safety Coordinator / Lab Lead to anchor our physical testing operations. Role Overview As the Safety Coordinator / Lab Lead , you will play a mission-critical role in scaling our physical testing infrastructure safely and efficiently. This is a high-impact position where your highest-priority responsibility will be owning the end-to-end execution of safety audits and incident documentation . Operating at the intersection of cutting-edge AI foundation models and complex robotics hardware, you will ensure our researchers, engineers, and autonomous systems interact in a secure, compliant, and highly organized environment. Core Responsibilities Priority Focus: Safety Audits & Incident Documentation Rigorous Safety Audits: Design, schedule, and execute routine safety audits across all physical testing environments, robot cells, and hardware workspaces to ensure continuous compliance with internal benchmarks and industrial safety standards. Incident & Near-Miss Documentation: Own the end-to-end incident management pipeline. Act as the primary point of contact for documenting, archiving, and analyzing any lab incidents, mechanical anomalies, or near-misses. Root-Cause Analysis (RCA): Lead structured post-incident investigations to identify systematic risks, authoring comprehensive RCA reports and implementing Corrective and Preventive Actions (CAPA). Data-Driven Risk Mitigation: Treat safety data as a core operational asset—tracking safety metrics and audit trends to proa

awsrestai
View job →
RS
15 days ago

OUR MISSION At Redwood, we empower our customers with lights-out automation for their mission-critical business processes. ABOUT US Redwood Software is the leader in full stack automation fabric solutions for mission-critical business processes. With the first SaaS-based composable automation platform specifically built for ERP, we believe in the transformative power of automation. Our unparalleled solutions empower you to orchestrate, manage and monitor your workflows across any application, service or server — in the cloud or on premises — with confidence and control. CORE VALUES One Team. One Redwood Make Your Own Weather Obsess over Customer Success Work the Problem Be Curious Own the Outcome Respect Each Other YOUR IMPACT We are looking for an Engineering Manager, Products & Platforms to join our Product engineering team to lead a high-performing software engineering team while driving technical strategy across Redwood’s Workload Automation Platform. You will be instrumental in mentoring engineers, managing team deliverables, and guiding the design, development, and enhancement of scalable, secure software that powers enterprise data exchange for more than 1,000 customers worldwide. As an Engineering Manager, you will: People Leadership & Mentorship: Manage, coach, and grow a team of talented software engineers, supporting career development, conducting performance reviews, and fostering an inclusive, collaborative team culture. Technical Strategy & Architecture: Provide hands-on technical guidance, participate in design reviews, and define technical roadmaps for Java/Spring Boot microservices while ensuring high standards for architecture, security, and observability. Delivery & Platform Ownership: Oversee team execution, Sprint planning, and delivery timelines to ensure resilient, scalable core platform features and infrastructure. AI Integration: Research and apply AI/ML concepts and their usage to innovate and enhance the MFT (Managed

javaawskubernetes
View job →
M
Mongodb
📍 United States• Full-time• From $151K/yr
1mo ago

We’re looking for a Senior Engineering Manager who is ready to lead through ambiguity and improve how software gets built at MongoDB. This role leads teams focused on developer productivity, with an emphasis on measurable improvements to the software development lifecycle. This role can be based remotely in the United States. The Team The AXIS team (AI, X-functional tools, Insights, and Signals) sits within Developer Productivity and is responsible for overseeing the metrics and observability infrastructure of our expansive developer environment to help build a strong data-driven culture. You’ll also be a key partner in building the agentic ecosystem for AI-driven development across engineering. Candidate Profile We’re looking for an experienced leader with a passion for solving the big challenge of measuring developer productivity and providing the actionable signals that help teams improve their performance. They should be comfortable working collaboratively with other leaders and partners across our Engineering and Data teams in maximizing the use of data for insights and AI enablement. The right candidate for this role will have 4+ years of experience managing software engineers, including hiring, performance management, growth planning, and compensation; required for external candidates and preferred for internal candidates 8+ years of hands-on software engineering experience building and operating production systems; experience in developer tooling, platform engineering, observability, or data engineering is a strong plus Demonstrated the ability to lead through ambiguity, work across team boundaries, and deliver outcomes without close supervision Strong customer orientation and sound judgment in finding practical, high-leverage solutions Experience working with systems involving analytics, data pipelines, and metrics platforms Experience with AI tools development and enablement efforts Strong technical judgment, including the ability to evaluate t

mongodbawsazure
View job →
M
Mongodb
📍 Alberta• Full-time• From C$191K/yr
1mo ago

We’re looking for a Senior Engineering Manager who is ready to lead through ambiguity and improve how software gets built at MongoDB. This role leads teams focused on developer productivity, with an emphasis on measurable improvements to the software development lifecycle. This role can be based remotely in Canada. The Team The AXIS team (AI, X-functional tools, Insights, and Signals) sits within Developer Productivity and is responsible for overseeing the metrics and observability infrastructure of our expansive developer environment to help build a strong data-driven culture. You’ll also be a key partner in building the agentic ecosystem for AI-driven development across engineering. Candidate Profile We’re looking for an experienced leader with a passion for solving the big challenge of measuring developer productivity and providing the actionable signals that help teams improve their performance. They should be comfortable working collaboratively with other leaders and partners across our Engineering and Data teams in maximizing the use of data for insights and AI enablement. The right candidate for this role will have 4+ years of experience managing software engineers, including hiring, performance management, growth planning, and compensation; required for external candidates and preferred for internal candidates 8+ years of hands-on software engineering experience building and operating production systems; experience in developer tooling, platform engineering, observability, or data engineering is a strong plus Demonstrated the ability to lead through ambiguity, work across team boundaries, and deliver outcomes without close supervision Strong customer orientation and sound judgment in finding practical, high-leverage solutions Experience working with systems involving analytics, data pipelines, and metrics platforms Experience with AI tools development and enablement efforts Strong technical judgment, including the ability to evaluate tradeoffs, i

mongodbawsazure
View job →

The mission of OrgStore is to provide an easy-to-use, fully-managed platform to store and search data. With Postgres as our cornerstone technology, we focus on transactional (OLTP) data use-cases while providing a custom data plane and a flexible administrative layer. Our users are the thousands of engineers representing hundreds of teams producing products at Datadog. As a platform infrastructure team, our goal is to accelerate product teams by helping them rapidly stand up new features and make step change improvements to their application's performance. Data storage is ubiquitous and the service we are providing is critical for Datadog as the business scales and expands its product offerings. As the Engineering Manager for the OrgStore Blueprint Team, you will lead a high-performing team of software engineers, guiding them in building and maintaining these mission-critical systems. Collaboration is key in this role, needing to work closely with other engineering, product, and support teams across Datadog. Your strategic vision will not only guide your team's day-to-day operations but also influence the long-term roadmap of support within Datadog. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Guide and mentor a diverse team of 4-8 software engineers, fostering their career growth while ensuring high team performance. Drive the technical roadmap in collaboration with your team, product managers, and support teams ensuring it aligns with company objectives and directly contributes to improving customer satisfaction. Engage strategically with complex technical problems and work with your team to produce well-defined and actionable plans. Be a core contributor and reviewer of code, lead design decisions, and participate i

aigorust
View job →
P
Point72
📍 Bengaluru• Full-time
19 days ago

AI/ML – Investment Services A Career with Point72's AI/ML – Investment Services Team The AI/ML – Investment Services team at Point72 spearheads the development of cutting-edge AI solutions that seek to transform our business processes and enhance enterprise intelligence. The team aims to bridge the gap between business challenges and technological innovation, collaborating with stakeholders across the firm and leveraging expertise in generative AI, data engineering, and machine learning. WHAT YOU'LL DO Build and scale core backend services and platforms that power generative AI applications and data infrastructure used across the firm’s investment workflows Design and implement high-throughput, low-latency data pipelines to ingest, normalize, and serve both structured and unstructured data Develop robust APIs and microservices to support model inference, feature serving, and downstream applications Integrate generative AI tools and model-serving workflows into production, including embedding stores, retrieval components, and fine-tuning pipelines Optimize system performance, cost, and reliability through profiling, capacity planning, and architectural improvements Implement automated testing, continuous delivery pipelines, monitoring, and incident response practices to maintain production health Partner with data scientists, AI engineers, product owners, and operations to translate models and prototypes into scalable, production-grade solutions Mentor engineers, lead code reviews, and establish engineering best practices for maintainability, security, and observability Own end-to-end delivery, operational runbooks, and metrics-driven measurement of feature impact and system reliability WHAT'S REQUIRED Bachelor’s degree in computer science, software engineering, or a related technical field Minimum 5+ years of professional experience building backend systems and production services Demonstrated experience designing and operating large-scale data engineering pipelines

pythonjavakubernetes
View job →
S
Snowflake
📍 Bellevue• Full-time• Remote
19 days ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Engineering Manager, Cloud Efficiency Snowflake runs large scale cloud infrastructure to deliver its own service — production and internal deployments, Kubernetes fleets, CI/CD, etc. Our cloud spend is in billions of dollars per year. We are looking for an experienced Engineering Manager to lead the Cloud Efficiency engineering team. In this role, you will own the technical vision and execution for building a unified, self-serve cloud efficiency platform along with AI skills and agents that makes resource usage and spend attributable and governable while driving insights and optimization of our cloud spend. AS AN ENGINEERING MANAGER IN CLOUD EFFICIENCY, YOU WILL: Lead and grow our talented team of software engineers, fostering a culture of technical excellence, ownership, and continuous learning. Drive the roadmap for Cloud Efficiency — translating company-level spend objectives into engineering systems: authoritative cost data, resource ownership registry, attribution pipelines, cost and unit economics modeling, observability, governance policies, and optimization workflows — in partnership with Product, Engineering, Finance, and Data Science. Set technical strategy for backend systems, data pipelines, and APIs that measure, attribute and surface cost and usage insights at

REMOTEawsazuregcp
View job →

Observability Pipelines (OP) is Datadog's on-premise, vendor-agnostic telemetry pipeline product. As an Engineering Manager on the team, you'll own people management and engineering execution for one of OP's core missions, spanning areas like Integrations (ingesting from and routing to the many source and destination systems customers rely on), streaming insights, cost control, or pipeline capabilities, reliability and scalability. You'll partner directly with Product to help shape the roadmap, and work closely with your peer EMs and senior ICs to define how OP operates and grows. This is an opportunity to build your management craft while having real influence over the technical direction of a fast-growing product area. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Own people management and engineering execution Establish a strong operating rhythm for the team Drive high standards for on-call rotations and incident response Partner with Product on the roadmap, balancing product priorities with technical realities Lead, coach, and grow the careers of engineers on your team Who You Are: Experienced managing engineers directly, comfortable owning a team’s operating rhythm end-to-end, from planning through execution and stakeholder communication to incident and on-call ownership Have a technical background in distributed systems and data infrastructure Have experience with on-premises or customer-installed software concepts A product-minded partner to have on the team — you enjoy working with Product on strategy Experience with high-performance or Rust-based data pipeline systems Datadog values people from all walks of life. We understand not everyone will meet all the above qualifications o

aigorust
View job →

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team... Here at GoDaddy, the ML Engineering (MLE) team exists as the backbone of our machine learning infrastructure, enabling ML scientists and product teams across Domains to ship models to production reliably, efficiently, and at scale. This team owns the full lifecycle of ML systems — from CI/CD pipelines and model serving infrastructure to GPU workload orchestration and observability. Through disciplined engineering practices, thoughtful system design, and close collaboration with ML scientists, data engineers, and product teams, we deliver the platform that powers domain search, pricing, recommendations, and emerging AI experiences for millions of customers worldwide. We are currently looking for an experienced, highly motivated Senior Engineering Manager to lead our ML Engineering team based in India. This is an established team with existing engineers — we expect the candidate to ramp up quickly on our ML infrastructure stack, build strong relationships with the team, and partner with both India-based teams and US-based teams to drive execution and grow the team further. This individual will join us on our journey to build and scale ML infrastructure that serves real-time predictions at low latency, automates model deployment and promotion, and provides the observability and reliability guarantees that production ML systems demand. Become part of a team that bridges the gap between ML research and production engineering — shipping systems that directly impact GoDaddy's core revenue. What you'll get to do... Lead a team o

typescriptpythonaws
View job →
R
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role We're hiring a hands-on Engineering Manager to build and lead Replit's Anti-Abuse team from the ground up. This is a foundational 0-to-1 role: you'll define the anti-abuse roadmap, hire a small team of engineers and data analysts, and ship the systems that protect Replit's platform, users, and economics from adversarial actors. You'll partner across Support, Legal, Security, Infrastructure, and the Money and Growth teams to make abuse economically unviable while keeping friction low for legitimate users. Replit sits at the frontier of AI-native abuse. Our platform is a target for phishing and scam hosting, cryptomining, LLM token farming, card and coupon fraud, and increasingly, abuse driven by AI agents themselves. The team you build will define how Replit defends against all of it. What You'll Do Build the anti-abuse roadmap from scratch : Define the threat model, prioritize across abuse vectors (phishing/scam hosting, cryptomining, token farming, payment fraud, AI agent exploitation), and translate it into a shipping plan with clear sequencing and tradeoffs. Design progressive verification and identity infrastructure : Build the "ladder of trust" that gates increasing platform capabilities (referrals, additional credits, access to powerful agent features, Missions) behind escalating verification. This includes a humanity/identity layer that's distinct from user accounts, integrations with KYC-grade verification providers, and the policy engine that decides what level of trust unlocks what behavior. This infrastructure is core not just to promo integrity but to how Replit safely expands agent capabilities over time. Ship as a hands-on EM : Stay in the code. Use the latest AI coding tools (including Rep

E
19 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As the Engineering Manager for the FlashArray (FA) Foundation Quality Engineering team, you will build and lead a high-performing team of Quality Engineers and Software Test Developers in Bangalore. You will combine people leadership, technical judgment, and execution discipline to drive the quality, test automation, and release readiness of foundational FlashArray capabilities. Working at the intersection of software, hardware, firmware, and cloud infrastructure, you will partner closely with Development, Product Management, Release, and Support teams. Your focus will be translating complex roadmap requirements into modern test strategies, establishing strong shift-left CI/CD signals, and ensuring our products consistently deliver industry-leading reliability to our customers. WHAT YOU’LL DO Build & Lead a High-Performing Team: Hire, mentor, and coach a newly forming team of quality and automation engineers in Bangalore, fostering a culture of technical excellence, continuous learning, and shared ownership. Drive Quality Strategy & Shift-Left Automation: Define and execute the FA Foundation quality roadmap—including risk-based coverage, automated CI/CD pipeline integration, fault injection, and release-readiness criteria across software, firmware, and hardware. Partner Cross-Functionally for Delivery: Collaborate early in the design cycle with Development, Architecture, Product, and Release teams to impro

pythonawsci/cd
View job →
🔔

Get new lead software engineer data infrastructure jobs by email

Daily job updates · Unsubscribe anytime