The Infrastructure Engineering team is responsible for building and maintaining a self-service internal development platform that enables MongoDB engineering teams to reliably deploy and operate their own production services and products. We work with numerous engineering teams across the company to understand their infrastructure requirements and development workflows, develop broadly applicable self-service platform services and tooling, continuously monitor how platform services are being utilized, and look for ways to improve developer productivity through automation and education. We are big open source enthusiasts and use a number of open source tools in our stack (contributing upstream whenever possible). Some of the tools we use regularly include Go, AWS, Kubernetes, Crossplane, Terraform, Helm, Drone, Prometheus, and Grafana. However, technology is nothing without a stellar team of engineers that are focused on doing high quality work and working as a team to solve complex distributed computing and platform engineering problems. This is where you come in! We are looking to speak to candidates who are based in Gurugram for our hybrid working model. Our ideal candidate 2+ years of experience managing and mentoring a team of 3+ engineers Has 5+ years of experience owning the design and implementation of large software/infrastructure projects Has built and operated large-scale distributed systems in cloud providers (AWS strongly preferred) Has a strong backend programming background. Fluency in Go is strongly preferred; deep experience with another compiled or strongly-typed backend language is acceptable Pragmatic, detail-oriented, self-motivated, and understands the benefits of collaboration Strong experience operating production Kubernetes clusters, not just deployed to it Has practical experience defining and operating against SLI/SLOs for services they owned Strong experience with observability tooling: metrics, logging, traces, Prometheus, Grafana, OpenTe
Jobiba hiring network
Lead Software Engineer Data Infrastructure Jobs
6,753 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current lead software engineer data infrastructure jobs. Use filters to narrow by work mode, employment type, experience and date posted.
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Role We are seeking a seasoned Manager, Software Engineering with 12+ years of experience to lead our Database Engineering and Cloud Infrastructure team. In this role, you will lead a team of high-performing engineers responsible for architecting, scaling, and optimizing multi-cloud relational and in-memory database platforms. You will bridge technical execution, engineering leadership, and strategic infrastructure planning across AWS and Azure environments. Key Responsibilities Technical Leadership & Architecture Lead the architectural design and operations of enterprise-grade, multi-cloud relational databases across AWS (RDS PostgreSQL, MySQL, Aurora) and Azure (Database for PostgreSQL/MySQL, Azure SQL Managed Instance). Drive high-availability architecture strategies, including Multi-AZ deployments, auto-failover groups, read replica scaling, and cross-region disaster recovery (DR). Oversee zero-downtime operations, including major-version engine upgrades, schema migrations, and blue/green deployment strategies. In-Memory Infrastructure & Open-Source Strategy Manage scale operations for in-memory datastores (AWS ElastiCache, Azure Cache for Redis), focusing on cluster mode operations, eviction policies, and persistence tuning. Spearhead open-source caching initiatives and migration pathways from Redis to Valkey (e.g., AWS ElastiCache for Valkey) using zero-downtime tools like RedisShake to ensure open-source license compliance and optimize cloud spend. Aut
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers. We’re looking for an EM to lead a team of exceptional ML infrastructure engineers, build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You Will: Lead strategic planning and roadmap execution of scalable production-ready ML systems including model training, data pipelines, feature engineering and model inference. Own the architecture, establish engineering best practices of scalability, reliability, and cost-effectiveness of ML infrastructure (e.g., training, serving, feature). Work closely with data scientists, ML engineers, platform teams, and product stakeholders to design, implement, and operate robust ML platforms that accelerate model development and deployment. Recruit, mentor, and grow a high-performing team of ML infrastructure engineers. You Have: 5+ years of experienc
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity Are you ready to step into a pivotal leadership role where your engineering depth directly shapes the future of our core platform? As our new Engineering Manager, you will lead a talented, distributed team across US and EU time zones, acting as the critical manager bridging regional collaboration. Our Cloud Foundation team is the backbone of the New Relic platform. In this role, you won't just manage tasks; you will mentor and empower engineers, transitioning our operational framework from a reactive state to a culture of proactive ownership and engineering excellence. You will oversee critical global initiatives, including major regional expansions into FedRAMP High / IL4, India, and Australia. If you thrive on solving complex multi-cloud challenges at an exabyte scale while helping engineers grow in their careers, this is your opportunity to make a lasting impact. What you'll do Empower & Mentor: Lead and nurture a high-performing engineering team across the US and EU, facilitating career development, performance growth, and a collaborative team culture. Drive Strategic Ownership: Champion a shift from reactive delivery to proactive technical ownership, establishing best practices for platform reliability and cross-regional alignment. Lead Regional Expansions: Architect and execute key global infrastructure expansions across complex environments (including FedRAMP High / IL4, India, and Australia). Architect for Extreme Scale: Guide decisions around micr
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's App Primitives team owns the shared building blocks that span data structures to UI: comments, notifications, event logs, feature flags, and task management. We own the general infrastructure; product teams own the logic built on top. Getting this layer right unlocks velocity for every team at Vanta. As the Engineering Manager, App Primitives at Vanta, you'll lead the team building the shared product primitives that every Vanta product team depends on, owning the infrastructure layer that makes collaboration, communication, and core workflows possible across the entire platform. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow the App Primitives team, owning hiring, team health, delivery, and the development of senior engineers and technical leads Own the strategy and roadmap for Vanta's shared product primitives: comments infrastructure, notifications platform, event log, feature flag system (Statsig), and task management Define and steward the interface model between App Primitives systems and product teams, ensuring product teams can build on top of shared infrastructure quickly and safely, without owning the underlying systems themselves Partner closely with product engineering leaders across Vanta to surface developer ne
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's App Primitives team owns the shared building blocks that span data structures to UI: comments, notifications, event logs, feature flags, and task management. We own the general infrastructure; product teams own the logic built on top. Getting this layer right unlocks velocity for every team at Vanta. As the Engineering Manager, App Primitives at Vanta, you'll lead the team building the shared product primitives that every Vanta product team depends on, owning the infrastructure layer that makes collaboration, communication, and core workflows possible across the entire platform. Our Engineering Managers develop and grow high-performing teams that deliver significant value to our customers and enable our business to scale. This role sits at the intersection of technical architecture and team development, with real authority to set direction and grow a world-class platform team. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as an Engineering Manager at Vanta: Lead and grow the App Primitives team, owning hiring, team health, delivery, and the development of senior engineers and technical leads Own the strategy and roadmap for Vanta's shared product primitives: comments infrastructure, notifications platform, event log, feature flag system (Statsig), and task management Define and steward the interface model between App Primitives systems and product teams, ensuring product teams can build on top of shared infrastructure quickly and safely, without owning the underlying systems themselves Partner closely with product engineering leaders across Vanta to surface developer ne
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. SHOULD YOU ACCEPT THIS CHALLENGE... In this role as an Engineering Manager, you will lead a team of engineers located in Bengaluru, India. You will focus on driving and shaping the direction of our observability software and enabling product engineers to deliver high-quality, reliable software to our customers. You will provide technical leadership and direction, mentor engineers on your team, and collaborate with product, engineering, and cross-functional stakeholders to deliver successful outcomes. The successful candidate must understand the dynamics of global R&D, possess deep knowledge of local culture, and have the ability to champion Pure values and leadership attributes. This role requires the ability to lead and influence multiple stakeholders across cross-functional teams and drive alignment across complex, distributed engineering environments. The team will help build and evolve observability capabilities that provide actionable insights into the health, performance, capacity, and reliability of Pure's products and infrastructure. WHAT YOU'LL NEED TO BRING TO THIS ROLE... 12+ years of combined experience as a software developer and manager 3+ years of technical management experience while staying hands-on 7+ years of hands-on software development experience Strong exposure to one or more of the following areas: distributed systems, systems programming, observability/telemetry, data platforms, or solving prob
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a Delivery Director, Capacity programs for our on-premises data center builds and neo cloud (GPU cloud) delivery programs. This is a high-visibility, execution-critical role sitting at the intersection of infrastructure engineering, capacity planning, vendor/partner management, and customer delivery. You will own the end-to-end delivery lifecycle for large-scale compute infrastructure — from initial site/capacity commitments through power, networking, and hardware bring-up, to production-ready GPU/compute capacity landing in the hands of internal teams or customers. You'll be the person who turns ambitious infrastructure roadmaps into predictable, on-time, delivery. RESPONSIBILITIES Own delivery of on-prem infrastructure builds — colocation expansions, power/cooling readiness, rack-and-stack, network fabric bring-up, and hardware acceptance testing — coordinating across colo providers and partners, network engineering, hardware ops, and vendor teams. Drive neo cloud delivery programs — manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps, SLAs, and go-live readiness. Build and maintain master delivery schedules across concurrent, multi-site, multi-vendor programs, integrating power/shell timelines, hardware lead times, logistics, and software/platform readiness into a single critical path.
About Vercel: Vercel is the agentic infrastructure company, freeing people and agents to ship what's next. For more than a decade we've helped builders move from idea to production with speed, security, and exceptional developer experience. Now we're scaling our products for both agents and people to ship and run software, built in the open and trusted by OpenAI, PayPal, Ramp, Supreme, and millions of developers worldwide. About Vercel Vercel's Frontend Cloud provides the developer experience and infrastructure to build, scale, and secure faster, more personalized web applications. Customers like Under Armour, Nintendo, The Washington Post, Porsche, and Zapier rely on Vercel to create dynamic, seamless user experiences on the web. Our mission is to empower the world to deliver the best digital products. That mission extends to building a supportive, innovative workplace where you can do the best work of your career. About the Role As an Engineering Manager on the Dashboard team, you will lead a team of engineers responsible for the primary interface developers use to build, deploy, and manage applications on Vercel. The Dashboard is where customers experience the platform every day, from onboarding new projects to managing deployments, observability, billing, security, AI, and team collaboration. You'll partner closely with Product, Design, Data, and other engineering teams to deliver intuitive, high-quality product experiences while helping shape the long-term architecture of one of Vercel's most important customer-facing surfaces. This role requires balancing technical leadership, product thinking, and strong people management to help the team ship quickly without compromising quality. What You Will Do Build and grow a high-performing team of engineers through hiring, coaching, mentorship, and thoughtful career development. Foster a culture of ownership, collaboration, and high engineering standards while creating an environment where engineers can do their best w
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role We're looking for a Technical Program Manager who can operate at the intersection of engineering, product, and business — someone deeply technical, trusted instinctively by engineers, and sharp enough to drive clarity and momentum across complex, cross-functional programs. This is a high-agency role with real executive visibility and direct impact on how Ramp's engineering organization scales. We're looking for someone who is energized by complexity, deeply curious about what AI can unlock for engineering teams, and eager to apply it hands-on in their work. You should be someone who experiments with AI tools regularly, thinks about how they change the way software gets built, and brings that perspective into how you run programs. What You’ll Do Lead large-scale technical programs across engineering and adjacent teams—from CI/CD and infrastructure scaling to incident response, and driving other strategic projects across the engineering organization Own Ramp's engineering incident response program, improving processes, running retrospec
About the Team OpenAI’s Governance, Risk, and Compliance team helps ensure security and privacy are grounded in how our products and systems actually operate. Assurance Operations partners with Security, Engineering, Infrastructure, Product, Privacy, and Legal to make controls provable, risk decisions explicit, and audit readiness a result of well-designed systems. About the Role We are hiring a technical, product-minded GRC builder who can own consequential audits while improving the control and evidence systems behind them. You will build a reusable common control framework, use Codex to automate assurance work, validate changing system scope, and turn repeated audit friction into measurable improvements. We are looking for someone who questions inherited assumptions, solves novel problems creatively, works closely with engineers, and makes the next audit easier by improving the underlying system. You’ll be responsible for: Lead external, internal, customer, and certification audit work from scoping through evidence review, fieldwork, remediation, and closeout. Build a common control framework linking risk, control intent, implementation, owner, system, environment, evidence, and applicable frameworks. Validate actual scope and ownership instead of assuming last year's controls, product boundaries, or evidence remain accurate. Use Codex to build and test evidence checks, control mappings, request triage, owner workflows, monitoring, and remediation reporting. Partner with engineers on cloud architecture, identity, logging, data flows, software changes, vulnerabilities, and control effectiveness. Design maintainable, permission-aware tools that preserve source provenance, human review, and evidence integrity. Reduce repeated requests and operational burden for control owners through measurable workflow improvements. Define roadmaps, decision rights, milestones, success metrics, and clear cross-functional escalations. We’re looking for someone with: Direct ownership
About the Team Security is at the foundation of OpenAI's mission to ensure that artificial general intelligence benefits all of humanity. The Identity Infrastructure Engineering team sits at the core of this effort, designing and building the identity and access management solutions that protect model weights, customer data, and critical systems across multiple cloud environments. The team partners across OpenAI, including Applied Engineering, Research, IT, Security, Infrastructure, and Engineering, to provide secure and scalable platforms for identity, access management, permissioning, orchestration, and safe AI research. About the Role We’re looking for an engineering leader to lead Identity Infrastructure Engineering, the team building the systems that govern and scale access across OpenAI’s research, engineering, and internal platforms. This role sits at the center of cloud infrastructure, identity, software engineering, and security-critical operations. You’ll lead engineers building control planes, policy systems, workload and agent authorization patterns, infrastructure-as-code, and operational foundations that help OpenAI move quickly while keeping access reliable, auditable, least-privileged, and safe under failure. The ideal candidate has led teams responsible for large-scale, mission-critical infrastructure. They can go deep into code and architecture when needed, while giving engineers and technical leads the clarity and ownership to do their best work. They set technical direction, grow strong teams, make durable architecture decisions, and turn ambiguous 0-to-1 problems into platforms OpenAI can trust and build on for years. In this role, you will: Build and lead a high-performing Identity Infrastructure team, going deep enough technically to set direction while empowering the team to own delivery. Define the strategy for identity platform as the policy plane for access across people, agents, workloads, services, clouds, and internal systems. Scale Acc
Job Details: Job Description: This is a high-visibility, commissioned sales leadership role within Intel's US Sales organization, specifically focused on our most disruptive AI-Native and Strategic CSP accounts. You will be the primary architect of Intel's relationship with industry titans who are redefining the boundaries of AI model training, AIaaS solutions and deployment at scale. This is not a traditional sales role. You will operate at the intersection of deep technical engineering and executive business strategy, ensuring Intel's silicon and software roadmap aligns with the world's most demanding AI-as-a-Service and SaaS platforms across on-prem, Tier1 CSP and NeoCloud environments. Key Responsibilities Executive Orchestration: Act as the One Intel lead, building deep-rooted partnerships with C-suite executives and Principal Engineers at world-class AI and SaaS firms. Technical Value Synthesis: Translate complex hardware architectures (CPU, GPU, Accelerator, Networking, and Packaging) into business outcomes for customers running massive-scale distributed training and inference workloads. Strategic Growth: Drive Intel's data-centric growth strategy by identifying and securing design wins within the core infrastructure of the world's leading AI models and solution providers. Cross-Functional Leadership: Partner closely with Cloud Solution Architects (CSAs), Intel Business Units, Cloud and OEM partners to influence future product roadmaps based on the unique needs of AI-native disruptors. Market Evangelism: Serve as a technical and business evangelist, articulating Intel's vision for the future of AI and compute infrastructure in a highly competitive landscape. <p style="text-align:inhe
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We’re hiring a talented Software Engineering Manager to lead the Snowtrail infrastructure team at Snowflake. Snowtrail is the infrastructure that enables Snowflake to deliver dedicated coverage for customer-specific workloads. Its innovative approach allows Snowflake to precisely test and measure the impact of changes on individual customers, making it essential for ensuring the platform’s reliability, correctness, and performance. Through query replay, Snowtrail helps us catch regressions early. By leveraging machine learning models to intelligently sample queries and workloads, we continuously optimize for both cost and performance. Evolving Snowtrail to incorporate new engine features while improving scalability, efficiency, and reliability is central to our continued success OUR IDEAL MANAGER WILL HAVE : Strong passion and proven track record for shipping quality software in high code velocity environments 10+ years industry experience designing and building distributed data systems. Excellent problem solving skills, and strong CS fundamentals including data structures, algorithms, and distributed systems. Fluency in SQL, Java, C++, Python or Go. Ability to collaborate well across teams, build high-performing teams and mentor junior engineers. Excellent interpersonal co
This is Adyen Adyen provides payments, data, and financial products in a single solution for customers like Meta, Uber, H&M, and Microsoft - making us the financial technology platform of choice. At Adyen, everything we do is engineered for ambition. For our teams, we create an environment with opportunities for our people to succeed, backed by the culture and support to ensure they are enabled to truly own their careers. We are motivated individuals who tackle unique technical challenges at scale and solve them as a team. Together, we deliver innovative and ethical solutions that help businesses achieve their ambitions faster. Team Lead - Software Engineer As a Software Engineering Team Lead based in Singapore, you will lead a team of highly skilled software engineers responsible for building and evolving Adyen's Global Cards platform for the APAC region. Your team plays a critical role in developing scalable, reliable, and resilient payment capabilities that power card transactions for merchants across APAC. You'll work closely with Product Managers, Architects, and Engineering teams to deliver new functionality across the Cards domain, while continuously improving the performance, scalability, and reliability of our platform. As a people leader, you'll coach and develop engineers, foster a culture of ownership, and help shape the technical direction of one of Adyen's core payment domains. As part of Adyen's global engineering organisation, you'll collaborate closely with teams across Europe, APAC, and North America to build products that scale globally. What you'll do Lead, coach, and develop a team of Software Engineers, supporting both their technical and professional growth. Foster a high-performing engineering culture built on ownership, collaboration, and continuous learning. Partner closely with Product Managers to translate business priorities into scalable technical solutions. Drive the technical direction of the team, balancing product de
Get new lead software engineer data infrastructure jobs by email
Daily job updates · Unsubscribe anytime