Jobiba hiring network

Lead Infrastructure Software Engineer Jobs

6,876 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead infrastructure software engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

D
Datadog
📍 Lisbon• Full-time
1mo ago

As a Staff Engineer on Datadog's Compute – Disruption and Workload Placement team, you'll help define how our Kubernetes fleet scales to meet the demands of rapidly growing AI and cloud-native workloads. You'll work on the systems that ensure engineering teams have the right compute capacity, in the right region, at the right time across AWS, Google Cloud, and Azure. This is a highly technical, high-impact role where you'll shape the future of capacity orchestration, influence platform architecture, and solve infrastructure challenges that directly support Datadog's continued growth. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Lead the technical direction of capacity management and workload placement for Datadog's Kubernetes platform spanning 100,000+ virtual machines across multiple cloud providers. Design and build systems that optimize how engineering workloads are scheduled and deployed across regions while balancing capacity constraints, reliability, and performance. Partner across infrastructure teams to evolve multi-region and multi-cloud capacity orchestration as Datadog continues to scale. Develop production software in Go to improve Kubernetes platform capabilities, automation, and operational efficiency. Use data and capacity signals to influence infrastructure decisions, forecast growth, and improve workload placement strategies. Who You Are: You have significant experience designing and operating large-scale Kubernetes-based infrastructure or platform systems. You are an experienced software engineer with strong programming skills, ideally in Go or a comparable systems programming language. You have hands-on experience with at least one major cloud provider (AWS, Google Cloud, or Azure) and understand distributed cloud infrastructure. Yo

awsazurekubernetes
View job →
G
Godaddy
📍 Bulgaria• Full-time
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Our Global Sustaining Engineering team sits at the intersection of software engineering and infrastructure, ensuring the services our customers depend on are fast, resilient, and always available. As a Senior Site Reliability Engineer, you'll take direct ownership of production services — from initial design through day-to-day operation — while partnering with product, engineering, and security teams to build and maintain business-critical systems. In this role, you will deepen your technical expertise and grow your leadership presence by mentoring the next generation of SREs. You will also gain hands-on experience with intelligent tooling in real-world workflows. What you'll get to do... Design, implement, and operate scalable, highly available production services while diagnosing and resolving complex infrastructure, network, and application issues Build and maintain alerting pipelines, dashboards, and SLO-driven monitoring strategies using Icinga, Prometheus, and Grafana Lead incident response end-to-end — performing root-cause analysis, authoring blameless post-mortems, and driving corrective actions to closure Develop and extend Infrastructure as Code coverage and build internal tooling that eliminates manual, repetitive operational work Mentor SRE I and SRE II engineers through code reviews, debugging sessions, and knowledge-sharing talks Apply LLM-driven log analysis, anomaly detection, and generative AI tools to accelerate incident response and runbook creation — validating all outputs before use Your experien

pythondockerkubernetes
View job →
M
1mo ago

With a strong security engineering background, you’re looking for a role that gives you the freedom to increase MongoDB’s resonance with customers by strengthening our core database products. You’re passionate about solving hard security engineering problems while putting a strong emphasis on customer experience, leveraging your own significant experience. You enjoy collaborating with different teams to innovate and implement pragmatic solutions. Who We Are The MongoDB Product Security organization is a diverse collection of individuals working together to scale MongoDB’s security, both security of the products themselves and the security features we offer to customers. The team is responsible for the MongoDB Database Server ( Community and Enterprise editions). The MongoDB Product Security organization works with software engineers to design, implement, and operate systems in a manner that protects customer data. It is a multidisciplinary team that covers product, software, cloud, infrastructure, and operational security concerns. The team does the following: Build a developer driven security program where there is tight integration with engineering artifacts, process, and tooling. Use software architecture and coding patterns to reduce the impact of security issues. Be security subject matter experts for our tech stack and products. We are looking to speak to candidates who are based in Dublin for our hybrid working model. Responsibilities You will take ownership, define strategy, and drive improvement for parts of our program such as fuzzing, threat modeling, secrets management, or container security Advocate for and lead complex security projects from inception through completion Drive architecture, patterns, and processes across Server Engineering that make security the easiest path Partner closely with engineering teams to design and implement security controls across our software and systems Research and POC new attacks against our systems. Plan and per

mongodbawsazure
View job →
M
1mo ago

With a strong security engineering background, you’re looking for a role that gives you the freedom to increase MongoDB’s resonance with customers by strengthening our core database products. You’re passionate about solving hard security engineering problems while putting a strong emphasis on customer experience, leveraging your own significant experience. You enjoy collaborating with different teams to innovate and implement pragmatic solutions. Who We Are The MongoDB Product Security organization is a diverse collection of individuals working together to scale MongoDB’s security, both security of the products themselves and the security features we offer to customers. The team is responsible for the MongoDB Database Server ( Community and Enterprise editions). The MongoDB Product Security organization works with software engineers to design, implement, and operate systems in a manner that protects customer data. It is a multidisciplinary team that covers product, software, cloud, infrastructure, and operational security concerns. The team does the following: Build a developer driven security program where there is tight integration with engineering artifacts, process, and tooling. Use software architecture and coding patterns to reduce the impact of security issues. Be security subject matter experts for our tech stack and products. We are looking to speak to candidates who are based in Cork for our hybrid working model. Responsibilities You will take ownership, define strategy, and drive improvement for parts of our program such as fuzzing, threat modeling, secrets management, or container security Advocate for and lead complex security projects from inception through completion Drive architecture, patterns, and processes across Server Engineering that make security the easiest path Partner closely with engineering teams to design and implement security controls across our software and systems Research and POC new attacks against our systems. Plan and perfo

mongodbawsazure
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the team The Applied team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the role: We're seeking a Data Engineer to take the lead in building our data pipelines and core tables for OpenAI. These pipelines are crucial for powering analyses, safety systems that guide business decisions, product growth, and prevent bad actors. If you're passionate about working with data and are eager to create solutions with significant impact, we'd love to hear from you. This role also provides the opportunity to collaborate closely with the researchers behind ChatGPT and help them train new models to deliver to users. As we continue our rapid growth, we value data-driven insights, and your contributions will play a pivotal role in our trajectory. Join us in shaping the future of OpenAI! In this role, you will: Design, build and manage our data pipelines, ensuring all user event data is seamlessly integrated into our data warehouse. Develop canonical datasets to track key product metrics including user growth, engagement, and revenue. Work collaboratively with various teams, including, Infrastructure, Data Science, Product, Marketing, Finance, and Research to understand their data needs and provide solutions. Implement robust and fault-tolerant systems for data ingestion and processing. Participate in data architecture and engineering decisions, bringing your strong experience and knowledge to bear. Ensure the security, integrity, and compliance of data according to industry and company standards. You might thrive in this role if you: Have 3+ years of experience as a data engineer and 8+ years of any software engineering experience(including data engineering). Proficiency in at least one programming language commonl

pythonjavaaws
View job →
V
Verse
📍 San Francisco• Full-time• $150K – $210K/yr
15 days ago

Location: San Francisco, CA (Hybrid) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role We're seeking an experienced Senior Optimization Engineer to join our Data Science Team. In this role, you will lead the design, development, and deployment of optimization models that power our software platform across applications including electricity markets, renewable energy, and battery energy storage systems. You will be responsible for developing production-grade optimization engines that solve complex operational and planning problems at scale. This role requires deep expertise in mathematical optimization, strong software engineering skills in Python, and experience building optimization models that integrate with production systems. The ideal candidate has significant experience in the energy industry, particularly electricity markets and battery storage optimization. This position emphasizes technical leadership, ownership of complex optimization projects, and collaboration across engineering, product, and commercial teams to deliver high-impact optimization solutions. Key Res

pythonci/cdai
View job →

Location Details: Canada - BC or ON (remote) At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team GoDaddy - Global Production Engineering looks after GoDaddy's global infrastructure, in the cloud and on-premises. We are hiring an experienced Technical Program Manager, focused on our AWS cloud infrastructure, to plan, lead and deliver complex cross-team initiatives. This is a heavily coordination-focused role: you will own the execution of a portfolio of AWS cloud platform and cost-savings programs, working hands-on with software engineers, engineering managers and SREs to achieve outcomes aligned with the strategy. You will drive dependencies end-to-end, facilitate trade-off decisions, and give collaborators and leadership clear, reliable access to status and risk. You will be an integral part of the Technical Program Management team, partnering closely with engineering leads to ensure GoDaddy delivers on its planned objectives and global strategy. What you'll get to do... Own end-to-end delivery of a portfolio of concurrent cloud platform programs, coordinating across engineering and partner teams to manage scope, schedule and dependencies against the critical path. Drive cost-savings program coordination, including tracking, reporting and surfacing risks to goals and achievements proactively. Run intake and prioritization processes and keep priority pages and status sources current and trustworthy. Serve as the central coordination point across teams, facilitating trade-off and negotiation discussions, driving alignment, and resolving roadblocks with minimal issues. Build reports, scorecards and dashboards to c

awsaiproject management
View job →
S
1mo ago

Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we’re changing the way people think about and interact with personal finance. We’re a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we’re at the forefront. We’re proud to come to work every day knowing that what we do has a direct impact on people’s lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The role We are seeking a Staff Vulnerability Management Engineer to lead the most complex technical work in SoFi’s Vulnerability Management program. You will design and build scalable systems that identify, enrich, prioritize, route, and track vulnerabilities across applications, cloud and infrastructure, containers, software supply chains, and specialized hardware or firmware surfaces. This is a hands-on engineering role with broad technical influence: you will write production code, make architecture decisions, establish vulnerability management standards, and improve how teams understand and reduce vulnerability risk. You will partner with Engineering, Infrastructure, SRE, Compliance, Legal, and business stakeholders to accelerate remediation while protecting engineering velocity and customer trust. You will also serve as a senior technical responder for embargoed disclosures and zero-day events, lead root-cause analysis for high-impact vulnerability incidents, and mentor engineers. The ideal candidate combines deep vulnerability management expertise with strong software engineering judgment, systems thinking, and a bias for durable, measurable outcomes. What you’ll do Lead high-complexity vulnerability management initiatives and make architecture decisions for assigned program areas, from detecti

javascripttypescriptpython
View job →
D
Discord
📍 San Francisco Bay Area• Full-time• $248K – $310K/yr
1mo ago

Discord has a highly engaged community of millions of daily active users who use the platform for many different reasons, but there’s one thing that nearly everyone does: play video games. Discord plays a uniquely important role in the future of gaming, and we are focused on making it easier and more fun for people to hang out before, during, and after playing games. More broadly, Discord is about empowering people to find belonging in all kinds of communities, and those people trust us to keep their communications safe. Our Platform Security Engineering team protects the systems we use to create Discord, making the “secure way” the “easy way.” We’re looking for an Engineering Manager to lead a team of software engineers in articulating and pursuing the most leveraged opportunities to reduce security risk across Engineering. This team will design and build lovable “paved paths” for managing identities and access, shipping code, configuring cloud infrastructure, and operating services. If you’re an Engineering Manager who’s deeply curious, eager to own technically and socially complex projects, and excited to improve security and privacy at Discord, read on! What you'll do You’ll shape company-wide security strategy and lead a highly-autonomous and horizontally-integrated team of software engineers who will... Develop and apply best-in-class secure baselines for cloud infrastructure, owning the security of all cloud environments Manage infrastructure vulnerabilities while supporting a high-velocity engineering org with hundreds of developers Secure first- and third-party software supply chains, from the dev environment through CI/CD and into production Build and operate identity and access management (IAM) systems for humans and machines that are user-friendly and promote least privilege Consult on risk assessments, architectural designs, threat models, code reviews, and more—pragmatically balancing security with other business considerations What we look for 3+ year

typescriptpythonaws
View job →

The mission of OrgStore is to provide an easy-to-use, fully-managed platform to store and search data. With Postgres as our cornerstone technology, we focus on transactional (OLTP) data use-cases while providing a custom data plane and a flexible administrative layer. Our users are the thousands of engineers representing hundreds of teams producing products at Datadog. As a platform infrastructure team, our goal is to accelerate product teams by helping them rapidly stand up new features and make step change improvements to their application's performance. Data storage is ubiquitous and the service we are providing is critical for Datadog as the business scales and expands its product offerings. As the Engineering Manager for the OrgStore Blueprint Team, you will lead a high-performing team of software engineers, guiding them in building and maintaining these mission-critical systems. Collaboration is key in this role, needing to work closely with other engineering, product, and support teams across Datadog. Your strategic vision will not only guide your team's day-to-day operations but also influence the long-term roadmap of support within Datadog. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Guide and mentor a diverse team of 4-8 software engineers, fostering their career growth while ensuring high team performance. Drive the technical roadmap in collaboration with your team, product managers, and support teams ensuring it aligns with company objectives and directly contributes to improving customer satisfaction. Engage strategically with complex technical problems and work with your team to produce well-defined and actionable plans. Be a core contributor and reviewer of code, lead design decisions, and participate i

aigorust
View job →
P
Pinterest
📍 United States• Full-time• Remote• From $132.4K/yr
1mo ago

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Are you passionate about building impactful products for sales and finance teams? Come join the IT Enterprise Systems team at Pinterest where you will be responsible for advancing our sales and marketing systems. What you’ll do: Design, build, and operate full‑stack applications and services on Pinterest’s enterprise infrastructure to support our Sales, Marketing, and Finance teams, from backend services and APIs through integrations and user‑facing workflows. Lead the technical design and implementation of GenAI/ML‑powered services and pipelines that automate and augment enterprise workflows (for example, summarizing sales interactions, enriching account data, or surfacing intelligent recommendations), including clear evaluation frameworks, observability, and validation guardrails. Own the end‑to‑end software development lifecycle for th

REMOTEjavascriptpythonjava
View job →
D
Discord
📍 Remote• Full-time• Remote• $248K – $341K/yr
1mo ago

Discord has a highly engaged community of millions of daily active users who use the platform for many different reasons, but there’s one thing that nearly everyone does: play video games. Discord plays a uniquely important role in the future of gaming, and we are focused on making it easier and more fun for people to hang out before, during, and after playing games. We're looking for a technical, hands-on, and infrastructure-minded Engineering Manager to lead our Notifications team within the Growth organization. Notifications is a full-stack team owning the platform and infrastructure that powers every notification Discord sends, the orchestration layer that optimizes notifications globally across types, the in-app notifications center, and the user-facing settings and client experience. You will lead a team of full-stack engineers operating business-critical systems that fan out to hundreds of millions of users while shipping the re-engagement surfaces that bring them back. This role reports to the Senior Engineering Manager of Growth. What You'll Be Doing Lead a team of full-stack engineers building and operating the notifications platform — Elixir services, Python backends, pubsub pipelines, and the Kubernetes infrastructure underneath them Set technical direction for notifications infrastructure and ensure systems scale reliably as volume and complexity grow Partner with Product, Design, Data Science, and Marketing to define and execute the notifications roadmap Use experimentation and metrics to drive continuous improvement in notification engagement, retention, and long-term user trust Build a high-performing team through hiring, coaching, and instilling engineering best practices Partner with platform consumers across Discord to evolve the primitives other teams depend on to reach our users Partner with engineering leadership on strategic planning and organizational improvement What you should have 5+ years of experience as a Software Engineer, with signifi

REMOTEpythonkubernetesrest
View job →
S
Snowflake
📍 Bellevue• Full-time• Remote
15 days ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Engineering Manager, Cloud Efficiency Snowflake runs large scale cloud infrastructure to deliver its own service — production and internal deployments, Kubernetes fleets, CI/CD, etc. Our cloud spend is in billions of dollars per year. We are looking for an experienced Engineering Manager to lead the Cloud Efficiency engineering team. In this role, you will own the technical vision and execution for building a unified, self-serve cloud efficiency platform along with AI skills and agents that makes resource usage and spend attributable and governable while driving insights and optimization of our cloud spend. AS AN ENGINEERING MANAGER IN CLOUD EFFICIENCY, YOU WILL: Lead and grow our talented team of software engineers, fostering a culture of technical excellence, ownership, and continuous learning. Drive the roadmap for Cloud Efficiency — translating company-level spend objectives into engineering systems: authoritative cost data, resource ownership registry, attribution pipelines, cost and unit economics modeling, observability, governance policies, and optimization workflows — in partnership with Product, Engineering, Finance, and Data Science. Set technical strategy for backend systems, data pipelines, and APIs that measure, attribute and surface cost and usage insights at

REMOTEawsazuregcp
View job →

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. A Career within Software Engineering We're always looking for great talent to join our growing team in EMEA. Although we may not have an open position right now, if you are interested in this opportunity we encourage you to apply. What happen’s next.. We'll carefully review your application and if we believe we may have a role in the near future that could be a good match with your experience and qualifications, we will get in touch and may invite you to an introductory call with a member of our talent acquisition team. ENGINEERING IN PRAGUE We’ve built the world’s first all-flash storage, the introduction of the first evergreen storage model, the first AI-ready infrastructure, became the first fastest-growing private company (prior to IPO), and now we are the first to tackle the modern data experience. Why? Our mission is to constantly innovate to enable our customers to innovate. As an Engineering Manager, you will lead a team of software developers, located in Prague, Czech Republic. You will lead the team in the creation of our product and will need to partner with other cross functional teams, You will also provide technical direction and mentoring to employees on your team and others to achieve successful project outcomes. The successful candidate must understand the dynamics of global R&D while at the same time have the ability to adapt to Pure values and leadership attributes. This role

pythonjavaaws
View job →

Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. This role is open to people who are currently in a product leadership role or someone in an engineering leadership role that is interested to make the switch into product. This role combines experienced product management leadership with the pivotal responsibilities of a Group Product Lead, offering an opportunity to shape the product strategy and the missions of several teams. You will oversee the product management execution working closely with the product managers for several cross-functional software engineering teams around the world. This position reports to the Head of Product for the Platform Engineering Strategic Business Unit, which is responsible for Appian Cloud offering. An excellent candidate will come with a wide range of experience in a product leadership role in the cloud infrastructure domain at a software-as-a-service firm. The product portfolio owned by this position will cover a wide range of cloud native infrastructure and services – from the highly available, resilient, scalable, and efficient infrastructure components on which all Appian Cloud services operate to the the scalable, enterprise-grade managed services that power the backend data persistence and event streams for customer sites in Appian Cloud. You will provide product management leadership for initiatives that deliver efficient and scalable infrastructure, core compute capacity, and managed data planes as multi-tenant services to accelerate the innovations of Engineering teams throughout the department. Your portfolio will include aspects that intersect with AWS accou

awskubernetesrest
View job →
🔔

Get new lead infrastructure software engineer jobs by email

Daily job updates · Unsubscribe anytime