Jobiba hiring network

Lead Infrastructure Software Engineer Jobs

6,876 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current lead infrastructure software engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

G
1mo ago

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role You will lead the product strategy for GitLab's next generation of engineering insight, helping software leaders, engineering managers, CTOs, and CPOs understand not just what their teams shipped, but why, and what to do next. As Principal Product Manager, Engineering Intelligence & Insights, you will shape how GitLab turns engineering data into decision-ready intelligence: the core data infrastructure and knowledge graph integrations that connect it, the out-of-the-box dashboards customers rely on today, and the conversational AI experiences that let leaders simply ask a question and get an answer in

REMOTEgitrestai
View job →
O
OpenAI
📍 London• Full-time
1mo ago

About the Team OpenAI’s Network Security team designs and operates the secure, reliable connectivity behind our offices, labs, campuses, cloud environments, people, and devices. We combine strong network fundamentals with automation, observability, and close partnership across IT, Security, Research, Applied, and business teams. About the Role As a Network Engineer, you will design, operate, troubleshoot, and automate secure, reliable networks across offices, labs, cloud connectivity, and production services. You will balance strategic platform work—architecture, standards, roadmaps, lifecycle planning, and automation—with responsive operations such as incidents, escalations, break/fix, and time-sensitive delivery. We’re looking for broad network engineers who meet users where they are, lead with curiosity, own outcomes end-to-end, move with urgency grounded in security, and iterate with purpose. You will turn operational signals and recurring reactive work into durable systems and standards. In this role, you will: End-to-end ownership of secure enterprise routing, switching, wireless, WAN, network services, and cloud connectivity. A deliberate balance of strategic platform improvement and responsive troubleshooting, change safety, incident response, and operational delivery. Purposeful iteration through software, APIs, Infrastructure-as-Code, Git workflows, testing, and CI/CD that reduces recurring reactive work. You might thrive in this role if you have: End-to-end ownership of secure enterprise routing, switching, wireless, WAN, network services, and cloud connectivity. A deliberate balance of strategic platform improvement and responsive troubleshooting, change safety, incident response, and operational delivery. Purposeful iteration through software, APIs, Infrastructure-as-Code, Git workflows, testing, and CI/CD that reduces recurring reactive work. Compensation, Benefits and Perks This is a position with OpenAI UK Ltd., which controls the hiring and manageme

reactawsci/cd
View job →
O
1mo ago

About the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme level of severity. Our work involves: Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping misuse safeguards, alignment tools, and security measures on track to adequately address extreme threats that might arise in the future. Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness framework , and partnering with other staff to achieve these targets. This is urgent, fast-paced work that has far-reaching implications for the company and for society. About the role The stakes of securing OpenAI increases as our internal coding and research becomes increasingly driven by autonomous AI agents. Compromising these agents could allow a cyber threat actor to compromise many other parts of the company. In this role, you would lead Preparedness work defending the security of our internal AI agents against insiders, Advanced Persistent Threats (APTs), or powerful AI agents. We’re looking for a strong hands-on technical executor with experience working directly with advanced cyber threat actors. In this role, you will: Develop and maintain threat models via which advanced attackers could compromise our coding assistants and automated security systems. Identify security investments that are especially critical to make in advance; for example, prioritizing by implementation lead-times, costs, and benefit. Partner with Security, Infrastructure, Research, Legal, and Preparedness to align on implementation plans and tradeoffs. Lead technical execution directly when needed, including prototyping controls, writing and reviewing software, and coordinating engineers across teams. Work with penetration testers to close gaps in defenses. You might thrive in this role if you: Are an exceptional hands-on technical executor. Have worked with advanced

awsrestai
View job →
S
Snowflake
📍 Menlo Park• Full-time
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake is expanding the boundaries of the Data Cloud to support mission-critical transactional workloads. Our goal is to deliver OLTP capabilities with the performance, reliability, simplicity, and scale customers expect from Snowflake, while creating a seamless experience across transactional and analytical data. We are looking for a Senior Engineering Manager – OLTP to lead engineering teams and leaders building core transactional database technology and the cloud infrastructure required to operate it at scale. You will help define the architecture and roadmap, grow the organization, and drive technology from design through production. AS A SENIOR ENGINEERING MANAGER – OLTP AT SNOWFLAKE, YOU WILL: Set technical and execution strategy for key areas of Snowflake's OLTP platform, translating product goals into architecture, roadmaps, and team plans. Lead and grow multiple engineering teams, developing managers and senior technical leaders while fostering a culture of ownership, technical excellence, and execution. Drive adoption of AI and agentic development practices to improve engineering velocity, quality, and productivity across the software development lifecycle. Drive architecture and technical decisions in areas such as transactions, concurrency control, low-latenc

awsazuregcp
View job →
PE
1mo ago

Security Architect - AI-Powered Security Engineering About the Team The security team at Meesho is like the Avengers to Meesho's S.H.I.E.L.D. When 5% of Indian households shop with us, it’s important to build resilient systems to manage millions of orders every day. We’ve done this, with zero downtime! We value speed over perfection, and see failures as opportunities to become better. By 2030, Meesho will be the world’s most trusted and secure digital ecosystem, operating on an AI-native, Zero-Trust, self-healing architecture. As we shift toward an era where autonomous agents generate and deploy code, we want to ensure our security scales seamlessly with engineering velocity. We place special emphasis on the continuous growth of each team member, fostering a strong 'Founder’s Mindset' that helps us move fast. About the Role We are looking for a highly technical, hands-on Security Architect (Individual Contributor) to guide and lead our AI-Powered Security Engineering charter. In this role, you will bridge the gap between conventional DevSecOps, infrastructure security, and the rapidly evolving landscape of AI. You will optimize our AI-driven Software Development Lifecycle (SDLC) by designing secure-by-default guardrails for both human developers and AI coding agents. Your mission is to ensure near-zero vulnerabilities from application and data security failures reach production. A significant portion of your focus will involve advanced Red Teaming, utilizing LLMs to improve Meesho's security posture, and architecting defenses for our autonomous AI systems against prompt injection, data leakage, and jailbreaking attacks.

gitairust
View job →
O
29 days ago

About the Team OpenAI's Industrial Compute organization is building and operating the infrastructure foundation for the next generation of AI. Infrastructure Operations works across facilities, hardware, network operations, incident management, data center engineering, delivery teams, and external partners to bring capacity online safely, understand its operational state, and improve it over time. As OpenAI's data center portfolio grows across first-party and partner-delivered capacity, the organization needs clear goals, trusted data, repeatable processes, and systems that make ownership, risk, readiness, and performance visible. This role will help build the operating mechanisms that allow Infrastructure Operations to scale with rigor. About the Role We are seeking a Technical Program Manager to own the systems, data, reporting, governance, and program-management backbone for Infrastructure Operations. Reporting to the Delivery & Operations Lead, you will translate strategy into executable goals and operating cadences, turn operational needs into software and data solutions, and create the mechanisms that keep a rapidly evolving organization aligned and accountable. This role will also own the current 1P+3P delivery-tracking layer within Operations: milestones, delivery timelines, quantity forecasts, risks, decisions, and executive reporting. You will partner closely with 1P Delivery Program Management, Compute TPMs, Data Center Engineering, construction, commissioning, and operations leaders to ensure that delivery information becomes complete, usable input for readiness, handover, and ongoing operations. You will own program health and the operating system around it: the goals, data definitions, workflows, reporting, decision paths, and follow-through that help functional DRIs execute. The ideal candidate is comfortable in ambiguity, technically fluent enough to implement real systems, and relentless about converting scattered information into durable mechan

REMOTEsqlawsrest
View job →

Observability Pipelines (OP) is Datadog's on-premise, vendor-agnostic telemetry pipeline product. As an Engineering Manager on the team, you'll own people management and engineering execution for one of OP's core missions, spanning areas like Integrations (ingesting from and routing to the many source and destination systems customers rely on), streaming insights, cost control, or pipeline capabilities, reliability and scalability. You'll partner directly with Product to help shape the roadmap, and work closely with your peer EMs and senior ICs to define how OP operates and grows. This is an opportunity to build your management craft while having real influence over the technical direction of a fast-growing product area. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Own people management and engineering execution Establish a strong operating rhythm for the team Drive high standards for on-call rotations and incident response Partner with Product on the roadmap, balancing product priorities with technical realities Lead, coach, and grow the careers of engineers on your team Who You Are: Experienced managing engineers directly, comfortable owning a team’s operating rhythm end-to-end, from planning through execution and stakeholder communication to incident and on-call ownership Have a technical background in distributed systems and data infrastructure Have experience with on-premises or customer-installed software concepts A product-minded partner to have on the team — you enjoy working with Product on strategy Experience with high-performance or Rust-based data pipeline systems Datadog values people from all walks of life. We understand not everyone will meet all the above qualifications o

aigorust
View job →
R
Replit
📍 Foster City• Full-time• Remote
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role We are looking for an AI Agent Security Architect to function as the primary technical authority for Replit’s autonomous and AI agent security blueprint. In this critical role, you will design, implement, and maintain the runtime defense systems, guardrail frameworks, and sandboxing architectures that govern AI agents executing code, invoking tools, and reasoning across our platform. You will be a key technical contributor—leading high-impact AI security initiatives and bridging the gap between non-deterministic AI behavior and rigorous cybersecurity controls for both engineering and executive leadership. What You'll Do AI Agent Security Strategy & Technical Execution AI Agent Security Blueprint: Define the long-term vision and architectural patterns for securing autonomous agent workflows, Model Context Protocol (MCP) integrations, multi-turn reasoning loops, and multi-agent coordination. Runtime Guardrails & Policy Enforcement: Architect and deploy dynamic input/output guardrail systems, semantic firewalls, and real-time intent verification filters to prevent goal hijacking, system prompt leaks, and indirect prompt injections. Agent Execution & Tool Sandboxing: Partner with Infrastructure and AppSec teams to design secure, short-lived, micro-isolated environments (e.g., microVMs, WebAssembly, container sandboxes) where agents can dynamically execute code, run shell commands, and interact with host operating systems safely. Agentic Threat Modeling & Red Teaming: Conduct specialized threat modeling against non-deterministic systems. Lead automated and manual AI red-teaming initiatives to uncover vulnerabilities in RAG context pipelines, vector stores, and tool-calling interfaces. Identity

REMOTEjavascripttypescriptpython
View job →
R
Ramp
📍 New York City• Full-time• From $10K/yr
1mo ago

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role We're looking for a Technical Program Manager who can operate at the intersection of engineering, product, and business — someone deeply technical, trusted instinctively by engineers, and sharp enough to drive clarity and momentum across complex, cross-functional programs. This is a high-agency role with real executive visibility and direct impact on how Ramp's engineering organization scales. We're looking for someone who is energized by complexity, deeply curious about what AI can unlock for engineering teams, and eager to apply it hands-on in their work. You should be someone who experiments with AI tools regularly, thinks about how they change the way software gets built, and brings that perspective into how you run programs. What You’ll Do Lead large-scale technical programs across engineering and adjacent teams—from CI/CD and infrastructure scaling to incident response, and driving other strategic projects across the engineering organization Own Ramp's engineering incident response program, improving processes, running retrospec

ci/cdrestai
View job →
T
Tenstorrent
📍 Austin• Full-time• $100K – $500K/yr
16 days ago

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. We are seeking an experienced Project Manager to lead our ERP implementation and drive delivery across technical and business teams. In this role, you will bridge Solution Architects with stakeholders across Sales, Supply Chain, Finance, Trade, Infosec, and Engineering to keep the program aligned, accountable, and moving forward. This role is hybrid, based out of Santa Clara, CA, or Austin, TX. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are An experienced project manager who has led complex ERP implementations, such as NetSuite or SAP. Skilled in project management methodologies and tools, including Jira and Asana. A strong interpersonal communicator who can influence, align, and mediate across technical and business teams. Able to synthesize complex technical discussions into clear, actionable reporting for varied stakeholders. What We Need Define project scope, milestones, deliverables, and governance for the ERP implementation. Manage external consultants, monitor adherence to statements of work, and drive accountability for delivery. Coordinate Sales, Supply Chain, Finance, Trade, Infosec, Engineering, Solution Architecture, and o

O
Okta
📍 San Francisco• Full-time• From $162K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The ISV Ecosystem team sits within the Okta Products and Technology organization and is focused on expanding our breadth and depth of SaaS integrations across both Okta Workforce and Auth0 product areas, a C-level priority for the company. Our team ensures SaaS ISVs (independent software vendors) have the knowledge, tools, and incentives to build integrations that extend our core product functionality. We are driving the formation and adoption of new identity standards focused on building an interconnected and secure AI ecosystem. Okta’s ecosystem is a core differentiator, and developing a scalable approach to ecosystem development will be a critical driver of our next stage of growth. As Program Manager, you will maximize platform value through integration breadth and depth. You will lead a high-visibility, cross-functional program spanning Okta and Auth0 product investments, marketing, technical enablement, and open standards. What you'll be doing: Lead end-to-end program governance including defining operational rhythms, milestones, success metrics, and risk mitigation strategies Scale program impact using sophisticated program management methodologies, frameworks, and tools to drive efficiency across the organization Leverage AI to enhance our program management methodology by defining, building, and documenting AI-driven processes to increase overall program efficiency and effectiveness Own reporting for C-level executives and cross-f

awsrestmachine learning
View job →
GR
16 days ago

Who we are Graviton Research Capital is a privately funded quantitative trading firm. We trade across a multitude of asset classes and trading venues using a diverse range of concepts, from time series analysis and stochastic models to machine learning and statistical inference. We analyse terabytes of data to identify pricing anomalies and drive innovation in financial markets. Role Overview We are looking for a Program Manager who thrives at the intersection of rigorous engineering and predictable delivery. You will not just "manage tasks" — you will orchestrate the development lifecycle for mission-critical systems. Your goal is to ensure that our elite engineering teams can focus on high-performance code while you own the execution strategy, dependency mapping, and release discipline. Key Responsibilities Lead Agile ceremonies (Sprint Planning, Stand-ups, Retrospectives) tailored for deep-tech engineering teams. Transform high-level trading requirements into granular, executable backlogs. Own capacity planning and burn-down metrics to provide high-visibility delivery timelines. Navigate the complex interplay between engineering teams (e.g., Connectivity, Core Infrastructure, Simulation) to prevent bottlenecks. Build and maintain advanced Jira dashboards, automated roadmaps, and Confluence documentation that serve as the "single source of truth" for stakeholders. Proactively identify technical debt, architectural blockers, or resource gaps that threaten release stability. Continuously refine Agile methodologies to suit low-latency, performance-sensitive development cycles (where "Definition of Done" includes rigorous performance benchmarking). Eligibility & Required Skills 5+ years of experience as a TPM, Program Manager, or Scrum Lead in a product-engineering environment (HFT, FinTech, Networking, or Kernels/Systems). A strong grasp of the software development lifecycle for high-performance systems. While you won't write code, you must understand concepts li

ci/cdagilescrum
View job →

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. This is a fully on-site role with mandatory relocation. Competitive relocation package provided. THE ROLE As a Technical Program Manager in the Core Platform Business Unit (CPBU), you will define and lead the program execution strategy for Everpure’s™ flagship FlashArray and FlashBlade hardware and software initiatives. You will align multi-stream, highly complex engineering programs with overarching business objectives, bridging hardware, software, and cloud-native developments for AWS and Azure. Serving as a strategic catalyst across global site leads, executive stakeholders, and infrastructure teams, you will tackle intangible architectural dependencies and drive high-impact initiatives to market on schedule without downtime. If you thrive on navigating multi-system complexity and driving organizational clarity, this role offers an exceptional platform for leadership. WHAT YOU'LL DO Drive Strategic Program Execution: Architect and manage multi-stream engineering roadmaps for FlashArray, FlashBlade, and hybrid cloud releases, ensuring on-time delivery across the full software development lifecycle. Lead Cross-Organizational Alignment: Direct end-to-end communication, status transparency, and dependency tracking across engineering, product management, infrastructure, and executive leadership teams. Optimize Operations & Capital Expenditure: Manage CapEx allocation, budget tracking, and resource planning for eng

awsazurerest
View job →
E
16 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Portworx team to build and deliver our highest-quality product suite. In this role, you will write clean, scalable code with a strong focus on quality, reliability, and user-centric design. You will directly contribute to building a new SaaS platform that delivers a secure, consistent, and best-in-class experience for customers purchasing and managing Portworx offerings. As a core developer, you will take ownership of designing and implementing critical features across the entire Portworx portfolio. WHAT YOU’LL DO Design & Scale SaaS Microservices: Develop, test, and integrate high-performance microservices and features into the Portworx product suite, ensuring high availability in distributed systems. Drive End-to-End Delivery: Lead software lifecycle activities including architectural design, code reviews, unit/functional testing, documentation, and continuous integration and deployment (CI/CD). Partner Across Teams: Collaborate with product managers, cross-functional engineering peers, and early-adopter customers to transform requirements into production-ready software. Own Product Quality & Iteration: Take full ownership of feature stability by proactively incorporating customer feedback and rapidly resolving issues identified during testing and deployment. Innovate & Experiment: Research emerging technologies and cloud infrastructure tools to push performance boundaries and continuously i

javaawskubernetes
View job →
E
16 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Join the Portworx team to build and deliver our highest-quality product suite. In this role, you will write clean, scalable code with a strong focus on quality, reliability, and user-centric design. You will directly contribute to building a new SaaS platform that delivers a secure, consistent, and best-in-class experience for customers purchasing and managing Portworx offerings. As a core developer, you will take ownership of designing and implementing critical features across the entire Portworx portfolio. WHAT YOU’LL DO Design & Scale SaaS Microservices: Develop, test, and integrate high-performance microservices and features into the Portworx product suite, ensuring high availability in distributed systems. Drive End-to-End Delivery: Lead software lifecycle activities including architectural design, code reviews, unit/functional testing, documentation, and continuous integration and deployment (CI/CD). Partner Across Teams: Collaborate with product managers, cross-functional engineering peers, and early-adopter customers to transform requirements into production-ready software. Own Product Quality & Iteration: Take full ownership of feature stability by proactively incorporating customer feedback and rapidly resolving issues identified during testing and deployment. Innovate & Experiment: Research emerging technologies and cloud infrastructure tools to push performance boundaries and continuously i

javaawskubernetes
View job →
🔔

Get new lead infrastructure software engineer jobs by email

Daily job updates · Unsubscribe anytime