Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The Product Security Data Platforms team is a newly established engineering team within Stripe Security. Our mission is to build the foundational infrastructure that provides our users with unprecedented visibility into the security posture of their Stripe integration. While Stripe is renowned for industry-leading payment protection, we are expanding our focus to provide a comprehensive security telemetry platform that helps businesses protect their entire digital ecosystem on Stripe. As a founding member of this team, you'll architect a large-scale customer-facing security data pipeline and presentation layer. Much like modern security observability platforms and data lakes that have transformed cloud infrastructure, we're building an API-first service that transforms massive streams of behavioral data into actionable security intelligence. This team operates at the intersection of high-throughput data engineering and cybersecurity, creating the systems that will allow the world’s most sophisticated companies to monitor, detect, and respond to threats in real time. What you’ll do As a Senior Software Engineer on this founding team, you'll lead the technical design and implementation of our core security data pipelines. You'll define how we capture security signals, process them at scale, and deliver them to our users through robust, developer-friendly interfaces. If you have security domain knowledge, you'll have opportunities to help shape
Jobiba hiring network
Lead Infrastructure Software Engineer Jobs
6,876 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current lead infrastructure software engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers. We’re looking for an EM to lead a team of exceptional ML infrastructure engineers, build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You Will: Lead strategic planning and roadmap execution of scalable production-ready ML systems including model training, data pipelines, feature engineering and model inference. Own the architecture, establish engineering best practices of scalability, reliability, and cost-effectiveness of ML infrastructure (e.g., training, serving, feature). Work closely with data scientists, ML engineers, platform teams, and product stakeholders to design, implement, and operate robust ML platforms that accelerate model development and deployment. Recruit, mentor, and grow a high-performing team of ML infrastructure engineers. You Have: 5+ years of experienc
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Observability team builds the infrastructure that empowers engineers to understand, operate, and improve the Roblox platform and ecosystem. Our team owns the end-to-end observability stack across telemetry, distributed tracing, logging, profiling, storage systems, and developer-facing visualization tools. We are looking for an Engineering Manager to lead the next generation of AI-powered observability platforms. In this role, you will help build intelligent systems that leverage AI to revolutionize CI/CD, testing, and DevOps workflows — enabling engineers to move faster, improve reliability, and operate large-scale distributed systems with greater efficiency and confidence. This is a highly impactful leadership role at the center of Roblox infrastructure. Your work will directly improve developer productivity, platform reliability, and operational excellence across the company. You will partner closely with infrastructure, product engineering, and AI platform teams to shape the future of developer tooling and autonomous operations at scale. You Have 3+ years of engineering management experience with a proven track record of hiring, mentoring, and growing high-performing teams. Strong ex
You’ll shape the future of a business‑critical platform as the technical lead across both product engineering and cloud infrastructure. You’ll modernize a mature .NET application running on AWS today, while steering its evolution toward a cloud‑native, React/Node.js, AI‑enabled architecture. If you enjoy owning architecture end‑to‑end, from backend and frontend through CI/CD, DevOps, and AWS infrastructure, this role gives you real influence at Staff Engineer level and the opportunity to set engineering standards that others follow. You’ll spend your time leading complex .NET and React features, designing scalable AWS infrastructure with Infrastructure as Code, and building automation that makes releases fast, safe, and repeatable. You’ll work on performance, reliability, and modernization in equal measure—fixing what’s slowing the platform down today and designing what it will look like in the next generation. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead the architecture and development of enterprise .NET services and APIs that power a business‑critical platform. Design and operate AWS infrastructure (using AWS CDK in TypeScript) to support secure, scalable, multi‑environment deployments. Build and optimize CI/CD pipelines (AWS CodePipeline, CodeBuild, Windows build agents) to make shipping .NET and React changes fast and reliable. Drive modernization initiatives across the stack, including clean architecture, refactoring legacy components, and reducing technical debt. Design and tune PostgreSQL and MSSQL database solutions for performance, scalability, and reliability. Mentor engineers and influence engineering practices across teams, raising the bar on cloud, DevOps, and software design. These are the essentials you’ll need to get an interview Significant experience (typically 8+ years) delivering and operating scalable enterprise software, owning both application code and cloud infrastructure. Deep hands‑on expertise with C
Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,
Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,
Squarespace is looking for a Backend Staff Software Engineer to lead the technical direction of our Communications Platform team. The Communications Platform is a critical engine powering customer engagement across Squarespace — providing internal teams with a scalable, reliable, and secure infrastructure for delivering communications via email, push notifications and in-product messaging. The team owns multiple interconnected production platforms that collectively send over 1.75+ million notifications per day, strategic initiatives including platform modernization, In-Product Placements, and multi-channel communication capabilities. In this role, you’ll serve as the senior technical voice for the team — driving architecture, shaping roadmap execution, and elevating the engineering quality of a small but high-impact team. This is a hands-on role with real technical depth, paired with broad cross-functional influence. This is a hybrid role based in our NYC office (3 days per week), reporting to the Engineering Manager of Communications Platform. You’ll Get To… Drive the architecture, design, and implementation of the Communications Platform’s platforms — including Email Delivery, In-Product Notifications and Push Notifications. Define and own the technical strategy for platform modernization, reliability hardening, and engineering standards across backend services. Provide hands-on technical leadership and mentorship to a team of backend and frontend engineers, helping them grow in system design, decision-making, and ownership. Write high-quality Java code and stay close to execution — leading design reviews, code reviews, and architectural decisions that raise the bar across the codebase. Proactively identify and address systemic risks in performance, reliability, security, and operability before they become incidents. Serve as a key technical partner on cross-team and company-wide architecture initiatives. Balance short-term delivery with long-term platform health,
About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu
About the Role: As a Staff Software Engineer on the ML Infrastructure team, you will collaborate closely with the Machine Learning and Product teams to build world-class machine learning inference platforms. These platforms power essential services like personalized recommendations, search, and content understanding across Tubi. A core responsibility of this team is developing and maintaining low-latency ML model serving systems that support Deep Learning, LLM, and Search models. This involves building self-service infrastructure and critical components such as the inference engine, feature store, vector store, and experimentation engine. You will improve the way we deploy and operate our services and even contribute to open-source projects. This role grants the architectural freedom to explore new frameworks, lead critical cross-functional projects, and transform the capabilities of our ML and Product teams. Responsibilities: Design and build scalable, high throughput, and low latency distributed systems using Scala Build reusable components and services that serve various ML applications like Personalization, Search, Ads and Exploration Partner closely with ML engineers to understand their challenges and limitations and develop scalable solutions to address them. Proactively recommend solutions to keep our ML Inference stack state of the art. Take a data driven approach to identifying & optimizing latency, cost, and efficiency of our infra. Lead large scale cross functional refactorings if necessary Mentor other engineers on the team on system design, effective incident management, interviewing, leveraging LLMs for work, etc. Collaborate with ML, Product, and cross functional engineering teams to define the long term vision and architecture for ML Infrastructure at Tubi. Your Background: Experience designing and building scalable, distributed systems in any modern backend language (e.g., Scala, Java, Python, Go, C++); experience with Scala or JVM b
Opportunity Overview: We are seeking a Lead Software Engineer to join our Integrations team. In this role, you will be designing, developing, and scaling highly available healthcare integration systems supporting prior authorization workflows across providers, payers, and delegated entities. You'll direct a fast-paced, autonomous,agile team of software engineers in the design, development, and operational support of a growing enterprise integration platform. This is an opportunity to drive technical excellence at the intersection of healthcare interoperability and modern distributed systems. What you’ll do: Technical Leadership: Provide technical leadership across architecture, system design, platform scalability, reliability, and operational excellence. Platform Engineering: Design and build scalable, resilient, and high-performing systems that support critical business workflows and enterprise integrations. Integration Solutions: Lead the development and maintenance of secure integrations with internal and external platforms, partners, and third-party systems. Cloud & Automation: Drive cloud infrastructure, deployment automation, and software delivery practices that enable reliable and efficient releases. Distributed Systems: Design and support event-driven and distributed architectures that enable scalable and fault-tolerant processing. Operational Excellence: Establish monitoring, observability, and incident response practices to ensure system reliability, performance, and availability. Quality Engineering: Champion automated testing, quality assurance, and engineering best practices throughout the software development lifecycle. Production Support: Lead the resolution of complex production issues and drive continuous improvement in platform stability and operational efficiency. Cross-Functional Collaboration: Partner with product, operations, data, security, and business stakeholders to deliver solutions aligned with organizational goals. Agile Delive
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team We're experts in data, working to make it cost-effective, understandable, and trustworthy. We build pipelines processing billions of events a day and are stewards of canonical data warehouses and datasets delivering products for Stripe Users while embedding with teams to build their data products. We're experts in using the Stripe Data Platform and to scale we lead the data culture and data education to enable product teams to own their data. We invest in AI Data Ops to scale incident handling and serve as an escalation path for data incidents to minimize their impact. The Data Engineering Solutions team works closely with product teams delivering trustworthy data, backend code, and innovative AI tools, platforms, and services for data. What you'll do We're looking for a person who drives the Data Engineering Solutions Team in solving high-impact, cutting-edge data problems. The ideal candidate is someone who has built data pipelines for large-scale volume, is deeply knowledgeable of Data Engineering tools including Airflow, Spark, Kafka, and Flink, is empathetic, excels at building strong relationships, and collaborates effectively with other Stripe teams to understand their use cases and unlock new capabilities. • Lead the technical outcomes for a team of ambitious, talented engineers, providing mentorship, guidance, and support to ensure their success • Partner with our recruiting team to attract and hire top talent. • Deliver cutting-edg
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team Data Quality and Governance owns the infrastructure that makes Stripe's data trustworthy and findable - the Data Catalog, the Knowledge Graph Service, dataset tiering and governance standards, lineage tracking, and the quality scoring system (DQPD) that every engineering team reports against. They're the team that defines what "good data" means at Stripe and then builds the enforcement and measurement tools to drive adoption across the company. What you’ll do Responsibilities Lead the technical outcomes for a team of ambitious, talented engineers, providing mentorship, guidance, and support to ensure their success Build and operate large-scale data discovery, metadata, or catalog platform Develop strong subject matter expertise and manage the SLAs of data pipelines and full stack web applications that support critical stakeholders Collaborate with product managers and peers across the company to create/improve canonical datasets and data warehouses, use golden paths, and ensure Stripes and customers are using trustworthy data Leverage AI/LLM and Agents at scale to produce and analyze high-quality data on ambiguous problems Have the opportunity to drive the execution of key data initiatives for Stripe, overseeing the entire development lifecycle from planning to delivery while maintaining high standards of quality and timely completion Foster a collaborative and inclusive work environment, promoting innovation, knowledge sharing, and continuo
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a Software Engineer on the DX Security Components team , you will architect the backend services that safeguard Pure’s core authentication and remote access infrastructure. You’ll act as a security-minded engineer within a high-impact platform team, ensuring our cloud offerings and customer appliances remain auditable and resilient. Collaborating across the DX organization, you will bridge the gap between robust security protocols and seamless developer integration to protect our global customer base. WHAT YOU’LL DO Engineer Security Infrastructure: Design and operate high-availability backend services that manage authentication, authorization, and certificate lifecycles to ensure secure access across all Pure1 cloud and appliance environments. Drive End-to-End Ownership: Lead the full service lifecycle—from initial architectural design and threat modeling (STRIDE) to deployment, observability, and long-term cost efficiency. Champion Secure Integration: Partner with Security Governance and product teams to streamline remote-access flows, translating complex security requirements into pragmatic, automated workflows for other engineering squads. Ensure System Resilience: Maintain the integrity of security-sensitive systems by participating in a global follow-the-sun on-call rotation, performing root-cause analysis, and hardening infrastructure against emerging threats. Automate Trust: Evolve Infrastructure as Cod
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Storage Abstractions (STAX) team builds the software layer through which Stripe services access stored data. We own the database SDKs that hundreds of Ruby and Java services use to read and write data safely, reliably, and efficiently without needing to understand the underlying database, routing, or operational complexity. Our work creates leverage across Stripe: by providing stable interfaces and safeguards at the storage layer, we help teams build products with confidence while enabling Stripe’s data architecture to evolve. Alongside engineers, we are designing for AI agents as customers of this foundation, making storage capabilities discoverable, interoperable, and safe to use across languages and storage backends. What you’ll do As a Staff Software Engineer on STAX, you will set technical direction and lead multi-year initiatives at the intersection of developer infrastructure, data access, and AI. You will work hands-on with engineers across Stripe to make storage access simpler, safer, and more interoperable, while helping product and infrastructure teams evolve their systems without fleet-wide migrations. You will help turn Stripe’s AI strategy into practical developer infrastructure by treating AI agents as customers of the storage layer. This is an opportunity to build frameworks and interfaces that make complex storage operations discoverable, interoperable, and safe for both human developers and agentic workflows, while rais
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Events Analytics Platform (EAP) team is responsible for the infrastructure that powers all of Sentry's time-series data and searching capabilities across billions of events with sub-second latency. We started this initiative by building Snuba, the primary storage and query service for Sentry's event data powered by ClickHouse, and we are now focused on unlocking deeper visibility and reporting across the terabytes of event data our users generate. As a Senior Software Engineer, you will lead efforts to push the boundaries of data visibility at Sentry. You will do this by expanding the capabilities of our search infrastructure, building new capabilities on top of our state-of-the-art storage layer and increasing the performance and integrity of Sentry’s core data services. You will also help shape Infrastructure's technical direction at Sentry and collaborate with Product and other Engineering teams to turn that vision into a reality. If you want to solve the hard problems that come with scaling event data into the petabyte range, this could be the job for you. In this role you will: Expand EAP's ability to deliver data at world-class speed and reliability. Architect and automate services and systems to scale reliably under growing demand. Make architectural trade-offs that balance product requirements with engineering constraints. Maintain and grow the team's code quality initiatives by regularly reviewing code and contributing to design decisions. Lead design and discussions around deliverables the team is working towards. Improve the maintainability and developer experience of the codebases EAP owns. Exa
Get new lead infrastructure software engineer jobs by email
Daily job updates · Unsubscribe anytime