Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role We're looking for an AI-native product growth lead to own Replit's entire paid acquisition and lifecycle marketing engine as the single DRI. This is an IC role for someone who builds systems, not just campaigns. Someone who uses AI and automation aggressively to do what traditionally requires an entire team. You don't need an engineering background. You need the instinct to build: when you see a repetitive task, you reach for Replit or an API before you reach for a spreadsheet. You'll have data science and data engineering partners for infrastructure and modeling, and brand marketing will set creative direction. Your job is to turn that direction into a closed-loop optimization system: a self-improving system that turns performance data into better creative, optimizes spend across channels in near-real-time, and compounds every insight so nothing learned is ever lost. Each cycle, the system gets smarter. That's the engine you'll build and own. The ideal candidate has deep channel expertise across paid search, paid social, and lifecycle, but their real edge is building AI-powered workflows that scale creative production, automate measurement, and compound institutional knowledge. You Will Own full-funnel performance marketing across paid search & social, ASO, and lifecycle Build a self-improving creative engine: AI-driven ad generation, testing, and iteration that scales without scaling headcount Maintain a persistent knowledge layer so every experiment, result, and creative insight compounds automatically into the next cycle Own conversion signal quality end-to-end: right events, right audiences, right attribution, partnering with DE on the infrastructure Run a structured experimentation program wher
Jobiba hiring network
Lead Infrastructure Software Engineer Jobs
6,876 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current lead infrastructure software engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. We are looking for a Security Operations Lead (SOC Lead) to build, mature, and operate our 24/7 detection and response capabilities across a modern cloud-native and AI-driven environment. This role leads the global SOC function—monitoring, SIEM ownership, detection engineering, alert triage, and operational readiness—while also evaluating and integrating emerging AI-based SOC products and autonomous response platforms . You will oversee monitoring across multi-cloud environments (GCP primary, AWS/Azure secondary), Kubernetes, SaaS services, endpoints, developer tools, and AI workloads . You’ll collaborate closely with Cloud Security, Compliance/GRC, SRE, Platform Engineering, IT/Endpoint teams, and AI Infrastructure to ensure our detection strategy scales and stays ahead of evolving threats. This is a hands-on leadership role perfect for someone who wants to shape the SOC of the future while solving complex challenges in a high-scale AI setting. What You’ll Do SOC Leadership & 24/7 Monitoring Lead, mentor, and scale a global SOC team responsible for 24/7 monitoring, alert intake, triage, correlation, and escalation. Build operational rigor: processes, runbooks, SLAs, metrics, and quality standards for high-scale environments. Cover monitoring across: Cloud infrastructure (GCP, AWS, Azure) Kubernetes/GKE/EKS/AKS clusters SaaS platforms (Google Workspace, GitHub, Slack, Okta, etc.) Endpoints (macOS, Linux, Windows) including EDR/XDR telemetry Developer platforms + CI/CD pipelines AI/ML systems and model-serving workflows AI-Based SOC Integration & Innovation Evaluate, adopt, and integrate AI-native SOC technologies for triaging, detection, and correlation Identify opportunities to automate triage, investigations,
The IT Systems Engineering team builds the systems and processes that keep Asana running: endpoint management, identity, and the enterprise applications every Asana employee uses every day. As Manager of IT Applications and Client Platform Engineering, you’ll lead four engineers: two on client platform, two on applications. You report directly to the Head of IT Systems Engineering. This is a player/coach role: you set direction, run planning, and step in on escalations and architecture decisions while your engineers own most of the day-to-day. Your primary depth is in endpoint engineering; you have enough SaaS administration background to lead the apps side without micromanaging it. We’re looking for someone who knows endpoint work deeply and is ready to grow as a manager. Your title history matters less than technical foundation and how you work with people. This role is based in our San Francisco office with an office-centric hybrid schedule. The standard in-office days are Monday, Tuesday, and Thursday. Most Asanas have the option to work from home on Wednesdays. Working from home on Fridays depends on the type of work you do and the teams with which you partner. If you’re interviewing for this role, your recruiter will share more about the in-office requirements. What you’ll achieve Lead and develop your team. Manage and mentor four engineers (two on CPE, two on Apps). Set priorities, clear blockers, and help each person grow. Own the client platform program. Set strategy and lead execution for Asana’s endpoint fleet: macOS, Windows, and mobile. This means MDM infrastructure, zero-touch deployment, software catalog, and OS lifecycle. We’re primarily an Apple shop with a growing Windows footprint. Strengthen endpoint security. Work with the Corporate Security team to harden our endpoint posture, respond to incidents, and build initiatives that connect security priorities and endpoints. Oversee the applications engineering function. Your Apps engineers manage the
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations Austin, US New York, US San Francisco, US Responsibilities Lead and Mentor: Manage a high-performing team of kernel and systems engineers, fostering a culture of "collaboration-first" and continuous knowledge sharing both internally and with the global open-source community. Own the OS Lifecycle: Oversee the end-to-end delivery of the Linux Kernel and core OS components. Your goal is a fully automated continuous delivery pipeli
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Architect Team The Architecture team is a small group of very senior engineers reporting to our VP of Engineering Excellence, working broadly across the organization in collaboration with Engineering, Product, and Security. We partner deeply with other Engineering teams for large projects, and provide direction and architectural guidance for smaller initiatives. We have a dual-pronged charter to “level up the tech stack and level up the people stack” via both technical contributions and partnerships/mentoring. In this role, you will have the opportunity to significantly contribute to Auth0’s future technology direction. Through your experience, knowledge of industry trends, and technical abilities you will provide guidance, build proof of concepts, and deliver production software implementations that help Auth0 Engineering teams move faster by using and developing standard patterns and technologies. You will also help advance the engineering culture and help uplevel other engineers. Note that while this role involves a lot of guidance, documentation, and leadership, it also requires substantial hands-on coding and development of both applications and systems. What you’ll be doing You will collaborate with Product, Security, and Engineering teams to define and continually improve Auth0’s technology stack and architecture. Foster and lead innovation in the IAM space, with a strong focus on Agentic Identity Lead initiat
Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you’ll be part of a passionate team dedicated to accomplishing hard things, together. Job Description: ABOUT THE TEAM Our Engineering team is revolutionizing how Appian resolves customer issues by building AI-powered diagnostic tools and internal platforms that dramatically reduce resolution time. We combine deep technical expertise with AI innovation to empower support engineers to identify root causes faster, ensuring our customers experience uninterrupted business-critical operations. This team operates with high autonomy and directly impacts customer satisfaction and retention across Appian's enterprise client base. THE OPPORTUNITY Appian is at an inflection point where intelligent automation meets customer success, and we need a technical leader who can architect and ship the next generation of diagnostic platforms. This role exists because our growing enterprise customer base demands faster, smarter issue resolution, and AI agents are the key to unlocking that scale. You will shape both the technical direction and the team culture while delivering tools that become essential to how Appian supports thousands of mission-critical deployments worldwide. WHAT YOU'LL DO Lead the design and delivery of a scalable AI-powered triage platform that reduces mean time to resolution by leveraging AI agents, LangChain frameworks, and intelligent diagnostics Lead the engineering team building AI-powered diagnostic tools that cut customer issue resolution time in half while scaling a platform that empowers support engineers across the globe. Architect quarterly releases of diagnostic tools on AWS infrastructure, balancing innovation velocity with pro
WHO ARE WE? We are a bunch of super enthusiastic, passionate, and highly driven people, working to achieve a common goal! We believe that work and the workplace should be joyful and always buzzing with energy! CloudSEK , one of India’s most trusted Cyber security product companies, is on a mission to build the world’s fastest and most reliable AI technology that identifies and resolves digital threats in real-time. The central proposition is leveraging Artificial Intelligence and Machine Learning to create a quick and reliable analysis and alert system that provides rapid detection across multiple internet sources, precise threat analysis, and prompt resolution with minimal human intervention. Founded in 2015, headquartered at Singapore, we are proud to say that we’ve grown at a frenetic pace and have been able to achieve some accolades along the way, including: CloudSEK’s Product Suite: CloudSEK XVigil constantly maps a customer’s digital assets, identifies threats and enriches them with cyber intelligence, and then provides workflows to manage and remediate all identified threats including takedown support. A powerful Attack Surface Monitoring tool that gives visibility and intelligence on customers’ attack surfaces. CloudSEK's BeVigil uses a combination of Mobile, Web, Network and Encryption Scanners to map and protect known and unknown assets. CloudSEK’s Contextual AI SVigil identifies software supply chain risks by monitoring Software, Cloud Services, and third-party dependencies. CloudSEK’s AIVigil is an AI-native Attack Surface Monitoring platform that continuously discovers, monitors, and secures exposed AI infrastructure, MCP servers, leaked AI credentials, vector databases, agentic workflows, and shadow AI across the internet. Key Milestones: 2016 : Launched our first product. 2018 : Secured Pre-series A funding. 2019 : Expanded operations to India, Southeast Asia, and the Americas. 2020 : Won the NASSCOM-DSCI Excellence Award for Security Product Company
The Anyscale Technical Program Management (TPM) team is expected to play a critical role executing high impact programs while continuously improving processes to sustainably grow and increase the effectiveness of the Tech organization spanning Design, Engineering and Product teams. As a Technical Program Manager focused on Anyscale’s core product solution, you’ll help lead complex application development in service of enhancing our product platform. In this role, you'll support and help scale the technical solutions that make Anyscale’s products and services possible. As part of the overall development life cycle you’ll plan requirements, identify risks, manage schedules, and communicate clearly with project stakeholders on complex projects with significant bottom line impact. Your Program management contributions will span prioritization, planning of projects and features, stakeholder management, tracking of external commitments while contributing to the organization's technical culture by highlighting and espousing best practices. You’ll learn and grow alongside talented teammates who share your commitment to excellence and appetite for innovative problem-solving. This is a rare opportunity to join in the leadership of a team that will be responsible for building a successful commercial ML/AI oriented solution from the ground up! As part of this role, you will: Help us build, track and ship our Commercial / OSS product Will work closely with the software development and product teams to deliver high quality, scalable products used by customers around the world Collaborate with the product teams and align all the stakeholders to assemble project teams, assign responsibilities, identify appropriate resources needed, and develop schedules to ensure timely completion of projects by meeting project milestones Assess risks, anticipate bottlenecks, provide escalation management, make tradeoffs, balance the business needs versus technical constraints and encourage risk ta
ABOUT THE ROLE We are a leading streaming global fitness content company with studios around the world including London, revolutionizing the way people access and engage with fitness workouts. Our platform offers a wide range of interactive, live and on-demand fitness content that caters to users of all fitness levels, empowering them to stay fit and healthy from the comfort of their homes. As the Senior Manager of Broadcast Engineering, you will play a pivotal role in our mission to deliver high-quality, seamless, and engaging fitness content to our global audience. You will lead the Broadcast Engineering team based in London, ensuring the smooth operation and optimization of our broadcast infrastructure, content delivery systems, and broadcast equipment. This position reports to the Director of Global Production Technology. YOUR DAILY IMPACT AT PELOTON Oversee and guide the Broadcast Engineering team in designing, implementing, and maintaining an efficient and reliable broadcast studio facility to deliver the best member experience possible Collaborate with global broadcast engineering leads to maintain parity and system wide connectivity between facilities Manage the procurement, installation, and maintenance of all broadcast equipment, ensuring their proper functioning and readiness for live and on-demand fitness classes Collaborate with cross-functional teams, including Content Production Operations, IT, and Product, to streamline content workflows, improve efficiency, and enhance the overall broadcast transmission process Stay up-to-date with the latest trends, advancements, and emerging technologies in broadcast engineering and streaming to propose and implement cutting-edge solutions Lead the team in promptly addressing technical issues and incidents, minimizing downtime and disruptions to the streaming service Mentor and guide the Broadcast Engineering team members, fostering a culture of learning, growth, and innovation YO
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our Site Reliability Engineering (SRE) team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of developers worldwide. As a Staff Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability. We are seeking Staff SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to proactively find and analyze reliability problems across our stack, then design and implement software and systems to create step-function improvements. You will design robust observability solutions, lead incident response, automate operational tasks, and continuously improve our infrastructure's reliability, all while mentoring and educating the broader engineering team to make reliability a core value at Replit. You Will: Architect and Implement Observability: Design, build, and lead the implementation of comprehensive monitoring, logging, and tracing solutions. Create dashboards and metrics that provide real-time visibility into system health and performance, enabling proactive issue detection. Define and Drive Reliability Standards: Work with product and engineering teams to define, implement, and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to monitor and report on these metrics, holding teams accountable and ensuring we maintain high reliability standards while balancing innovation speed. Lead Incident Management and Response: Act as a senior leader during high-impact incidents, guiding the team to rapid resolution
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the team The Billing team sits at the intersection of product, finance, and infrastructure. They're responsible for ensuring every observable event—errors, logs, traces, tokens—gets accurately measured, priced, and billed. Their work directly impacts company revenue and customer trust, requiring distributed systems expertise, attention to financial accuracy, and deep understanding of product usage patterns. The team works cross-functionally with product, engineering, BizOps, marketing, and sales to build systems that enable new products and pricing models. As an Engineering Manager, you’ll lead a team of engineers owning critical workflows such as checkout and invoicing, while also developing new features to help customers manage their spend growth. In this role, you’ll partner across the organization to ensure our customers redeem everything Sentry has to offer and budget for future expansion. In this role you will Strategic Planning & Roadmap: Define and drive the team's roadmap. Align team goals with organizational objectives and contribute to the overall platform strategy. Technical Guidance & Operational Excellence: Provide technical leadership and guidance on complex distributed systems and design. Ensure the team is proactively identifying areas for improvement. Cross-functional Collaboration: Partner closely with business and technical teams to translate business goals into actionable objectives and scalable solutions. Team Leadership & Development: Lead, mentor, and grow a team of talented engineers, including Staff-level engineers. Build a culture of technical excellence, collaboration, continuous
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our Site Reliability Engineering team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of developers worldwide. As a Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability. We are seeking SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to design and implement robust monitoring solutions, automate operational tasks, and continuously improve our infrastructure's reliability and performance. You will: Design and Implement Observability Solutions : Develop comprehensive monitoring and alerting systems using modern observability tools. Create dashboards and metrics that provide real-time visibility into system health and performance. Implement logging strategies that enable quick problem identification and resolution. Drive Automation and Infrastructure as Code : Architect and implement infrastructure automation solutions using tools like Terraform, Ansible, or Pulumi. Design and maintain CI/CD pipelines that enable reliable and consistent deployments. Create self-healing systems that can automatically respond to common failure scenarios. Establish SLOs and SLIs : Work with product and engineering teams to define and implement Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to track and report on these metrics, ensuring we maintain high reliability standards while balancing innovation speed. Incident Management and Response : Lead incident response efforts, conducting thorough post-morte
The data management software market is transforming how organisations build and run applications. MongoDB is the leading developer data platform and the first database provider to IPO in more than 20 years. Join us at the forefront of data and application development. MongoDB Technical Services Engineers combine deep technical expertise with exceptional problem-solving and customer-service skills. You’ll advise customers and resolve complex challenges across MongoDB Core, drivers, Atlas, Cloud Manager, cloud platforms, and infrastructure. We’re looking for candidates based in Dublin to join our vibrant office and collaborative in-office team. This is a five-day role with one of the following schedules: Tuesday–Saturday, Sunday–Thursday, or a five-day pattern covering both Saturday and Sunday. Under our hybrid model, employees on weekend schedules are expected to work from the office two days per week. Cool things you’ll do You’ll help customers troubleshoot complex issues and run critical MongoDB workloads with confidence. You’ll: Solve customer challenges across architecture, performance, recovery, and security Lead investigations from diagnosis to resolution, providing clear, actionable guidance Partner with Product Management and Engineering to advocate for customers and improve MongoDB Build tools, documentation, and training while mentoring peers and raising technical excellence What you need We value curiosity, adaptability, strong technical foundations, and a genuine desire to help customers. You should bring many of the following: 5–6 years of experience in technical support, systems engineering, database administration, SRE, or a related field Experience running complex, mission-critical production database systems Strong Linux and systems engineering skills, including performance, memory, I/O, storage, networking, security, clustering, and troubleshooting A solid understanding of networking concepts and protocols, including DNS, TCP/IP, and SSL/TLS Ability
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role We're hiring a hands-on Engineering Manager to build and lead Replit's Anti-Abuse team from the ground up. This is a foundational 0-to-1 role: you'll define the anti-abuse roadmap, hire a small team of engineers and data analysts, and ship the systems that protect Replit's platform, users, and economics from adversarial actors. You'll partner across Support, Legal, Security, Infrastructure, and the Money and Growth teams to make abuse economically unviable while keeping friction low for legitimate users. Replit sits at the frontier of AI-native abuse. Our platform is a target for phishing and scam hosting, cryptomining, LLM token farming, card and coupon fraud, and increasingly, abuse driven by AI agents themselves. The team you build will define how Replit defends against all of it. What You'll Do Build the anti-abuse roadmap from scratch : Define the threat model, prioritize across abuse vectors (phishing/scam hosting, cryptomining, token farming, payment fraud, AI agent exploitation), and translate it into a shipping plan with clear sequencing and tradeoffs. Design progressive verification and identity infrastructure : Build the "ladder of trust" that gates increasing platform capabilities (referrals, additional credits, access to powerful agent features, Missions) behind escalating verification. This includes a humanity/identity layer that's distinct from user accounts, integrations with KYC-grade verification providers, and the policy engine that decides what level of trust unlocks what behavior. This infrastructure is core not just to promo integrity but to how Replit safely expands agent capabilities over time. Ship as a hands-on EM : Stay in the code. Use the latest AI coding tools (including Rep
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Senior Identity Security Engineer on Palantir's Identity Security team, you will own the security posture of the identity infrastructure that Palantirians, customers, and services rely on every day. The Identity Security team is responsible for all identity types at Palantir - workforce, customer, workload, and agentic - giving you the rare ability to architect, threat model, and drive security outcomes across the full identity surface. You will help shape the technical direction for identity security at Palantir, reduce standing access, lead identity threat modeling, and contribute to the next generation of identity primitives including agent identity, JIT-native governance, and unified policy enforcement across workforce and customer IAM. As part of Palantir's best-in-class Information Security organization, you will research, architect, and scale solutions that help Palantir stay ahead of a dynamic identity threat landscape.
Get new lead infrastructure software engineer jobs by email
Daily job updates · Unsubscribe anytime