The mission of The New York Times is to seek the truth and help people understand the world. That means independent journalism is at the heart of all we do as a company. It’s why we have a world-renowned newsroom that sends journalists to report on the ground from nearly 160 countries. It’s why we focus deeply on how our readers will experience our journalism, from print to audio to a world-class digital and app destination. And it’s why our business strategy centers on making journalism so good that it’s worth paying for. About the Role, Mission or Department Overview As a Senior Engineer, Marketing, you'll be embedded with the Marketing team, providing us with technical leadership, consulting, and systems thinking. You'll assemble the orchestration systems, integrations, and AI capabilities that help teams work faster, and in more data‑driven ways. You will lead end‑to‑end design, implementation, and management of the marketing campaign lifecycle system, integrations, internal tools, and learning/experimentation infrastructure. You will report to our Director, Advertising Systems. Responsibilities: You will write high-quality, secure, and well-tested code, contributing to standards and documentation for Marketing infrastructure, data models, and tools You will build the campaign lifecycle orchestration backbone You will build integrations and internal tools that connect project management, collaboration, creative, media, and email/lifecycle platforms You will work with Marketing partners as a technical advisor, translating needs into designs You will integrate AI agents and automations (including LLM-powered flows) into Marketing workflows according to company GenAI guidance You will design and operate data pipelines and models that ingest marketing and performance data into a data layer, supporting experimentation and analytics workflows You will deliver abstractions and APIs that ensure AI systems and our users to create campaign wrap-ups, insights, next-t
Jobs in United States
Lead Software Engineer Infrastructure in United States
2,434 active opportunities · Updated October 2026
Showing
15 jobs
Explore current lead software engineer infrastructure jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
From $243.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Enterprise Security Engineer, you will play a critical role in executing Roblox’s Enterprise Security Strategy. You will design, deploy, and manage security solutions to protect Roblox’s corporate infrastructure and ensure secure, compliant operations across the organization. Working closely with Corporate Engineering and Trust & Safety teams, you will translate business requirements into robust security implementations that enable secure productivity while mitigating risk. You will be reporting directly to the Senior Manager of Enterprise Security Engineering. You'll partner with security professionals across the Information Security organization, and work cross-functionally with teams throughout Roblox to drive security initiatives that scale with our business. You will: Evaluate and implement security technologies and vendor solutions to ensure alignment with enterprise security requirements, compliance standards, and overall risk management strategy Lead and drive initiatives across core security domains, including Endpoint Security, SaaS Security, Identity & Access Management (IAM), Agentic AI Governance, and Supply Chain Security. Collaborate closely with IT, engin
From $293.8K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Enterprise Security Engineer, you will advance Roblox’s Enterprise Security strategy by shaping and evolving security architecture in alignment with business objectives. You will lead the design, deployment, and governance of security solutions that safeguard Roblox’s corporate infrastructure while enabling scalable, secure operations. Partnering cross-functionally with Corporate Engineering and Trust & Safety, you will translate organizational priorities into resilient security capabilities that balance risk, compliance, and productivity. You will join the Platform, Enterprise, and Application Security group, reporting directly to the Senior Manager of Enterprise Security Engineering. You'll partner with security professionals across the InfoSec team, and work cross-functionally with teams throughout Roblox to drive security initiatives that scale with our business. You will: Define and maintain enterprise-wide security standards and principles that guide how security is implemented across business workflows, ensuring consistency, scalability, and alignment with organizational risk posture. Lead and drive initiatives across core security domains, including Endpoint Secur
From $397.5K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. At Roblox , we’re building the tools and platform that empower a global community of creators and developers to build immersive experiences and a dynamic virtual economy. Our Economy ML team sits at the heart of this mission, delivering scalable machine learning systems that power personalization, pricing, search, and content understanding across all Economy surfaces: Marketplace, Developer Monetization, Payments, and Avatar. We’re looking for a Distinguished Engineer/Technical Director to lead the strategy and technical direction for ML systems , with a focus on large-scale recommendations, infrastructure, and emerging Generative AI applications. You’ll help build the systems that support retrieval, ranking, generative modeling, and LLM-powered personalization, all at massive scale. This role requires deep systems thinking, hands-on ML expertise, and a vision for how traditional ML and GenAI come together to power the future of the Roblox economy. Why Roblox for ML Systems AI/ML is a top company priority , with long-term investment. Real-world scale : Power millions of daily economic interactions across ranking, pricing, fraud, and search. Full-system ownership : Build and optimize end-to-
From $196.8K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Security Engineer on the Detection and Response (D&R) team at Roblox, you’ll protect our user community alongside the underlying platform infrastructure. You’ll design high-fidelity detections, engineer security data platforms, and respond alongside the team during incidents. This is a hybrid in-office role in San Mateo. You Will: Deliver robust D&R capabilities: Engineer high-fidelity detections end-to-end. Lead partners through threat modeling and logging, to deploying actionable alerts, while keeping false positives low. Build security data pipelines: Develop security data pipelines and actively contribute to internal software and data platforms, collaborating across engineering teams. Ensure service reliability: Participate in an on-call rotation to keep detection and response services healthy. Embody security culture: Serve as a trusted security partner across Roblox, helping protect our community and enterprise while fostering a culture grounded in trust, ownership, and shared responsibility. You Have: 3+ years of experience in Security Data Engineering: You have built services that are efficient, reliable, and scalable using programming languages like Golang or Py
We are investing in agentic AI and need a Senior AI Engineer to lead the design and delivery of these systems. This is a foundational hire: you will own both the agent-facing workstreams — pipelines, orchestration, conversational interfaces — and the underlying context layer that makes them reliable, including memory management, knowledge graph integration, and retrieval infrastructure. You will work closely with data engineers, project leads, and client stakeholders, and play a key role in shaping how Lynx builds and ships AI solutions at scale. What This Involves: Lead the architecture and delivery of agentic AI systems end-to-end: agents, orchestration, tool use, and multi-step reasoning workflows. Own the context layer: design and implement memory architectures (episodic, semantic, working memory) and integrate GraphRAG and knowledge graph retrieval into agentic pipelines. Build robust RAG systems — including vector retrieval, graph traversal, and hybrid search — and ensure retrieval quality through evaluation frameworks. Translate client requirements into technical designs, presenting approaches and trade-offs to both technical and non-technical stakeholders. Define standards and reusable patterns for agentic AI development that other engineers at Lynx can build on. Set up observability, evaluation, and monitoring pipelines to ensure AI systems perform correctly in production. Requirements: 5–8 years of software or ML engineering experience, with at least 2–3 years building LLM-based or agentic AI systems in production. Deep hands-on experience with agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI, or similar) and LLM APIs (OpenAI, Anthropic, etc.). Strong understanding of agent design patterns: ReAct, planning loops, tool use, multi-agent coordination, and memory architectures. Practical experience with GraphRAG or knowledge graph-based retrieval (e.g., Neo4j, Microsoft GraphRAG) and vector databases (Pinecone, Weaviate, Qdrant, etc.). Proficiency in
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and support next-generation infrastructure design. About the Role We are seeking an Performance Modeling Engineer to support the development and application of modeling tools used to evaluate AI system performance and inform architectural decisions. In this role, you will partner closely with Senior Performance Modeling Engineers and the Performance Modeling Lead to analyze system behavior, run simulations and analytical models, and help evaluate tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks while developing a strong foundation in system architecture and AI infrastructure. This role is ideal for early-career engineers with 1–2 years of experience in software engineering, systems analysis, or performance modeling who are excited to grow in large-scale infrastructure and hardware/software systems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Support the development and maintenance of performance modeling tools and frameworks Assist in building models to evaluate system behavior across compute, memory, networking, and interconnect subsystems Help analyze distributed system scaling behavior and identify performance bottlenecks Run simulations and analytical models to support architecture and infrastructure decisions Partner with senior engineers to evaluate design tradeoffs across hardware and system components Interpret modeling outputs and help translate findings into clear recommendations Vali
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Security Systems Engineer at Palantir, you are responsible for implementing, designing, and maintaining the physical security systems that ensure the protection of Palantir’s people, assets, intellectual property, and reputation. You lead research and evaluations for the latest technology, security infrastructure, policies, procedures, systems, and applications, looking for easy-to-use capabilities that reduce button clicks and ultimately give time back to the operator. You will be collaborating with technical teams to develop and maintain end-to-end project plans and ensure on-time delivery. Your technical expertise is second only to your integrity and genuine passion for security and technology. As a member of the Global Security and Investigations team, you work with security professionals, architects, lawyers, and software and network engineers to deliver a safe and secure environment for your fellow employees. Palantir is a global company, so you'll lead physical security projects all over the world. You'll work hard to break down barriers between teams by communicating your knowledge and goals effectively across time zones and other physical and virtual barriers. *Please note this role is NOT for cyber security or information security positions* Core Responsibilities Design and oversee the technical aspects of your projects from conception to full deployment. Systems naturally break and need maintenance. You'll work with vendors and end users to find the best solutions for a problem, in the shortest possible time. Leverage both system and software engineering skills in order to address the needs of all teams within Globa
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for Forward Deployed Engineers on our engineering team who want to work at the intersection of deep infrastructure work and direct customer impact. As an FDE, you'll partner with leading AI companies and foundation labs on cloud architecture, networking, storage, containerization, sandboxing, and more — helping them design and ship production infrastructure on Modal's platform. The FDE team today includes world-class software engineers, computational scientists, ML engineers, and former founders. We're looking for people with strong engineering fundamentals, deep curiosity across the infrastructure stack, and energy for working directly with customers on hard problems. You will: Work hands-on with companies like Suno, Lovable, Cognition, and Meta to architect and deploy massive-scale production workloads on Modal Lead technical discovery and architect
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with architecture, infrastructure, and vendor teams to evaluate system performance and guide critical design decisions. Our team focuses on building and applying performance modeling frameworks to understand system behavior, quantify tradeoffs, and inform next-generation infrastructure design. About the Role We are seeking Performance Modeling Engineers to develop and apply modeling tools that evaluate AI system performance and inform architectural decisions. In this role, you will work closely with the Performance Modeling Lead and partner teams to analyze system behavior, run simulations or analytical models, and help quantify tradeoffs across compute, memory, networking, and storage. You will contribute to building modeling frameworks and applying them to real-world questions that impact system design and vendor decisions. This role is well-suited for engineers with strong software or modeling backgrounds who are interested in developing deeper expertise in system architecture and AI infrastructure. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Develop and maintain performance modeling tools and frameworks. Build models to evaluate system behavior across: compute, memory, and interconnect subsystems distributed system scaling and bottlenecks. Run simulations and analytical models to support architectural tradeoff analysis. Collaborate with performance modeling lead and system architects to answer forward-looking design questions. Analyze and interpret modeling outputs, translating results into actionable insights. Validate models against real system measurements and workload behavior. Contribute to improving modeling fidelity, usability, and scalability. Qualifications Strong software engineeri
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary We are seeking an accomplished Principal Cloud Storage Engineer to lead the design, engineering, and evolution of our private cloud storage platforms. This role will focus on large-scale storage architecture, data protection, cyber recovery, and resiliency technologies across complex enterprise environments. The ideal candidate will combine deep technical expertise in storage systems with strong leadership, architectural vision, and the ability to influence technical direction across the organization. Key Responsibilities Architect and engineer enterprise storage platforms that ensure data integrity, availability, security, and disaster recovery readiness Design and implement end-to-end storage solutions, including Software Defined Storage, SAN, NAS, and object storage across private cloud and data center environments Drive strategic technology decisions by evaluating emerging products, tools, and standards supporting storage, data protection, cloud, and compute platforms Lead infrastructure initiatives involving storage modernization, data protection, cyber recovery, data migration, and resilience engineering Develop and execute enterprise strategies for backup, recovery, cyber vaulting, and business continuity Create and maintain comprehensive documentation of storage architectures, configurations, policies, and operation
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role The Security Engineering team helps make Ramp the most secure place for our customers to collect, manage, and put to work their business’ financial information Our work centers in three areas: Ramp builds products with an eye for security Ramp detects and responds to threats before they cause harm Security powers Ramp’s growth Check out our Engineering Blog for more on our tech stack, mission and values! What You’ll Do Drive our cloud security roadmap: review our cloud deployments to identify opportunities for improvement Design and build security-focused infrastructure primitives and integrate them into our existing products and development processes Lead remediation of prioritized issues across our technology stack Partner with infrastructure, data, and devops teams to design and deploy solutions that are inherently secure What You Need Minimum 5 years of experience building software Minimum 3 years of experience building in AWS (with Terraform) A strong sense of ownership: you need to drive projects from inception to scaling it in
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role As an Enterprise/Strategic Field Engineer (L5) , you'll be the technical cornerstone for Replit's largest and most strategic accounts. This is a hybrid role: high-impact pre-sales (closing complex technical evaluations) and post-sales (driving adoption, expansion, and retention). You'll own the end-to-end technical relationship—from pre-sales architecture discussions through multi-year expansion—ensuring our enterprise customers don't just use Replit, but become Replit-powered companies. You'll partner with Account Executives and Account Managers in a high-accountability Pod structure . This is not a reactive support role—this is a proactive, strategic technical leader who identifies blockers before they become problems, champions new use cases, and directly influences $5M+ in annual recurring revenue. In this role you will: Pre-Sales Strategic Technical Discovery: When you are pulled into complex deals, you join as the expert closer. You run deep discovery on their stack and constraints, then design the winning technical strategy. Proof of Value (POV) & Live Building: You build live, functional applications on the fly during executive meetings to prove immediate value and technical feasibility to VPs and C-suite stakeholders. Context & Connectivity (MCP): You write and deploy Model Context Protocol (MCP) servers to securely connect Replit Agents to customer-specific data, making Replit the central hub for their internal development. Enterprise Governance: You own the "Guardrails" mission. You configure workspace policies and AI governance templates that solve for data safety, compliance, and CISO approval. Infrastructure Strategy: You lead deep-dive reviews for Single-Tenant/VPC deployments, ens
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role As an Enterprise/Strategic Field Engineer (L5) , you'll be the technical cornerstone for Replit's largest and most strategic accounts. This is a hybrid role: high-impact pre-sales (closing complex technical evaluations) and post-sales (driving adoption, expansion, and retention). You'll own the end-to-end technical relationship—from pre-sales architecture discussions through multi-year expansion—ensuring our enterprise customers don't just use Replit, but become Replit-powered companies. You'll partner with Account Executives and Account Managers in a high-accountability Pod structure . This is not a reactive support role—this is a proactive, strategic technical leader who identifies blockers before they become problems, champions new use cases, and directly influences $5M+ in annual recurring revenue. In this role you will: Pre-Sales Strategic Technical Discovery: When you are pulled into complex deals, you join as the expert closer. You run deep discovery on their stack and constraints, then design the winning technical strategy. Proof of Value (POV) & Live Building: You build live, functional applications on the fly during executive meetings to prove immediate value and technical feasibility to VPs and C-suite stakeholders. Context & Connectivity (MCP): You write and deploy Model Context Protocol (MCP) servers to securely connect Replit Agents to customer-specific data, making Replit the central hub for their internal development. Enterprise Governance: You own the "Guardrails" mission. You configure workspace policies and AI governance templates that solve for data safety, compliance, and CISO approval. Infrastructure Strategy: You lead deep-dive reviews for Single-Tenant/VPC deployments, ens
About the Team We bring OpenAI's technology to the world through products like ChatGPT and the OpenAI API. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role OpenAI is looking for an experienced Performance Engineer to help us scale the performance, reliability, and efficiency of our systems. In this role, you'll apply deep technical expertise to optimize infrastructure and application-level performance across mission-critical products like ChatGPT and our developer API. You’ll work cross-functionally with teams building core services, training models, and developing real-time user experiences to push our latency, throughput, and cost-efficiency to the next level. We are looking for engineers who thrive in ambiguous environments, value deep systems understanding, and are motivated by delivering measurable impact. This is a highly technical, individual contributor role focused on root-cause analysis, profiling, instrumentation, and architecture-level performance improvements across our stack. In this role, you will: Analyze and optimize performance across application, middleware, runtime, and infrastructure layers—networking, storage, Python runtime, GPU utilization, and beyond. Develop tooling and metrics that provide deep observability into system performance. Collaborate closely with infra, platform, training, and product teams to identify key performance goals and drive systemic improvements. Influence architecture and design decisions to prioritize latency, throughput, and efficiency at scale. Lead investigations into high-impact performance regressions or scalability issues in production. Drive performance testing strategies and help define SLAs/SLOs around latency and throughput for critical systems. You might thrive in this role if you: Have 7+ years of experience in software engineering with a strong tr
Higher-paying openings
Jobs with higher listed pay
Staff Software Engineer - Fern
Postman · New York, California, United States
Staff Software Engineer, Business Platform
Postman · San Francisco, California, United States
Principal Software Engineer
Roblox · San Mateo, CA, United States
Principal Software Engineer, Game Safety
Roblox · San Mateo, CA, United States
Staff Software Engineer- Codegen
Postman · Austin, Texas, United States
Sr. Staff Software Engineer, Merchants
Pinterest · San Francisco, CA, US
Related career options
Similar roles with stronger pay
Demand 46/100 · 8 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 5 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 6 jobs
$382.5K – $382.5K/yr
Salary →Demand 43/100 · 8 jobs
$300K – $300K/yr
Salary →Demand 42/100 · 7 jobs
$300K – $300K/yr
Salary →Demand 43/100 · 22 jobs
$278.9K – $278.9K/yr
Salary →Other cities to consider
More places hiring for this role
Get new lead software engineer infrastructure jobs in United States by email
Daily job updates · Unsubscribe anytime