Jobs in United States

Lead Infrastructure Software Engineer in United States

2,434 active opportunities · Updated October 2026

Explore current lead infrastructure software engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

51/100

steady · 543 related jobs

Hiring trend

-77.4%

Job postings compared with the previous 30 days

Remote options

15.1%

Share of matching jobs listed as remote

Typical salary

$177.2K – $177.2K/yr

Based on 29 salary observations

O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -82%

About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. In partnership with leading cloud providers, hardware manufacturers, utilities, construction partners, and internal engineering organizations, we are delivering hyperscale AI campuses that power the next generation of frontier AI models. Infrastructure Delivery Operations sits at the center of this effort. Our team develops the operating model that connects infrastructure strategy, supply planning, manufacturing operations, and delivery into a single, integrated system that enables OpenAI to deploy AI infrastructure predictably at scale. We partner across Hardware Engineering, Network Engineering, Capacity Delivery, Hardware Operations, Security, Finance, Strategic Sourcing, and external infrastructure partners to create a single, integrated view of program health. Through governance, operational analytics, executive reporting, and scalable operating mechanisms, we enable leaders to proactively manage risk, optimize capacity, and deliver infrastructure predictably at Industrial Compute speed. About the Role We are seeking a Technical Program Manager, Infrastructure Delivery Operations to drive integrated strategy and delivery across OpenAI's rapidly expanding AI infrastructure portfolio. This role sits at the intersection of infrastructure strategy, New Product Introduction (NPI), supply planning, manufacturing operations, and infrastructure delivery. You will lead highly cross-functional programs spanning engineering, supply planning, manufacturing, logistics, construction, commissioning, and operations, ensuring technical and operational dependencies remain synchronized from planning through production readiness. Beyond driving program execution, you will leverage operational insights to improve capacity planning, infrastructure strategy, and deployment readiness. You will also help operationalize new technologies and suppliers by partnering w

AWSRestAgileAI
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this team? The GPU Clusters team builds and operates the superclusters that train Cohere’s frontier models. We sit at the intersection of hardware, distributed systems, and AI research. We work with cloud providers, researchers, and other infrastructure teams on problems few companies get to take on. As an Engineering Manager, you’ll lead a team of engineers who care deeply about GPU infrastructure. You’ll set technical direction, grow people, and help the company scale a rapidly growing compute footprint. As an Engineering Manager, you will: Hire, mentor, and grow a team of GPU infrastructure engineers , including performance, career development, and technical guidance on hard infrastructure problems Own the technical roadmap for the fleet: how we deploy, operate, and scale Kubernetes clusters, including workload scheduling, hardware fault detection, and performance Partner with researchers and ML engineers so the training and inference stack works well on new GPU architectures Work with cross-functional stakeholders such as Capacity, Finance, Legal, Security, and other infrastructure teams on planning, cost, compliance, an

KubernetesGitAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Role As a Director, Compute & Infrastructure FP&A, you will own and drive the monthly forecasting process for the Compute & Infrastructure org by partnering with various stakeholders across Finance, Accounting, Tax and Engineering. You will play a critical role in planning and forecasting the company’s largest and most complex cost center ( Compute & Infrastructure ). You will collaborate cross-functionally to develop long-range infrastructure investment plans, evaluate build vs. buy decisions, and ensure capital is deployed efficiently to support rapid growth. You will also provide strategic financial guidance through scenario modeling, ROI analysis, and performance tracking, enabling leadership to make high-stakes decisions under uncertainty. What You’ll Do Own compute financial planning & Forecasting. Build and manage consolidation models for GPU/CPU capacity, storage, networking, and data center investments. Translate infrastructure roadmaps into short- and long-term financial forecasts (LRP, annual planning) Coordinate closely with Corporate FP&A on timelines and process Present insights on a monthly basis to senior management. Drive infrastructure investment decisions. Evaluate build vs. buy, vendor vs. owned infrastructure, and capacity allocation tradeoffs. Develop frameworks for investment trade-offs to guide executive decision making. Build scalable tooling & reporting. Implement stakeholder-facing dashboards to track compute spend, utilization, and efficiency metrics. Improve visibility into unit economics (e.g., cost per training run, cost per inference, cost per customer). Drive forecasting accuracy & accountability. Lead budget vs. actual analysis for compute and infrastructure spend. Identify key cost drivers (utilization, pricing, efficiency gains) and reduce forecast variance. Support close & financial reporting. Partner with Accounting to ensure accurate classification of infrastructure spend (OpEx vs C

SQLAWSAzureGCP
O
📍 United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role OpenAI is seeking a Principal Security Engineer to join our Infrastructure Security (InfraSec) team. InfraSec protects the foundations of OpenAI’s research and production environments, spanning GPU supercomputing clusters, multi-cloud infrastructure, datacenters, networking, storage, and the critical services that power our frontier AI models. Our charter includes securing everything from bare-metal hardware and firmware, to Kubernetes clusters and service meshes, to data storage and access pathways for highly sensitive model weights and user data. As a principal engineer, you will set technical direction and drive execution on high-impact infrastructure security programs, partnering across various orgs at OpenAI to deliver durable controls that raise the security bar at OpenAI scale. In this role, you will: Own end-to-end security outcomes for one or more critical infrastructure areas, including multi-quarter strategy, roadmap, and delivery. Design and build security controls across diverse layers (e.g., physical hardware, firmware/BMC, OS, Kubernetes, networks, and CI/CD) to defend against sophisticated adversaries and insider threats. Lead cross-functional programs to deploy security enhancements and control changes across broad-scale infrastructure, balancing security guarantees with reliability and velocity. Take a generalist approach to building security controls, balancing a mix of security expertise and broad technical skillsets

AWSAzureKubernetesCI/CD
P
📍 New York, United States
✓ High-confidence listingCompany trend +671.4%

$230.9K – $384.8K/yr

Quick readStrong listing-quality and freshness signals

ROLE SUMMARY Pfizer Commercial Oncology is introducing the world to the next era of cancer care. With a growing portfolio of novel therapies, industry-leading R&D, and a goal of delivering eight breakthroughs by 2030 across major cancer types, we're translating cutting-edge science into market-shaping impact. Here, you'll partner with exceptional colleagues across scientific, medical, and manufacturing teams, backed by advanced digital and AI-enabled infrastructure and the authority to accelerate medicines from discovery to delivery. Guided by our values of courage, excellence, equity, and joy, you'll have the opportunity to stretch your skills and build a career that evolves with you—across teams, roles, and the Pfizer enterprise. Join us to make history — for patients, for their families, for the future. The US Precision Medicine Thoracic Franchise is entering its most consequential — and most complex — growth chapter. Our in-line thoracic portfolio spans two biomarker-defined populations: LORBRENA (lorlatinib) for adults with ALK-positive metastatic non-small cell lung cancer (NSCLC), and the BRAFTOVI + MEKTOVI (encorafenib + binimetinib) regimen for BRAF V600E-mutant metastatic NSCLC. Both sit inside a molecularly driven treatment landscape where the right patient is identified by a biomarker signature. Winning here means mastering precision targeting: reaching narrow, high-value, genomically defined patient populations through the exact clinicians who test for and treat them. This is a people-management leadership role that demands a strong enterprise leader — one fluent in both commercial strategy and AI-enabled execution, and able to think and operate at the enterprise level. The colleague who leads this team will set a multi-year strategic vision, make thoughtful trade-off and resource-prioritization decisions across two brands, and influence senior leadership across the wider organizati

Machine LearningAIExcelRecruitment
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -82%

$216K – $240K/yr

Quick readStrong listing-quality and freshness signals

About the Team OpenAI Finance is responsible for ensuring the organization is set up for success in pursuit of its mission. The Technical Accounting team plays a crucial role in helping OpenAI navigate complex, judgmental, and rapidly evolving accounting matters with rigor and clarity. We aim to bring both technical excellence and strong business partnership to some of the most novel accounting questions in the industry. About the Role As Senior Manager, Technical Accounting, Compute Infrastructure, you will lead the evaluation, documentation, and operationalization of complex accounting matters related to OpenAI's compute infrastructure, strategic investments, and other non-routine business activities.. This role sits at the intersection of U.S. GAAP technical accounting, infrastructure strategy, financial reporting, controls, and cross-functional execution. Key areas may include cloud compute arrangements, data center and colocation arrangements, lease accounting under ASC 842, power purchase agreements, strategic investments, consolidation evaluations under ASC 810, financial instruments, and other emerging or non-standard arrangements. This role is based in San Francisco, CA or remote. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead technical accounting analysis for complex, judgmental, and non-routine transactions under U.S. GAAP. Evaluate accounting implications for compute infrastructure arrangements, including cloud compute, data center, colocation, lease, PPA, infrastructure procurement, and related commercial arrangements. Partner with Controllership, Tax, Legal, FP&A, Procurement, Infrastructure, and other cross-functional teams to assess the accounting implications of new products, commercial arrangements, strategic transactions, and business initiatives. Prepare and review technical accounting memoranda, position papers, and other auditor-ready documentation.

AWSRestAIGo
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking an experienced AI Systems Reliability Engineer to help define, build, and maintain the infrastructure and processes that ensure the reliability, scalability, and performance of Postman’s AI-powered API and agentic systems in production. This role focuses on monitoring, availability, incident response, and automation to support AI services and tools trusted by millions of developers globally. What You’ll Do Develop and manage reliability metrics (SLOs) for AI-driven API services and agentic AI platform features Implement comprehensive observability and monitoring systems for real-time performance and fault detection Design and drive automated failover, recovery, and incident response strategies for high-availability AI infrastructure Optimize resource utilization, particularly GPU/accelerator efficiency, ensuring cost-effective AI system operation Collaborate closely with engineering, platform, and product teams to align reliability efforts with broader organizational goals Lead efforts to build internal tooling and automation focused on AI system stability and operational excellence Drive continuo

AIGoRustDevOps
P
📍 San Francisco, CA, United States· Remote
✓ High-confidence listingCompany trend -86.3%
Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . The Team Pinterest's Ads Delivery Infrastructure team builds and operates the real-time serving, indexing, retrieval, and budgeting systems that power Shopping and Ads at Pinterest. We're in the middle of a significant modernization effort: unifying our catalog and index platforms to scale to billions of items, moving core services onto our next-generation compute platform, and re-architecting our budgeting systems for real-time accuracy. We work closely with engineering, product, and monetization teams to keep Pinterest's advertising systems fast, reliable, and ready to scale. What You'll Do: Lead or support programs that scale Pinterest's Shopping and Ads catalog, indexing, and retrieval systems to handle billions of items, partnering with engineering leads across catalog, search infrastructure, and retrieval. Drive cross-team execution on inf

C
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -79.2%

£215K – £260K/yr

Quick readStrong listing-quality and freshness signals

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About the Role Cohere is seeking a Global Public Policy Manager to lead policy engagement on compute infrastructure, export controls, AI competitiveness, and sovereign AI strategies. Governments increasingly view AI as critical national infrastructure and are investing heavily in compute capacity, energy resources, and domestic AI ecosystems. This role will help position Cohere as a trusted partner in emerging discussions around AI infrastructure, national competitiveness, and sovereign AI deployment. Key Responsibilities Monitor and analyze developments related to AI infrastructure, data centers, energy policy, semiconductor policy, export controls, and national AI strategies. Develop policy positions on sovereign AI, compute access, digital sovereignty, and AI competitiveness. Support engagement with governments developing AI infrastructure investment programs and national AI initiatives. Collaborate with commercial, product, and corporate development teams on strategic opportunities involving public-private partnerships. Represent Cohere in policy discussions related to AI infrastructure, energy requirements, and technology c

GitRestAIGo
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $128K/yr

Quick readStrong listing-quality and freshness signals

We’re looking for an experienced sourcing leader to join the Datadog Procurement team and help grow the Strategic Sourcing group. Make an impact by partnering closely with senior technology and business leaders, owning our fastest-growing AI, neocloud, and inference spend end-to-end, and continuing to prove the value that Strategic Sourcing brings to the organization. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the AI spend category end-to-end, covering foundation-model and AI APIs, AI development and productivity tools, and AI-enabled SaaS (with neocloud and inference compute as an emerging area), while developing and executing category strategies aligned to business objectives Champion AI and automation adoption across the sourcing function by identifying tools (e.g. Claude, ChatGPT) and building repeatable workflows that make sourcing faster, smarter, and more scalable Lead high-value AI vendor negotiations spanning foundation-model and API agreements, AI development and productivity tools, and AI SaaS, structuring seat- and consumption-based pricing across net-new purchases and strategic renewals, and building neocloud and GPU-capacity capability as the category grows Partner with Engineering and Finance to turn architectural and consumption trade-offs into commercial business cases, forecasts, and savings targets Build pricing and consumption models and apply FinOps discipline to quantify buying scenarios and uncover savings across a fast-moving spend base Work alongside executive and senior technology leadership as a trusted advisor on AI and neocloud investment decisions, influencing strategy and commercial trade-offs at the leadership level Track category KPIs (savings, pipeline, cycle time), monitor AI market, vendor, and pricing trends,

H
📍 Boston, Massachusetts, United States· Full-time
✓ High-confidence listing

From $119.6K/yr

Quick readStrong listing-quality and freshness signals

We take play seriously. We’re looking for curious adventurers ready to find their party, fueled by imagination and drive to build what’s never been built before. At Hasbro and Wizards of the Coast, you’ll collaborate with passionate teams to reimagine our iconic brands and create experiences that spark joy, connection, and community through the magic of play. This is your chance to shape legendary play that lasts a lifetime. The Senior Network Engineer leads Wizards of the Coast's enterprise network — datacenters, corporate offices, studios, and AWS cloud. Our stack runs on a Juniper/Mist campus fabric, a Palo Alto Networks security edge, and cloud-native AWS connectivity. You'll set technical direction, drive complex initiatives end to end, and mentor the broader team. Come help us build the future of network operations. What You'll Do: Own the architecture, build, and roadmap for our Juniper/Mist campus and branch infrastructure, and lead its evolution across sites and business units. Lead the Palo Alto Networks security stack, from policy architecture to secure-by-design standards across teams. Own end-to-end AWS cloud network implementation — VPC, Transit Gateway, Direct Connect/VPN, Route 53 — and hybrid connectivity, including BGP, OSPF, and SD-WAN traffic engineering at scale. Drive automation and AI adoption across network operations, from config deployment to AI-powered monitoring, observability, and root-cause analysis. Be the go-to for critical issues, and mentor less-experienced engineers through code review, troubleshooting, and skill-building. What You'll Bring: 10+ years in enterprise network engineering — routing, switching, wireless — with deep expertise in Juniper (EX/QFX/SRX, Mist) and/or Cisco (Catalyst, Nexus), plus hands-on work with intelligent ops tools like Mist/Marvis at scale. Expert-level BGP, OSPF, SD-WAN, and load balancing chops. You've built these solutions from scratch, not just maintained them! Extensive Palo Alto Net

AWSAIGoTerraform
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%
Quick readStrong listing-quality and freshness signals

Datadog's integrations are the connective tissue between our platform and the technologies our customers run in the real world. As a Sr. PM on the Agent Integrations team, you will own the vision, prioritization, and execution for 100+ integrations that run directly inside the Datadog Agent from foundational infrastructure (MySQL, Kafka, Kubernetes) to the rapidly growing landscape of self-hosted AI and on-premise enterprise technologies. This is a high-impact, breadth-first role at the intersection of infrastructure observability and the frontier of AI-native workloads. At Datadog, we place value in our office culture; the relationships it builds, the creativity it brings, and the collaboration of being together. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own the Agent Integrations roadmap. Determine which new integrations to build and which existing ones to improve, balancing customer demand, business impact, and engineering capacity across a catalog of 100+ technologies. Drive the expanding AI integration surface. Lead product strategy for self-hosted AI workloads, including LLM inference frameworks (e.g., Hugging Face TGI, BentoML), AI agents, MCP servers, and model orchestration tools, so Datadog customers can monitor every layer of their AI stack. Expand on-prem and hybrid coverage. Prioritize and execute new integrations for on-prem technologies including storage systems, HPC schedulers, network devices, and legacy enterprise platforms where customers run critical workloads. Build observability for ERP systems. Define and drive Datadog's strategy for monitoring enterprise ERP platforms (SAP, Oracle EBS/Fusion, Microsoft Dynamics) covering performance, job execution health, and integration layer telemetry so enterprise customers can observe their ERP stack alongside the rest of their infrastructure. Analyze adoption and customer feedback at scale. Use data from multiple sources to

SQLPostgreSQLMySQLMongoDB
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

Build the infrastructure that keeps every NVIDIA chip aligned from first spec to final shipment. NVIDIA's Silicon Co-Design Group sits at the convergence of architecture, silicon, systems, and manufacturing. The System–Manufacturing Architecture (SMAC) team coordinates between system specifications and manufacturing test specifications from pre-silicon POR through production release across GPU, SoC, and CPU programs. When that alignment drifts, silicon faces the consequences: escapes, yield loss, and performance loss. We're hiring a Senior Manufacturing & System Co-Design Workflow Engineer to lead the methodology and infrastructure that maintains holistic, systematic alignment, at scale across the full portfolio. The strongest candidates in this role design the workflow before being asked to fix a program, and build the checks and automation that confirm alignment holds long after they've moved on to the next problem. What you’ll be doing: SMAC Workflow Methodology: Define manufacturing spec types, including schema and semantics, derived from system PORs and features. Own the methodology that governs how specification work gets structured, versioned, and validated across the program lifecycle. Production Python Pipelines & Automated Checks: Develop production-grade Python pipelines and automated checks that catch specification drift between system POR and manufacturing test programs ,ATE, SLT, BLT, L10+, before silicon exposes the discrepancy. The goal is that misalignments surface in the workflow, not on the tester. E2E Program Integration & TPM Attestation: Wire SMAC work into the end-to-end program spine, milestones, gates, and artifacts, and define explicit TPM-driven attestation when checks lag. Alignment can't be assumed; it must be proven at every stage. Agent-Ready Tooling & CI Infrastructure: Integrate tooling into an agent-ready

G
📍 Austin, Texas, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About the job Lead the team that proves our AI silicon performs reliably before it reaches customers. You will build and lead Graphcore's characterisation capability for next generation silicon and system platforms. Your work will help ensure our products perform consistently across real world conditions. You will define bring up and characterisation strategies, lead technical execution, and shape the lab infrastructure needed for success. You'll work across silicon, hardware, manufacturing, architecture and product teams to solve complex engineering challenges. This is a hands on leadership role with the opportunity to influence both product design and how Graphcore validates future AI systems. The team and culture This is a newly formed team within Manufacturing Operations. You'll have the opportunity to establish how the team works while building strong partnerships across engineering and operations. Day to day, you'll work closely with architecture, silicon, hardware, production test and product teams. Decisions are driven by data, technical evidence and close collaboration across disciplines. We value ownership and clear communication. You'll be trusted to lead technical direction, remove blockers and help teams make progress with confidence. What we're looking for Essential Proven track record of delivering complex technical projects as an individual contributor, manager, or project manager, with the ability to work independently and drive execution. Strong expertise in silicon digital device design, bring-up, characterisation, and silicon process technologies, with an understanding of their impact on transistor- and system-level performance. In-depth knowledge of high-performance processors, system-on-chip (SoC) architectures, and high-speed digital interfaces such as PCIe, Ethernet, and DDR. Experience with measurement automation, data analysis, and scripting/coding to develop automated test and analysis workflows, with familiarity of ATE systems and t

GitAIGoRust
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As the Head of IT at Baseten, you will build, scale, and secure our internal technology function to support our rapid growth. Reporting to our Chief Information Security Officer, you will lead and mentor a team of 5+ IT engineers, leading the charge to transition Baseten from startup-era IT to a highly automated, enterprise-ready IT organization. You will take full ownership of corporate IT infrastructure, Helpdesk operations, corporate identity management, device lifecycles, and vendor procurement. As we scale to support the world’s most dynamic AI companies, you will ensure our internal systems scale seamlessly with our headcount, providing a secure, frictionless, and world-class technology experience for all Baseten employees. RESPONSIBILITIES Team Leadership: Manage, mentor, and grow a team of IT engineers, fostering a high-performance culture focused on technical excellence and end-user satisfaction. Helpdesk Operational Excellence: Build a fast-response support function by establishing clear response SLAs, tracking employee satisfaction metrics, and formalizing on-call and incident response processes. Zero-Touch Automation: Architect and implement automated employee onboarding, offboarding, and role-based access changes through deep integrations across HRIS, MDM, and IAM systems. SaaS Management & Procurement: Establish comprehensive SaaS management processes to eliminate shadow IT, automate access w

Machine LearningAIGoRust

Related career options

Similar roles with stronger pay

Client Service Associate

Demand 46/100 · 8 jobs

$840K – $840K/yr

Salary →

$840K – $840K/yr

Salary →
Director of Product

Demand 43/100 · 6 jobs

$382.5K – $382.5K/yr

Salary →
Physical Design Engineer

Demand 43/100 · 8 jobs

$300K – $300K/yr

Salary →
Sr. Engineer

Demand 42/100 · 7 jobs

$300K – $300K/yr

Salary →
Senior Director

Demand 43/100 · 22 jobs

$278.9K – $278.9K/yr

Salary →
🔔

Get new lead infrastructure software engineer jobs in United States by email

Daily job updates · Unsubscribe anytime