Clear all

Jobiba hiring network

Technical Program Manager Infrastructure Jobs

10,000 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current technical program manager infrastructure jobs. Use filters to narrow by work mode, employment type, experience and date posted.

F
Figma
📍 San Francisco, CA • New York, NY • United States• Full-time
1mo ago

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us! The mission of the Engineering TPM team is to drive Figma's most important cross-company engineering efforts, and we are looking for a Technical Program Manager (TPM) to partner with our Infrastructure team. The TPM provides oversight of the most important efforts that require coordinated technical execution across the Org to succeed. This is a role focused on enabling Figma's infrastructure teams to scale, improve performance, and deliver on critical projects. These large-scale efforts will involve collaboration across numerous backend, infrastructure, and security teams and cross-functional stakeholders, prioritization, decision-making, tracking execution, and driving operational excellence. We're looking for someone that can work in a TPM greenspace environment and is passionate about people, technology, and program management. Progress over process is our mantra. This is a full-time role that can be held from one of our US hubs or remotely in the United States. What you'll do at Figma: Lead the execution, coordination, and risk management of Figma's infrastructure projects, ensuring seamless integration with minimal performance impact Drive key infrastructure initiatives, including reliability, storage, distributed systems, cloud-native perfo

redisawsgcp
View job →
O
28 days ago

Technical Program Manager – Applied Infrastructure About the Team The Applied team safely brings OpenAI’s technology to the world, powering products like ChatGPT, and the APIs for GPT and more. Behind these products is a complex and rapidly evolving infrastructure platform that enables scale, performance, and safety. The Applied Infrastructure TPM team partners across engineering to lead foundational programs that ensure OpenAI’s infrastructure can meet current and future demand. About the Role We’re looking for a seasoned Technical Program Manager to drive critical infrastructure programs across the Applied organization. This TPM will focus on cross-cutting initiatives such as general compute capacity planning, process transformation, cost and quota attribution and optimization, and coordination across infrastructure and product stakeholders. There will also be focus on evolving OpenAI’s infrastructure to support growth, scale and new products. This work is core to how OpenAI manages and grows its infrastructure footprint in a disciplined, scalable way. Location: San Francisco, CA (Hybrid – 3 days/week in-office) In this role, you will: Serve as the DRI for complex infrastructure programs spanning CPU planning, orchestration, and other resource management domains (e.g. networking, storage). Build and operationalize systems to capture demand signals, model future capacity needs, and align infrastructure planning across internal teams and partners external to the company. Partner closely with Infrastructure, Product and Finance teams to forecast infrastructure usage patterns and ensure supply/demand alignment. Lead cost attribution and quota enforcement programs to promote stability and ensure equitable access to resources across teams. Drive simplification and standardization of infrastructure tooling and processes across Applied and Infra organizations. Drive cross functional programs to evolve our infrastructure to support new growth and scale Work with external v

awsazurerest
View job →
S
Stripe
📍 San Francisco• Full-time• Remote
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The infrastructure teams provide efficient and optimized infrastructure for Stripe to build secure, reliable, and differentiated products, while enabling Stripe developers to achieve their highest potential. Stripe makes it easy for any developer to access and manage the capabilities of the financial system while maintaining the least regulatory friction. We work to enable developers to have the most productive results of their entire career from the very first days they join Stripe through years of developing new systems and products. What you'll do As a Technical Program Manager in Infrastructure, you'll drive programs that span multiple Stripe engineering organizations with a focus on improving the internal platforms that power all of our products. In partnership with engineering and product management leaders, you're responsible for planning, comms, and steering execution of large-scale technical programs that solve complex problems and enable product engineering teams across Stripe. You'll deliver outstanding results by implementing solutions that scale to the entire company, minimize disruption to product teams, and are aligned with other engineering efforts. You'll work closely with Service Infrastructure, which enables engineering teams at Stripe to build, ship, and operate products that are efficient, reliable, and performant. They are responsible for the frameworks, async platforms, and tooling used to write and operate all Stripe

REMOTEaigo
View job →

Location Details: Canada - BC or ON (remote) At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team GoDaddy - Global Production Engineering looks after GoDaddy's global infrastructure, in the cloud and on-premises. We are hiring an experienced Technical Program Manager, focused on our AWS cloud infrastructure, to plan, lead and deliver complex cross-team initiatives. This is a heavily coordination-focused role: you will own the execution of a portfolio of AWS cloud platform and cost-savings programs, working hands-on with software engineers, engineering managers and SREs to achieve outcomes aligned with the strategy. You will drive dependencies end-to-end, facilitate trade-off decisions, and give collaborators and leadership clear, reliable access to status and risk. You will be an integral part of the Technical Program Management team, partnering closely with engineering leads to ensure GoDaddy delivers on its planned objectives and global strategy. What you'll get to do... Own end-to-end delivery of a portfolio of concurrent cloud platform programs, coordinating across engineering and partner teams to manage scope, schedule and dependencies against the critical path. Drive cost-savings program coordination, including tracking, reporting and surfacing risks to goals and achievements proactively. Run intake and prioritization processes and keep priority pages and status sources current and trustworthy. Serve as the central coordination point across teams, facilitating trade-off and negotiation discussions, driving alignment, and resolving roadblocks with minimal issues. Build reports, scorecards and dashboards to c

awsaiproject management
View job →
O
1mo ago

About the Team The compute infrastructure team runs the GPU fleet and large-scale compute clusters that serve the models backing ChatGPT and the API, while also supporting training workloads for our next generation models. We operate a large, modern GPU fleet and provide a unified platform for other OpenAI teams to seamlessly run production Applied AI and Research training workloads. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role You will be part of an engineer-first TPM team as a Technical Program Manager for Compute Infrastructure who owns the end-to-end delivery of large-scale GPU clusters, partnering with engineers to bring clusters online across external providers and partners. You’ll run a broad, parallel portfolio spanning hardware, networking, power, and cooling—driving execution, risk management, and crisp alignment from working teams through leadership to deliver production-ready capacity at scale. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead end-to-end delivery of both New Compute SKUs and large-scale GPU clusters across an external partner ecosystem while supporting capacity planning for training and inference. Ability to contextually drive multi-threaded bring-up programs spanning hardware, networking, power, and cooling—owning plans, dependencies, and critical paths. Interface with chip providers to derisk long-term onboarding to new hardware platforms by working across kernels, comms, hardware, and scheduling engineering teams. Build and operationalize program mechanisms (roadmaps, milestones, risk registers, runbooks) that make delivery predictable at massive scale. Partner with engineering to improve cluster turn-up reliability, repeatability, and automation

awsrestai
View job →
S
Stripe
📍 Seattle• Full-time• Remote
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The infrastructure teams provide efficient and optimized infrastructure for Stripe to build secure, reliable, and differentiated products, while enabling Stripe developers to achieve their highest potential. Stripe makes it easy for any developer to access and manage the capabilities of the financial system while maintaining the least regulatory friction. We work to enable developers to have the most productive results of their entire career from the very first days they join Stripe through years of developing new systems and products. What you'll do As a Technical Program Manager in the Infrastructure team, you'll play a key role within engineering and drive programs that span across Stripe in the core infrastructure of Stripe's payment systems. You're responsible for the successful definition, cross-functional strategy, planning, and execution of large-scale technical programs that help to solve complex problems and enable products and infrastructure at scale. You'll deliver outstanding results by building and implementing solutions that scale while protecting our users and serving the business. Responsibilities Work with teams across the organization to understand pain points in their infrastructure usage to find common ideas and work to create solutions that span multiple domains. Define and produce high-quality written proposals, communications, and documentation. Execute on technical programs that require deep systems and engineering-l

REMOTEsqlaigo
View job →

About the Team The Industrial Compute team is responsible for building the physical infrastructure that powers OpenAI’s largest-scale AI systems. We design, deploy, and operate next-generation compute infrastructure across a rapidly expanding global footprint, combining OpenAI-owned infrastructure with strategic cloud and infrastructure partners to support frontier AI workloads. As our infrastructure footprint grows, operational excellence across third-party providers becomes increasingly critical. Our team ensures external infrastructure partners consistently deliver the reliability, performance, and operational maturity required to support OpenAI’s rapidly expanding compute environment. About the Role We are seeking a Hardware Technical Program Manager, Infrastructure Partner Operations to lead operational delivery across OpenAI’s third-party infrastructure partners, including major cloud service providers and strategic compute vendors. In this role, you will serve as the primary operational program manager for external infrastructure partners, driving accountability for service delivery, operational readiness, incident management, performance reporting, and continuous operational improvement. You will work closely with partner engineering and operations teams while coordinating internally across Hardware Engineering, Infrastructure Operations, Capacity Planning, Networking, Supply Chain, Deployment, Reliability Engineering, and executive leadership. Success in this role requires someone who understands how hyperscale infrastructure organizations operate, can establish strong operational governance with external partners, and is comfortable driving complex technical programs without direct ownership of the underlying infrastructure. Key Responsibilities Own operational engagement with third-party infrastructure providers, ensuring consistent execution against operational commitments, service-level agreements (SLAs), and performance expectations. Develop operationa

awsazuregcp
View job →
R
Roblox
📍 San Mateo• Full-time• From $277.4K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Technical Program Manager for Compute Infrastructure, you will lead technical programs to develop and rollout next generation technology solutions to process a diverse spectrum of infrastructure workloads across Roblox services to enable users around the world to efficiently, reliably, and securely enjoy the Roblox Platform. You Have: Been a technical leader with domain expertise in systems and software used to process and manage Cloud and on-prem Infrastructure 7+ years of experience in software industry driving the build of large-scale infrastructure Experienced with establishing work relationships across multi-disciplinary teams and earning trust as a technical leader with all partners Experienced with identifying critical technical problems and opportunities Experienced with delivering end to end technical programs through roadmapping and reliable execution Able to turn specific solutions into systems that benefit the larger community in the long run Knowledgeable about user needs, scoping, planning, execution and delivery Flexible around process, using process as a tool when it makes sense for a team Experience with Kubernetes, Cloud platforms (e.g. AWS), and AI/ML infrastru

awskubernetesgit
View job →
O
24 days ago

About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.

awsazurerest
View job →
O
27 days ago

About the Team OpenAI's Industrial Compute organization is building and operating the infrastructure foundation for the next generation of AI. Infrastructure Operations works across facilities, hardware, network operations, incident management, data center engineering, delivery teams, and external partners to bring capacity online safely, understand its operational state, and improve it over time. As OpenAI's data center portfolio grows across first-party and partner-delivered capacity, the organization needs clear goals, trusted data, repeatable processes, and systems that make ownership, risk, readiness, and performance visible. This role will help build the operating mechanisms that allow Infrastructure Operations to scale with rigor. About the Role We are seeking a Technical Program Manager to own the systems, data, reporting, governance, and program-management backbone for Infrastructure Operations. Reporting to the Delivery & Operations Lead, you will translate strategy into executable goals and operating cadences, turn operational needs into software and data solutions, and create the mechanisms that keep a rapidly evolving organization aligned and accountable. This role will also own the current 1P+3P delivery-tracking layer within Operations: milestones, delivery timelines, quantity forecasts, risks, decisions, and executive reporting. You will partner closely with 1P Delivery Program Management, Compute TPMs, Data Center Engineering, construction, commissioning, and operations leaders to ensure that delivery information becomes complete, usable input for readiness, handover, and ongoing operations. You will own program health and the operating system around it: the goals, data definitions, workflows, reporting, decision paths, and follow-through that help functional DRIs execute. The ideal candidate is comfortable in ambiguity, technically fluent enough to implement real systems, and relentless about converting scattered information into durable mechan

REMOTEsqlawsrest
View job →

About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. In partnership with leading cloud providers, hardware manufacturers, utilities, construction partners, and internal engineering organizations, we are delivering hyperscale AI campuses that power the next generation of frontier AI models. Infrastructure Delivery Operations sits at the center of this effort. Our team develops the operating model that connects infrastructure strategy, supply planning, manufacturing operations, and delivery into a single, integrated system that enables OpenAI to deploy AI infrastructure predictably at scale. We partner across Hardware Engineering, Network Engineering, Capacity Delivery, Hardware Operations, Security, Finance, Strategic Sourcing, and external infrastructure partners to create a single, integrated view of program health. Through governance, operational analytics, executive reporting, and scalable operating mechanisms, we enable leaders to proactively manage risk, optimize capacity, and deliver infrastructure predictably at Industrial Compute speed. About the Role We are seeking a Technical Program Manager, Infrastructure Delivery Operations to drive integrated strategy and delivery across OpenAI's rapidly expanding AI infrastructure portfolio. This role sits at the intersection of infrastructure strategy, New Product Introduction (NPI), supply planning, manufacturing operations, and infrastructure delivery. You will lead highly cross-functional programs spanning engineering, supply planning, manufacturing, logistics, construction, commissioning, and operations, ensuring technical and operational dependencies remain synchronized from planning through production readiness. Beyond driving program execution, you will leverage operational insights to improve capacity planning, infrastructure strategy, and deployment readiness. You will also help operationalize new technologies and suppliers by partnering w

REMOTEawsrestagile
View job →
B
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE At Baseten, we’re looking for a Technical Program Manager to drive our most complex, cross-cutting infrastructure programs. This role will operate across all domains of AI infrastructure, from the GPUs up to the multi-cluster orchestration layer. This is an execution-first role. The work is less about owning a single system and more about imposing order on ambiguity: standing up the right structures, driving decisions to closure, and making sure nothing falls through the cracks across dozens of stakeholders. If you take satisfaction in turning a chaotic, half-defined initiative into a predictable, well-governed program, this role is for you. RESPONSIBILITIES Own complex migrations end to end. Lead large-scale infrastructure migrations across teams and domains. This will involve scoping the work, sequencing dependencies, managing risk, and driving them to completion without surprises. Drive process across infrastructure. Establish and run the operating rhythms that keep programs healthy: planning cadences, status reporting, decision logs, risk reviews, and escalation paths. Make the process light enough that teams adopt it and rigorous enough that it actually works. Help managers build the right structures. Partner with engineering managers and leads to design the team structures, ownership boundaries, and working models a program needs to succeed. Spot gaps in accountability before they become problems. Own fo

machine learningaigo
View job →
S
Stripe
📍 South San Francisco• Full-time• $173.4K – $260.2K/yr
1mo ago

Who we are About Stripe Stripe, LLC. is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. What you’ll do Responsibilities Working with teams across the organization to understand pain points in their infrastructure usage to find common ideas and work to create solutions which span multiple domains; Define and produce high quality written proposals, communications and documentation. Execute on technical programs that require deep systems and engineering level knowledge; Partner with Engineering Managers, Tech Leads, Engineers and other Technical Program Managers to define, scope and drive large programs to conclusion; Play a key part in shaping the technical design, predicting technical roadblocks by collaborating with engineers, and identifying trade-offs; Develop, implement, and iterate on program management techniques, frameworks, and KPIs to achieve goals with well defined success criteria; Elevate the execution muscle of engineering teams around you; Train them to be better at delivery where needed; Help influence peers / stakeholders and build consensus while dealing with ambiguity; Leverage data and acquired knowledge to drive strategic decisions at an engineering leadership level; Create widely circulated plans, driving consistency, clarity and building alignment across teams; Operationalize and execute critical cross functional programs spanning multiple engineering organizations for Infrastructure (Developer Infrastructure, Core Infrastructure, Service Infrastructure); Shape technical design, predicting tech

sqlawsazure
View job →

About the Team OpenAI's Industrial Compute organization is building the world's most advanced AI infrastructure ecosystem. In partnership with leading cloud providers, hardware manufacturers, utilities, construction partners, and internal engineering organizations, we are delivering hyperscale AI campuses that power the next generation of frontier AI models. Infrastructure Delivery Operations sits at the center of this effort. The team ensures that large, highly complex infrastructure programs execute predictably across planning, design, construction, hardware deployment, commissioning, and operational handoff. We partner across Hardware Engineering, Network Engineering, Capacity Delivery, Hardware Operations, Security, Finance, Supply Chain, and our external infrastructure partners to keep programs aligned, risks visible, and execution moving at Industrial Compute speed. About the Role We are seeking a Technical Program Manager, Infrastructure Delivery Operations to drive execution across large-scale AI infrastructure deployments. This role is responsible for orchestrating cross-functional delivery programs spanning multiple organizations, ensuring dependencies remain synchronized from early planning through production readiness. You will develop operational mechanisms that allow Industrial Compute to scale infrastructure delivery across multiple campuses simultaneously. Rather than owning any individual engineering discipline, you will own program health—bringing together engineering, construction, operations, supply chain, and partner organizations into a single coordinated execution model. Success in this role requires exceptional program management, executive communication, systems thinking, and operational rigor. You should be comfortable operating amid ambiguity while bringing structure to highly technical, multi-year infrastructure programs. Candidates should have experience leading complex infrastructure, cloud, data center, semiconductor, networking, manuf

awsrestagile
View job →

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . What We're Looking For: 2–4 years of program or project management experience (or equivalent), ideally touching infrastructure, platform, or cloud environments. Working knowledge of cloud platforms (AWS or similar) — compute, storage, and basic usage/cost concepts. Interest in, and some exposure to, capacity management and infrastructure efficiency (right-sizing, utilization, waste reduction); deep FinOps expertise is not required. Strong organizational and execution skills: able to track a program's moving pieces, follow up, and keep things on schedule. Clear written and verbal communication; comfortable presenting status and asks to engineering partners. Collaborative and coachable — works well with engineers and more senior TPMs, seeks input, and takes feedback well. Experience at a large-scale consumer tech or infrastructure organization is

REMOTEsqlawsai
View job →
🔔

Get new technical program manager infrastructure jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More technical program manager infrastructure opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.