Technical Program Manager – Applied Infrastructure About the Team The Applied team safely brings OpenAI’s technology to the world, powering products like ChatGPT, and the APIs for GPT and more. Behind these products is a complex and rapidly evolving infrastructure platform that enables scale, performance, and safety. The Applied Infrastructure TPM team partners across engineering to lead foundational programs that ensure OpenAI’s infrastructure can meet current and future demand. About the Role We’re looking for a seasoned Technical Program Manager to drive critical infrastructure programs across the Applied organization. This TPM will focus on cross-cutting initiatives such as general compute capacity planning, process transformation, cost and quota attribution and optimization, and coordination across infrastructure and product stakeholders. There will also be focus on evolving OpenAI’s infrastructure to support growth, scale and new products. This work is core to how OpenAI manages and grows its infrastructure footprint in a disciplined, scalable way. Location: San Francisco, CA (Hybrid – 3 days/week in-office) In this role, you will: Serve as the DRI for complex infrastructure programs spanning CPU planning, orchestration, and other resource management domains (e.g. networking, storage). Build and operationalize systems to capture demand signals, model future capacity needs, and align infrastructure planning across internal teams and partners external to the company. Partner closely with Infrastructure, Product and Finance teams to forecast infrastructure usage patterns and ensure supply/demand alignment. Lead cost attribution and quota enforcement programs to promote stability and ensure equitable access to resources across teams. Drive simplification and standardization of infrastructure tooling and processes across Applied and Infra organizations. Drive cross functional programs to evolve our infrastructure to support new growth and scale Work with external v
Technical Program Manager - Infrastructure
Market pay estimate
$133,144–$190,993 / year for comparable Program Manager roles. Not employer-provided.
Role overview
Job description
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us!
The mission of the Engineering TPM team is to drive Figma's most important cross-company engineering efforts, and we are looking for a Technical Program Manager (TPM) to partner with our Infrastructure team. The TPM provides oversight of the most important efforts that require coordinated technical execution across the Org to succeed. This is a role focused on enabling Figma's infrastructure teams to scale, improve performance, and deliver on critical projects. These large-scale efforts will involve collaboration across numerous backend, infrastructure, and security teams and cross-functional stakeholders, prioritization, decision-making, tracking execution, and driving operational excellence. We're looking for someone that can work in a TPM greenspace environment and is passionate about people, technology, and program management. Progress over process is our mantra.
This is a full-time role that can be held from one of our US hubs or remotely in the United States.
What you'll do at Figma:
- Lead the execution, coordination, and risk management of Figma's infrastructure projects, ensuring seamless integration with minimal performance impact
- Drive key infrastructure initiatives, including reliability, storage, distributed systems, cloud-native performance improvements, and compliance programs (e.g., encryption key management, FedRAMP, SOC 2)
- Partner closely with engineering, security, compliance, and legal teams to ensure alignment and on-time delivery
- Track program milestones and ensure seamless delivery across multiple infrastructure teams, including data, caching, observability, and security engineering
- Provide regular updates to executive leadership, external partners, and internal teams on the status of infrastructure programs, including risks, blockers, and dependencies
- Develop and drive best practices in infrastructure program management, improving visibility into progress, risks, and technical dependencies
- Facilitate large-scale testing and rollout strategies for critical infrastructure changes to minimize downtime and ensure high system reliability
- Translate technical constraints and risks into executive-level communications for senior leadership and external stakeholders
We’d love to hear from you if you have:
- 8+ years in Infrastructure TPM or related roles (cloud engineering, SRE, or infra-focused software engineering), with hands-on experience in cloud-native architectures and distributed systems (AWS, GCP, or similar)
- Deep technical expertise across infrastructure components - storage (S3, RDS, DynamoDB), caching (Redis), search (OpenSearch), and event-driven systems (Kafka) - with the ability to reason across org-level goals, architectural trade-offs, and component-level details
- Track record driving large-scale infrastructure programs such as migrations, cost optimization, encryption/key management, and reliability initiatives for high-availability, business-critical services
- Strong cross-functional program management skills, including coordinating multi-team engineering orgs, tracking performance bottlenecks, optimizing infrastructure SLAs, and establishing PM best practices in ambiguous environments
- Proven stakeholder management, with experience engaging external partners (cloud providers, enterprise customers) and translating complex technical challenges into clear executive-level communications
While it’s not required, it’s an added plus if you also have:
- A track record of setting up and scaling TPM functions or program management practices within an engineering organization
- Experience working on high-visibility enterprise security and compliance programs (SOC 2, FedRAMP, encryption key management)
- Infrastructure incident management and analysis experience
- Experience working with external enterprise partners on technical programs (e.g., Apple, AWS, Google)
- Background in software engineering or site reliability engineering (SWE/SRE-to-TPM career path)
Pay Transparency Disclosure
Job level and actual compensation will be decided based on factors including, but not limited to, individual qualifications objectively assessed during the interview process (including skills and prior relevant experience, potential impact, and scope of role), market demands, and specific work location.
Figma offers equity to employees, as well as a competitive package of additional benefits, including health, dental, and vision coverage; retirement benefits with company contributions; parental leave and reproductive or family planning support; mental health and wellness benefits; and paid time off. Figma provides paid sick leave, holidays, and other leave benefits in compliance with applicable federal, state, and local laws, including the requirements of the Washington Minimum Wage Act and related regulations. Exempt employees are eligible for employer‑provided paid flexible PTO in addition to flexible paid sick leave. PTO is subject to manager approval. Additional benefits may include company recharge days, cell phone and home internet reimbursements, and a number of lifestyle spending accounts. Figma also offers sales incentive compensation for most sales roles and an annual bonus plan for eligible non-sales roles. All compensation and benefits are subject to applicable plan terms and may be modified by Figma at any time, consistent with applicable law.
At Figma we celebrate and support our differences. We know employing a team rich in diverse thoughts, experiences, and opinions allows our employees, our product and our community to flourish. Figma is an equal opportunity workplace - we are dedicated to equal employment opportunities regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity/expression, veteran status, or any other characteristic protected by law. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements.
We will work to ensure individuals with disabilities are provided reasonable accommodation to apply for a role, participate in the interview process, perform essential job functions, and receive other benefits and privileges of employment. If you require accommodation, please reach out to [email protected]. These modifications enable an individual with a disability to have an equal opportunity not only to get a job, but successfully perform their job tasks to the same extent as people without disabilities.
Examples of accommodations include but are not limited to:
- Holding interviews in an accessible location
- Enabling closed captioning on video conferencing
- Ensuring all written communication be compatible with screen readers
- Changing the mode or format of interviews
To ensure the integrity of our hiring process and facilitate a more personal connection, we require all candidates keep their cameras on during video interviews. Additionally, if hired you will be required to attend in person onboarding.
By applying for this job, the candidate acknowledges and agrees that any personal data contained in their application or supporting materials will be processed in accordance with Figma's Candidate Privacy Notice.
What they are looking for
Skills & requirements
Hiring company
Figma
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us!
Keep exploring
Similar active roles
Fresh roles matched to this title and market.
About the Team The compute infrastructure team runs the GPU fleet and large-scale compute clusters that serve the models backing ChatGPT and the API, while also supporting training workloads for our next generation models. We operate a large, modern GPU fleet and provide a unified platform for other OpenAI teams to seamlessly run production Applied AI and Research training workloads. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role You will be part of an engineer-first TPM team as a Technical Program Manager for Compute Infrastructure who owns the end-to-end delivery of large-scale GPU clusters, partnering with engineers to bring clusters online across external providers and partners. You’ll run a broad, parallel portfolio spanning hardware, networking, power, and cooling—driving execution, risk management, and crisp alignment from working teams through leadership to deliver production-ready capacity at scale. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead end-to-end delivery of both New Compute SKUs and large-scale GPU clusters across an external partner ecosystem while supporting capacity planning for training and inference. Ability to contextually drive multi-threaded bring-up programs spanning hardware, networking, power, and cooling—owning plans, dependencies, and critical paths. Interface with chip providers to derisk long-term onboarding to new hardware platforms by working across kernels, comms, hardware, and scheduling engineering teams. Build and operationalize program mechanisms (roadmaps, milestones, risk registers, runbooks) that make delivery predictable at massive scale. Partner with engineering to improve cluster turn-up reliability, repeatability, and automation
Location Details: Canada - BC or ON (remote) At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team GoDaddy - Global Production Engineering looks after GoDaddy's global infrastructure, in the cloud and on-premises. We are hiring an experienced Technical Program Manager, focused on our AWS cloud infrastructure, to plan, lead and deliver complex cross-team initiatives. This is a heavily coordination-focused role: you will own the execution of a portfolio of AWS cloud platform and cost-savings programs, working hands-on with software engineers, engineering managers and SREs to achieve outcomes aligned with the strategy. You will drive dependencies end-to-end, facilitate trade-off decisions, and give collaborators and leadership clear, reliable access to status and risk. You will be an integral part of the Technical Program Management team, partnering closely with engineering leads to ensure GoDaddy delivers on its planned objectives and global strategy. What you'll get to do... Own end-to-end delivery of a portfolio of concurrent cloud platform programs, coordinating across engineering and partner teams to manage scope, schedule and dependencies against the critical path. Drive cost-savings program coordination, including tracking, reporting and surfacing risks to goals and achievements proactively. Run intake and prioritization processes and keep priority pages and status sources current and trustworthy. Serve as the central coordination point across teams, facilitating trade-off and negotiation discussions, driving alignment, and resolving roadblocks with minimal issues. Build reports, scorecards and dashboards to c
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. About the team The infrastructure teams provide efficient and optimized infrastructure for Stripe to build secure, reliable, and differentiated products, while enabling Stripe developers to achieve their highest potential. Stripe makes it easy for any developer to access and manage the capabilities of the financial system while maintaining the least regulatory friction. We work to enable developers to have the most productive results of their entire career from the very first days they join Stripe through years of developing new systems and products. What you'll do As a Technical Program Manager in the Infrastructure team, you'll play a key role within engineering and drive programs that span across Stripe in the core infrastructure of Stripe's payment systems. You're responsible for the successful definition, cross-functional strategy, planning, and execution of large-scale technical programs that help to solve complex problems and enable products and infrastructure at scale. You'll deliver outstanding results by building and implementing solutions that scale while protecting our users and serving the business. Responsibilities Work with teams across the organization to understand pain points in their infrastructure usage to find common ideas and work to create solutions that span multiple domains. Define and produce high-quality written proposals, communications, and documentation. Execute on technical programs that require deep systems and engineering-l
About the Team The Industrial Compute team is responsible for building the physical infrastructure that powers OpenAI’s largest-scale AI systems. We design, deploy, and operate next-generation compute infrastructure across a rapidly expanding global footprint, combining OpenAI-owned infrastructure with strategic cloud and infrastructure partners to support frontier AI workloads. As our infrastructure footprint grows, operational excellence across third-party providers becomes increasingly critical. Our team ensures external infrastructure partners consistently deliver the reliability, performance, and operational maturity required to support OpenAI’s rapidly expanding compute environment. About the Role We are seeking a Hardware Technical Program Manager, Infrastructure Partner Operations to lead operational delivery across OpenAI’s third-party infrastructure partners, including major cloud service providers and strategic compute vendors. In this role, you will serve as the primary operational program manager for external infrastructure partners, driving accountability for service delivery, operational readiness, incident management, performance reporting, and continuous operational improvement. You will work closely with partner engineering and operations teams while coordinating internally across Hardware Engineering, Infrastructure Operations, Capacity Planning, Networking, Supply Chain, Deployment, Reliability Engineering, and executive leadership. Success in this role requires someone who understands how hyperscale infrastructure organizations operate, can establish strong operational governance with external partners, and is comfortable driving complex technical programs without direct ownership of the underlying infrastructure. Key Responsibilities Own operational engagement with third-party infrastructure providers, ensuring consistent execution against operational commitments, service-level agreements (SLAs), and performance expectations. Develop operationa
About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.
🔔 Get job alerts
New Technical Program Manager - Infrastructure jobs in San Francisco, CA • New York, NY • United States, straight to your inbox.
No spam · Unsubscribe anytime