About the Role As a Sales Manager, Energy, you will build and lead a team of Account Directors focused on strategic growth across utilities, oil and gas, renewables, power generation, and energy services. The team will partner with complex organizations modernizing operations, improving reliability, accelerating the energy transition, and adopting enterprise AI responsibly at scale. You’ll help the team navigate regulated enterprise sales cycles, deepen relationships with business, technology, operations, engineering, security, and risk leaders, and drive adoption of OpenAI’s platform across safety-conscious, asset-intensive organizations. Key Responsibilities Recruit, develop, and lead a high-performing team of Energy Account Directors. Create a strong coaching culture through deal reviews, account strategy sessions, ride-alongs, and structured 1:1s. Define the Energy GTM strategy, including subsector segmentation, account prioritization, partner strategy, executive engagement, and territory planning. Drive disciplined pipeline generation, forecast accuracy, and operational rigor. Guide multi-stakeholder opportunities involving operations, engineering, digital, data, security, legal, risk, procurement, and executive leadership. Help customers translate AI and API capabilities into measurable outcomes across asset and field operations, grid and generation planning, engineering knowledge, customer service, commercial workflows, and enterprise productivity. Partner with Product, Solutions Architecture, Technical Success, Legal, Security, Finance, and policy experts to support responsible deployment. Provide structured feedback on customer requirements, integration blockers, reliability and governance needs, and emerging industry trends. What We’re Looking For 15+ years of enterprise sales, GTM, or sales leadership experience. Proven experience building and scaling enterprise sales teams responsible for complex strategic accounts and large revenue targets. Deep underst
Jobs in United States
Reliability Engineer Iii in San Francisco
227 active opportunities · Updated October 2026
Showing
15 jobs
Explore current reliability engineer iii jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team The Industrial Compute team is responsible for building the physical infrastructure that powers OpenAI’s largest-scale AI systems. We design, deploy, and operate next-generation compute infrastructure across a rapidly expanding global footprint, combining OpenAI-owned infrastructure with strategic cloud and infrastructure partners to support frontier AI workloads. As our infrastructure footprint grows, operational excellence across third-party providers becomes increasingly critical. Our team ensures external infrastructure partners consistently deliver the reliability, performance, and operational maturity required to support OpenAI’s rapidly expanding compute environment. About the Role We are seeking a Hardware Technical Program Manager, Infrastructure Partner Operations to lead operational delivery across OpenAI’s third-party infrastructure partners, including major cloud service providers and strategic compute vendors. In this role, you will serve as the primary operational program manager for external infrastructure partners, driving accountability for service delivery, operational readiness, incident management, performance reporting, and continuous operational improvement. You will work closely with partner engineering and operations teams while coordinating internally across Hardware Engineering, Infrastructure Operations, Capacity Planning, Networking, Supply Chain, Deployment, Reliability Engineering, and executive leadership. Success in this role requires someone who understands how hyperscale infrastructure organizations operate, can establish strong operational governance with external partners, and is comfortable driving complex technical programs without direct ownership of the underlying infrastructure. Key Responsibilities Own operational engagement with third-party infrastructure providers, ensuring consistent execution against operational commitments, service-level agreements (SLAs), and performance expectations. Develop operationa
About the Team OpenAI Finance is responsible for ensuring the organization is set up for success in pursuit of its mission. OpenAI’s Tax and Trade team sits at the center of OpenAI’s global growth—shaping how cutting-edge AI products, partnerships, and infrastructure scale across borders while navigating complex tax, trade, and regulatory regimes. We operate as strategic operators, not just compliance experts, embedding early in product, finance, policy, and infrastructure decisions to manage risk, unlock incentives, and enable OpenAI to grow responsibly and competitively worldwide. As OpenAI’s products, monetization models, and global infrastructure expand, our tax systems must scale with the same rigor and reliability as our product architecture. The team partners deeply with Financial Engineering, Product Engineering, and Finance Systems to build a modern tax technology stack that enables accurate, real-time tax determination across OpenAI’s global business. About the Role We're hiring a Director, Tax Infrastructure & Incentives to lead OpenAI's global tax infrastructure strategy supporting AI infrastructure expansion programs. You will own the tax strategy supporting data center development, site selection, infrastructure investments, and government incentive programs across. This role extends far beyond traditional property tax planning and reporting - you will help shape where and how OpenAI invests billions of dollars in AI infrastructure by partnering with executive leadership, infrastructure teams, governments, utilities, and external stakeholders. You will build scalable frameworks for negotiating and managing tax incentives, property tax, indirect tax, and infrastructure-related tax matters while developing repeatable processes that enable OpenAI to expand rapidly across multiple jurisdictions. You will also establish the governance, reporting, and operational infrastructure required to support long-term compliance, financial reporting, and executive
About the Team Training Runtime builds the distributed systems that power OpenAI's largest model training runs - most recently GPT-5.5! The Data Movement area owns the infrastructure that keeps training jobs supplied with the right data at the right time, and keeps model state moving safely and efficiently across large clusters. Our work spans machine learning systems, distributed storage, high-throughput data loading, reliability engineering, and developer experience. Success means researchers can move quickly while training runs remain fast, reproducible, debuggable, and resilient at scale. About the Role We are looking for a deeply hands-on Technical Lead Manager to own datasets throughout our training infrastructure. This person will set the direction for how training jobs read data: the APIs, storage contracts, versioning model, benchmarks, debugging tools, and reliability guarantees that make data access consistent across current and future training frameworks. You will begin as the primary technical owner for dataset reads, working directly in the code while aligning researchers, training framework owners, storage teams, and infrastructure partners around a durable platform. The problem is deceptively hard at frontier scale: make enormous, heterogeneous datasets easy to consume, correct across distributed workers, observable when something goes wrong, and flexible enough to support pretraining, reinforcement learning, and multimodal training. In this role, you will Design and build a unified dataset read platform for multiple current and future training frameworks. Define dataset APIs, storage-format expectations, registration/versioning, and migration paths that make data access reproducible and maintainable. Build reliability into the read path, including stateful iteration, caching, fast restart, recovery, and clear operational contracts. Build terminal and web-based visualizers that let teams inspect text, multimodal, and reinforcement learning data late
$342K – $445K/yr
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking a Technical Lead to lead deployment and operations for OpenAI’s Silicon & Systems team. This person will become the Directly-Responsible Individual responsible for bringing OpenAI’s custom silicon and associated systems into data center environments, ensuring successful deployment, bring-up, validation, operational readiness, and ongoing reliability at scale. This role sits at the intersection of silicon, systems, infrastructure, data center operations, and software. You will lead a team focused on taking new hardware platforms from lab validation into production data center deployment. You will be responsible for building the operational processes, technical workflows, tooling, and cross-functional alignment required to deploy and operate custom AI hardware reliably in OpenAI’s supercomputing infrastructure. The ideal candidate is both a strong leader and a deeply technical operator. You should be comfortable staying close to the technical details of hardware bring-up, fleet deployment, debugging, system validation, data center integration, and production operations. This role requires strong execution, excellent cross-functional judgment, and the ability to drive clarity in ambiguous, fast-moving environments. In this role, you will: Lead a team responsible for deployment and operations of OpenAI’s custom silicon and systems in data center environments Own the path from hardware bring-up and validation through production deployment, operati
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Pinterest is seeking a Sr. Manager to lead our Capacity Engineering team. The team ensures that Pinterest’s cloud infrastructure has the capacity it needs while operating reliably, efficiently and with clear financial accountability. You’ll lead the full portfolio across forecasting and supply, capacity-management systems, compute and GPU efficiency, infrastructure data and governance and capacity operations. What you’ll do: Lead the Capacity Engineering team and establish its 12–18 month functional and technical strategy, roadmap and success measures tied to Infrastructure and company goals. Develop CPU and GPU forecasts and supply plans that account for workload demand, delivery constraints, cost and reliability requirements. Guide the design and delivery of capacity requests, reservations, entitlements, allocation policy and infra
About the Team Employee Tech & Experience (ETX) helps people at OpenAI do their most ambitious work. Across Helpdesk, Executive Support, Systems Operations, Logistics and AV, we make technology simple, reliable and secure. Employee needs guide what we build, improve and choose to eliminate. About the Role Reporting to the Head of Global IT, you’ll lead ETX globally, building on the team’s capabilities and customer-zero work to continually advance the employee experience. You’ll shape ETX’s strategy, investment priorities and operating model in partnership with leadership across the company, turning new capabilities into measurable amplification. You’ll develop leaders and strengthen teams where people feel valued, own meaningful work and enjoy working together. This role is based at our San Francisco headquarters and requires an in-office presence. In this role, you will: Lead the next stage of ETX’s global growth across Helpdesk, SysOps, Logistics and AV, with a shared strategy and accountability for employee outcomes. Continually elevate the employee experience through research and design, directing investment to simplify entire user journeys, remove unnecessary effort and amplify what employees can accomplish. Accelerate ETX’s agent-led and customer-zero work as capabilities advance: continually challenge which workflows need to exist, extend what agents can own end to end and evolve the operating model to deliver measurable amplification. Develop leaders who earn trust, grow others and sustain an environment where people feel valued, take pride in their work and enjoy working together. Give people meaningful ownership and opportunities to stretch and grow, with clear priorities and sustainable workloads. Scale global services to support company growth, with clear commitments to reliability, security, responsiveness, and effective controls. Measure performance through employee effort, service quality, and time to resolution. Shape priorities and investment wi
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. About the Team Our Network Enablement and Access team works to unlock the potential of Plaid's network by broadening and deepening our connections with data partners. We build the capabilities that help data providers participate in the network, strengthen the quality and reliability of those connections, and enable great products and experiences for Plaid's customers. Within Network Enablement and Access, the Data Supply Traffic and Health team owns how Plaid's requests flow to the data providers we depend on. We manage the load placed on each provider, the constraints that shape our access, and the fair allocation of capacity across Plaid's products and new initiatives. As paid access expands across the network, we also work to keep that traffic reliable, efficient, and cost-effective. As the Product Manager for Data Supply Health and Traffic, you will establish and lead a new product area at the foundation of every Plaid product. You will define how Plaid allocates constrained provider capacity, scales traffic across products, and manages the economics of paid data access. You will also optimize for data freshness, balancing timeliness with provider capacity and cost so Plaid's
About the Team The Applied organization brings OpenAI’s most advanced technology to the world through products like ChatGPT and the APIs that power a growing ecosystem of developer and enterprise applications. Data Engineering builds and operates the trustworthy, secure, and reliable data systems that power decisions across OpenAI. About the Role We’re looking for a Data Engineering Manager to lead the Growth & Revenue data engineering team. This leader will own the data strategy and execution for the data subject areas spanning growth accounting across all product surfaces, product partnerships, checkout, billing, payments, revenue, and monetization, helping OpenAI understand how people adopt, engage with, and pay for our products. You will partner closely with several Data Science, Business, and Engineering partners to connect product behavior to trustworthy subscriber, payment, and revenue measurement. In this role, you will: Build, manage, and grow a high-performing, inclusive team across the Growth & Revenue data subject areas. Define the data strategy for all the data subject areas you own. Deliver durable, well-modeled data products that connect product behavior, subscription state, checkout events, payment outcomes, and revenue. Establish trusted metric definitions and data quality standards so product, growth, finance, and executive leaders can make fast, consistent decisions. Partner with Data Science and Product teams to support experimentation, causal measurement, funnel analysis, and scalable self-serve analytics. Partner with Finance and Financial Engineering to ensure analytical revenue views reconcile to financial truth and production billing systems. Raise operational excellence for critical pipelines, including reliability, observability, privacy, governance, and incident response. Set a clear roadmap, make principled tradeoffs, and communicate progress and risk across technical and business stakeholders. You might thrive in this role if yo
About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team builds next-generation AI-native silicon and systems while working closely with software, research, and manufacturing partners to co-design hardware tightly integrated with AI models. In addition to delivering systems for OpenAI’s supercomputing infrastructure, the team develops the tools, methodologies, and strategic partnerships needed to accelerate hardware innovation. About the Role We’re seeking an experienced Hardware Strategic Sourcing Manager to own sourcing strategy and supplier partnerships for fiber and optical interconnect components across OpenAI’s next-generation AI infrastructure. Reporting to the Head of Partnerships & Strategic Sourcing, you will lead sourcing across fiber cable assemblies, internal optical harnesses, fiber shuffles, optical backplane assemblies, connectorized and standalone passive optical assemblies, fiber-array units (FAUs), fiber-to-chip and coupling interfaces, detachable connectors, optical routing, and assigned optical packaging, assembly, and test services. You will work closely with electrical engineering, optical engineering, systems engineering, mechanical and packaging engineering, quality, rack integration, data-center deployment,manufacturing, supply chain, finance, legal, and program management teams to translate demanding bandwidth, signal integrity, reliability, and scale requirements into resilient supplier partnerships and scalable commercial strategies. Your work will directly support the performance, reliability, manufacturability, and scale of the high-speed optical connectivity required for OpenAI’s next-generation AI systems. In this role, you will: Develop and execute a comprehensive sourcing strategy for fiber and optical interconnect components supporting high-bandwidth AI systems and infrastructure. Own sourcing across optical fiber cable assembli
About the Team The Statsig team is responsible for the experimentation, feature rollout, dynamic configuration, and analytics systems that help OpenAI ship products with speed, safety, and evidence. Teams across ChatGPT, Codex, model measurement, monetization, business subscriptions, developer products, and shared infrastructure rely on Statsig to introduce capabilities safely, measure their impact, and make high-confidence product decisions. About the Role As a Product Lead on the Statsig team, you will define how experimentation, rollout, configuration, and analytics become a simple, reliable, and trusted part of how every OpenAI product team ships. You will set strategy across multiple product and platform workstreams, translate company-wide needs into durable capabilities, and help Statsig become a core part of OpenAI’s product development system. We’re looking for a product leader who combines strong product judgment, technical fluency, and deep analytical thinking. You should be comfortable navigating ambiguous customer needs, influencing teams across the company, and balancing rapid adoption with reliability, usability, and measurement quality. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. Travel requirements should be confirmed with the recruiter before publishing. In this role, you will: Define the product vision, strategy, and roadmap for experimentation, feature management, dynamic configuration, rollout safety, and analytics. Partner with product, engineering, research, data, design, and infrastructure leaders to turn recurring launch and measurement needs into reusable platform capabilities. Develop a deep understanding of workflows across ChatGPT, Codex, model measurement, monetization, subscriptions, and developer products, then establish clear priorities across competing needs. Drive adoption by making sophisticated experimentation and analytics c
About the Team OpenAI, in close collaboration with our capital partners, is building the world’s most advanced AI infrastructure ecosystem. The Power Execution team owns the strategy and execution required to secure reliable, scalable, and economically resilient power for OpenAI’s global data center portfolio. The team sits at the intersection of commercial, technical, policy, legal, and operational work, partnering across OpenAI and with utilities, grid operators, regulators, counterparties, and public-sector stakeholders. About the Role The Energy Regulatory Lead will own energy regulatory strategy and execution for OpenAI’s infrastructure growth. This role will be the primary bridge between the Power Execution team and Public Policy and Government Affairs on energy regulatory matters, ensuring that OpenAI’s external engagement is grounded in project realities and that changing policy and regulatory conditions are translated into actionable infrastructure decisions. This is an individual contributor lead role and does not have direct reports initially. The role combines portfolio-level regulatory positioning with transactional regulatory work: evaluating jurisdictional pathways, supporting utility and energy transactions, coordinating approvals and filings, and helping project teams navigate tariffs, interconnection, load-service requirements, market rules, and regulatory risk from diligence through execution. In this role, you will: Develop and maintain OpenAI’s energy regulatory strategy across priority U.S. markets and, as needed, emerging geographies for infrastructure expansion. Coordinate closely with Public Policy and Government Affairs to shape energy regulatory priorities, engagement plans, messaging, and positions before utilities, public utility commissions, grid operators, state energy offices, and other relevant policymakers. Translate project requirements—load size, timing, reliability, cost, carbon, and expansion needs—into clear regulatory objectiv
About the Team OpenAI’s People team hires, engages, and retains world-class talent to safely build and deploy AGI that benefits all of humanity. The People Analytics team helps leaders make rigorous, evidence-based talent decisions and ensures that the systems supporting those decisions are valid, reliable, fair, and accountable. About the Role As a People Data Scientist focused on AI fairness and bias testing, you will help establish how OpenAI evaluates AI-assisted People systems and high-impact talent processes. You will design and conduct rigorous assessments to identify, measure, and mitigate potential bias across the lifecycle of models, agents, decision-support tools, and automated workflows. Your work will span the entire employee life-cycle, such as hiring, performance, promotion, employee development, workforce planning, etc. You will evaluate both technical systems and the broader human-AI decision processes in which they operate, examining not only model performance but also data quality, measurement validity, differential outcomes, human oversight, and unintended consequences. We’re looking for an experienced data scientist or applied researcher who can translate complex fairness questions into defensible evaluation strategies, scalable testing infrastructure, and clear recommendations for technical teams and senior leaders. This role is preferred to be based in San Francisco, CA. In this role, you will: Define and lead fairness and bias-testing strategies for AI-assisted People processes, models, agents, and decision-support systems from development through deployment and ongoing monitoring. Design rigorous algorithmic audits and validation studies, including adverse-impact analysis, subgroup and intersectional evaluation, error-rate analysis, calibration, measurement invariance, reliability, criterion-related validity, and sensitivity testing. Identify the appropriate fairness criteria for each use case, evaluate tradeoffs among competing definitions
About the Team OpenAI’s Industrial Compute team is building and productizing infrastructure capabilities that help organizations deploy and operate advanced AI systems at scale. The team works across AI hardware, systems engineering, physical infrastructure, and customer delivery to turn emerging technologies into reliable, repeatable infrastructure solutions. Our work sits at the intersection of technical strategy, product development, engineering, and deployment. We partner closely with customers and internal engineering teams to solve complex infrastructure challenges spanning compute, power, cooling, controls, and facility efficiency. About the Role We are seeking a senior, hands-on Data Center Infrastructure Architect to develop and optimize the physical infrastructure required for large-scale AI deployments. This is a broad technical role spanning data center architecture, electrical and mechanical systems, high-density compute, controls, telemetry, and digital modeling. You will use simulation, operational data, and digital-twin approaches to evaluate infrastructure designs, identify system-level constraints, and improve efficiency, reliability, cost, and speed of deployment. The ideal candidate can move fluidly between first-principles analysis, facility and equipment design, computational modeling, engineering review, and real-world implementation. You should be comfortable working across disciplines rather than operating solely within electrical, mechanical, or software boundaries. Key Responsibilities Define system-level architectures for high-density AI data centers across power, cooling, IT equipment, controls, and facility infrastructure. Develop digital twins and other computational models that represent the behavior of data center systems under changing workloads, environmental conditions, equipment configurations, and failure scenarios. Use design and operational data to identify constraints, improve PUE and related efficiency metrics, and optimize
About Team Our Robotics team is focused on unlocking general-purpose robotics and advancing toward AGI-level intelligence in dynamic, real-world environments. Working across the full model and systems stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the physical constraints of real-world systems to improve people’s lives. About the Role We are looking for a TPM to drive development and integration of a range of sensor systems for robotics. This role will drive cross-functional alignment across requirements, engineering design, integration, validation, manufacturing, supply chain, and release processes, helping turn complex sensor-system needs into clear plans, decisions, and milestones. Location and in-person expectations: This role is based in San Francisco, CA and requires in-person presence 4 days a week. In this role you will: Drive requirements alignment across engineering design, integration, testing, and validation for camera modules, LiDAR, IMUs, RADAR, proximity sensors, audio components and the systems they interact with. Coordinate the integration of modules including electrical, mechanical, harnessing, and software interfaces with the full robotic system with deep understanding of timelines to drive the respective PCBAs, enclosures, build and test fixtures, connectors and cables. Establish effective cadences for technical reviews, BOM readiness, change management, production releases, approvals, and decision tracking. Align harnesses, fasteners, assembly fixtures, test fixtures, and documentation so cross-functional teams can execute against a clear plan. Lead validation planning around functional, reliability, NVH failure modes, including testing needs, schedules, and exit criteria. Partner with manufacturing and supply chain to manage handoffs, lead times, dependencies, and production readiness. Drive tradeoff decisions across cost, qua
Other cities to consider
More places hiring for this role
Get new reliability engineer iii jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime