About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role At Sentry, Support is an engineering discipline. Our customers are the greatest technical minds in the world—developers at elite enterprises building the future of software—and they deserve answers that go deeper than a knowledge base link. We're looking for an APAC Technical Support Engineer based in San Francisco to join our global Support Engineering team. This role is designed to provide APAC coverage to our users; with the shift being Sunday through Thursday 4PM-12AM PST. We are architecting the Technical Support engine . We’re looking for an experienced engineer to help us redefine the standard of technical support by combining deep human expertise with autonomous agentic systems. You are a debugger of both code and systems. You will treat support volume as a data signal to build automated resolution paths, ensuring our human engineers only touch the most complex, high-impact architectural puzzles. Sentry Support Engineers aren't just clearing queues; they are Orchestrators . You will engage with our users across GitHub, Discord, and our internal systems, while acting as the Technical Lead for our Agentic Ops. You ensure that when a developer asks a complex question, our systems have the right context and a seamless "Human-in-the-Loop" path to you when deep, nuanced expertise is required. In this role you will Master the Sentry Ecosystem & Support Elite Developers Deep-Dive Debugging: Perform root-cause analysis on complex issues and distributed tracing gaps across polyglot environments. Support the Great Minds: Act as a strategic consultant for senior engineers at our largest enterprise customers, s
Jobs in United States
System Power Engineer in San Francisco
914 active opportunities · Updated October 2026
Showing
15 jobs
Explore current system power engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT THE TEAM Critical Harm Operations sits within User Safety & Risk Operations and builds enforcement systems for Frontier Risk and Material Harm that are accurate, fast, defensible, and built to scale. The Cyber vertical turns policy into reviewer standards, calibrated judgment, quality systems, escalation paths, and automation guardrails. ABOUT THE ROLE We are looking for a senior cybersecurity practitioner and operations strategist to raise the quality, scalability, and technical rigor of our Cyber Operations. You will combine hands-on cyber judgment with systems-level operating design: resolve the hardest dual-use questions, evolve SOPs, uplift reviewers and vendors, and build practical tools and automations. This is a senior IC role. Success is not primarily cases closed; it is durable improvement in the operating model and the reviewers who run it. IN THIS ROLE, YOU WILL: Drive the Cyber Operations operating model across domain priorities, SOPs, escalation paths, quality health, vendor capability, roadmap inputs and help inform trusted access strategies. Serve as the senior cyber expert for complex or high-risk decisions across ChatGPT, API, Codex, agents, and emerging product surfaces. Translate policy ambiguity, quality misses, appeals, and reviewer disagreement into clear decision rules, calibration examples, training, and tooling requirements. Build durable operating systems and quality loops: golden sets, holdouts, double-labeling, adjudication, error taxonomies, reviewer calibration, and automation evaluations. Raise FTE and BPO capability through onboarding, certification, coaching, recurring calibration, and vendor-performance partnership. Use quality, appeals, SLA, backlog, and disagreement signals to diagnose root causes and prioritize high-leverage fixes. Build hands-on solutions—SQL analyses, scripts, dashboards, LLM eval workflows, evidence enrichment, routing logic, and lightweight automations—that improve decision quality and reduce manua
About the Team OpenAI’s Industrial Compute organization is building the infrastructure required to support the next generation of frontier AI systems. Through a combination of strategic partnerships and self-built data center campuses, we are scaling the physical infrastructure needed to deliver compute at unprecedented scale. The Commissioning organization is responsible for ensuring this infrastructure is safely tested, validated, integrated, and transitioned into reliable operations. As the portfolio grows, the team is building common standards, processes, tools, and reporting systems that allow commissioning programs to operate consistently across projects while giving teams and leadership clear visibility into readiness, risk, and execution. About the Role We are seeking a Commissioning Program Manager to build and scale the operating systems behind OpenAI’s infrastructure commissioning programs. You will own the development and continuous improvement of commissioning standards, processes, tools, dashboards, and KPIs across the infrastructure portfolio. You will work closely with commissioning and construction teams to translate field execution needs into practical playbooks, workflows, templates, metrics, and reporting mechanisms that teams can use from construction readiness through testing and turnover. This role sits at the intersection of infrastructure delivery, program management, process design, and data. The ideal candidate understands how complex construction projects operate and can turn fragmented workflows and project data into repeatable systems that improve execution without creating unnecessary administrative burden. Key Responsibilities Develop and maintain commissioning program standards, playbooks, process maps, templates, checklists, stage gates, and acceptance criteria across infrastructure projects. Establish consistent workflows for commissioning planning, construction readiness, QA/QC, issue management, document control, testing evidence
About the Team Critical Harm Operations sits within User Safety & Risk Operations and builds enforcement systems for Frontier Risk and Material Harm that are accurate, fast, defensible, and built to scale. We turn policy intent into operational readiness, review standards, quality systems, escalation paths, automation guardrails, and durable cross-functional operating models. About the Role We are looking for an exceptional Program Manager to help build durable operating systems and run some of OpenAI’s most complex safety operations. The core need is a high-agency operator who can take an ambiguous problem, create the right structure, align cross-functional partners, and drive the work through execution. This role will move across Critical Harm priorities as needs evolve. You may step into operationalizing national security or violent-activities workflows, support wellbeing and Frontier Risk initiatives, or help scale programs such as Trusted Access. Deep domain expertise is helpful but not required; the strongest candidates will learn quickly, exercise excellent judgment, and make complex programs move. In this role, you will: Lead strategic operational builds across priority workflows from problem statement to implemented operating model, including scope, owners, milestones, risks, success measures, and execution cadence. Translate policy, safety, technical, legal, and operational constraints into workflows, requirements, playbooks, escalation paths, and decision-making structures that teams can execute. Coordinate with User Ops leadership, Product Policy, Integrity, Safety Systems, i2, Legal, Product, Engineering, Support, vendors, and other partners to resolve dependencies and keep critical work moving. Move in and out of workflows as priorities shift—standing up new programs, stabilizing operations, improving handoffs, and transitioning durable ownership to the right team. Use operational data and frontline signals to identify bottlenecks, quality gaps, ca
About the Team OpenAI’s Financial Engineering and Identity data science team owns how revenue flows through our products and builds the systems that enable people and organizations to access OpenAI products safely, seamlessly, and at global scale. Identity sits at the critical intersection of growth, trust, and user experience. The team owns the experiences and infrastructure behind sign up, sign in, account recovery, authentication, and identity integrations across both consumer and enterprise products. As OpenAI expands across products and markets, Identity plays an increasingly important role in helping more users get started quickly while protecting them from abuse, fraud, and account compromise. About the Role We're looking for the first dedicated Data Scientist to partner with the Identity organization. In this role, you will define how we measure success across the entire identity journey—from first-time sign up and onboarding through authentication, account recovery, and enterprise identity experiences. You'll develop the experimentation frameworks, metrics, and analytical approaches that guide product decisions while helping the team navigate one of Identity's core challenges: optimizing growth while maintaining trust and security. You'll work closely with Product, Engineering, Design, Abuse, Risk, and Go-to-Market teams to identify opportunities, quantify trade-offs, and influence strategy. Some questions can be answered through A/B tests. Others require observational analyses, causal inference, and judgment under uncertainty. This is an opportunity to shape the analytical foundations of a high-impact product area from the ground up. This role is based in San Francisco, CA. We use a hybrid model (3 days/week in office) and offer relocation support. In this role, you will Define the north-star metrics and measurement frameworks used to evaluate the identity experience across consumer and enterprise products. Design and analyze experiments to optimize top-of
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the Role We are looking for customer-focused software engineers to build effective custom software that leverages OpenAI’s APIs to solve real customer problems. As an FDSWE, you will work with our customers and OpenAI Forward Deployed Engineers to design and implement scalable solutions that solve their most difficult problems. You will design abstractions to solve customer problems, and then use them to scale our speed and quality of delivery across all Forward Deployed engagements. You will collaborate closely with Sales, Solutions Engineering, Solutions Architects, and Customer Success Managers who work on the same account. You will also work with our Research and Applied Product and Engineering teams to provide insightful customer feedback. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role, you will: Embed deeply with strategic customers to understand their business challenges and technical requirements in detail. Design, architect, and develop full-stack solutions using an experiment-driven, iterative approach. Prepare detailed scopes of work and project plans for both proof-of-concept prototypes and full production deployments. Work hands-on with customers' technical teams as a technical expert and trusted advisor, coding side-by-side to drive projects to completion on their infrastructure. Collaborate with Product, Research and Applied teams to ensure seamless customer experiences, project success and actionable product feedback Contribute to internal knowledge bases, codifying best practices and sharing insights gained from customer engagements to scale the Forward Deployed Engineering function. You’ll thrive in th
About the Team Industrial Compute is building the infrastructure ecosystem that enables OpenAI to train and deploy increasingly capable AI systems at unprecedented scale. The organization operates across compute supply, demand, infrastructure, partnerships, and the physical and commercial systems required to make large-scale compute available. Industrial Compute Strategy & Operations serves as the connective operating layer across this ecosystem. The team works directly with senior leadership across Scaling, Finance, Partnerships, Research, and Infrastructure to translate ambiguous, high-impact challenges into clear strategies, scalable operating mechanisms, and decisive execution. This team is responsible for ensuring that OpenAI’s compute strategy evolves into durable competitive advantage by identifying systemic constraints, aligning stakeholders around critical decisions, and driving the operating mechanisms required to execute at scale. About the Role We are seeking a highly experienced Strategic Operations leader to help shape and operationalize OpenAI’s compute strategy across supply, demand, infrastructure, partnerships, and commercial strategy. This is a senior individual contributor role operating at the intersection of strategy, operations, infrastructure, and executive decision-making. You will work closely with compute leadership to identify the most consequential problems facing the organization, develop structured approaches to solving them, align stakeholders across the company, and drive initiatives from ambiguous concepts through execution. The role will span both strategic and operational work. You may develop long-range compute strategies and investment frameworks, evaluate build-versus-buy decisions, shape major commercial transactions, establish organizational planning mechanisms, or take ownership of a cross-functional initiative that does not have a clear organizational home. Success in this role requires exceptional judgment, analytical
About the Team The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust. About the Role As a Research Program Manager, you will lead initiatives that test the safety and robustness of OpenAI’s models through creative experimentation and structured evaluation. You’ll coordinate efforts across research and engineering teams to transform ambiguous risks into concrete research programs and influence future model development and deployment. We’re looking for people who are technically savvy, comfortable with ambiguity, and excited about shaping the future of safe AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead programs that explore unexpected model behaviors and identify failure modes. Translate vague or emergent risk signals into clear priorities and actionable research plans. Design and run creative evaluations, experiments, and red-teaming campaigns. Collaborate with research, product, and deployment teams to integrate findings into model training and deployment cycles. Develop repeatable systems for tracking model performance and understanding emerging behavior patterns. You might thrive in this role if you: Have strong experience in technical program management, with excellent organizational and communication skills. Are familiar with large language models, prompt engineering, or model evaluation techniques. Are comfortable managing fast-paced, high-uncertainty projects and shaping them from the ground up. Are creative and resourceful in devising new methods for testing model behavior and performance. Can effectively coordinate across technical and non-technical stakeholders to drive alignment and execution. About OpenAI OpenAI is an AI resear
About the role We’re looking for an engineering manager to lead a team building software systems that detect and prevent harmful misuse of frontier AI models—before incidents occur. This is a builder’s role: you’ll lead engineers shipping production services, detection pipelines, and mitigation mechanisms that protect frontier model integrity and reduce high-severity misuse risk. While this work intersects with frontier model development, security and risk, we’re explicitly seeking someone with a software engineering foundation who is comfortable building reliable systems that can operate at billions of users scale. In this role you will: Lead a team of software engineers building detection + mitigation systems for frontier model misuse, with an emphasis on model IP protection / distillation detection and emerging risk surfaces from autonomous agents. Set the technical roadmap and execution strategy: prioritize, design, ship, iterate, measure impact. Build production systems: services, pipelines, tooling, instrumentation, and automation that scale with frontier model usage. Partner deeply with Research and Product to translate evolving model capabilities into concrete tests, signals, and mitigations that can be deployed at scale. Drive strong engineering fundamentals: architecture, reliability, monitoring, performance, and operational excellence. Hire and grow an exceptional team across backend, data systems, and applied ML engineering domains as needed. Anticipate what breaks at scale as agentic workflows become more capable. You might thrive in this role if you: Experience building systems in adversarial, fast-evolving environments Are comfortable with ambiguity and novelty Have experience adjacent to security (e.g., abuse prevention, fraud, integrity, platform defense, auth/identity, malware/spam, adversarial environments) Communicate clearly and build trust quickly with senior stakeholders—pragmatic, collaborative, and calm under scrutiny. Significant experience
About the Team Employee Tech & Experience (ETX) helps people at OpenAI do their most ambitious work. Across Helpdesk, Executive Support, Systems Operations, Logistics and AV, we make technology simple, reliable and secure. Employee needs guide what we build, improve and choose to eliminate. About the Role Reporting to the Head of Global IT, you’ll lead ETX globally, building on the team’s capabilities and customer-zero work to continually advance the employee experience. You’ll shape ETX’s strategy, investment priorities and operating model in partnership with leadership across the company, turning new capabilities into measurable amplification. You’ll develop leaders and strengthen teams where people feel valued, own meaningful work and enjoy working together. This role is based at our San Francisco headquarters and requires an in-office presence. In this role, you will: Lead the next stage of ETX’s global growth across Helpdesk, SysOps, Logistics and AV, with a shared strategy and accountability for employee outcomes. Continually elevate the employee experience through research and design, directing investment to simplify entire user journeys, remove unnecessary effort and amplify what employees can accomplish. Accelerate ETX’s agent-led and customer-zero work as capabilities advance: continually challenge which workflows need to exist, extend what agents can own end to end and evolve the operating model to deliver measurable amplification. Develop leaders who earn trust, grow others and sustain an environment where people feel valued, take pride in their work and enjoy working together. Give people meaningful ownership and opportunities to stretch and grow, with clear priorities and sustainable workloads. Scale global services to support company growth, with clear commitments to reliability, security, responsiveness, and effective controls. Measure performance through employee effort, service quality, and time to resolution. Shape priorities and investment wi
About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products—including next-generation ads experiences—that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, and Research to bring research breakthroughs into real-world systems at global scale. About the Role As a Research Engineer in OpenAI's Monetization Group, you will have the opportunity to work with some of the brightest minds in AI. You'll contribute to deploying state-of-the-art models in production environments, helping turn research breakthroughs into tangible solutions. If you're excited about making AI technology accessible and impactful, this role is your chance to make a significant mark. In this role, you will: Innovate and Deploy: Design and deploy advanced machine learning models that solve real-world problems. Bring OpenAI's research from concept to implementation, creating AI-driven applications with a direct impact. Collaborate with the Best: Work closely with researchers, software engineers, and product managers to understand complex business challenges and deliver AI-powered solutions. Be part of a dynamic team where ideas flow freely and creativity thrives. Optimize and Scale: Implement scalable data pipelines, optimize mod
$340K – $425K/yr
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Product at Brex The Product team is at the forefront of Brex's mission to empower employees anywhere to make better financial decisions. With a deep understanding of the business, we identify and scope out the most impactful opportunities for Brex to tackle. We are responsible for aligning cross-functional teams — such as Engineering, Legal, Compliance, and Design — on key decisions. We set strategy and drive products from inception to launch, enabling Brex to grow rapidly and help our customers reach their full potential. What you’ll do As a Product Leader at Brex, you will be the driving force behind our Growth Product team — overseeing the thoughtful strategy and execution of team and technical systems to drive customer acquisition and onboarding. You will also collaborate closely with our Go-to-Market (GTM) teams to ensure seamless acquisition and onboarding for customers, particularly those with significant and complex spending n
About the Team Compute Foundations builds the software that manages OpenAI’s GPU compute infrastructure across sites, data centers, and infrastructure providers, supporting model training and inference. Our systems turn large, heterogeneous fleets of machines into dependable compute for research and products. We build Kubernetes-based control planes, controllers, services, and APIs that coordinate the lifecycle of machines and clusters. We connect global infrastructure management with the realities of bare-metal systems, giving clients consistent interfaces across differences in hardware, topology, and provider behavior. About the Role You will build distributed systems that provision, configure, and manage compute throughout its lifecycle. Your work will connect global services and Kubernetes controllers with the systems that bring machines online, update them safely, and recover them when something goes wrong. This role combines software architecture with an understanding of how machines and data centers work. You might design a lifecycle API, improve controller performance under high concurrency and provider rate limits, or trace a provisioning failure from an API through reconciliation to network boot or host configuration. You will help these systems remain reliable as the fleet expands across sites and generations of GPU hardware. We value depth in relevant systems and the ability to connect layers. You do not need to arrive as an expert in every component of the stack. In this role, you will: Design, build, and operate Kubernetes-based controllers and distributed services that coordinate infrastructure across sites, isolate failures, and scale as GPU capacity grows. Define APIs and resource models that let clients request and track lifecycle operations through consistent interfaces across hardware platforms and providers. Build provisioning and configuration services that coordinate network boot, hardware management interfaces, and the deployment of firmware,
About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models. We’re looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience. You’ll transform complex operational data into actionable insights that directly influence how OpenAI allocates and scales one of the world’s largest AI compute environments. Key Responsibilities Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency. Develop forecasting models for inference demand across products, regions, and model families. Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities. Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies. Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs. Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions. Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps. Communicate technical findings clearly to both engineering teams and executive leadership. Qualifications MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience). 5+ years of experience working in the infrastructure data science space. Strong ex
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity Are you looking for a high-impact finance role that offers true flexibility? We are seeking a Revenue Accountant to join our team in a fully remote capacity . In this role, you will be a key contributor to our Revenue Accounting team, supporting day-to-day revenue operations and helping maintain our financial processes. This is an exciting opportunity to build your technical accounting skills in a fast-paced environment while working collaboratively across teams. If you thrive in a dynamic, remote setting and are eager to learn and grow, this is the role for you. What You'll Do: Support the Close: Assist with the revenue-related month-end close process, helping ensure financial reporting is accurate and completed on schedule. Contract Analysis: Help review customer contracts and non-standard terms to support proper revenue recognition in line with company policies. Financial Schedules: Prepare and reconcile revenue schedules, including deferred revenue and contract assets/liabilities, under guidance from senior team members. Process Support: Support continuous improvement efforts within Order-to-Cash processes and assist with SOX audit compliance tasks. Cross-Functional Collaboration: Collaborate on ad-hoc accounting tasks and projects, delivering reliable analysis and support to the team. This Role Requires: Solid Foundation: 5+ years of relevant accounting experience. ASC 606 Knowledge: Strong foundational understanding of US GAAP and ASC 606 principle
Other cities to consider
More places hiring for this role
Get new system power engineer jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime