Jobs in United States

Automation Senior Developer in San Francisco

333 active opportunities · Updated October 2026

Explore current automation senior developer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will work on the systems software strategy and execution that brings new AI silicon from first power-on to a fully integrated system running production-representative models at expected functionality and performance. You will define how software exercises and validates compute, memory, interconnect, and I/O subsystems, then build the diagnostics, automation, and observability needed to find issues quickly. This role sits at the center of silicon, firmware, platform, systems, and workload teams. You will turn hardware specifications and performance targets into an end-to-end bringup plan, drive cross-functional debug, and establish the stress and regression infrastructure that makes each new platform reliable across operating environments. In this role, you will: Contribute to the end-to-end software bringup and validation strategy for new silicon and first-party systems. Define software-driven test coverage across compute, memory, interconnect, I/O, and their system-level interactions. Build diagnostics, test automation, telemetry, and regression infrastructure that accelerate first-silicon learning and issue isolation. Lead bringup from initial silicon arrival through board and system integration, docking, runtime enablement, and model execution. Design stress tests that characterize reliability, performance, and stability across workloads and operating conditions. Translate architecture specifications and performance models into measurable acceptance crit

PythonAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Intelligence and Investigations team is dedicated to ensuring the safe, responsible deployment of AI by rapidly detecting and mitigating abuse. Our team leverages the latest testing methodologies to uncover vulnerabilities and emerging threats, helping safeguard OpenAI’s products and users. We work closely with cross-functional partners across product, policy, and engineering to drive a comprehensive defense strategy against evolving adversarial challenges. About the Role As a Red Team Specialist focused on cyber, you will help answer two practical questions: What cyber capabilities can our models provide to real-world attackers, and do our safeguards remain effective when those attackers use increasingly sophisticated techniques? The role combines scaled evaluation with expert-driven testing. You may bring deeper experience in cybersecurity and use that expertise to judge whether a model’s behavior meaningfully changes attacker capability. Alternatively, you may bring deeper experience in model evaluations, automation, or agentic harnesses and apply those skills to building rigorous cyber testing. We do not expect every candidate to be equally deep in both areas, but successful candidates will have a strong foundation in one and enough fluency in the other to work effectively across the boundary. Most of your work will focus on model cyber capabilities and safeguards; you will also spend a portion of your time testing novel abuse risks in agentic systems. This role is located in San Francisco, CA or Seattle, WA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design and run rigorous evaluations of model cyber capabilities and safeguards, including policy adherence, correct refusal, over refusal, and resilience to jailbreaking and other adversarial techniques. Conduct hands-on testing to understand what models can enable when used by experienced security practiti

AWSRestAIGo
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$220K – $450K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry's Infrastructure Engineering team is what makes operating Sentry simple, safe, and seamless for every other engineering team in the company. They build the internal control platforms, configuration systems, traffic routing, and automation that let product engineers operate services safely at scale without needing deep infrastructure expertise themselves. As the Engineering Manager for Infrastructure Engineering, you'll lead a team of engineers building the tools that power Sentry's growth: internal admin and change management tools, configuration automation, and the routing layer that underlies Sentry's architecture. You'll be responsible for technical vision, team health, system reliability, and partnership with engineering teams across the company who depend on your team's tools every day. You'll work closely with leaders across Infrastructure, Platform, and Production Engineering to shape how Sentry scales its operational model as the company grows. In this role you will Lead a team of engineers building the internal control platforms that every engineering team at Sentry relies on to operate services safely. Drive the evolution of Infrastructure Engineering's platform, including configuration management, traffic routing and environment controls Own the team's technical direction, contributing to key decisions on API architecture, internal tooling design, and automation frameworks. Nurture and grow engineers at different levels, providing support through coaching, mentorship, and career development. Foster an inclusive, high-performing team culture focused on ownership, learning, and delivery. Partne

PythonKubernetesAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s Network Engineering team within IT and Security advances the mission of deploying artificial general intelligence (AGI) for the benefit of all by delivering secure, scalable, and resilient network services. We build and operate the connectivity that supports OpenAI’s offices, labs, campuses, cloud environments, people, and devices. By combining strong network fundamentals with security, reliability, automation, and user-centered design, we enable impactful AI research, corporate operations, and product innovation. About the Role As a Network Engineer at OpenAI, you will design, operate, and continuously improve the global networks that connect our offices, labs, campuses, PoPs, cloud environments, people, and devices. The role spans strategic platform engineering and responsive production operations: you will shape architecture, standards, roadmaps, lifecycle plans, and automation while supporting incidents, escalations, and time-sensitive delivery. Operational signals will inform what we stabilize, simplify, standardize, or automate next. We work backward from user needs, investigate root causes, own outcomes end-to-end, and move quickly without compromising security. We are looking for a versatile engineer who can make pragmatic reliability and security tradeoffs, communicate clearly, and turn recurring operational work into durable platforms, tooling, and standards. You will partner across IT, Security, AppEng, Research, Applied, workplace teams, carriers, and vendors. In this role, you will: Design, implement, and operate secure, scalable enterprise networks across offices, labs, campuses, PoPs, cloud connectivity, and hybrid environments. Set strategic direction for network services through architecture, standards, roadmaps, lifecycle planning, capacity strategy, and measurable reliability outcomes. Own production operations, including on-call, incident response, escalations, and time-sensitive delivery, while protecting user experience,

PythonAWSAzureCI/CD
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team This team builds and operates the systems that enable OpenAI researchers to run reliable, scalable, and efficient research workflows. The team sits close to research and works across infrastructure, systems, and automation to make sure researchers have the tools and environments they need to move quickly. The work spans software engineering, infrastructure, systems administration, cluster operations, and reliability engineering. As OpenAI’s infrastructure evolves from bespoke bare-metal systems toward more standard, scalable platforms, the team needs engineers who can understand how systems work end-to-end and build the right abstractions without reinventing the wheel. About the Role As a Software Engineer on this team, you will build and operate the infrastructure that supports frontier research and critical research-facing systems. You will work on systems that sit close to the metal, but the role is not limited to classic operations or sysadmin work. We are looking for someone who can reason about networking, bootstrapping, Kubernetes, scalability, automation, and reliability - while also writing software to make these systems better over time. This role is a strong fit for an independent, high-ownership engineer who enjoys reliability-heavy infrastructure work but still wants to build. You do not need to come in as a kernel expert or highly algorithmic optimization engineer, but you should be deeply curious about infrastructure, comfortable debugging complex systems, and excited to support researchers doing novel work. We expect you to: Build and operate reliable infrastructure for research workloads and research-facing services. Support and improve systems across data infrastructure, processing, crawl and ingest, caching, search, observability, and clusterwide services. Improve cluster bootstrapping, provisioning, automation, and deployment workflows. Debug issues across networking, compute, storage, orchestration, and service reliability layers.

AWSKubernetesCI/CDGit
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team Critical Harm Operations sits within User Safety & Risk Operations and builds enforcement systems for Frontier Risk and Material Harm that are accurate, fast, defensible, and built to scale. We turn policy intent into operational readiness, review standards, quality systems, escalation paths, automation guardrails, and durable cross-functional operating models. About the Role We are looking for an exceptional Program Manager to help build durable operating systems and run some of OpenAI’s most complex safety operations. The core need is a high-agency operator who can take an ambiguous problem, create the right structure, align cross-functional partners, and drive the work through execution. This role will move across Critical Harm priorities as needs evolve. You may step into operationalizing national security or violent-activities workflows, support wellbeing and Frontier Risk initiatives, or help scale programs such as Trusted Access. Deep domain expertise is helpful but not required; the strongest candidates will learn quickly, exercise excellent judgment, and make complex programs move. In this role, you will: Lead strategic operational builds across priority workflows from problem statement to implemented operating model, including scope, owners, milestones, risks, success measures, and execution cadence. Translate policy, safety, technical, legal, and operational constraints into workflows, requirements, playbooks, escalation paths, and decision-making structures that teams can execute. Coordinate with User Ops leadership, Product Policy, Integrity, Safety Systems, i2, Legal, Product, Engineering, Support, vendors, and other partners to resolve dependencies and keep critical work moving. Move in and out of workflows as priorities shift—standing up new programs, stabilizing operations, improving handoffs, and transitioning durable ownership to the right team. Use operational data and frontline signals to identify bottlenecks, quality gaps, ca

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Ads Support Delivery team is responsible for helping successfully operate and grow on our Ads product. This includes technical guidance, troubleshooting complex delivery and monetization issues, and partnering closely with Product, Engineering, Trust & Safety and Go-To-Market teams to resolve customer-impacting problems and improve the platform over time. The team’s mission is to deliver a high-quality customer experience at scale by combining strong human support with automation, self-service, and AI-enabled workflows, while maintaining high operational rigor. About the Role: As a Support Delivery Lead for Ads, you will lead a team responsible for end-to-end support delivery across the ads ecosystem, including campaign setup, delivery, billing, measurement, and policy navigation. You will set the operational bar for quality, responsiveness, and consistency; coach and grow the team; and translate support signals into actionable improvements with Engineering, Product, and Go-To-Market partners. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead and support a team of Ads support engineers, ensuring they have the tools, clarity, and coaching needed to operate at a high bar in a technically complex domain. Set clear expectations and operating standards, run recurring performance reviews, and build development plans that grow both technical depth (ad tech fluency) and customer-facing excellence. Design and continuously improve support coverage for ad buyers, ensuring the team can diagnose delivery issues and monetization and integration issues with equal rigor. Act as the bridge between Support Delivery, Engineering, Product, and Go-To-Market teams. Drive alignment on priorities, escalation paths, launch readiness, tooling improvements and mechanisms to reduce repeated customer pain points. Partner with engineering teams

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Treasury team is responsible for protecting liquidity, enabling scale, maintaining strong controls, and helping the company operate with clarity and resilience. We work across finance and operational partners to ensure funds move safely, visibility remains high, and Treasury infrastructure keeps pace with a fast-moving business. As hands-on operators and builders, we combine financial judgment with AI, automation, and data to solve problems faster, strengthen controls, and continuously improve how Treasury operates. About the Role We’re looking for a Treasury Manager to own and execute critical activities across OpenAI’s global treasury operations, including cash management, payments, bank account management, forecasting support, controls, and reporting. Beyond day-to-day operations, this person will support Treasury leadership on high-impact projects such as M&A integrations, new legal entity formation, and geographic and currency expansion. They’ll also help build scalable, AI-enabled systems, workflows, and controls for a rapidly growing global business. This is an individual contributor role with broad scope, spanning operational ownership and building an AI-native Treasury function. This role is based in our San Francisco HQ. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own core global treasury operations across cash positioning, payments, bank account management, portal administration, forecasting support, and FBAR preparation. Coordinate day-to-day execution for intercompany funding, settlements, investment operations, letters of credit, guarantees, treasury close, and related accounting handoffs. Automate and scale treasury workflows for request intake, approvals, payment tracking, bank account changes, KYC follow-up, access reviews, evidence collection, and issue resolution. Identify recurring processes, manual pain points, duplicate work, control

AWSRestAIRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI’s Infrastructure Operations team is responsible for the availability, reliability, and operational excellence of one of the world’s largest AI infrastructure networks. The team owns day-to-day operations of production AI networks across Industrial Compute's data centers, working with colocation providers, deployment teams, and hardware vendors to deliver highly available GPU infrastructure for AI training and inference workloads. About the Role We are seeking an Infrastructure Operations Engineer to operate and improve the large-scale Ethernet fabrics that support GPU clusters, storage systems, and management infrastructure. This role combines hands-on production operations with automation, observability, and incident response across a global AI network. The ideal candidate has experience operating high-availability data center, cloud, AI, or HPC networks and can move comfortably from physical-layer troubleshooting to routing and fabric behavior, change execution, and root-cause analysis. You will partner closely with network architecture, systems engineering, GPU engineering, storage engineering, security, deployment, site operations, service providers, colocation partners, and hardware vendors to raise reliability and reduce operational toil. Key Responsibilities Own the operational health, availability, and reliability of production AI network infrastructure across Industrial Compute's data centers. Monitor, troubleshoot, and resolve network incidents while meeting service-level objectives (SLOs), reducing Mean Time to Detect (MTTD), and minimizing Mean Time to Recovery (MTTR). Operate and maintain large-scale Ethernet fabrics supporting GPU compute, storage, and management networks. Execute production network changes, maintenance windows, and capacity expansions with minimal customer impact. Manage the hardware lifecycle, including switch and optics replacements, RMA coordination, software upgrades, and preventive maintenance. Support new A

PythonAWSAzureGit
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The ChatGPT Model Flywheel team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement. Team Focus Areas Model Experimentation: Enable rapid, safe model validation for ChatGPT and Codex products through experiment automation and lifecycle management. Model Deployment: Ensure safe, scalable deployment of model capabilities with robust rollout and operational tooling. Automate capacity management and incorporate platform-wide health monitors. Model Measurement: Build comprehensive evaluation and measurement systems for model quality, from user signals to launch scorecards. Improve end-to-end feedback loops for continual model improvement. Key Partnerships Collaborate cross-functionally with teams including Model Measurement DS, Research, Codex, Fleet, Inference, and API. In this role, you will: Elevate and consolidate ChatGPT’s harness, context management, and system prompt frameworks. Drive expansion and improvement of multi-tier model experiences. Support and scale self-serve experiment capabilities and automated guardrails. Lead model rollout automation, capacity management, and health monitoring. Shape end-to-end measurement systems (evals, grader signals, user feedback, etc.). You might thrive in this role if you have: Proven experience leading engineering teams in complex, cross-functional environments. Demonstrated success shipping production systems at scale (ideally for AI or large backend services). Deep understanding of model-driven product development, deployment lifecycle, and measurement tooling. Excellent communication and collaboration skills—experience interfacing directly with engineering, research, and product stakeholders. Prior involvement with large language models, distributed infrastructure, or experimentation platforms is a plus. Why Work With Us Tackle highly impactful technical challenges at the cutting edg

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Finance Platform & Technology team at OpenAI builds and scales the future-proof systems and data architecture that power our core financial operations. We enable business agility, compliance, and operational excellence across quote-to-cash, procure-to-pay, inventory, and asset management for both B2B and B2C. Our focus is on modernizing workflows through strategic integrations, scalable automation, and seamless data flows empowering smarter decisions, reliable reporting, and sustainable growth as OpenAI evolves About the Role We are looking for a Supply Chain Transformation Architect to redesign and modernize our end-to-end supply chain operations supporting robotics, consumer hardware, and data center infrastructure. This role sits at the intersection of process, systems, and data. Your primary focus will be transforming supply chain processes across planning, procurement, manufacturing, logistics, and fulfillment—then enabling those processes with the right systems architecture, data foundation, and AI-driven automation. You will help move the organization from manual, reactive operations to intelligent, data-driven supply chain execution. In this role you will: Lead End-to-End Supply Chain Transformation Evaluate and redesign core supply chain processes across demand planning, supply planning, procurement, manufacturing operations, logistics, and fulfillment. Identify operational bottlenecks, fragmented workflows, and manual processes that limit scalability. Build standardized process frameworks and operating models that support rapid scaling of hardware programs. Drive Operational Excellence Implement structured supply chain practices such as: S&OP / Integrated Business Planning Supply risk management Inventory optimization Supplier collaboration frameworks Logistics visibility and execution models Establish operational KPIs and governance to improve predictability, responsiveness, and resilience. Architect the Digital Supply Chain Tra

ReactAWSGitRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team OpenAI's mission is to ensure that artificial general intelligence benefits all of humanity. The Consumer Devices team is building a new generation of AI-powered products that seamlessly integrate hardware and software to create intuitive, transformative experiences. We bring together experts across embedded systems, machine learning, hardware, design, and product engineering to develop products at the intersection of AI and consumer technology. About the Role OpenAI is seeking a System Power Engineer to characterize, measure, and optimize power consumption across our embedded hardware products. In this role, you will work closely with Electrical Engineering and system software teams to build power test automation, measure subsystem-level power usage, and drive improvements that directly impact battery life, thermal behavior, charging performance, and system reliability. You will help establish the methodologies and metrics used to understand and improve power efficiency across real-world product experiences, from controlled lab environments to representative day-in-the-life usage scenarios. This role requires hands-on experience with embedded hardware platforms, power instrumentation, and the analysis of power profiles and system behavior. This role is based in San Francisco, CA. We use a hybrid work model of four days per week in the office and one day working remotely. Relocation assistance is available for new hires. In this role, you will: Define and develop power testing automation to evaluate system behavior across a range of workloads and operating conditions. Measure subsystem-level power consumption using power breakout probes and other lab instrumentation. Develop and execute power characterization tests spanning basic workloads, complex mixed-use scenarios, and representative day-of-use experiences. Partner closely with Electrical Engineers to identify opportunities to improve system power efficiency. Collaborate with software engineering

PythonAWSRestMachine Learning
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

This role will support the fleet infrastructure team at OpenAI. The fleet team focuses on running the world’s largest, most reliable, and frictionless GPU fleet to support OpenAI’s general purpose model training and deployment. Work on this team ranges from Maximizing GPUs doing useful work by building user-friendly scheduling and quota systems Running a reliable and low maintenance platform by building push-button automation for kubernetes cluster provisioning and upgrades Supporting research workflows with service frameworks and deployment systems Ensuring fast model startup times though high performance snapshot delivery across blob storage down to hardware caching Much more! About the Role As an engineer within Fleet infrastructure, you will design, write, deploy, and operate infrastructure systems for model deployment and training on one of the world’s largest GPU fleet. The scale is immense, the timelines are tight, and the organization is moving fast; this is an opportunity to shape a critical system in support of OpenAI's mission to advance AI capabilities responsibly. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, implement and operate components of our compute fleet including job scheduling, cluster management, snapshot delivery, and CI/CD systems. Interface with researchers and product teams to understand workload requirements Collaborate with hardware, infrastructure, and business teams to provide a high utilization and high reliability service You might thrive in this role if you: Have experience with hyperscale compute systems Possess strong programming skills Have experience working in public clouds (especially Azure) Have experience working in Kubernetes Execution focused mentality paired with a rigorous focus on user requirements As a bonus, have an understanding of AI/ML workloads About OpenAI OpenAI is an AI resea

AWSAzureKubernetesCI/CD
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Employee Technology & Experience (ETX) team is responsible for delivering a world-class internal technology experience that enables employees to do their best work. We support and operate the employee-facing systems that keep the company moving quickly and efficiently. ETX spans support, logistics, AV, identity, endpoints, SaaS administration, automation, enterprise tooling, and internal infrastructure operations. We partner closely with Security, Engineering, Workplace, Finance, People, and other teams to keep OpenAI’s internal technology reliable, scalable, and moving at the pace of the company. About the Role We are hiring a Program Manager to help scale how IT operates across OpenAI. This role will lead complex cross-functional programs that improve operational maturity, streamline how teams work together, and turn high-impact initiatives into durable operational capabilities. You will work across IT, Security, Engineering, Workplace, and other functions to drive alignment, remove friction, and help build the operational foundation needed to support OpenAI’s rapid growth. You’ll be responsible for: Lead cross-functional operational programs that improve scalability, consistency, and operational maturity. Drive operational excellence initiatives across IT Support, employee lifecycle operations, meeting room and calendaring services, onsite support, vending, and research support environments. Build operating models, readiness plans, escalation paths, governance cadences, and success metrics for complex operational programs. Partner with technical teams to ensure new deployments, infrastructure investments, and internal platforms are operationally ready and sustainably supported at scale. Drive high-priority operational programs supporting company growth, including infrastructure expansion, operational integrations, and other emerging initiatives. Improve operational visibility, stakeholder alignment, and coordination across long-running cros

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are hiring a Sim Infrastructure Engineer to turn simulation systems into reliable, automated, production-quality pipelines that power model training, evaluation, and hardware-in-the-loop validation. This role owns the automation, orchestration, and tool integration that apply simulation to concrete robotics tasks: building CI/CD for SIL/HIL, presubmit checks, automatic model evaluation, metric computation and reporting, and the runtime infrastructure to run simulations at scale. You will collaborate closely with Sim Realism, Sim Environments, Research, and Ops to make simulation an integrated, reproducible, and measurable part of our ML and robotics workflows. This role is based in San Francisco, CA, and requires in-person 4 days a week. In this role, you will: Build and maintain presubmit checks, continuous integration and deployment pipelines for simulation code, environments, and tasks so simulation artifacts are testable, versioned, and reproducible. Implement end-to-end automation to run model evaluation in sim (SIL) and orchestrate HIL runs; compute realism and task metrics, generate dashboards and alerts, and ensure evaluation is repeatable and auditable. Create robust APIs and connectors so research, training, and data-collection systems can schedule, seed, and evaluate batches of simulations; support RL rollouts, imitation-data collection, and presubmit model checks. Build scheduling, batching and orchestration for running very large numbers of concurrent rollouts (target tens of thousands of rollouts / large RL workloads), sol

PythonAWSKubernetesCI/CD
🔔

Get new automation senior developer jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime