Jobs in United States

Hardware Operations Engineer in United States

408 active opportunities · Updated October 2026

Explore current hardware operations engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -85.4%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Lead at Baseten, you will lead the "engine room" of the company, architecting, securing, and optimizing the global GPU fleet that powers our customers' AI workloads. You’ll own the end-to-end journey of capacity management, from securing multi-million dollar GPU clusters to building the automation that ensures 99.9% uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering. You will act as the fleet orchestrator for the world's most advanced chips, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the next generation of hardware, like NVIDIA’s Blackwell (B200) architecture. EXAMPLE INITIATIVES The B200 Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's first Blackwell GPU clusters. Global Workload Orchestration: Building "Multi-cloud Capacity Management" systems to move customer workloads seamlessly across regions to optimize cost and latency. Precision GPU Triage: Developing automated Go-based operators to identify, cordon, and repair unhealthy H100 nodes in under an hour. The Supply Chain of Intelligence: Partnering with lead

PythonAWSAzureGCP
B
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -85.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a Delivery Director, Capacity programs for our on-premises data center builds and neo cloud (GPU cloud) delivery programs. This is a high-visibility, execution-critical role sitting at the intersection of infrastructure engineering, capacity planning, vendor/partner management, and customer delivery. You will own the end-to-end delivery lifecycle for large-scale compute infrastructure — from initial site/capacity commitments through power, networking, and hardware bring-up, to production-ready GPU/compute capacity landing in the hands of internal teams or customers. You'll be the person who turns ambitious infrastructure roadmaps into predictable, on-time, delivery. RESPONSIBILITIES Own delivery of on-prem infrastructure builds — colocation expansions, power/cooling readiness, rack-and-stack, network fabric bring-up, and hardware acceptance testing — coordinating across colo providers and partners, network engineering, hardware ops, and vendor teams. Drive neo cloud delivery programs — manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps, SLAs, and go-live readiness. Build and maintain master delivery schedules across concurrent, multi-site, multi-vendor programs, integrating power/shell timelines, hardware lead times, logistics, and software/platform readiness into a single critical path.

ReactMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -85.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. Product at Baseten Product at Baseten is a nascent function. Our company today has a strong engineering culture, is heavily customer-obsessed, and moves fast. We're building the product function now, and you'd be one of the first people who will help define it. You'll work directly with our founders and with some of the best systems and infrastructure engineers in the world, and you'll set the standard for building great AI Infrastructure. PMs at Baseten don't sit above engineers - you earn ownership by being technical, finding the truth in front of customers, building great cross-functional relationships, and shipping great product experiences. The role Getting a model into production still takes real expertise — choosing a serving engine, sizing hardware, tuning it, wiring it into an app. We want a developer to go from "it runs on my laptop" to "it's serving production traffic" in minutes, on their own. You'll own the entire experience a developer touches to deploy and iterate: the CLI and SDKs, the console, onboarding, model discovery, deployment configuration, truss, and the increasingly agent-driven ways developers build. Your job is to make Baseten synonymous with Great DevEx and make it effortless to drive and self-serve deploy models on Baseten for far more developers than it is today. Impact and outcomes you'll drive You will collapse time-to-production — take a developer from first sign-up to a running, maint

Machine LearningAIGoRust
M
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -100%

What you’ll do Execute weekly system-level exploratory testing across the scanner and supporting software; log and triage issues with clear reproduction steps. Work with engineering to debug root cause and validate fixes. Help maintain the DHF and traceability between user needs, design requirements, tests, and results. Own practical test execution logistics (fixtures, test data, environments, calibration artifacts) and keep things repeatable. Help build the continuous testing strategy: automated tests where feasible, plus structured manual and system tests. Support V&V activities, including coordination with external partners as needed. What we’re looking for Strong hands-on testing instincts for complex electromechanical systems with substantial software. Ability to write clear bug reports and communicate risk/impact. Experience building and maintaining test plans/protocols; comfort operating lab equipment and debugging across layers. Useful experience Experience testing complex systems end-to-end (automation where it pays off, plus hands-on hardware/instrumentation). Medical device or other safety-critical environments and comfort translating risk into practical test coverage.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About OpenAI OpenAI is dedicated to ensuring that artificial general intelligence (AGI) benefits all of humanity. Our mission requires building not only world-class AI models, but also the infrastructure that enables those models to be deployed reliably, efficiently, and at global scale. As demand for AI continues to grow, we are expanding the ways OpenAI can bring high-performance inference capacity online across a diverse hardware ecosystem. About the Team The GPT Infrastructure team builds software that turns advanced inference and optimization research into production products. One focus is enabling strategic infrastructure partners and accelerator vendors to qualify and onboard new compute without a bespoke porting and optimization effort for every hardware platform. We build the control planes, APIs, secure partner-side execution environments, evaluation systems, artifact pipelines, and operational tooling that make these workflows repeatable and trustworthy. The work sits at the intersection of distributed systems, AI inference, compilers and runtimes, performance engineering, security, and external partnerships. About the Role We are seeking an experienced systems generalist who can work comfortably across the stack to help build an automated inference optimization platform. Given a workload, target hardware profile, compiler and runtime context, and a trusted verifier, the system runs durable optimization campaigns that generate, compile, execute, grade, and improve candidate kernels, runtime configurations, and serving-stack changes. You will design both the OpenAI-hosted control plane and the partner-side software that evaluates candidates on real accelerator hardware. The product must keep long-running workflows reliable, make performance results reproducible, and maintain clear trust boundaries around sensitive model and hardware information. This is a deeply cross-stack role, combining strong software engineering fundamentals with systems thinking and

PythonAWSLinuxRest
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team OpenAI’s Infrastructure organization builds the systems that power frontier AI workloads at global scale. As compute demand accelerates, our ability to rapidly convert infrastructure investments into usable production capacity has become mission critical. The CPU / Storage / PoP / WAN team is responsible for the end-to-end infrastructure layers required to bring compute online: server and cluster activation, storage platforms, Points of Presence (PoPs), backbone connectivity, and global network expansion. We operate across first-party facilities, colocation environments, and strategic cloud partners to ensure OpenAI can scale reliably and quickly. About the Role We are seeking a highly technical Program Manager to lead execution across CPU, Storage, PoP, and WAN infrastructure programs that directly unlock OpenAI’s next generation compute capacity. In this role, you will own complex cross-functional programs spanning compute cluster activation, storage deployment, PoP bring-up, and backbone expansion. You will coordinate hardware readiness, site readiness, network pathing, storage availability, vendor execution, and engineering dependencies required to turn contracted infrastructure into live training and inference capacity. This role requires strong technical fluency across hardware systems, network infrastructure, storage architecture, and deployment execution. You should be comfortable operating from rack-level implementation details through executive-level capacity planning discussions. This role is based in San Francisco, CA, with travel as needed. Key Responsibilities Lead end-to-end execution of CPU / GPU cluster activation programs across OpenAI’s global infrastructure footprint Drive readiness to convert contracted compute capacity into schedulable production clusters Own deployment programs for new PoPs, backbone nodes, WAN expansion, and interconnection initiatives Build integrated schedules spanning procurement, logistics, installation, st

AWSAzureRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team OpenAI’s Central Procurement – Hardware Operations team is building scalable, AI-enabled operating foundations for a rapidly growing hardware footprint. We help the business move quickly while maintaining the data quality and controls needed to manage financial and operational risk responsibly. As OpenAI scales across R&D, new-product introduction, manufacturing, and third-party custody models, we are building the foundations to absorb complexity without adding unnecessary friction. That means establishing practical standards, strengthening discipline where it matters, using AI thoughtfully, and continuously improving how teams manage hardware operations. About the Role We are hiring an Asset Compliance Program Lead to establish the cross-business framework that gives OpenAI reliable lifecycle visibility over hardware-related financial assets and capital equipment. You will define the controls, systems roadmap, and evidence practices needed to manage those assets within OpenAI’s financial-control scope. This is a senior individual-contributor role in Central Procurement – Hardware Operations within Finance. Distinct from sourcing and transaction execution, you will partner with hardware business units, Accounting, Financial Risk Management, Procurement, contract manufacturers, and other third-party custodians to deploy practical operating mechanisms. These mechanisms will connect how assets are purchased, built, received, moved, held, verified, and retired with the ownership, data, reporting, and evidence needed for financial governance. The role covers those assets regardless of location or custody, including manufacturing equipment and tooling, supplier- and contract-manufacturer-held assets, leased assets, and other third-party-held equipment. You will stay close to the operational details—how assets move, where records diverge, which controls are not working, and where evidence is incomplete—and use that view to strengthen processes, accountab

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving areas of work. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are seeking a Device Safety & Risk Operations Specialist to build the safety operating model for a new category of consumer hardware. This is a senior individual-contributor role for someone who can turn emerging product risks and incomplete requirements into practical workflows, controls, launch plans, and durable systems. You will define how product-safety incidents, critical escalations, regulated cases, and privacy-sensitive issues should be identified, investigated, escalated, resolved, and learned from. You will also establish operational requirements for case management, data access, decision logging, quality assurance, monitoring, and cross-functional response. You will stand up priority workflows through launch and early operations, then help transition them into durable homes across USRO and partner teams. The right person combines deep operational judgment with strong technical and hardware product fluency. They can move from executive-level risk framing to detailed workflow design, tabletop exercises, launch readiness, frontline guidance, and post-launch improvement. Location / work model: San Francisco, CA; hybrid, 3 days/week in-office. Please note: This role may involve exposure to sensitive or concerning material. Strong discretion, judgment, and resilience are essential. In This Role, You Will:

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team At OpenAI, our User Safety & Risk Operations (USRO) team helps protect our products and users from abuse, fraud, safety risks, and other forms of misuse. We operate at the front line of real-world safety and risk management, translating user and operational signals into timely decisions, effective interventions, and improvements to our systems. This role sits on a team focused on building operational capacity for new, ambiguous, and fast-moving areas of work. The team defines what needs to be built, creates the operating model to support it, and works with partner teams to make the work scalable and durable over time. About the Role We are seeking a Device Safety & Risk Operations Specialist to build the safety operating model for a new category of consumer hardware. This is a senior individual-contributor role for someone who can turn emerging product risks and incomplete requirements into practical workflows, controls, launch plans, and durable systems. You will define how product-safety incidents, critical escalations, regulated cases, and privacy-sensitive issues should be identified, investigated, escalated, resolved, and learned from. You will also establish operational requirements for case management, data access, decision logging, quality assurance, monitoring, and cross-functional response. You will stand up priority workflows through launch and early operations, then help transition them into durable homes across USRO and partner teams. The right person combines deep operational judgment with strong technical and hardware product fluency. They can move from executive-level risk framing to detailed workflow design, tabletop exercises, launch readiness, frontline guidance, and post-launch improvement. Location / work model: San Francisco, CA; hybrid, 3 days/week in-office. Please note: This role may involve exposure to sensitive or concerning material. Strong discretion, judgment, and resilience are essential. In This Role, You Will:

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team Tax and Trade at OpenAI shapes business strategy by embedding critical tax, export control, customs, and cross-border considerations into how the company builds, sources, scales, and operates in support of the mission. We combine deep expertise with practical systems thinking to look around corners, identify emerging risks and opportunities early, and help teams make smarter decisions at the point where strategy becomes execution. Across procurement, hardware operations, manufacturing, logistics, finance, legal, supplier onboarding, and operator workflows, we build robust, scalable support services leveraging cutting-edge technology—including governed AI and automation—to make complex regulated work more durable, more efficient, and easier to scale. About the Role We’re hiring a Senior Manager, Export Controls to lead OpenAI’s export controls strategy and operating model. This is a senior role with broad scope across advanced computing, semiconductors, software, hardware, manufacturing, and high technology partnerships. You will refine how OpenAI classifies controlled technology, software, and hardware, structures access-controlled environments, manages licensing and supplier commitments, and scales export-control operations in a way that supports the company’s pace of innovation. You will also shape how OpenAI applies AI and agentic workflows to policy-heavy operational work, building systems that make complex rules easier to navigate and easier to execute. In this role, you will: Refine the strategy and operating model for OpenAI’s export controls program across advanced computing, semiconductors, software, hardware, manufacturing, and high technology partnerships. Own export classification and licensing strategy for controlled technical data, software, hardware, and research environments. Lead the design and operation of compliant controlled environments and related governance processes. Partner with Research and Infrastructure to support efficient

AWSRestAIGo
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team OpenAI is building Procurement of the future: one that uses AI-enabled systems, clean data, and scalable processes to help the business absorb increasing complexity, speed, volume, and throughput — while enabling teams across OpenAI to operate more strategically and at greater scale. The Hardware Procurement Operations team supports the operational flow of goods and hardware purchasing across OpenAI's hardware business units. We help teams turn business needs into controlled, accurate, and executable procurement activity, from request intake through purchase orders, supplier coordination, change management, receiving readiness, and downstream invoice support. As OpenAI scales, we are building procurement operations that can support changing hardware needs, new business unit priorities, and increasing operational complexity. That means creating practical standards where they help, preserving flexibility where the business requires it, and continuously improving how requests, suppliers, POs, data, and controls move through our systems. About the Role As a Senior Manager of Procurement Operations focused on Hardware, you will own procurement operations for one or more assigned hardware business units. You will be the day-to-day DRI for hardware and goods procurement activity, leading work from intake through purchase requests, POs, supplier needs, changes, receiving readiness, and operational escalations. This is a hands-on operator-builder role for someone who understands the practical complexity of hardware and goods procurement and can both run the work and redesign how the work gets done. You will keep work moving while improving the operating model through better systems, workflows, automation, AI-enabled processes, controls, metrics, and operating leverage. You will evaluate how people, process, policy, systems, data, fiscal controls, and accounting practices come together, and you will work closely with other Central Procurement and BU Procurement

AWSRestAIGo
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -88%

About the Team The Marketing team helps OpenAI bring products, research, and company priorities to the world with clarity, creativity, and impact. We partner across Product, Research, Design, Creative, Communications, Finance, Legal, and Go-to-Market teams to translate complex technology into work that people can understand, trust, and use. Marketing Strategy & Operations builds the operating systems that help this work move with focus and speed. We connect priorities to plans, plans to resourcing, and resourcing to execution, so teams can make better decisions earlier and stay aligned as priorities evolve. About the Role We’re looking for a Strategy & Operations Lead, Hardware Marketing to help build and run the operating system for OpenAI’s Hardware Marketing work. This role will partner closely with Hardware Marketing leadership, Product Marketing, Product, Design, Creative, Production, Communications, Finance, Legal, and external partners to bring structure to a fast-moving, highly cross-functional product area. You will help the team clarify priorities, build plans, manage intake, track decisions, surface risks, frame tradeoffs, and keep critical workstreams moving from planning through operating readiness. This is a strong fit for someone who can operate independently in ambiguity, build practical systems from scratch, and move fluidly between strategy and execution. Experience with complex product launches or sensitive, cross-functional product areas is preferred. This role is based in San Francisco. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Build and run the core operating rhythm for Hardware Marketing, including planning cycles, planning reviews, weekly priorities, decision forums, risk tracking, and leadership-ready updates. Partner with Hardware Marketing leadership to translate business and product priorities into clear workstreams, owners, milestones,

ReactAWSRestAI
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -88%

$342K – $445K/yr

Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking a Technical Lead to lead deployment and operations for OpenAI’s Silicon & Systems team. This person will become the Directly-Responsible Individual responsible for bringing OpenAI’s custom silicon and associated systems into data center environments, ensuring successful deployment, bring-up, validation, operational readiness, and ongoing reliability at scale. This role sits at the intersection of silicon, systems, infrastructure, data center operations, and software. You will lead a team focused on taking new hardware platforms from lab validation into production data center deployment. You will be responsible for building the operational processes, technical workflows, tooling, and cross-functional alignment required to deploy and operate custom AI hardware reliably in OpenAI’s supercomputing infrastructure. The ideal candidate is both a strong leader and a deeply technical operator. You should be comfortable staying close to the technical details of hardware bring-up, fleet deployment, debugging, system validation, data center integration, and production operations. This role requires strong execution, excellent cross-functional judgment, and the ability to drive clarity in ambiguous, fast-moving environments. In this role, you will: Lead a team responsible for deployment and operations of OpenAI’s custom silicon and systems in data center environments Own the path from hardware bring-up and validation through production deployment, operati

AWSRestAIGo
R
📍 New York, California, United States· Full-time
✓ High-confidence listing

$100K – $150K/yr

Quick readStrong listing-quality and freshness signals

Revivn is a profitable and rapidly growing company that helps enterprises manage their technology through our end of life software platform. We take electronic recycling one step further by repurposing hardware that still has remaining life and providing it to people who lack dedicated computer access and make it more affordable for people who may not be able to purchase new technology. Working with companies like Instacart, Lyft, Qualtrics, X, Gensler, and Spotify, we are changing the way companies view used technology with a new model that focuses on repurposing instead of recycling. We're looking for an Assistant Director of Operations to help scale the systems, teams, and processes behind our IT asset disposition (ITAD) business. This role is ideal for a hands-on operator who loves building teams, improving processes, and solving problems on the warehouse floor—not just from behind a desk. You'll work closely with the Sr. Director of ITAD Operations to drive execution across our facilities, raise the bar for operational excellence, and help prepare the business for its next stage of growth. This position is based in Fresno, CA or New York, NY and requires up to 50% travel to Revivn facilities and partner sites. What you’ll do Team Leadership & Development Lead, coach, and develop a high-performing operations organization that delivers consistent execution every day. Hire, onboard, and grow frontline leaders and team members as Revivn continues to scale. Set clear expectations, provide frequent coaching and feedback, and create opportunities for people to grow. Foster a culture of accountability, safety, ownership, and continuous improvement. Operations Management Own day-to-day execution across receiving, processing, data destruction, testing, and outbound logistics. Partner with the Sr Director of ITAD Operations to set throughput, quality, and SLA targets — and hold the team accountable for hitting them. Manage staffing plans, schedules, and cap

AIGoRustExcel
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -88%

About the Team OpenAI’s acquisition of io marks our entry into consumer hardware and our ambition to define the next human–computer interface. Success in hardware requires strong financial stewardship across the full product cost stack—from early design and sourcing decisions through manufacturing, logistics, inventory, returns, and warranty. Hardware Finance works across Product, Supply Chain, Operations, Accounting, Systems/Data, and Finance to connect business decisions to product cost, inventory, cash, COGS, and margin. About the Role We are seeking a Hardware Finance Manager to own an assigned area of hardware COGS and inventory end to end. The initial assignment will depend on business priorities and the successful candidate’s expertise. It may include BOM and product cost, manufacturing variance analysis, inventory planning, logistics, returns and warranty, customer support, or another connected set of hardware-finance responsibilities. This is an individual-contributor role with broad scope. Prior hardware experience and deep, hands-on expertise in at least two relevant domains are required. The person will be expected to operate independently, build reusable processes and analytical workflows, and remain accountable for the analysis, judgment, and recommendations. In this role, you will: Own an assigned area of hardware COGS and inventory end to end. Own forecasting, close, and business variance analysis for the assigned scope. Provide hardware leadership with clear variance explanations, trend analysis, and forward-looking signals that connect business and supplier decisions to inventory, cash, COGS, and margin. Partner with business teams and Finance Platforms to establish the financial data, systems, and dashboards needed to support analysis. Ensure data integrity and governance through clear definitions, ownership, validation checks, controls, and review processes. Improve forecasting, reporting, systems, and finance processes so they remain reliable an

AWSRestAIGo
🔔

Get new hardware operations engineer jobs in United States by email

Daily job updates · Unsubscribe anytime