Jobiba hiring network

Hardware Operations Engineer Jobs

1,283 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current hardware operations engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

DU
DoorDash USA
📍 San Francisco• Full-time• From $1.1M/yr
21 days ago

About the Role As the Fleet Coordinator, you will be the single source of truth for the fleet: tracking what configuration each aircraft is flying and where it is, with no aircraft cleared to fly outside its approved configuration.You'll oversee all aircraft configuration that supports engineering and Part 135 operations, while keeping aircraft audit-ready. You will play a key role in supporting safe, compliant, and reliable operations by dispatching aircraft, tracking configuration and compliance records, and acting as the connective tissue between engineering, operations, maintenance control, and the FAA. You're excited about this opportunity because... Fleet Configuration & Testing: Maintain and verify compliance across the Part 107 test and Part 135 Commercial aircraft inventory: registration, airworthiness status, maintenance logs, and battery/component life-cycling." Maintain accurate software, hardware, and firmware configuration records for every aircraft, including designated testing candidates, so results are always traceable to a known config. Track aircraft location across the fleet ensuring one hundred percent accountability. Initiate and track parts requests tied to configuration changes, testing, and scheduled maintenance. Operations & Compliance: Maintain OpSpecs and the general maintenance manual, ensuring maintenance procedures and standards align with Part 135 requirements. Maintain the maintenance, training, and flight-log records the FAA expects an air carrier to produce on demand. Act as point of contact during FAA audits, representing the accuracy of aircraft configuration records and documentation against the maintenance program's compliance requirements. Fleet Dispatch & Availability: Decide which aircraft flies which mission (dispatch-style), verifying maintenance status before an aircraft is released. Track each aircraft's real-time location and availability, including movements between bases, hangars, and maintenance fac

awsgitrest
View job →
SA
Scale AI
📍 San Francisco• Full-time• From $134.4K/yr
22 days ago

At Scale, we believe that the next frontier of artificial intelligence is embodied. The Physical AI team is focused on building general AI that can reason and act in the physical world. By leveraging Scale’s massive, industry-leading data infrastructure, we are partnering with frontier labs to build Foundation Models for Physical AI that will redefine the future of automation. To support our rapid hardware-software iteration cycles and ensure a world-class R&D environment, we are looking for a Safety Coordinator / Lab Lead to anchor our physical testing operations. Role Overview As the Safety Coordinator / Lab Lead , you will play a mission-critical role in scaling our physical testing infrastructure safely and efficiently. This is a high-impact position where your highest-priority responsibility will be owning the end-to-end execution of safety audits and incident documentation . Operating at the intersection of cutting-edge AI foundation models and complex robotics hardware, you will ensure our researchers, engineers, and autonomous systems interact in a secure, compliant, and highly organized environment. Core Responsibilities Priority Focus: Safety Audits & Incident Documentation Rigorous Safety Audits: Design, schedule, and execute routine safety audits across all physical testing environments, robot cells, and hardware workspaces to ensure continuous compliance with internal benchmarks and industrial safety standards. Incident & Near-Miss Documentation: Own the end-to-end incident management pipeline. Act as the primary point of contact for documenting, archiving, and analyzing any lab incidents, mechanical anomalies, or near-misses. Root-Cause Analysis (RCA): Lead structured post-incident investigations to identify systematic risks, authoring comprehensive RCA reports and implementing Corrective and Preventive Actions (CAPA). Data-Driven Risk Mitigation: Treat safety data as a core operational asset—tracking safety metrics and audit trends to proa

awsrestai
View job →

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. This is a fully on-site role with mandatory relocation. Competitive relocation package provided. THE ROLE As a Technical Program Manager in the Core Platform Business Unit (CPBU), you will define and lead the program execution strategy for Everpure’s™ flagship FlashArray and FlashBlade hardware and software initiatives. You will align multi-stream, highly complex engineering programs with overarching business objectives, bridging hardware, software, and cloud-native developments for AWS and Azure. Serving as a strategic catalyst across global site leads, executive stakeholders, and infrastructure teams, you will tackle intangible architectural dependencies and drive high-impact initiatives to market on schedule without downtime. If you thrive on navigating multi-system complexity and driving organizational clarity, this role offers an exceptional platform for leadership. WHAT YOU'LL DO Drive Strategic Program Execution: Architect and manage multi-stream engineering roadmaps for FlashArray, FlashBlade, and hybrid cloud releases, ensuring on-time delivery across the full software development lifecycle. Lead Cross-Organizational Alignment: Direct end-to-end communication, status transparency, and dependency tracking across engineering, product management, infrastructure, and executive leadership teams. Optimize Operations & Capital Expenditure: Manage CapEx allocation, budget tracking, and resource planning for eng

awsazurerest
View job →
T
22 days ago

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is building the next generation of AI and RISC-V compute. As a Sr. Staff NPI Global Supply Planner, you will own end-to-end demand, materials, and production planning across our AI hardware product lines, including Blackhole, Galaxy, and future platforms. You will translate engineering release schedules, BOM structures, and production forecasts into executable supply and capacity plans with our contract manufacturers, while keeping planning data accurate and reliable across ERP and PLM systems. This role is hybrid, based out of Toronto, ON. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A strategic, detail-oriented supply chain planner who can turn engineering and program ambiguity into clear, executable plans. Experienced in NPI planning within semiconductor, AI hardware, or high-tech manufacturing environments. Comfortable working across BOM structures, engineering change processes, ERP and PLM systems, and system-of-record data. A clear, collaborative communicator who can align engineering, operations, planning teams, and contract manufacturers. What We Need Lead end-to-end materials and production planning across NPI r

awsaisem
View job →

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Notion is seeking a Workplace Technology Operation contractor for our Tokyo office. As part of our growing Workplace Technology Operations team, you'll be a resourceful, security-focused, and people-centered professional who will help us continue scaling and establish best-in-class support for our end users. The ideal candidate is expected to provide exceptional technical support, contribute to automation opportunities, and make a meaningful impact on our day-to-day workplace operations. This is a 12–18 month fixed-term role, with longer-term opportunities possible depending on business needs and performance. In-person collaboration is essential to Notion's culture. We require all team members to work from our offices on Mondays, Tuesdays, and Thursdays, our designated Anchor Days. Certain teams or positions may require additional in-office workdays. What You'll Achieve: End User & AV Support: Provide reliable onsite and remote technical support for Tokyo/APAC employees across hardware, software, connectivity, SaaS tools, conference room AV, and internal events, with clear communication and practical judgment. Emp

gitrestai
View job →

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Notion is seeking a Workplace Technology Operation contractor for our Dublin office. As part of our growing Workplace Technology Operations team, you'll be a resourceful, security-focused, and people-centered professional who will help us continue scaling and establish best-in-class support for our end users. The ideal candidate is expected to provide exceptional technical support, contribute to automation opportunities, and make a meaningful impact on our day-to-day workplace operations. This is a 12–18 month fixed-term role, with longer-term opportunities possible depending on business needs and performance. In-person collaboration is essential to Notion's culture. We require all team members to work from our offices on Mondays, Tuesdays, and Thursdays, our designated Anchor Days. Certain teams or positions may require additional in-office workdays. What You'll Achieve: End User & AV Support: Provide reliable onsite and remote technical support for Dublin/EMEA employees across hardware, software, connectivity, SaaS tools, conference room AV, and internal events, with clear communication and practical judgment. E

gitrestai
View job →
G
22 days ago

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to the Quality leadership within Manufacturing Operations, the Senior Reliability Scientist is responsible for leading reliability activities across complex, high-performance systems. Working closely with established reliability experts and cross-functional teams, this role uses experimental data and advanced modelling to inform design decisions, validate product reliability and optimise serviceability strategies, including spares provisioning. The Team The Quality team within Manufacturing Operations is responsible for ensuring product robustness, reliability and lifecycle performance across Graphcore’s hardware portfolio. The team includes experienced reliability specialists and works closely with technology research, chip, board, system design, platform and operations teams to translate reliability insights into actionable improvements across the product lifecycle. Responsibilities and Duties: · Define and refine reliability requirements across silicon, board and system levels, working in partnership with research and design teams · Apply ad

aigoexcel
View job →
O
OpenAI
📍 San Francisco• Full-time• $226K – $285K/yr
1mo ago

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As a Supply Chain Program Manager, you will own material readiness and supply chain execution for critical hardware programs spanning custom silicon, systems, memory, storage, networking, and rack infrastructure. You will work cross-functionally with Engineering, Strategic Sourcing, Manufacturing Operations, Finance, Planning, Quality, and external suppliers to develop and execute scalable supply strategies that support aggressive product development and deployment timelines. This role requires deep understanding of hardware supply chains, material planning, NPI execution, supplier management, and operational scaling in constrained and rapidly evolving environments. In this role you will: Material Readiness & Supply Planning - Own end-to-end material readiness across NPI and production phases, including building the necessary framework and processes for enablement. Drive supply planning and execution for long lead-time and constrained commodities including ASICs, HBM, DDR, SSDs, networking, optics, power, thermal, and mechanicals. Build and manage material readiness plans aligned to proto/pre-EVT, EVT, DVT, PVT, and mass production schedules. Monitor supply health, lead times, inventory positions, allocation risk, and capacity constraints. Drive shortage management, allocation mitigation, and recovery planning. Coordinate supply commits, forecast alignment, and supply continuity planning with suppliers and manufacturing partners. Cross-Functional Program Ma

awsrestai
View job →

About Graphcore Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors and data center hardware that deliver the specialized processing power needed to advance AI while improving the efficiency required for broad adoption. As part of SoftBank Group, Graphcore belongs to a family of companies developing some of the world's most transformative technologies. Our AI Engineering Campus in Austin plays an important role in building the future of AI computing. The Opportunity As Technical Services Director, you will lead the teams that operate and evolve Graphcore's engineering labs, high-performance computing (HPC) platforms, and data center environments globally. You will be accountable for reliable, secure, cost-effective infrastructure that supports demanding engineering, AI, silicon-development, and validation workloads. This role combines people leadership, infrastructure strategy, operational excellence, capacity and financial planning, procurement, and program delivery. You will partner with Engineering, Information Technology, Security, Finance, Facilities, Supply Chain, customers, and external suppliers. The position is based onsite in Austin and requires travel to company facilities, data centers, and supplier locations, including international travel. What You'll Do Lead, recruit, mentor, and develop the systems administration, lab operations, and technical services teams responsible for the facility supporting global Engineering and Research and Development. Own the reliability, efficiency, protection, safety, supportability, and continuous improvement of engineering labs, HPC systems, and infrastructure facilities. Establish service levels, operating standards, escalation paths, performance measures, monitoring, observability, automation, ticketing, and configuration-management practices. Translate engineering and customer requirements into infrastructure roadmaps, capacity p

linuxaigo
View job →
D
1mo ago

As a Network Engineer II on the Office Technology team, you will implement and support office network solutions across Datadog’s global offices. You will own well-scoped deployment and upgrade work, troubleshoot connectivity issues across LAN, WAN, and Wi-Fi, and partner with senior engineers on larger designs and cross-office initiatives. You will collaborate closely with vendors, Facilities, AV, IT Support, Security, and Infrastructure teams to keep Datadog offices connected, reliable, and observable. This is an IC2 role with a mix of project work and day-to-day operations. You will take ownership of clearly defined areas of the office network stack, execute production changes with care, and grow your ability to design, automate, and improve office network services with support from senior engineers. You will also use AI-assisted tools responsibly to improve troubleshooting, documentation, automation, knowledge discovery, and operational follow-through. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Deploy and configure network solutions for new and existing offices, including switches, wireless access points, routers, and firewalls Own scoped portions of office rollouts and upgrade projects, from planning through execution and validation Partner with ISPs, cabling providers, hardware vendors, and internal teams to turn up new services and resolve issues Troubleshoot LAN, WAN, and Wi-Fi issues using structured debugging methods, packet capture, logs, and performance data Manage routing, VLAN, firewall, and wireless configurations following team standards and change-management processes Contribute to repeatable deployment patterns through templates, documentation, scripting, or infrastructure-as-code Monitor office netwo

pythonrestai
View job →
M
Mongodb
📍 Bengaluru• Full-time
1mo ago

MongoDB Technical Services Engineers use their exceptional problem solving and customer service skills, along with their deep technical experience, to advise customers and to solve their complex MongoDB problems. Technical Service Engineers are experts in the entire MongoDB ecosystem - database server, drivers, cloud, and infrastructure. This also includes services such as Atlas (database as a service), or Cloud Manager (which helps customers with automation, backup and monitoring of their MongoDB systems). Our engineers combine their MongoDB expertise with passion, initiative, teamwork, and a great sense of humor to help our customers be successful with MongoDB. We are looking to speak to candidates who are based in Bengaluru for our hybrid working model. Cool things you’ll do You'll be working alongside our largest customers, solving their complex challenges - resolving questions on architecture, performance, recovery, security, and everything in between. You'll be an expert resource on best practices in running MongoDB at scale, whatever that scale may be. You'll be an advocate for customers' needs - interfacing with our product management and development teams on their behalf. And you'll contribute to internal projects, including software development of support tools for performance, benchmarking, and diagnostics. What you need We consider all candidates with an eye for those who are self taught, insatiably curious, and multi-faceted. The ideal candidates should have strong technical experience in more than one of the following areas Systems administration Distributed systems Network administration Database architecture and administration Application architecture Experience with authentication systems such as OIDC/SAML, LDAP, Kerberos, AD etc. Experience with virtualization - docker, kubernetes If you have an operations background, we prefer experience administering large-scale production environments, including hardware, operating systems (e.g. Linux, Windows),

javascriptpythonjava
View job →
O
1mo ago

About the Team OpenAI's Industrial Compute organization is building and operating the infrastructure foundation for the next generation of AI. Infrastructure Operations works across facilities, hardware, network operations, incident management, data center engineering, delivery teams, and external partners to bring capacity online safely, understand its operational state, and improve it over time. As OpenAI's data center portfolio grows across first-party and partner-delivered capacity, the organization needs clear goals, trusted data, repeatable processes, and systems that make ownership, risk, readiness, and performance visible. This role will help build the operating mechanisms that allow Infrastructure Operations to scale with rigor. About the Role We are seeking a Technical Program Manager to own the systems, data, reporting, governance, and program-management backbone for Infrastructure Operations. Reporting to the Delivery & Operations Lead, you will translate strategy into executable goals and operating cadences, turn operational needs into software and data solutions, and create the mechanisms that keep a rapidly evolving organization aligned and accountable. This role will also own the current 1P+3P delivery-tracking layer within Operations: milestones, delivery timelines, quantity forecasts, risks, decisions, and executive reporting. You will partner closely with 1P Delivery Program Management, Compute TPMs, Data Center Engineering, construction, commissioning, and operations leaders to ensure that delivery information becomes complete, usable input for readiness, handover, and ongoing operations. You will own program health and the operating system around it: the goals, data definitions, workflows, reporting, decision paths, and follow-through that help functional DRIs execute. The ideal candidate is comfortable in ambiguity, technically fluent enough to implement real systems, and relentless about converting scattered information into durable mechan

REMOTEsqlawsrest
View job →
H
Hp
📍 Texas• $147.1K – $230.9K/yr
19 days ago

Solution Architect Program Manager Description - Job Summary The Solution Architect Program Manager is responsible for driving advanced technology solutions from early concept through architecture definition, proof of concept, productization planning, system integration, validation, and launch readiness. This role combines deep technical capability with disciplined program execution to ensure that innovative concepts are translated into clear scope, committed deliverables, product milestones, and launch timelines. This position requires strong solution architecture judgment, structured planning, and cross-functional leadership across engineering, product, hardware, firmware, software, validation, security, operations, and external partners. The role owns the bridge between concept feasibility and production execution by defining requirements, aligning stakeholders, managing schedules, resolving technical dependencies, tracking risks, and maintaining readiness documentation required to deliver solutions on product milestones and market launch commitments. Responsibilities Drive end-to-end solution programs from concept exploration, business and technical feasibility, architecture definition, POC execution, productization planning, system integration, validation, and launch readiness. Translate ambiguous innovation concepts into clear solution scope, use cases, requirements, architecture assumptions, success criteria, milestones, resource needs, and execution plans. Own integrated schedules that align engineering deliverables with product development checkpoints, POR decisions, validation gates, manufacturing readiness, and launch timeline commitments. Partner closely with product management, system architecture, hardware, firmware, software, security, validation, operations, and external suppliers to manage dependenci

pythonlinuxai
View job →

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Everpure is seeking an Operational Excellence Technical Program Manager to join the Global Information Security Office and help drive execution for a strategic security program. This role is designed for a strong technical program leader who can bring structure, accountability, and operational rigor to complex work that spans security engineering, security architecture, and cross-functional delivery partners. This person will serve as a key operational driver for ticketing, escalations, dependency management, and execution governance tied to a strategic Everpure program. You will partner closely with security engineering and architecture leaders to ensure work is clearly defined, tracked, escalated appropriately, and moved forward with urgency and discipline. This is a highly cross-functional role for someone who is comfortable operating in ambiguity, translating technical context into actionable program structure, and building repeatable operational mechanisms that improve visibility, decision-making, and outcomes. WHAT YOU'LL DO Own day-to-day program operations, including ticketing workflows, intake, triage, prioritization, issue resolution, escalation management, and execution tracking. Partner with Security Engineering and Security Architecture leaders to translate technical priorities into defined workstreams, milestones, dependencies, accountable owners, and decision paths. Build and run operating mechanism

awsrestai
View job →
O
1mo ago

About the Team OpenAI’s Compute organization turns ambitious AI research into real-world capability by delivering the compute infrastructure behind our most advanced models. The team works across software, hardware, facilities, operations, and engineering disciplines to make enormous amounts of compute available, reliable, and efficient. As the demand for frontier AI grows, so does the complexity of the systems required to support it. Scaling this infrastructure means solving problems that cut across distributed systems, ML infrastructure, GPU fleets, power, cooling, networking, manufacturing, supply chain, and data center delivery. Our work is focused on expanding the compute foundation that enables OpenAI to train more capable models, including systems like GPT-5.6, and make frontier AI available to more people, products, and workflows. We’re looking for exceptional people across many disciplines to help build the next generation of AI infrastructure at a scale few organizations have attempted. About the Role We are hiring across a broad range of roles to help design, build, scale, and operate OpenAI’s compute infrastructure. Depending on your background, you may work on large-scale distributed systems, ML infrastructure, hardware systems, manufacturing, supply chain, data center development, or the physical engineering systems required to bring massive compute capacity online. You’ll work with teams across research, engineering, hardware, operations, and infrastructure to solve high-impact problems at extraordinary scale. This may include improving system reliability, accelerating deployment timelines, increasing operational efficiency, designing new infrastructure, or helping bring new compute platforms and facilities from concept to production. This is an opportunity to work on one of the most important infrastructure challenges in AI: building the compute foundation required to train and serve increasingly capable frontier models. Key Responsibilities Help bui

awsrestai
View job →
🔔

Get new hardware operations engineer jobs by email

Daily job updates · Unsubscribe anytime