←Jobiba.Me
E
Active1w ago

Hardware Lead Engineer, Hyperscale Line of Business

Everpure·📍 Santa Clara, California

Employment

Full-Time

Work mode

On-site

Experience

SENIOR LEVEL

Salary

Not disclosed

Salary not disclosed

Check market pay for comparable Hardware Lead Engineer roles before applying.

Salary →

Role overview

Job description

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Drive the hardware engineering lifecycle for Everpure’s Hyperscale Line of Business, shaping high-performance, energy-efficient storage architectures tailored directly to hyperscale partners . You will serve as the technical lead bridging customer requirements and internal engineering execution, ensuring custom hardware solutions mee

…

What they are looking for

Skills & requirements

Qualification

Working closely with cross-functional teams in validation, diagnostics, manufacturing test, and operations, you will lead hardware qualification projects, solve critical escalations, and advance system robustness; WHAT YOU'LL DO Lead Hardware Qualification & System Hardening: Plan, execute, and automate comprehensive x86 hardware validation cycles—including electrical, signal integrity, and protocol testing—to guarantee system reliability for enterprise hyperscale deployments; Mentor and Elevate Engineering Standards: Provide technical guidance and mentorship to engineering peers, refining qualification processes and stress-testing methodologies (using tools like PTU, Mprime, and Iometer) to continuously improve team output; Validation & Automation Capability: Proven ability to design, automate, and execute hardware qualification frameworks using scripting languages (such as Python or BASH) alongside industry-standard stress and diagnostic tools

Department · Engineering

E

Hiring company

Everpure

Technology

We’re in an unbelievably exciting area of tech and are fundamentally reshaping the data storage industry. Here, you lead with innovative thinking, grow along with us, and join the smartest team in the industry. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us.

Keep exploring

Similar active roles

Fresh roles matched to this title and market.

View all →
G
📍 Santa Clara, California, United States· Full-time

From $79.5K/yr

Location Details: Santa Clara, CA At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is an in-office position and you’ll be expected to work full-time in an office location, and therefore must live within commuting distance from your assigned office. You will work from this office beginning on your first day. Join our team Join GoDaddy's Global IT Support team as a Desktop Support Technician in our Santa Clara office, where you'll be the face of IT for employees who rely on technology to do their best work every day. In this hands-on, 100% on-site role, you'll provide walk-up and escalated technical support, troubleshooting hardware, software, workplace technology, and connectivity issues while delivering an exceptional customer experience.You'll collaborate closely with a globally distributed team of support professionals across North America, EMEA, and APAC, helping resolve both everyday technical requests and high-priority incidents in a fast-paced environment. Success in this role is driven as much by your communication, empathy, and problem-solving skills as your technical knowledge, making it an excellent opportunity for someone who enjoys helping people and learning new technologies.If you're passionate about customer service, curious about how technology works, and looking to grow your career in enterprise IT, we'd love to hear from you. What you'll get to do... Provide front-line technical support to employees by diagnosing and resolving hardware, software, operating system, peripheral, and account-related issues, ensuring minimal disruption to productivity. Manage and prioritize incidents and service requests through the IT ticketing system, taking ownership from initial intake through troubleshooting, resolution, and follow-up communication. Configure, deploy, maintain,

E
📍 Bengaluru, India· Full-time

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As an Escalation Engineer on our FlashBlade team, you will serve as the premier technical authority driving customer trust and operational stability across complex, enterprise-scale storage environments . You will collaborate closely with front-line Support, Engineering, and product leaders to rapidly resolve high-impact technical challenges and transform complex system failures into long-term product reliability. By bridging real-world customer insights with engineering solutions, you will elevate team performance and ensure our enterprise customers achieve flawless platform availability. WHAT YOU'LL DO Drive High-Stakes Escalation Resolution: Take end-to-end ownership of critical, multi-platform system issues—evaluating hardware, software, networking, and environmental factors—to rapidly restore service, perform root-cause analysis, and protect customer business continuity. Elevate Engineering Talent & Knowledge: Mentor and coach support team members through joint case triage, structured technical trainings, and internal documentation, accelerating technical capabilities and resolution velocity across the organization. Bridge Product Engineering & Customer Insights: Partner directly with Product Engineering to relay real-world system behavior, ensuring critical customer feedback, feature enhancements, and bug fixes trickle back into core product design. Lead Strategic Customer Communications: Facilitate

AWSLinuxRestAI
E
📍 Bengaluru, India· Full-time

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As an Escalation Engineer on our FlashBlade team, you will serve as the premier technical authority driving customer trust and operational stability across complex, enterprise-scale storage environments . You will collaborate closely with front-line Support, Engineering, and product leaders to rapidly resolve high-impact technical challenges and transform complex system failures into long-term product reliability. By bridging real-world customer insights with engineering solutions, you will elevate team performance and ensure our enterprise customers achieve flawless platform availability. WHAT YOU'LL DO Drive High-Stakes Escalation Resolution: Take end-to-end ownership of critical, multi-platform system issues—evaluating hardware, software, networking, and environmental factors—to rapidly restore service, perform root-cause analysis, and protect customer business continuity. Elevate Engineering Talent & Knowledge: Mentor and coach support team members through joint case triage, structured technical trainings, and internal documentation, accelerating technical capabilities and resolution velocity across the organization. Bridge Product Engineering & Customer Insights: Partner directly with Product Engineering to relay real-world system behavior, ensuring critical customer feedback, feature enhancements, and bug fixes trickle back into core product design. Lead Strategic Customer Communications: Facilitate

AWSLinuxRestAI
G
📍 Austin, Texas, United States· Full-time

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of a best-in-class family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from a diverse group of backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a senior validation lead engineer to lead at-scale rack validation efforts for next-generation AI hyperscale systems. This role focuses on post-silicon system validation across the full lifecycle, ensuring functional, electrical, and thermal performance meets product objectives. You will own end-to-end blade and rack validation including planning, development, execution, and debug while collaborating across firmware, systems, and hardware teams. The Team The Rack Validation team is responsible for ensuring system readiness and quality at scale. The team works cross-functionally with firmware, silicon, and system engineering teams to validate complex AI compute platforms. Responsibilities and Duties Lead post-silicon validation of AI compute blades and racks including test planning, development, and automation. Drive provisioning and integration of system components (SoC FW, BMC, RMC, OS) for rack-level readiness. Own execution against program achievements and report validation progress and risks. Triage test failures, collect debug data, and collaborate on root cause analysis. Track

PythonCI/CDLinuxAI
G
📍 Austin, Texas, United States· Full-time

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Senior Principal Network Engineer to help design, deploy, and optimize next‑generation AI data center networks. AI training and inference workloads require extremely high bandwidth, deterministic low latency, and zero‑packet‑loss networking environments. In this role, you will partner closely with the Network Architecture Lead to design and scale high‑performance computing (HPC) network fabrics supporting GPU clusters. You will work across hardware, networking, and AI application layers to ensure Graphcore’s large‑scale AI infrastructure operates at peak performance. The ideal candidate brings deep experience operating hyperscale or HPC data center networks and has expertise in high‑speed Ethernet fabrics, RDMA technologies, advanced automation, and telemetry systems. The Team The Data Center Network Engineering team designs and operates the high‑performance network fabrics that power Graphcore’s AI compute platforms. The team collaborates closely with hardware engineering, AI researchers, and infrastructure teams to build scalable networking environments optimized for distributed training and infe

PythonAIGoDevOps
O
📍 San Francisco, California, United States· Full-time

About the Team The compute infrastructure team runs the GPU fleet and large-scale compute clusters that serve the models backing ChatGPT and the API, while also supporting training workloads for our next generation models. We operate a large, modern GPU fleet and provide a unified platform for other OpenAI teams to seamlessly run production Applied AI and Research training workloads. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role You will be part of an engineer-first TPM team as a Technical Program Manager for Compute Infrastructure who owns the end-to-end delivery of large-scale GPU clusters, partnering with engineers to bring clusters online across external providers and partners. You’ll run a broad, parallel portfolio spanning hardware, networking, power, and cooling—driving execution, risk management, and crisp alignment from working teams through leadership to deliver production-ready capacity at scale. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead end-to-end delivery of both New Compute SKUs and large-scale GPU clusters across an external partner ecosystem while supporting capacity planning for training and inference. Ability to contextually drive multi-threaded bring-up programs spanning hardware, networking, power, and cooling—owning plans, dependencies, and critical paths. Interface with chip providers to derisk long-term onboarding to new hardware platforms by working across kernels, comms, hardware, and scheduling engineering teams. Build and operationalize program mechanisms (roadmaps, milestones, risk registers, runbooks) that make delivery predictable at massive scale. Partner with engineering to improve cluster turn-up reliability, repeatability, and automation

AWSRestAIGo

🔔 Get job alerts

New Hardware Lead Engineer, Hyperscale Line of Business jobs in Santa Clara, California, straight to your inbox.

No spam · Unsubscribe anytime