Jobiba hiring network

Data Center Ssd Performance Validation Engineer Jobs

8,341 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current data center ssd performance validation engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role OpenAI's Hardware organization builds supercompute platforms from silicon and boards to full rack-scale systems to power advanced AI workloads. This role owns end-to-end quality for high-speed interconnect hardware across the product lifecycle: early design influence, supplier/contract manufacturer readiness, qualification, ramp, and fleet quality in lab and data center environments. You will be the quality lead for advanced interconnect components and assemblies, including high-speed copper cables, cable cartridges, patch panels, backplane/cable-backplane solutions, high-speed connectors, and related electro-mechanical interfaces. You will partner closely with electrical, mechanical, SI/PI, systems, reliability, operations, and external vendors to prevent escapes and drive rapid, data-driven containment and corrective action. In this role you will: Own quality for advanced interconnect components and assemblies: high-speed connectors, high-speed copper cables, cable cartridges (e.g., cable cassette style assemblies), patch panels & optics, and backplane/cable-backplane interconnect solutions. Drive quality-by-design: participate in design reviews, DFM/DFx, tolerance stacks, material and plating selections, connector mating strategy, strain relief, and assembly methods to reduce variation and field failures. Define and track quality and reliability metrics (DPPM, yield, escapes, RMA/FRACAS trends, Cpk/Ppk where applicable) for interconnects across NPI and m

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI’s Industrial Compute team is responsible for building and scaling large-scale compute capacity across first-party data centers, strategic partners, and industrial infrastructure environments. We focus on converting power, land, hardware, and operational execution into reliable compute capacity that can support frontier AI training and inference workloads. This team operates at the intersection of infrastructure delivery, hardware systems, utilities, supply chain, and capacity strategy—ensuring OpenAI can scale compute faster than traditional models allow. About the Role We are seeking a Tokens-as-a-Service (TaaS) Lead to drive the end-to-end conversion of industrial-scale infrastructure investments into usable token capacity for OpenAI workloads. In this role, you will own execution across complex compute programs where raw infrastructure capacity must be transformed into operational GPU throughput. You will coordinate across data center delivery, power, networking, hardware deployment, workload enablement, finance, and external partners to ensure capacity becomes productive tokens as quickly and efficiently as possible. This role is ideal for someone who can bridge physical infrastructure delivery with compute utilization outcomes. Success requires strong systems thinking, elite program leadership, and the ability to drive accountability across internal teams and strategic partners. In this role, you will Lead Tokens-as-a-Service programs across industrial compute environments, including first-party and partner-owned capacity. Convert delivered power, space, and hardware capacity into production-ready token throughput. Build integrated execution plans spanning construction, power energization, rack deployment, networking, cluster readiness, and workload onboarding. Partner with infrastructure engineering, hardware, networking, finance, supply chain, and operations teams. Drive external providers, EPCs, OEMs, utilities, and strategic partners t

awsrestai
View job →

About the Team At Trendyol Tech, our mission is to create a positive impact in our ecosystem by enabling commerce through technology. We solve complex problems with data, creativity, and agility — always driven by real outcomes. With a culture built on learning, collaboration, and ownership, we grow together while building what’s next. About the Role As a leading and fast-growing e-commerce company, our physical infrastructure is the backbone of our massive logistics and digital ecosystem. We operate high-capacity, multi-site data center environments supporting our continuous digital operations. In addition, our physical footprint spans a highly distributed network of nationwide logistics centers, fulfillment hubs, and corporate offices. We are seeking a highly resilient, strategically minded, and hands-on Infrastructure & Field Operations Manager to lead our Field Operations Team. In this role, you will oversee all physical infrastructure deployments, structured cabling, rack installations, hardware maintenance, and field operations across all our data centers, supply chain networks, and offices.

O
1mo ago

About the Team The compute infrastructure team runs the GPU fleet and large-scale compute clusters that serve the models backing ChatGPT and the API, while also supporting training workloads for our next generation models. We operate a large, modern GPU fleet and provide a unified platform for other OpenAI teams to seamlessly run production Applied AI and Research training workloads. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role You’ll own the hands-on and automation work that brings WAN, fiber, carrier, and cloud-interconnect circuits into service. Partner with network engineers, fiber providers, cloud service providers, colocation teams, and data-center technicians to move each connection from ordered and patched to verified, stable, and ready for handoff. You’ll own Layer 1 troubleshooting and circuit bring-up while building workflows that translate reliable system or model output into precise, approved technician actions, capture field feedback, and drive each connection to a green-port handoff. The right person combines strong physical-networking judgment with practical automation skills: patch-panel and port mappings, optics and light levels, provider coordination, structured operational data, API or scripting workflows, and human-in-the-loop LLM tooling. Responsibilities Own Layer 1 activation and restoration for carrier circuits, dark fiber, wavelengths, Ethernet handoffs, and dedicated cloud interconnects across data centers and points of presence. Reconcile complete A-side/Z-side as-builts: circuit IDs, LOAs/CFAs, carrier demarcations, MMR/ODF/MDF and patch-panel positions, fiber pairs, cross-connects, optics, and device ports. Investigate no-light, low-light, wrong-port, link-flap, and error-rate issues across providers and CSPs; isolate continuity, dirty connectors, polarity, incorrect patching

awsazurerest
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Finance Platform & Technology team at OpenAI builds and scales the future-proof systems and data architecture that power our core financial operations. We enable business agility, compliance, and operational excellence across quote-to-cash, procure-to-pay, inventory, and asset management for both B2B and B2C. Our focus is on modernizing workflows through strategic integrations, scalable automation, and seamless data flows empowering smarter decisions, reliable reporting, and sustainable growth as OpenAI evolves About the Role We are looking for a Supply Chain Transformation Architect to redesign and modernize our end-to-end supply chain operations supporting robotics, consumer hardware, and data center infrastructure. This role sits at the intersection of process, systems, and data. Your primary focus will be transforming supply chain processes across planning, procurement, manufacturing, logistics, and fulfillment—then enabling those processes with the right systems architecture, data foundation, and AI-driven automation. You will help move the organization from manual, reactive operations to intelligent, data-driven supply chain execution. In this role you will: Lead End-to-End Supply Chain Transformation Evaluate and redesign core supply chain processes across demand planning, supply planning, procurement, manufacturing operations, logistics, and fulfillment. Identify operational bottlenecks, fragmented workflows, and manual processes that limit scalability. Build standardized process frameworks and operating models that support rapid scaling of hardware programs. Drive Operational Excellence Implement structured supply chain practices such as: S&OP / Integrated Business Planning Supply risk management Inventory optimization Supplier collaboration frameworks Logistics visibility and execution models Establish operational KPIs and governance to improve predictability, responsiveness, and resilience. Architect the Digital Supply Chain Tra

reactawsgit
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Stargate team is responsible for building the physical infrastructure that powers large-scale AI systems. We design and deliver next-generation data centers optimized for dense compute clusters, advanced networking, and rapidly evolving hardware platforms. This work sits at the intersection of hardware engineering, systems architecture, and infrastructure execution—translating cutting-edge compute roadmaps into scalable, production-ready environments. Our teams partner across silicon vendors, server and storage OEMs, networking teams, and data center engineering organizations to bring new capacity online quickly, reliably, and at global scale. About the Role We are seeking a CPU & Storage Technical Lead to define and drive the server compute and storage architecture strategy for Stargate infrastructure. In this role, you will own technical direction across CPU platforms, memory configurations, local and disaggregated storage systems, and their integration into large-scale AI clusters. You will evaluate vendor roadmaps, lead platform tradeoff decisions, and ensure compute and storage systems are optimized for training, inference, and supporting services. You will work cross-functionally with hardware engineering, performance modeling, networking, supply chain, and deployment teams, as well as external partners such as AMD, Intel, OEMs, ODMs, and storage vendors. This is a highly strategic role for someone who can operate deeply at the component level while also driving long-range infrastructure decisions. Key Responsibilities Own CPU and storage technical strategy for Stargate compute infrastructure across current and future generations. Evaluate CPU platforms across performance, efficiency, memory bandwidth, PCIe topology, cost, and roadmap alignment. Define storage architectures for AI environments, including boot media, local NVMe, shared storage, caching tiers, metadata services, and high-performance data pipelines. Drive server platform de

awsrestai
View job →
C
Cvshealth
📍 United States• Remote
10 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary We are seeking an accomplished Principal Cloud Storage Engineer to lead the design, engineering, and evolution of our private cloud storage platforms. This role will focus on large-scale storage architecture, data protection, cyber recovery, and resiliency technologies across complex enterprise environments. The ideal candidate will combine deep technical expertise in storage systems with strong leadership, architectural vision, and the ability to influence technical direction across the organization. Key Responsibilities Architect and engineer enterprise storage platforms that ensure data integrity, availability, security, and disaster recovery readiness Design and implement end-to-end storage solutions, including Software Defined Storage, SAN, NAS, and object storage across private cloud and data center environments Drive strategic technology decisions by evaluating emerging products, tools, and standards supporting storage, data protection, cloud, and compute platforms Lead infrastructure initiatives involving storage modernization, data protection, cyber recovery, data migration, and resilience engineering Develop and execute enterprise strategies for backup, recovery, cyber vaulting, and business continuity Create and maintain comprehensive documentation of storage architectures, configurations, policies, and operation

REMOTEkubernetesproject management
View job →
C
Cvshealth
📍 Woonsocket 1 Cvs Drive
10 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary CVS Health is seeking a highly skilled and visionary Sr. Manager to lead and manage our enterprise Cisco Application Centric Infrastructure (ACI) environment and the transition to Cisco Nexus Dashboard Fabric Controller (NDFC). This role sits within the Network Engineering team and is pivotal to driving the modernization and scalability of our next-generation data center fabric. This role will help set and drive the network technology strategy for ISTS, ensuring that ISTS delivers on our mission to transform technology and provide an agile, cost optimized and resilient set of network infrastructure services to meet the evolving needs of all the CVS Health lines of business. The ideal candidate will have experience supporting complex enterprise network environments and demonstrate deep expertise in Cisco ACI and NDFC technologies. As a strategic leader, you will be responsible for setting the direction, leading a team of engineers, and ensuring the operational integrity, performance, and evolution of our fabric-based data center architecture. This position will also mentor staff members in an effort to develop excellent enterprise networking skills. Key Responsibilities Lead the deployment, administration, and lifecycle management of the Cisco ACI environment Oversee the strategic transition to and ongoing management of the Cisco

awsazuregcp
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. At Micron, we are transforming how the world uses information to enrich life. The High Bandwidth Memory (HBM) Design team develops industry-leading memory solutions that enable advances in Artificial Intelligence, high-performance computing, graphics, and data-center applications. Our engineers collaborate across global teams to deliver innovative memory architectures and semiconductor technologies that power next-generation computing systems. As an HBM Design Engineer Intern, you will work alongside experienced memory engineers and gain hands-on experience in semiconductor design, simulation, verification, and analysis. This internship provides exposure to industry-standard design methodologies, EDA tools, and cross-functional collaboration throughout the product development lifecycle. You will also have opportunities to apply AI-Assisted and AI-Enabled workflows to improve engineering productivity, debug efficiency, and design quality. Responsibilities Assist with the design, simulation, analysis, and verification of HBM memory and logic circuits using industry-standard semiconductor design tools. Support RTL development, circuit implementation, timing analysis, power analysis, and functional validation activities. Develop scripts, automation solutions, and AI-Assisted workflows to improve design productivity, debug efficiency, and design-flow quality. Collaborate with Design, Verification, Physical Design, CAD, and Product Engineering teams on technical reviews, debug activities, and project deliv

pythonlinuxartificial intelligence
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. In this role, your primary focus will be on Microsoft Windows technologies as well as VMware/Nutanix/Dell/Lenovo/Splunk/Cisco server Platforms. This role helps define and implement Windows Server technologies and methodologies, which will have a heavy emphasis on automation and a hybrid cloud environment, while maintaining operational excellence in multiple world class Data Center environments. Responsibilities and Tasks: You will work regularly with Micron’s Business Units to ensure solutions meet or exceed business requirements. You will be expected to suggest, promote, and leverage published standards to minimize environment complexity and ensure regulatory as well as license compliance. You will have the opportunity to travel to other Micron sites as well as technology conferences and training. As part of your responsibilities you will be required to: Provide day to day technology directions and operations within Infrastructure of VMware/Nutanix HCI platform and Dell/Lenovo/Splunk/Cisco hardware. IT Incident and Request handling and management within service level agreement Maintain IT operations by monitoring of system performance metrics and proactive actions Site expansion and global project execution Propose system or architectural design to improve performance, high availability, scalability and disaster recovery by innovative technologies and roadmap managem

dockerkuberneteslinux
View job →
MT
10 days ago

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Responsibilities and Tasks You will work regularly with Micron’s Business Units to ensure solutions meet or exceed business requirements. You will be expected to suggest, promote, and leverage published standards to minimize environment complexity and ensure regulatory as well as license compliance. You will have the opportunity to travel to other Micron sites as well as technology conferences and trainings. As part of your responsibilities you will be required to • Storage platform management - NetApp filer, EMC VMAX, Isilon. • Backup platform management - IBM TSM, Cohesity. • SAN Switch management - Cisco MDS. • Data Center management. • Routine jobs, ticketing and troubleshooting. • Installation, maintenance and improvement for storage devices. • Need to participate in on-call duty. Specific responsibilities include (but are not limited to) Continually improves operations through routine monitoring of system performance metrics and proactive actions. Reviewing of system log files, trend analysis, configuration, incident and problem management processes. Recommends and works with the Solution and Enterprise Architect on system or architectural changes when needed to improve performance, high availability, scalability, or propose innovative technologies. Ensures architectural consistency and

artificial intelligenceairecruitment
View job →
MT
10 days ago

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Responsibilities and Tasks You will work regularly with Micron’s Business Units to ensure solutions meet or exceed business requirements. You will be expected to suggest, promote, and leverage published standards to minimize environment complexity and ensure regulatory as well as license compliance. You will have the opportunity to travel to other Micron sites as well as technology conferences and trainings. As part of your responsibilities you will be required to • Storage platform management - NetApp filer, EMC VMAX, Isilon. • Backup platform management - IBM TSM, Cohesity. • SAN Switch management - Cisco MDS. • Data Center management. • Routine jobs, ticketing and troubleshooting. • Installation, maintenance and improvement for storage devices. • Need to participate in on-call duty. Specific responsibilities include (but are not limited to) Continually improves operations through routine monitoring of system performance metrics and proactive actions. Reviewing of system log files, trend analysis, configuration, incident and problem management processes. Recommends and works with the Solution and Enterprise Architect on system or architectural changes when needed to improve performance, high availability, scalability, or propose innovative technologies. Ensures architectural consistency and

artificial intelligenceairecruitment
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. In this role, your primary focus will be on Microsoft Windows technologies as well as VMware/Nutanix/Dell/Lenovo/Splunk/Cisco server Platforms. This role helps define and implement Windows Server technologies and methodologies, which will have a heavy emphasis on automation and a hybrid cloud environment, while maintaining operational excellence in multiple world class Data Center environments. Responsibilities and Tasks: You will work regularly with Micron’s Business Units to ensure solutions meet or exceed business requirements. You will be expected to suggest, promote, and leverage published standards to minimize environment complexity and ensure regulatory as well as license compliance. You will have the opportunity to travel to other Micron sites as well as technology conferences and training. As part of your responsibilities you will be required to: Provide day to day technology directions and operations within Infrastructure of VMware/Nutanix HCI platform and Dell/Lenovo/Splunk/Cisco hardware. IT Incident and Request handling and management within service level agreement Maintain IT operations by monitoring of system performance metrics and proactive actions Site expansion and global project execution Propose system or architectural design to improve performance, high availability, scalability and disaster recovery by innovative technologies and roadmap managem

dockerkuberneteslinux
View job →
M
11 days ago

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. At Micron, we are transforming how the world uses information to enrich life. The High Bandwidth Memory (HBM) Design team develops industry-leading memory solutions that enable advances in Artificial Intelligence, high-performance computing, graphics, and data-center applications. Our engineers collaborate across global teams to deliver innovative memory architectures and semiconductor technologies that power next-generation computing systems. As an HBM Design Engineer Intern, you will work alongside experienced memory engineers and gain hands-on experience in semiconductor design, simulation, verification, and analysis. This internship provides exposure to industry-standard design methodologies, EDA tools, and cross-functional collaboration throughout the product development lifecycle. You will also have opportunities to apply AI-Assisted and AI-Enabled workflows to improve engineering productivity, debug efficiency, and design quality. Responsibilities Assist with the design, simulation, analysis, and verification of HBM memory and logic circuits using industry-standard semiconductor design tools. Support RTL development, circuit implementation, timing analysis, power analysis, and functional validation activities. Develop scripts, automation solutions, and AI-Assisted workflows to improve design productivity, debug efficiency, and design-flow quality. Collaborate with Design, Verification, Physical Design, CAD, and Product Engineering teams on technical reviews, debug activities, and project deliv

pythonlinuxartificial intelligence
View job →
N
11 days ago

We are seeking a highly skilled and hard-working Senior Test Developer / test engineer to join our multifaceted Enterprise Software QA team. This role offers an outstanding opportunity to leave your mark on the design, construction, optimization and testing of large-scale infrastructure for various foundational NVIDIA unified cloud services and data center offerings. If you are a dedicated engineer with strong expertise in cloud infrastructure and distributed systems and want to apply your skills with AI tools, this role could fit you perfectly. You will thrive in an exciting, innovative environment. What you'll be doing: Work with development teams on test plans for all layers of SW stack for cloud infrastructure, execution, reviews, failure analysis and assessing overall quality and risk. Work with customer PMs on software issues including technical feedback from OEMs and CSPs. Develop key benchmarks to track execution and deploy process improvements to improve efficiency Leverage AI skills to expedite the test scope, test plan, execution and automation workflows. Lead NVIDIA Cloud and Data Center bring up activities which will involve validation, reporting, working with engineering to debug issues, providing design input at times, adding coverage in different areas. Design, develop and maintain CI/CD pipelines for continuous testing in cloud environments when needed. Perform performance, scalability, and reliability testing of cloud services. Implement and maintain test environments in cloud platforms such as AWS, Azure, or Google Cloud. Supervise the infrastructure to alert on significant events, ensuring the highest level of system performance and reliability. Work with various different partner teams to ensure availability of clusters to test on and take the lead in resolve all issues. Working with tea

awsazuredocker
View job →
🔔

Get new data center ssd performance validation engineer jobs by email

Daily job updates · Unsubscribe anytime