ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a Delivery Director, Capacity programs for our on-premises data center builds and neo cloud (GPU cloud) delivery programs. This is a high-visibility, execution-critical role sitting at the intersection of infrastructure engineering, capacity planning, vendor/partner management, and customer delivery. You will own the end-to-end delivery lifecycle for large-scale compute infrastructure — from initial site/capacity commitments through power, networking, and hardware bring-up, to production-ready GPU/compute capacity landing in the hands of internal teams or customers. You'll be the person who turns ambitious infrastructure roadmaps into predictable, on-time, delivery. RESPONSIBILITIES Own delivery of on-prem infrastructure builds — colocation expansions, power/cooling readiness, rack-and-stack, network fabric bring-up, and hardware acceptance testing — coordinating across colo providers and partners, network engineering, hardware ops, and vendor teams. Drive neo cloud delivery programs — manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps, SLAs, and go-live readiness. Build and maintain master delivery schedules across concurrent, multi-site, multi-vendor programs, integrating power/shell timelines, hardware lead times, logistics, and software/platform readiness into a single critical path.
Salary not disclosed
Check market pay for comparable Rack Scale Software Architecture Director roles before applying.
Role overview
Job description
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.
NVIDIA has a rapidly expanding ecosystem of data center platform & node designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each bringing together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We're searching for a highly technical, motivated manager to lead & manage the team responsible for rack-scale system software architecture. From firmware, kernel drivers, operating systems, networking, fabrics and associated user mode drivers + manageability software. You will work with component leads internally and engage with industry leading hyperscalar / cloud service providers on taking these products to market.
What you’ll be doing:
Drive the software end-to-end architecture for NVIDIA's rack-scale products
Maintain deep understanding of the product portfolio and roadmap; translate forward-looking plans into clear, formal software requirements that anchor execution across the organization.
Ensure high quality & reliable software; serving as a trusted architectural partner to teams requiring guidance or oversight.
Work directly with major customers to understand their requirements and work to align their roadmap with NVIDIA’s roadmap.
Using strong communication skills, present the team vision to senior NVIDIA and external leaders.
Provide technical leadership and career mentorship to the team
Make key technical decisions even when faced with ambiguity
What we need to see:
BS or MS degree in Computer Engineering, Computer Science, or related degree or equivalent experience.
15+ overall years of experience in the area of System architecture and design with 8+ yrs of proven experience in management
Deep experience in designing architecture for scalable and performant server systems, particularly at the SW/HW interface.
Proven leadership skills and strong ownership on past projects involving a large scale sophisticated code base
Previous experience working with complex system software for accelerators such as GPUs, DPUs, or FPGAs
Possess strong managerial, problem solving and critical thinking skills.
Comfortable operating in highly matrixed organizations while holding a leadership position
Known for your strong interactive, verbal and written communications skills
Ways to stand out from the crowd:
Knowledge of large-scale cloud and cluster level deployment and management systems. Experience with designing robust, resilient and performant scale-up fabrics
Demonstrated track record of leading data center products across the entire lifecycle, spanning inception, pre-silicon development, post-silicon bring-up, manufacturing, and deployment.
Strong understanding of networking technology & protocols (e.g. Ethernet, Infiniband). Familiarity with CXL, UCIE and other C2C technology architectures. Knowledge in storage and networking technologies.
We are widely considered to be one of the technology world’s most desirable employers, and as a result have some of the most forward-thinking and hardworking people in the world working for us. So if you're clever, creative, and driven, we'd love to have you join the team.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 320,000 USD - 488,750 USD.You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.What they are looking for
Skills & requirements
Qualification
Provide technical leadership and career mentorship to the team Make key technical decisions even when faced with ambiguity What we need to see: BS or MS degree in Computer Engineering, Computer Science, or related degree or equivalent experience; 15+ overall years of experience in the area of System architecture and design with 8+ yrs of proven experience in management Deep experience in designing architecture for scalable and performant server systems, particularly at the SW/HW interface
Hiring company
Nvidia
Explore this employer's active roles, salary signals and company profile on Jobiba.
Keep exploring
Similar active roles
Fresh roles matched to this title and market.
$193.4K – $276.3K/yr
From $341.5K/yr
$230K – $250K/yr
🔔 Get job alerts
New Director, Rack Scale Software Architecture jobs in Ca, Santa Clara, United States, straight to your inbox.
No spam · Unsubscribe anytime