We are looking for a Senior Software Engineer to become part of our storage management plane team. The management plane is a web-based application crafted to provide our storage customers the capabilities to handle and supervise our distributed storage infrastructure. Our team is continually dedicated to acquiring and implementing ground breaking technologies to overcome obstacles and innovate solutions for improving our ability to handle large clusters of machines efficiently. What You Will Be Doing: Maintain and develop Kubernetes operators and our Container Storage Interface (CSI) plugin. Develop a web-based solution that manages, operates and monitors our distributed storage. Work closely with other teams to define and implement new APIs. What We Need to See: B.Sc., M.Sc. or Ph.D. in Computer Science, or related discipline, or equivalent experience. 8+ years of experience in web development ( both client and server ) Proven experience with Kubernetes (K8s), including developing or maintaining operators and/or CSI plugins. Experience scripting with Python, Bash or similar. Experience with nodejs is a must At least 5 years of experience working in a Linux OS environment Youβre smart and a quick learner You do what it takes to get the job done Passionate about coding and big challenges Ways to stand out from the crowd: NodeJS for the server side: dominant modules are async & express . Kafka, MongoDB, K8s JavaScript frameworks: React, jQuery, c3j
Jobs in United States
Software Tpm Director in United States
2,142 active opportunities Β· Updated October 2026
Showing
15 jobs
Explore current software tpm director jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload. What you will be doing: Build storage technologies, client libraries, and filesystem frameworks that help AI workloads access data across object stores, file systems, and hybrid cloud infrastructure. Develop high-performance storage paths for training and inference workflows, including data loading, checkpointing, caching, POSIX-style access, and object-store integration. Build observability systems that diagnose storage bottlenecks, attribute GPU idle time to I/O behavior, and expose actionable telemetry through production monitoring stacks. Improve performance, scalability, and reliability of storage systems serving massive datasets, deep directory trees, and high-concurrency AI workloads. Work closely with internal AI teams, platform teams, SRE, and operations to validate storage behavior against real workloads and production environments. Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, performance, and verification. What we need to see: BS in Computer Science, Information Sys
NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. We are the GPU Communications Libraries and Networking team at NVIDIA. We deliver libraries like NCCL, NVSHMEM, UCX for Deep Learning and HPC. We are looking for a motivated Performance engineer to influence the roadmap of our communication libraries. The DL and HPC applications of today have a huge compute demand and run on scales which go up to tens of thousands of GPUs. The GPUs are connected with high-speed interconnects (eg. NVLink, PCIe) within a node and with high-speed networking (eg. Infiniband, Ethernet) across the nodes. Communication performance between the GPUs has a direct impact on the end-to-end application performance; and the stakes are even higher at huge scales! This is an outstanding opportunity for someone with HPC and performance background to advance the state of the art in this space. Are you ready for to contribute to the development of innovative technologies and help realize NVIDIA's vision? What you will be doing: Conduct in-depth performance characterization and analysis on large multi-GPU and multi-node clusters. Study the interaction of our libraries with all HW (GPU, CPU, Networking) and SW components in the stack Evaluate proof-of-concepts, conduct trade-off analysis when multiple solutions are available Triage and root-cause performance issues reported by our customers Collect a lot of performance data; build tools and infrastructure to visualize and analyze the information <li
Mission Systems Software Developer (Experienced & Senior) β USA - Daytona Beach, FL. Apply via Workday.
Mission Systems (MS) Software Platform Mission Computing (PMPC) Manager β USA - Berkeley, Macau S.A.R.. Apply via Workday.
Enterprise Operations Software Internship β Texas, United States of America. Apply via Workday.
Lead DevSecOps Software Engineer β USA - El Segundo, CA. Apply via Workday.
Senior System Software Engineer - Halos Core and Robotics Platform β US, CA, Santa Clara. Apply via Workday.
Senior System Software Engineer - Halos Core and Robotics Platform β US, CA, Santa Clara. Apply via Workday.
Senior Systems Software Engineer - Advanced Technology Group β US, OR, Hillsboro. Apply via Workday.
Senior/Staff Software Engineer with C++ - Drivers, Diagnostic, & Embedded Software (San Diego, CA)
Philips IndiaSenior/Staff Software Engineer with C++ - Drivers, Diagnostic, & Embedded Software (San Diego, CA) β San Diego, California, United States. Apply via Workday.
Mission Systems (MS) Software Test and Infrastructure Manager β USA - Berkeley, MO. Apply via Workday.
Principal Automation Software Test Engineer: SDET β Lafayette, Colorado, United States of America. Apply via Workday.
Senior Tegra Software Engineer β US, CA, Santa Clara. Apply via Workday.
Senior System Software Test Engineer, Networking β US, CA, Santa Clara. Apply via Workday.
Other cities to consider
More places hiring for this role
Get new software tpm director jobs in United States by email
Daily job updates Β· Unsubscribe anytime