Jobiba hiring network

Nvidia Manager Jobs

561 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current nvidia manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

We’re looking for a technical architecture expert with hands-on development experience to define the architecture of the Retail and Consumer Packaged Goods Platform and collaborate with engineering to build the platform. This is a technical engineering role and requires familiarity with NVIDIA’s GPU acceleration libraries, experience in Retail, CPG, logistics, or marketplaces, and knowledge of enterprise AI architectures will make you an ideal candidate. Join us and take the lead in building the retail platform and help us develop the future of AI in retail! Your mission is to leverage NVIDIA’s existing retail and supply chain blueprints, along with proven workflows and best practices codified through customers and ISV engagements, to define the architecture and build the Retail & CPG platform. You will be responsible for architecting and collaborating with engineering teams across NVIDIA, to build the platform. Built on NVIDIA’s full stack—including Agentic AI, Physical AI, and accelerated computing—this platform will power our three strategic pillars: digital commerce, supply chain, and intelligent stores. What you'll be doing: You will work with Retail & CPG product management, business development, and developer relations teams to harness their deep expertise in supply chain and retail to architect and build a platform that makes it easier for our ecosystem and customers to build AI solutions at scale. The platform will use our NVIDIA Nemo microservices and Nemotron open-source models, with initial priority on Agentic AI for supply chain, digital commerce and employee productivity. This role will initially be an individual contributor with developer experience building AI agents and will collaborate with a strong technical retail and supply chain team, as well as horizontal engineering team(s). Define product, technical vision and architecture Develop platform s

artificial intelligenceaisupply chain
View job →

NVIDIA is seeking a Senior Software Engineer to help us develop distributed storage services for AI/ML. In this role you will work closely with the broader NVIDIA team to design and build a reliable, scalable, and efficient storage-as-a-service tailored to AI applications that can be deployed anywhere and scale without limitations. This service supports the whole NVIDIA critical business from graphics drivers to autonomous vehicles to deep learning frameworks. To achieve this goal, we are looking for an engineer with a deep understanding of distributed systems, outstanding design skills, and a track record in building and delivering large-scale distributed services. What you will be doing: Leading the overall architecture and design of our distributed storage service optimized for AI/ML Develop and maintain distributed, robust and scalable Go programs deployed to state of the art open-source ecosystems, including Kubernetes. Develop and maintain user-space applications, containers, Go-bindings, and CLI tools. Building features for a distributed storage service to enhance availability and reliability for large-scale deployments Engaging and collaborating with NVIDIA Research, Computing, Product teams, cross-functional teams, and external customers to deliver Cloud services. Automating distributed storage service end-to-end, including deployment, management, and monitoring What we need to see: Bachelor’s of Science in Computer Science, or related field (or equivalent experience) with 8+ years of industry experience Strong background in developing distributed systems involving Golang, Kubernetes, and Cloud Service Provider integrations Strong track record of delivering distributed services in a variety of distributed computing environments Experience in i

kubernetesartificial intelligenceai
View job →

NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by phenomenal technology – and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are seeking a Senior Site Reliability Engineer – Storage, you will own the reliability, performance, and scalability of our global NAS, SAN, and Object Storage platforms that power critical internal and external services. You will combine deep storage expertise with strong automation and SRE practices to design, build, and operate highly available storage systems at scale. What you will be doing: Lead design, deployment, and operations of production NAS, SAN, and Object Storage platforms, ensuring reliability, performance, and security. Capture requirements from partner teams, architect storage solutions, and drive end‑to‑end implementation for new and existing services. Develop, maintain, and improve automation for provisioning, configuration, monitoring, incident response, and lifecycle management of storage infrastructure. Participate in on‑call and incident response, lead troubleshooting of complex storage and performance issues, and drive root cause analysis and preventive actions. Define and track SLOs/SLIs and error budgets for storage services, using observability and analytics to continuous

pythondockerkubernetes
View job →
N
1mo ago

We are looking for a creative and independent Production Engineer. NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing — with the GPU acting as the brains of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and build our teams with the smartest people in the world. Join us at the forefront of technological advancement. Make the choice to join us today. We need a creative individual who will help design and operationalize manufacturing plans for GPU Server products. Because of the increasing complexity of our GPU servers, we need your manufacturing experience to forecast equipment and capacity, as well as ensuring we think about possible trouble spots ahead of time. You will be exposed to various aspects of building and testing NVIDIA server products, from GPUs to full system testing. In addition, your responsibilities will include working with overseas manufacturing teams to increase yields, test coverage, and capacity, and reduce production costs. What you'll be doing: Identify the factory key performance index & metrics measurement for periodic review Work with ODM/CM engineers for continuous process improvement Manage manufacturing and quality issues on the production line Forecast equipment, tooling, and production capacity Escalate critical quality issues to engineering teams and management Assist, develop, and provide feedback on server design and DFx Collaborate with

We are looking for a creative and independent Production Engineer for NVIDIA server products. NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing — with the GPU acting as the brains of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and build our teams with the smartest people in the world. Join us at the forefront of technological advancement. Make the choice to join us today. We need a creative individual who will help design and operationalize manufacturing plans for GPU Server products. Because of the increasing complexity of our GPU servers, we need your manufacturing experience to forecast equipment and capacity, as well as ensuring we think about possible trouble spots ahead of time. You will be exposed to various aspects of building and testing NVIDIA server products, from GPUs to full system testing. In addition, your responsibilities will include working with overseas manufacturing teams to increase yields, test coverage and capacity, and reduce production costs. What you'll be doing: Identify the factory key performance index & metrics measurement for periodic review Work with ODM/CM engineers for continuous process improvement Manage manufacturing and quality issues on the production line Forecast equipment, tooling, and production capacity Escalate critical quality issues to engineering teams and management Assist, develop, and provide feedback on server design and DFx Collaborate with server engineering and product engineering teams

We are looking for an innovative thermal solutions integration Engineer. NVIDIA offers you to be a part of the System Product Engineering group and be responsible for assuring the best quality products to be a sale and deliver to NVIDIA’ s customers. The job provides deep knowledge of NVIDIA systems, a system-level view of our solutions, and a dynamic and positive working environment and offers the candidates the opportunity to take a major role in our testing strategy by leading our thermal solutions integration and testing for the company production. NVIDIA Networking unit has continuously reinvented itself over two decades. Our high-speed buses & network products are leading in the markets with innovative ways to improve speed and bandwidth from one generation to another. Today, we are increasingly known as the place for getting “End-to-End High-Speed Ethernet and InfiniBand Solutions” We're looking to grow our company and build our teams with smart people who can join us at the forefront of technological advancement. If you are passionate about enabling the highest quality Network products that will change the world, we want to hear from you! What You’ll Be Doing: Integration and testing of next-generation, large-scale thermal, pressure, and liquid management solutions. Perform qualification tests for cutting edge cooling and sensing solution. Analyse and summarize thermal performance data and fluid dynamics results to support design reviews and decision-making. Primary onsite focal point for malfunctions in thermal liquid cooling stations. Diagnose and resolve hardware/software issues to maintain continuous development labs activity Maintain and update hardware (manifolds, connectors, sensors) and software versions across all thermal systems according to engineering specifications. Perform initial RCA on thermal system failures. Extract detailed fail reports and corrective/

N
Nvidia
📍 Santa Clara, United States
1mo ago

We are seeking a Director of Events to lead the planning and delivery of NVIDIA's most significant and high-profile events. You will act as a strategic collaborator throughout the company — partnering with business units, product teams, marketing, communications, and executive leadership. NVIDIA is widely considered to be one of the technology world’s most desirable employers! We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you! What you'll be doing: Own and drive the overall strategy for assigned flagship events — defining goals, structure, and success metrics in partnership with senior leadership. Work across NVIDIA's business units, product teams, marketing, and communications organizations to align events with company priorities and ensure cross-functional agreement and implementation. Function as the chief strategic and operational authority — guiding and managing teams through every workstream stage, from idea development to on-site delivery and subsequent event evaluation. Drive the creative and experiential vision in close partnership with creative, content, and brand teams. Manage large, complex budgets — tracking costs, forecasting, and driving financial accountability across a multi-million dollar event portfolio. Serve as the primary executive point of contact, providing leadership with regular updates, critical issues, and strategic recommendations. Lead all vendor, agency, and venue relationships — holding external partners accountable to the highest standards of quality and delivery. Develop and manage master event timelines, ensuring all workstreams stay on track and interdependencies are clearly mapped. What we need to see: 12+ overall years of ev

awsaisalesforce
View job →

NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by phenomenal technology – and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are seeking a Senior Site Reliability Engineer – Storage, you will own the reliability, performance, and scalability of our global NAS, SAN, and Object Storage platforms that power critical internal and external services. You will combine deep storage expertise with strong automation and SRE practices to design, build, and operate highly available storage systems at scale. What You Will Be Doing: Lead design, deployment, and operations of production NAS, SAN, and Object Storage platforms, ensuring reliability, performance, and security. Capture requirements from partner teams, architect storage solutions, and drive end‑to‑end implementation for new and existing services. Develop, maintain, and improve automation for provisioning, configuration, monitoring, incident response, and lifecycle management of storage infrastructure. Participate in on‑call and incident response, lead troubleshooting of complex storage and performance issues, and drive root cause analysis and preventive actions. Define and track SLOs/SLIs and error budgets for storage services, using observability and analytics to continuously improve reliability and efficiency. Build and maintain runbooks, standard operating procedures, and comprehensive documentation for storage services and automation.<

pythondockerkubernetes
View job →

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. At NVIDIA, as a Principal Rack Scale Systems Infrastructure Engineer, you will build and guide the development of software systems. These systems support our upcoming rack-scale infrastructure products and services. This exceptional role sits where software meets hardware. You will work on control planes, state machines, orchestration systems, firmware, OS lifecycle, and networking fabrics. Your task is to compose infrastructure-as-a-service control plane software that converts complex rack-scale hardware into dependable, manageable, and programmable infrastructure for NVIDIA, partners, and leading cloud and enterprise clients globally. What You Will Be Doing: Define the complete software architecture for rack-scale infrastructure products and services, covering control plane services, infrastructure management, firmware, operating systems, kernel drivers, networking fabrics, accelerator software, and user-mode manageability software. Use Kubernetes and cloud-native primitives as an infrastructure fabric when appropriate. This includes controllers, operators, reconciliation loops, and open source components. These components can operate safely at rack and fleet scale. Build open source infrastructure software that can b

kuberneteslinuxai
View job →
N
1mo ago

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. NVIDIA has a rapidly expanding ecosystem of data center platform & node designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each bringing together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We're searching for a highly motivated, technical leader to drive the engineering roadmap and innovation for our rack system software architecture. From firmware, kernel drivers, operating systems, networking, fabrics and associated user mode drivers &#43; manageability software. You will work with component leads internally and engage with industry leading hyperscalar / cloud service providers on taking these products to market. What you’ll be doing: Drive the software end-to-end architecture for NVIDIA's rack-scale products Maintain deep understanding of the product portfolio and roadmap; translate forward-looking plans into clear, formal software requirements that anchor execution across the organization. Ensure high quality & reliable software; serving as a trusted architectural partner to teams requiring

V
Verse
📍 San Francisco• Full-time• $110K – $150K/yr
18 days ago

Location: San Francisco, CA (hybrid) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role We're looking for a highly analytical Senior Business Operations Analyst to support the growth and execution of our Dispatch Intelligence product. This role sits at the intersection of business operations, customer, product, engineering, and data science and has two core areas of responsibility. First, you will help drive overall program execution for Dispatch Intelligence: bringing structure to complex cross-functional initiatives, improving processes, tracking progress, and ensuring teams stay aligned on priorities and timelines. Second, you will help build and operate the processes through which flexible energy assets are onboarded onto the Verse platform and continuously improve their operational performance. The ideal candidate combines strong analytical problem-solving with exceptional project management and is comfortable working across both technical and commercial teams. You will play a critical role in helping Verse scale Dispatch Intelligence from individual projects and assets

aigorust
View job →
V
Verse
📍 San Francisco• Full-time• $150K – $240K/yr
18 days ago

Location: San Francisco, CA (Remote/Hybrid Available) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role You will be a member of the technical staff developing product experiences for our Dispatch Intelligence users – customers who want and have battery energy storage systems for additional energy cost savings or faster interconnection times. In this role, you will serve in a “full stack” capacity designing, building, and maintaining frontend and backend components of our energy storage suite of applications. We use Typescript, React, Next.js, Tailwind CSS, Radix/ShadCN, Jest, Cypress, Playwright, Vitest, and Storybook with Echarts and D3/Observable for data visualization for our frontend, and Cloudflare Pages for hosting and content delivery. We rely on identity and auth platforms like Clerk for sign-in flows. Our backend is written in Go and Python with Postgres/AlloyDB and blob storage for data persistence. Key Responsibilities Foster a culture and mindset of well-designed systems, test-driven software, and proactive communication with a high degree of transparency, mutual resp

typescriptpythonreact
View job →
S
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies - from the world’s largest enterprises to the most ambitious startups - use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone's reach while doing the most important work of your career. Metronome is the leading usage-based billing platform built for modern software companies. With Metronome, companies can launch products faster, offer any pricing model, and streamline finance workflows without writing code. Our platform computes millions of invoices per billing period and is scaling rapidly to accommodate new customers, saving them hours of development time and manual invoicing and enabling them to use consumption data to better serve their customers. Our customers love our product and approach, and we’re humbled to work with amazing companies like OpenAI, Databricks, NVIDIA, Confluent, and Anthropic. You'll be joining an experienced team that includes founders who have successfully built and sold startups before. Our founders and employees have direct experience building and scaling teams through massive growth at companies like Dropbox, Clever, and New Relic. On the back of this experience and our success-to-date, we’ve raised over $128M from leading investors including NEA, Andreessen Horowitz, General Catalyst, Elad Gil, and Workday Ventures. We’re also proud to have founders and executives of companies like Segment, Plaid, Looker, Gitlab, Confluent, HashiCorp, and Snowflake, as investors who have experienced the pain we're solving firsthand. About the team The Solutions Architecture team at Metronome is a technical group that sits at the intersection of sales, growth, product, and R&D. In simple terms, we own the technical as

gitrestai
View job →
S
Synthesia
📍 United States• Full-time
1mo ago

Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. Remote (US East Coast preferred, for timezone coverage) About the team Cloud Infrastructure owns the platform every Synthesia product runs on — AWS, Kubernetes, MongoDB, Temporal, our observability stack, and the vendor and cost relationships underneath them. We're a small, high-leverage team scaling toward a domain-ownership model: small groups that both build and operate the systems they're accountable for. The role We're hiring a dedicated SRE to take real ownership of operational excellence across Cloud Infrastructure. Today, too much critical operational knowledge — vendor relationships, cost management, and incident response — lives with one or two people. Your mission is to take genuine ownership of those domains, make them resilient to any single person, and raise the bar on how reliably we run. This is not simply a ticket-queue or keep-the-lights-on role. You'll own domains end to end: understand them deeply, operate them well, and build the automation and tooling that make them boring . We deliberately pair operational and engineering work so the role grows rather than narrows. What you'll own Incident management & operational excellence — take custody of the incident process: on-call quality, resp

pythonmongodbaws
View job →
V
Verse
📍 San Francisco• Full-time• $150K – $240K/yr
18 days ago

Location: San Francisco, CA (Remote/Hybrid Available) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role As a Software Engineer focusing on Distributed Systems at Verse, you will work in collaboration with some of the brightest industry experts in the field building cloud-native applications that scale to trillions of data points collected from electricity markets globally. You will be a part of a dynamic, robust team primarily supporting the backend needs of our Aria software product spanning hundreds of data sources, sinks, services, and jobs. Your expertise will not only have a direct impact on product decisions, but you also be well-positioned to drive the development and trajectory of our entire platform and infrastructure and influence important architectural decisions that affect the whole organization. Key Responsibilities Foster a culture and mindset of well-designed systems, test-driven software, and transparent communication with a high caliber of mutual respect and consideration for stakeholders Read and write a lot of Go, Python, and Protobuf Build, test, debug, maint

pythonjavakubernetes
View job →
🔔

Get new nvidia manager jobs by email

Daily job updates · Unsubscribe anytime