We are seeking a Senior Software Engineer with strong infrastructure expertise to design, build, and operate the next generation of our enterprise Observability, Automation, and AI-driven Reliability Platform. This role will build highly scalable distributed systems and platform services spanning Storage, Compute, Network, VMware, OpenShift, and bare-metal infrastructure. The engineer will help transform infrastructure operations from reactive monitoring and manual remediation to proactive, predictive, and AI-driven autonomous operations. What You Will Be Doing: Design, build, and operate distributed software platforms for enterprise observability, telemetry, automation, and infrastructure reliability at large scale. Develop reusable platform services, APIs, automation frameworks, and control planes that enable self-service, reduce operational toil, and automate infrastructure operations across multiple engineering teams. Build scalable telemetry and event-processing systems spanning metrics, logs, traces, events, topology, and alerts, with the performance and efficiency to process billions of infrastructure signals. Build intelligent and AI-native reliability capabilities, including agentic workflows for anomaly detection, forecasting, root-cause analysis, automated debugging, and closed-loop remediation. Drive technical architecture and engineering direction across Storage, Compute, Network, and Platform domains, solving complex and ambiguous problems that span multiple teams. Engineer for production at scale, with strong focus on software quality, scalability, security, performance, observability, maintainability, and operational readiness. Provide technical leadership and mentorship, influence engineerin
Jobiba hiring network
Nvidia Jobs
568 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current nvidia jobs. Use filters to narrow by work mode, employment type, experience and date posted.
We're looking for a Principal Software Engineer to join our CSP Engagements team as the technical focal point for rack-scale system SW/FW, working with CSP engineering teams to ensure they can deploy, monitor, and operate these systems reliably at fleet scale. In this role, you will collaborate with NVIDIA's cross-functional rack-scale system SW/FW engineering teams with dedicated CSP-facing technical leadership. Your focus is on the system-level software that manages, monitors, and recovers the rack as a whole — fabric management, GPU/NVSwitch error handling and recovery, health telemetry APIs, firmware update orchestration, and SW-driven serviceability. You will drive work streams with CSP engineering teams to build shared understanding of the architecture, incorporate their operational feedback, and ensure integration readiness. What you'll be doing: Drive rack-scale SW/FW architecture alignment across CSP engagements — including fabric management software, link health monitoring, GPU/NVSwitch error handling, SW/FW serviceability features (e.g., hot-plug support, component isolation, firmware-driven recovery), and multi-component firmware orchestration Drive technical work streams with CSP engineering teams on rack-scale system software — ensuring they deeply understand fabric management, NVSwitch behavior, error handling and recovery policies, health telemetry APIs, and SW/FW-controlled recovery operation Capture and synthesize CSP engineering feedback on rack-scale system software — health monitoring APIs, SW-driven serviceability workflows, firmware update orchestration, and error recovery behavior — champion that feedback into NVIDIA's architecture decisions Collaborate with multi-functional teams to ensure customer operational requirements are reflected in system software and firmware development Identify cross-CSP patterns in rack-scale SW/FW iss
We are seeking a highly skilled and experienced Staff Network Site Reliability Engineer (SRE) to join our Enterprise Network Operations and SRE team. In this role, you will be pivotal in implementing our vision for a reliable and efficient network infrastructure. The ideal candidate is passionate about network operations and committed to enhancing the user experience. You'll have the opportunity to solve complex network challenges using hands-on debugging and by focusing on network automation, observability, documentation, and operational excellence. This is a critical position focused on ensuring user satisfaction and brilliance in network operations. What you'll be doing: Owning the operational aspect of the network infrastructure, ensuring its high availability and reliability, actively working on network incidents and service requests. Partnering with architecture and deployment teams to guarantee that new implementations are supportable and align with production standards. Advocating for and implementing automation to reduce toil and improve operational efficiency. Minimizing manual operational tasks to achieve and maintain Service Level Objectives (SLOs). Monitoring network performance, identifying areas for improvement, and collaborating with relevant teams to implement refinements. Proactively identifying and mitigating network risks to promote continuous improvement. Collaborating with domain experts across functions to resolve production issues swiftly and effectively, ensuring customer happiness. Conducting blameless postmortems and following through on Root Cause Analyses (RCAs). Discovering opportunities for operational improvements and teaming up with colleagues to devise solutions that enhance excellence and sustainability in network operations. Developing knowledge base articles for automa
NVIDIA is looking to hire a Test Engineer in Roskilde, Denmark for the development of silicon photonics and semiconductor devices that will link-up the datacenters of tomorrow. The successful candidate will work on electrical and optical device test program development, including technical management of development projects with multiple partners from across the organization. This role is an opportunity to work on NVIDIA core technology under development. What you'll be doing: Develop and implement electrical semiconductor test programs, collaborating closely with international teams to build solutions that ensure flawless execution. Perform device characterization and conduct data analysis to provide quality feedback to designers. Partner with internal teams to ramp new silicon from development to product and hand over to production teams. Improve yield and reduce test time through innovative solutions. What we need to see: A B.Sc/M.Sc in Electrical Engineering or Physics from a reputable university, or equivalent experience. Demonstrated experience in electrical engineering and programming. Over 2 years of experience in the semiconductor industry. Experience in photonic or semiconductor testing or validation, along with programming languages like Python and VBA. New outstanding graduates will also be considered. Ways to stand out from the crowd: Strong knowledge of probability and statistics tools, such as JMP. A new way of problem-solving that drives success. Experience leading cross-organizational activities with a high sense of responsibility. Join us in our mission to pioneer the next generation of computing and leave a lasting mark on the world. At NVIDIA, you will be part of a
NVIDIA is seeking a Senior Staff SRE to build and operate reliable, scalable compute platforms that support global engineering workloads. This role spans Kubernetes, KubeVirt, bare-metal infrastructure, automation, observability, and AI-enabled operations. Join a team that solves complex infrastructure challenges, builds durable automation, and improves the reliability and operational experience of critical compute services. What you’ll be doing: Build, operate, and improve large-scale Kubernetes, KubeVirt, Linux, container, and bare-metal compute platforms, with a focus on performance, capacity, reliability, and operational scale. Lead bare-metal provisioning and lifecycle management in data centers, including PXE boot, DHCP, DNS, OS provisioning, hardware validation, and fleet automation. Develop automation, self-service capabilities, and observability solutions using APIs, Python or Go, Infrastructure as Code, configuration management, metrics, logs, traces, and service-health data. Define and operate SLOs, SLIs, error budgets, alerting, and incident-response practices; lead complex incident investigations, corrective actions, and blameless postmortems. Partner with infrastructure, security, hardware, data-center, and application teams to deliver global platform initiatives, and participate in an on-call rotation. What we need to see: BS in Computer Science, Engineering, a related technical field, or equivalent experience, plus 10+ years operating production infrastructure or platform services. Strong expertise in Kubernetes administration, KubeVirt, Docker, containerization, microservices, Linux systems, and resolving distributed-system challenges. <l
We're looking for a Principal Engineer to join our CSP Engagements team as the technical focal point for end-to-end performance, working directly with engineering teams of key CSP/hyperscale customers to ensure they achieve various performance targets on NVIDIA platforms. In this role, you will augment NVIDIA's performance and benchmark teams with a dedicated CSP-facing focus. You will drive work streams with CSP engineering teams to build shared understanding of platform performance characteristics, gather and incorporate their workload-specific feedback into NVIDIA's optimization priorities, and validate that performance targets are met in customer-representative configurations. Your cross-CSP visibility enables you to identify patterns and drive systemic improvements in documentation, configuration guidance, and tooling. What you'll be doing: Drive performance characterization work streams with engineering teams of key CSP/hyperscale customers — ensuring they understand platform performance expectations, profiling methodology, and tuning options for their specific workloads Gather and synthesize CSP performance feedback — identify gaps between expected and actual throughput, and champion optimization priorities back into NVIDIA's CUDA, NCCL, driver, and firmware teams Ensure key open-source performance and stress tools (e.g., STREAM, GPU Burn, GPU BLAST) are updated and validated for the latest NVIDIA rack-scale systems, GPU architectures, and CPU platforms — so customers and internal teams have reliable baseline measurements from day one Work closely with CSPs to ensure their own performance and validation tooling reflects the latest GPU capabilities, memory hierarchy changes, and platform-specific tuning parameters Conduct cross-CSP performance comparison and pattern analysis — identify configuration, software, or workload differences that explai
We are looking for a highly motivated AI/ML Software Engineer to join the Enterprise Agentic AI Platform team within IT. You will work closely with Business Analysts, and Engineering teams to design, develop, and deploy enterprise AI solutions that improve productivity and automate business workflows across Engineering, Operations, and Manufacturing. What you'll be doing: Design, develop, and deploy Agentic AI applications using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and AI orchestration frameworks. Build scalable AI services and reusable components integrated with enterprise applications such as PLM, SAP, and other business systems. Collaborate with business and IT teams to translate business requirements into AI-driven solutions. Develop secure, scalable APIs and enterprise integrations to enable intelligent workflows and automation. Improve AI solution quality, performance, and reliability through prompt engineering, evaluation, and continuous optimization. Partner with cross-functional teams throughout the Software Development Lifecycle (SDLC), from solution design through deployment and production support. What we need to see: Bachelor's or Master's degree in Computer Science, Information Technology, AI/ML, or a related field. 6+ years of software engineering experience with strong proficiency in Python and backend application development. Hands-on experience with Generative AI, LLMs, RAG, AI agents, REST APIs, and cloud-native application development. Experience integrating enterprise applications and building scalable, production-ready software solutions. Strong analytical, problem-solving, communicatio
Join NVIDIA's outstanding team and contribute to our legacy of innovation in the Senior Product Quality role! At NVIDIA, we are dedicated to pushing the boundaries of pioneering technology. The Senior Product Quality Engineer will provide leadership and management for the Networking system product quality within the organization. Establish, maintain, and optimize an effective quality management methods and target to achieve high quality performance. As part of our team, you will work on networking IC products and ensure their readiness for production stage. Leading and mentoring root cause analysis for issues related to design, manufacturing or customers. Enhance and improve products performance by proactive risk assessment and other prevention methods. This is an outstanding opportunity to be part of a company that is crafting the future of AI and computing. What You'll Be Doing Integrated Circuits (IC) quality manager will coordinate and lead the quality function and sustain the quality philosophy for the organization. Establish, maintain and optimize an effective quality management system. Build and implement a quality plan to achieve the levels of quality established through organizational goals, customer expectations, and related partners. Improve products by working towards problem-solving tools, prevention methods, quality-at-the-source and continual improvement techniques. Maintain and improve product build development, verification and validation processes, to allow high quality of interconnects and systems products throughout the entire product life cycle. What We Need to See Bachelor’s degree or master’s degree or equivalent experience in electrical engineering or material science. Other technical degrees will be considered Experience: 5+ years preferred in leadership and manufacturing Knowledge in semiconductor
NVIDIA is seeking a Senior Technical Program Manager to join the CSP Engagements team, focused on deep technical engagement with hyperscale cloud service providers for NVIDIA’s next‑generation datacenter systems such as Vera Rubin NVL72. This role is intended for experienced systems and embedded software leaders—including software engineering managers, technical leads, or senior architects—who have led datacenter server and platform software programs and can operate as a trusted technical partner to hyperscale CSP engineering teams. As a member of the CSP Engagements team, you will act as the primary technical engagement leader between NVIDIA’s system software organizations and CSP platform, system software, and AI teams, ensuring alignment, readiness, and successful large‑scale deployment of NVIDIA‑based datacenter solutions. What you will be doing: Lead deep technical engagements with hyperscale CSPs as the primary NVIDIA point of contact for system software, firmware, and platform readiness for NVIDIA datacenter products. Partner directly with CSP system software, firmware, and infrastructure engineering leaders to align on software architecture, bring‑up plans, deployment readiness, and production requirements for NVIDIA‑based server and rack‑scale platforms. Represent CSP technical priorities internally, advocating for customer requirements and tradeoffs across NVIDIA’s system software, firmware, hardware, silicon, and product teams are aligned to customer needs, timelines, and constraints. Own the end‑to‑end CSP engagement lifecycle, from early technical alignment and pre‑production readiness through large‑scale deployment, escalation management, and sustained production support. Drive bi‑directional technical communication: translating CSP system‑level requirements into actionable focus areas for NVIDIA engineering teams, while clearly communicating N
As a Senior Technical Program Manager at NVIDIA you will act as the bridge between strategy, execution, and technical alignment, driving complex programs end-to-end. This role demands strong technical acumen, exceptional program management skills, and the ability to influence without authority across a dynamic, matrixed organization. You will collaborate across software engineering, product management, OEMs and other teams to deliver solutions that shape the future of accelerated computing, with focus on Multimedia technologies (audio and video). What you’ll be doing: Program Ownership. Lead multimedia programs from inception to delivery, ensuring timelines, milestones, and quality standards are met. Manage release checklists, post-release feedback, and continuous improvement initiatives. Maintain and continuously improve program plans; ensure project focus and execution against commitments. Create and maintain metrics that track execution effectiveness and program health. Technical Leadership: Deep dive into product architecture and design; maintain curiosity and technical proficiency. Collaborate with engineering teams to resolve technical challenges and optimize solutions. Stakeholder Management: Communicate decisions, progress, and risks clearly to stakeholders, customers at all levels. Build trust and influence across cross-functional teams. Strategic Thinking & Risk Management: Identify and mitigate risks proactively to prevent delays and cost overruns. Drive RCCA (Root Cause and Corrective Action) and ensure gaps are closed effectively. Communication & Information Management: Deliver precise, tailored information for diverse audiences and higher management. Consume complex data and bring clarity to discussions for actionable outcomes. Complex programs with sign
NVIDIA Inception is our free global programme for start-ups building with AI, data science and accelerated computing — now more than 40,000 companies worldwide, with over 2,000 of them based in India alone. South Asia is one of the fastest-moving corners of that network: a market whose density sits in applied and vertical AI — medical imaging, drug discovery, edge vision, agentic and sovereign AI, robotics, agri-tech, AI governance — rather than in frontier model labs. We are looking for an Inception Regional Lead to own that motion end to end. This is a builder’s role with a commercial edge. You will set the regional strategy, lead a team of Inception Partner and Community managers across India, Bangladesh, Sri Lanka and neighbouring markets, and be personally accountable for how much NVIDIA hardware, software and cloud the region’s start-ups design in and consume. You will spend real time on the ground in founders’ offices, because that is where stalled accounts turn into concrete asks. And you will act as the connective tissue between start-ups and the rest of NVIDIA — worldwide field operations, solution architects, developer marketing, the reseller and cloud partner ecosystem, and NVentures. What you’ll be doing: Own the South Asia Inception strategy and number. Set regional priorities, coverage model and account tiering across strategic, member, community and prospect accounts; forecast and report on pipeline, design wins and consumption to programme and field leadership. Lead and grow the team. Manage, coach and develop a team of Inception Partner and community programme managers. Set territory and vertical coverage, run the operating cadence, and raise the bar on technical fluency and commercial rigour across the team. Drive commercial outcomes with start-ups. Move accounts from awareness to technical discovery to a named
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are looking for an AI Native Startup Partner Managers to join NVIDIA for Startups — NVIDIA’s business got startups building on NVIDIA technologies across AI, accelerated data science, high-performance computing, and advanced visual computing. What you’ll be doing: Identify, engage, technically assess, and recruit the most relevant and promising startups across Europe building on NVIDIA-accelerated platforms, with a focus on Media & Entertainment, LegalTech, AI Developer Tools, Data Science, and Generative AI Evaluate startup technical maturity, architecture decisions, and platform alignment, including usage of CUDA libraries, NVIDIA SDKs, AI frameworks, GPU-accelerated data pipelines, and training/inference stacks Share ecosystem, technology, and developer-level insights with internal partners, translating startup feedback into actionable input for NVIDIA platform, SDK, and product teams Act as the voice of NVIDIA Inception members by developing a deep understanding of their technical roadmaps, research foundations, system architectures, and platform dependencies Provide AI Native priori
We are seeking software engineers to work on next-generation graphics and computing products. Our charter is to build the most stressful set of applications a GPU or high performance computing server would see in its life cycle. The best candidates will have strong C++ programming skills, thorough knowledge of graphics concepts and algorithms, a solid foundation of systems software with emphasis on OS fundamentals, and a deep understanding of current generation PC/hardware architecture. Excellent communication skills and a dedication to meticulous engineering practices are a requirement. As a system software engineer, you will extensively use your knowledge of operating systems, algorithms, and computer architecture to provide robust and efficient solutions to validate and test next generation processors. What you'll be doing: Working closely with architecture, hardware and driver teams through the product development lifecycle of computing and graphics processors, as well as compute products. Responsible for crafting software tools and infrastructure required for new chip development, validation, and productization. You will assess new hardware features and architect manufacturing diagnostic tests using pre-beta CUDA and OpenGL extensions. This job will require an understanding of our hardware and software architectures. What we need to see: BS or MS degree in one of the areas of Electrical Engineering, Computer Engineering, Computer Science or equivalent experience 3+ years experience in a related hardware/software position Strong C/C++ programming skills Familiarity with P
NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. Today, we are tapping into the unlimited potential of AI to define the next era of computing. As an NVIDIAN, you will address challenges spanning architecture, silicon, firmware, software, and production — and excellent judgment matters as much as technical depth! Every major NVIDIA silicon product family—from the chips powering AI and datacentre systems to gaming, professional, embedded, and automotive platforms—passes through our productization work on its way to production. NVIDIA’s Silicon Co-Design Productization team works from pre-silicon strategy and feature development through bring-up, characterization, correlation, and optimization. Our charter spans power & performance modelling, bring up & tuning of low-power features, power & thermal controllers , and system-level optimization that ultimately shape how NVIDIA products are configured, binned, specified, and shipped. What you'll be doing Drive silicon power productization from pre-silicon planning through bring-up and production, including test strategy, feature readiness, characterization, and optimization. Partner with architecture and design teams to identify improvements, validate features, and help translate them into production-ready solutions. Correlate measured silicon behaviour with pre-silicon expectations, investigate gaps, and drive complex issues to root cause. Build power and performance models and characterization methodologies that decide silicon binning, product specifications, productization decisions, and customer guidance. Use AI/ML and data-driven methods to analyse characterization & telemetry data, identify anomalies & trends , and accelerate issue debug across silicon, board, power delivery, firmware,
We are looking for a creative and independent Production Engineer. NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing — with the GPU acting as the brains of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and build our teams with the smartest people in the world. Join us at the forefront of technological advancement. Make the choice to join us today. We need a creative individual who will help design and operationalize manufacturing plans for GPU Server products. Because of the increasing complexity of our GPU servers, we need your manufacturing experience to forecast equipment and capacity, as well as ensuring we think about possible trouble spots ahead of time. You will be exposed to various aspects of building and testing NVIDIA server products, from GPUs to full system testing. In addition, your responsibilities will include working with overseas manufacturing teams to increase yields, test coverage, and capacity, and reduce production costs. What you'll be doing: Identify the factory key performance index & metrics measurement for periodic review Work with ODM/CM engineers for continuous process improvement Manage manufacturing and quality issues on the production line Forecast equipment, tooling, and production capacity Escalate critical quality issues to engineering teams and management Assist, develop, and provide feedback on server design and DFx Collaborate with
Get new nvidia jobs by email
Daily job updates · Unsubscribe anytime