←Jobiba.Me
N
Active1mo ago

Principal Systems Software Engineer, Semiconductor Systems Inspection

Nvidia·📍 Ca, Santa Clara, United States

Employment

Full-Time

Work mode

On-site

Experience

SENIOR LEVEL

Salary

$177,185–$208,592 / year · Jobiba est.

Market pay estimate

$177,185–$208,592 / year for comparable Software Engineer roles in United States. Not employer-provided.

Salary →

Role market pulse

How Software Engineer demand looks in United States

51/100 · steady

Live jobs

543

Posted 30d

103

30d movement

-76.6%

Remote share

15.1%

Salary listed

33.5%

Salary trend 1Y

Not enough history

Role overview

Job description

NVIDIA has transformed computer graphics, PC gaming, and accelerated computing for more than 25 years through exceptional technology and the people who build it. In semiconductor manufacturing, our role is to enable the ecosystem, not compete within it. We partner with fabs, equipment manufacturers, and software providers to make inspection, metrology, and manufacturing intelligence dramatically faster on the NVIDIA platform.

Our team builds the software that makes this possible: models, adaptation and evaluation workflows, and deployable inference capabilities that partners integrate into their own tools. We work in environments where labeled data is limited and proprietary, distributions shift across tools and fabs, production budgets are tight, and software must operate inside air-gapped facilities.

We’re seeking a Principal Systems Software Engineer for Semiconductor Inspection in Santa Clara. This is a hands-on architect role: you will define the approach, build it, evaluate it, and demonstrate the results. You will work across computer vision, time-series modeling, multimodal AI, anomaly detection, model adaptation, evaluation, and production inference. Success means technology that a fab or equipment vendor can integrate, operate, and trust—not only a successful internal demonstration. 

What you’ll be doing:

  • Define and prototype AI system architectures spanning optical and e-beam inspection, wafer and mask inspection, metrology, defect review, equipment signals, and process data.

  • Advance world foundation model capabilities for semiconductor manufacturing, including vision, time-series and multimodal representation learning, model adaptation, domain transfer, and data-scarce defect understanding.

  • Develop workflows for defect detection, classification, localization, segmentation, nuisance filtering, ADC, ADR, and equipment time-series anomaly detection.

  • Establish evaluation methods that measure model quality and robustness under tool-to-tool and fab-to-fab domain shift, diagnose failures, and define production-readiness thresholds.

  • Design fab-local evaluation and adaptation workflows that preserve customer data sovereignty and produce comparable evidence across partner deployments.

  • Connect inspection and metrology results with equipment signals, process history, wafer genealogy, SPC, and yield outcomes to support manufacturing-wide reasoning and root-cause analysis.

  • Design agentic workflows for air-gapped fabs that connect data triage, model inference, review assistance, root-cause analysis, human approval, and secure deployment.

  • Optimize and package inference capabilities to meet real production latency, memory, throughput, reliability, and security budgets, allowing partners to deploy without NVIDIA operating their systems.

  • Work directly with fabs and equipment manufacturers to move capabilities from prototype to partner-operated production.

  • Collaborate across research, software, process, metrology, inspection, and hardware teams to define platform capabilities and the technical roadmap for semiconductor AI.

What we need to see:

  • MS or PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related technical field, or equivalent experience.

  • 12+ years of professional software or algorithm engineering, including demonstrated technical ownership and architectural direction. Directly relevant PhD research may count toward this total; research alone cannot satisfy the domain and production-delivery requirements below.

  • 8+ years in semiconductor manufacturing, including inspection, metrology, defect review, process control, yield engineering, or fab-data systems.

  • Hands-on delivery of software that shipped in semiconductor equipment, operated in a fab, or was adopted by a manufacturing customer.

  • 4+ years of hands-on deep learning, machine learning, computer vision, or applied AI, including work within the last 24 months.

  • Strong Python skills and production experience with PyTorch or TensorFlow.

  • A track record of taking ambiguous technical problems through architecture, implementation, evaluation, and validated results.

  • Working GPU literacy, including the ability to reason about inference performance and evaluate hardware and software bottlenecks.

  • Direct engagement with fabs, manufacturing teams, or equipment vendors on technical requirements.

  • Strong analytical, communication, and cross-functional leadership skills.

Ways to stand out from the crowd:

  • Experience adapting large pretrained vision, multimodal, or world foundation models to specialized industrial domains.

  • Background in anomaly detection and generation, time-series modeling, FDC, virtual metrology, or advanced-packaging inspection.

  • Experience designing evaluation frameworks for sparse proprietary data and shifting production distributions.

  • Expertise in synthetic data, self-supervised or few-shot learning, domain adaptation, model compression, and inference optimization using NVIDIA technologies.

  • Experience building SDKs, platform capabilities, or agentic workflows adopted by external engineering teams.

With competitive salaries and a generous benefits package, NVIDIA is widely considered one of the technology industry’s most desirable employers. Our engineering teams are growing in some of the most consequential areas of manufacturing and semiconductor technology. If you are a hands-on Principal Systems Software Engineer who combines semiconductor expertise with architectural judgment and a drive to put advanced AI into production, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 20, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

What they are looking for

Skills & requirements

Qualification

What we need to see: MS or PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related technical field, or equivalent experience; Directly relevant PhD research may count toward this total; research alone cannot satisfy the domain and production-delivery requirements below

N

Hiring company

Nvidia

Explore this employer's active roles, salary signals and company profile on Jobiba.

Keep exploring

Similar active roles

Fresh roles matched to this title and market.

View all →
G
📍 Austin, Texas, United States· Full-time

About Graphcore At Graphcore, we’re building the future of AI compute.We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale.As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem.To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world.We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence. Job Summary We are looking for an experienced Silicon Test Engineer to join our Product Test and Diagnosis Department (PTD). This is a pivotal role and will involve building a team of engineers to develop System Level Test (SLT) capability within the company. Working closely with a cross-functional team you will implement SLT tests for a family of next generation AI Processors. The ideal candidate should have a focus on quality and demonstrate a good understanding of the importance of production test on the success of a product. T hey will have a proven Functional Test or ATE Test Engineering background, and will have a pragmatic, hands-on and flexible approach to a fast-changing environment. The Team The Product Test and Diagnostics team’s role is to detect and manage hardware defects that arise from the manufacture and use of our products. This covers chips, boards and finished systems and takes place both in the manufacturing sites and in the field. Responsibilities and Duties Managing a team of

PythonGitAIGo
T
📍 Austin, Texas, United States· Full-time

$100K – $500K/yr

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a highly skilled and experienced Engineer to lead post-silicon power characterization and correlation activities for cutting-edge semiconductor products. In this role, you will be responsible for developing and executing detailed power measurement strategies on silicon, correlating results with pre-silicon models, and driving improvements across power architecture, design, and modeling methodologies. You will serve as a key technical leader, interfacing across design, architecture, validation, and systems teams to ensure silicon meets power and performance specifications under all operating conditions. This role is hybrid, based out of Toronto, ON or Austin, TX or Santa Clara, CA. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who you are A Principal-level engineer with 8+ years in silicon power analysis and characterization, and a Master’s or PhD in EE, CE, or related field. Deep understanding of digital and mixed-signal power domains, including DVFS, leakage vs. dynamic power, and power gating. Highly proficient in lab-based power measurement using oscilloscopes, current probes, power analyzers, and SMUs, plus Python/Perl/MATLAB

PythonAWSGitAI
G
📍 Austin, Texas, United States· Full-time

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary Reporting to the Quality leadership within Manufacturing Operations, the Senior Reliability Scientist is responsible for leading reliability activities across complex, high-performance systems. Working closely with established reliability experts and cross-functional teams, this role uses experimental data and advanced modelling to inform design decisions, validate product reliability and optimise serviceability strategies, including spares provisioning. The Team The Quality team within Manufacturing Operations is responsible for ensuring product robustness, reliability and lifecycle performance across Graphcore’s hardware portfolio. The team includes experienced reliability specialists and works closely with technology research, chip, board, system design, platform and operations teams to translate reliability insights into actionable improvements across the product lifecycle. Responsibilities and Duties: · Define and refine reliability requirements across silicon, board and system levels, working in partnership with research and design teams · Apply ad

AIGoExcelSEM
G
📍 Austin, Texas, United States· Full-time

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking an experienced Principal Hardware Diagnostics Engineer to design and develop diagnostics software used to monitor hardware health and diagnose system-level issues across Graphcore’s AI infrastructure platforms. This role focuses on building diagnostics agents, tools, and analytics frameworks that enable engineers and automation systems to identify, isolate, and resolve hardware issues across blade-level servers and rack-scale clusters. The Team Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. The Systems Engineering and Platform Validation team ensures Graphcore’s AI compute platforms are reliable, diagnosable, and operationally robust at scale. The team co

PythonLinuxAIC++
N
📍 Santa Clara, United States
Demand 51/100

We are hiring senior engineers to work on the CUDA driver, a core component of our platform for accelerating general purpose computation on the GPU. Our team delivers features and improvements to better realize the potential of NVIDIA hardware for a growing range of computational workloads, ranging from deep learning, scientific computation, and self-driving cars to video games and virtual reality! CUDA defines a unified programming model across a range of system configurations and hardware capabilities. To accomplish this, the CUDA driver interacts with GPU hardware, kernel mode drivers, switches and the operating system. What you'll be doing: As a member of our team, you will use your design abilities, coding expertise, and creativity to deliver the best Compute platform in the world. You will craft elegant solutions to exciting problems and craft the future direction of CUDA as you collaborate with your peers across NVIDIA. You will evangelize, architect, and implement new CUDA features You'll oversee and drive development efforts across multiple teams Collaborate with members of hardware architecture teams Help define forward-looking improvements to the CUDA APIs and programming model Design and maintain performance and precision modeling Write effective, maintainable, and well-tested code Develop code for multiple operating systems What we need to see: Bachelor of Science or Master of Science degree in Computer Science, Electrical Engineering, or related field (or equivalent experience) 15&#43; years of relevant systems software development experience Strong C programming skills </

Artificial IntelligenceAI
H
📍 Texas, United States of America, United States
Demand 51/100

$147.1K – $230.9K/yr

Principal Embedded Firmware and Software Engineer Description - We are seeking a Principal Embedded Firmware & Software Engineer to lead the design, development, and debugging of embedded software and firmware for computer systems. In this role, you will combine deep, hands-on engineering expertise with system-level technical leadership to ensure seamless integration between software and hardware components, delivering reliable and efficient system performance. You will collaborate closely with cross-functional teams including hardware engineers, software developers, QA, and product managers to bring high-quality products to market. Responsibilities Provide technical leadership for the architecture, development, security, integration, debugging, validation, and deployment of embedded firmware and software including BIOS/UEFI, EFI applications and drivers, embedded controllers, and RTOS-based systems. Analyze hardware and system architectures to define firmware requirements, dependencies, interfaces, integration strategies, and validation approaches. Troubleshoot and resolve firmware issues by designing and implementing enhancements, updates, and programming changes across firmware subsystems. Define and drive firmware integration, verification, and validation strategies, including automated testing, regression testing, and continuous integration . Develop and improve engineering tools and automation using Python and other appropriate technologies for development, debugging, testing, analysis, and validation. Advance CI/CD and DevSecOps practices for embedded development to improve engineering velocity, quality, traceability, and release confidence. Evaluate and apply AI-assisted software development and engineering tools where they can improve developer productivity,

PythonGitLinuxAI

🔔 Get job alerts

New Principal Systems Software Engineer, Semiconductor Systems Inspection jobs in Ca, Santa Clara, United States, straight to your inbox.

No spam · Unsubscribe anytime