Applying for Principal System Power Management and Performance Architect at Nvidia? Compare this live job with your resume first.

🎯 Tailor my resume
←Jobiba.Me
N
Active1mo ago

Principal System Power Management and Performance Architect

Nvidia·📍 Ca, Santa Clara, United States

Employment

Full-Time

Work mode

On-site

Experience

SENIOR LEVEL

Salary

Not disclosed

Salary not disclosed

Check market pay for comparable Principal System Power Management and Performance Architect roles before applying.

Salary →

Role overview

Job description

NVIDIA builds the silicon behind AI, accelerated computing, and graphics. Every watt of performance and every degree of thermal headroom traces back to decisions made in power, performance, and thermal architecture. We are the Silicon Co-Design Group (SCG). We identify, own, and drive system-level co-design ideas. We start with initial concepts and advance to product differentiation across NVIDIA's roadmap.

We are hiring a Principal System Power Management and Performance Architect who operates at the ambiguous boundary where workload behavior, silicon capabilities, firmware policies, and platform constraints collide, and who turns that ambiguity into architecture that survives across multiple silicon generations. SCG scope spans architecture, design, software, operations, platforms, and productization. This role shapes system, platform, and data center features and behavior, and partners with teams across NVIDIA.

What You'll Be Doing:

The work here is rarely well-defined when it arrives. You will be given problems that appear to be performance gaps or power anomalies and encouraged to build a framework for solving them, not just tackle a single instance.

  • Define the multi-generation roadmap for system-level power and performance features, grounded in prototyping, use-case analysis, and cost/benefit trade-offs across segments. You will decide what problems are worth solving and why.

  • Own the architecture and integration strategy for HSIO power management, DVFS, P-states, and low-power features. Your decisions improve product performance, power, and reliability across product lines — not just the current program.

  • Lead system-level boot and IST architecture defining how power and clock domains initialize, sequence, and recover across complex multi-IP systems where the interaction space is large and the failure modes matter.

  • Drive power management strategy at datacenter scale, including rack-level power telemetry and platform state coordination in high-performance environments.

  • Identify and redesign system-level processes that break under new product requirements—boot/reset flows, low-power entry/exit sequences, and control-system policies across IPs. When a prior design no longer holds, you diagnose why and architect what comes next.

  • Serve as the multi-functional technical authority across architecture, ASIC, board/platform, and firmware teams. You improve the power-performance trade-offs by influencing decisions made by teams you do not control.

What We Need to See:

We are calibrating for candidates who can carry ambiguous product-level power and performance problems from concept through silicon correlation to release trade-off, and who can show the work!

  • BS or MS in EE/CE, or equivalent experience proven through the work itself

  • 15+ years of system architecture, development, and validation experience with a strong focus on power and performance-per-watt optimization in datacenter or high-performance platforms.

  • A verifiable history of guiding architecture decisions that shipped across multiple silicon programs. Be ready to walk through a case where you challenged a roadmap direction, explain your reasoning from first principles, and describe what changed.

  • Extensive knowledge of HSIO power management, DVFS, P-states, system boot flows, and low-power entry/exit architecture, including how they interact with IP, firmware, and platform layers. The expectation is not pattern-matching to prior solutions; it is knowing why the design space is built the way it is.

  • Strong fundamentals in low-power design, power management techniques, and rack-level performance architecture.

Ways to stand out from the crowd:

We are not looking for breadth of exposure. We are looking for depth of ownership in areas where the design space is hard, and the artifacts prove it!

  • Practical, hands-on depth in at least two of the following: Preference is given to candidates who have debugged these features in silicon and then architected the next-generation solution in any of the two areas: system power management, robust boot flow optimizations, HSIO, and memory power management.

  • A power management architecture or methodology you introduced that survived multiple silicon generations and was adopted beyond your immediate team. We want to understand the original problem, why prior approaches failed, how you structured the solution, and what the adoption path looked like.

  • A demonstrated AI/agentic workflow practice: Not just tool use, but a disciplined approach to compressing engineering velocity with AI while applying strong judgment to validate and refine results. Be prepared to describe where you've found AI dangerous, not just useful.

  • Visible technical leadership through patents, publications, conference presentations, standards participation, or broader industry recognition for work that influenced products, roadmaps, or engineering practice.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/ 

#LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 232,000 USD - 368,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 27, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

What they are looking for

Skills & requirements

Qualification

Every watt of performance and every degree of thermal headroom traces back to decisions made in power, performance, and thermal architecture

N

Hiring company

Nvidia

Explore this employer's active roles, salary signals and company profile on Jobiba.

Keep exploring

Similar active roles

Fresh roles matched to this title and market.

View all →
S
📍 Bellevue, Washington, United States· Full-time

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are the Snowflake Metadata team. We own Snowflake’s metadata systems that make it easy for customers to query, modify and manage their petabyte-scale data. We develop distributed systems that store and maintain metadata, transaction frameworks that power Snowflake’s query and DML capabilities, asynchronous systems that provide time travel and lifecycle management capabilities and entity metadata supporting DDL capabilities. We also build foundational capabilities that deliver global features like cross-region replication (Snowgrid), data sharing, and data marketplace. AS A PRINCIPAL SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real business needs at large scale by applying your software engineering and analytical problem solving skills. Design, develop and support fault-tolerant scalable distributed systems for our Snowgrid and Data Sharing teams. Create architecture and design, influence our product roadmap, and take ownership and responsibility over new projects. Analyze fault-tolerance and high availability issues, performance and scale challenges, and solve them. Mentor and grow junior engineers. Understand trade-offs between consistency, performance and costs to build solutions which can meet the demands of rapidly growing services. Ensure operational readiness of

JavaAIGoRust
MT
📍 Boise, ID - Main Site, United States

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron currently has an opening for a Staff or Principal level EUV Equipment Engineer for our Boise, Idaho location. Our Advanced Lithography and Equipment Engineering team enables the technologies that power the future of semiconductor innovation. We work at the leading edge of EUV and High-NA EUV development, partnering across process, manufacturing, facilities, and supplier organizations to deliver world-class equipment performance and support Micron’s technology roadmap. This EUV Equipment Engineer is a technical leadership role focused on maximizing the performance, reliability, and capability of EUV lithography systems. This position plays a critical role in advancing next-generation semiconductor technologies through equipment optimization, problem solving, supplier engagement, and collaboration across engineering disciplines. Responsibilities: Own the performance, reliability, availability, and productivity of assigned EUV lithography equipment Lead equipment installations, upgrades, qualifications, maintenance strategies, and continuous improvement initiatives Drive improvements in scanner performance, tool matching, imaging stability, defectivity reduction, contamination control, and recipe optimization Analyze equipment data and establish monitoring systems to improve overlay, focus, source performance, automation, and overall equipment health Provide technical leadership through supplier management, cross-functional collaboration, mentoring, a

AIRecruitment
M
📍 Minnesota, United States of America, United States

We anticipate the application window for this opening will close on - 5 Oct 2026 Careers that change lives start here. Medtronic is a global leader in healthcare technology with a Mission to alleviate pain, restore health, and extend life. Our 95,000 employees work across more than 150 countries to put patients first — developing innovative medical technologies that improve the lives of 72+ million patients each year. Your unique talents will help shape the future of healthcare while building a career grounded in purpose, growth, and impact. A Day in the Life We are seeking a highly motivated Finance Forward Deployed Engineer (FDE) to join our Finance Analytics and Transformation team. This role sits at the intersection of Finance, Data, Analytics, Automation, and Artificial Intelligence. Unlike a traditional data engineering or analytics role, the Forward Deployed Engineer will work alongside Finance teams to deeply understand business processes, identify high-value opportunities, and rapidly design, deploy, and scale solutions that improve how Finance operates and makes decisions. The ideal candidate combines strong Finance and business acumen with hands-on technical capabilities. This individual will be comfortable participating in forecasting, planning, performance reviews, management reporting, and other Finance processes while also working directly with data, building analytical solutions, automating workflows, and applying emerging AI capabilities. The role requires a strong builder mindset: someone who can move from an ambiguous business problem to a working solution, partner with users to iterate quickly, and ultimately transition successful solutions into scalable enterprise capabilities. At Medtronic, we bring bold ideas forward with speed and

PythonSQLMachine LearningArtificial Intelligence
G
📍 Austin, Texas, United States· Full-time

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives, spanning AI research specialists, silicon designers, software engineers and systems architects. Job Summary We are looking for an experienced Principal Engineer to join our System Management team and help lead the development of critical interfaces used by internal and external customers to manage system state. You will provide technical leadership within assigned areas of System Management, guide architecture and implementation choices, mentor engineers and translate broader technical direction into effective execution. This is a hands-on engineering role for someone who can lead complex technical work, improve reliability and operational readiness, and collaborate effectively across multiple engineering disciplines. The Team The System Management team sits within the Software Platform group and helps build Graphcore products into large-scale AI solutions for our customers. The team is responsible for developing the interfaces between hardware, AI software and frameworks, as well as providing interfaces for public and private cloud environments. This includes system management capabilities that abstract complex hardware administration and enable reliable deployment and operation at scale. As one of the first teams to work with new hardware and software, we regularly solve complex system-level problems

PythonKubernetesCI/CDGit
R
📍 San Mateo, CA, United States· Full-time

From $280.5K/yr

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. About the Role Inference Platform is Roblox's multi-tier orchestration system for microservices, powering AI Inference across Roblox. Through a simple developer API, we deliver a "deploy and forget" runtime, hiding the complexity of scheduling, scaling, and reliably running services across our on-prem and multi-cloud footprint at global scale. As Principal Product Manager, Jobs Platform, you'll take the helm at a defining moment - leading the charge as we scale the platform to become the default runtime for critical Roblox services worldwide. You Will Own Inference Platform end-to-end - set the multi-year vision for how Roblox engineers deploy, run, and scale AI and other services across our Core and Edge Datacenters, and cloud. Power Roblox's AI future - build the platform that brings frontier models and next-gen AI workloads to life, with the primitives, scheduling guarantees, and resource classes AI teams need to move fast. Evolve the platform's techni

ReactAWSAzureGCP

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but operational in how we execute, and we support every product and research effort at OpenAI. Our tenets include prioritizing for impact, enabling researchers and developers, preparing for future transformative technologies, and fostering a strong, collaborative security culture. About the Role OpenAI is seeking a Principal Software Engineer to join the Infrastructure Security (InfraSec) team. InfraSec safeguards the core of OpenAI’s research and production environments: GPU supercomputing clusters, multi-cloud infrastructure, datacenters, networking, storage, and the critical services that power our frontier AI models. Our charter spans everything from bare-metal hardware and firmware to Kubernetes clusters, service meshes, and the data pathways that carry highly sensitive model weights and user data. As a Principal Software Engineer, you will set technical direction and drive execution of critical foundational services, such as authentication systems, egress/ingress proxies, access brokers, and key management platforms, that demand high standards of reliability, scalability, and software craftsmanship. These systems form the security backbone of OpenAI’s customer and supercomputing environment and must remain robust under intense scale and adversarial pressure. In this role, you will: Own the architecture and roadmap for one or more core security services (e.g., authN/Z, policy enforcement, secure proxies, key management), taking them from design to rollout to long-term operation. Design and implement planet-scale security systems that provide strong guarantees across hardware, operating systems, Kubernetes, networks, and CI/CD: balancing security, reliability, latency, and developer ergonomics. Lead cross-functional launches

AWSAzureGCPKubernetes

🔔 Get job alerts

New Principal System Power Management and Performance Architect jobs in Ca, Santa Clara, United States, straight to your inbox.

No spam · Unsubscribe anytime