We're looking for a Principal Software Engineer to join our CSP Engagements team as the technical focal point for rack-scale system SW/FW, working with CSP engineering teams to ensure they can deploy, monitor, and operate these systems reliably at fleet scale. In this role, you will collaborate with NVIDIA's cross-functional rack-scale system SW/FW engineering teams with dedicated CSP-facing technical leadership. Your focus is on the system-level software that manages, monitors, and recovers the rack as a whole — fabric management, GPU/NVSwitch error handling and recovery, health telemetry APIs, firmware update orchestration, and SW-driven serviceability. You will drive work streams with CSP engineering teams to build shared understanding of the architecture, incorporate their operational feedback, and ensure integration readiness. What you'll be doing: Drive rack-scale SW/FW architecture alignment across CSP engagements — including fabric management software, link health monitoring, GPU/NVSwitch error handling, SW/FW serviceability features (e.g., hot-plug support, component isolation, firmware-driven recovery), and multi-component firmware orchestration Drive technical work streams with CSP engineering teams on rack-scale system software — ensuring they deeply understand fabric management, NVSwitch behavior, error handling and recovery policies, health telemetry APIs, and SW/FW-controlled recovery operation Capture and synthesize CSP engineering feedback on rack-scale system software — health monitoring APIs, SW-driven serviceability workflows, firmware update orchestration, and error recovery behavior — champion that feedback into NVIDIA's architecture decisions Collaborate with multi-functional teams to ensure customer operational requirements are reflected in system software and firmware development Identify cross-CSP patterns in rack-scale SW/FW iss
Jobs in United States
Principal Ai Engineer in United States
334 active opportunities · Updated October 2026
Showing
15 jobs
Explore current principal ai engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Mission: Building the Data Foundations for AI We are the Snowflake Interoperable Foundations organization - the foundational layer that powers Snowflake’s AI, Analytics and Data Engineering capabilities. We lead innovations across open table formats such as Apache Iceberg, helping customers build peta-byte scale multi-cloud data lakes on Snowflake. We deliver core Metadata capabilities that power Snowflake’s industry-leading performance, AI, governance and platform features. We are embarking on a 0->1 redesign of our core systems across Interoperable Foundations. While we already manage exabyte-scale data supporting Snowflake’s AI capabilities, the next frontier is providing the foundational data layer that accelerates agentic innovation in an open, multi-format data world, You will be setting the technical vision across our investments in metadata platforms, Apache Iceberg and AI-ready storage. Your Impact: From Redesign to Reality 0->1 Architectural Leadership: Lead the ground-up redesign of our core Metadata systems, influencing the transaction frameworks that power query, DML, and AI-driven data interactions in addition to extending our lead on platform capabilities such as Zero Copy Cloning and Cross-Region / Cross-Cloud Replication. Iceberg Innovation: Drive
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Principal ASIC Design Engineer About the Role At Micron, we transform how the world uses information to enrich life for all. As a global leader in memory and storage innovation, we develop technologies that accelerate intelligence and enable the next era of AI, ML, and advanced computing. Micron’s Interface Pathfinding team drives performance-scaling innovation across circuits, signaling, packaging, and interconnects with a 3–5 year technology horizon. A core part of that work is silicon-based validation of novel PHY solutions, and we are growing the team to execute. As the Principal ASIC Design Engineer , you will be the primary digital contributor on a deliberately small, senior team — united around the goal of carrying high-speed interface technologies from architecture to tape-out. The team’s analog and chip-level architecture is anchored by a deeply experienced analog custom design engineer; your role is to be the authoritative digital voice — owning RTL design and micro-architecture of the digital blocks, defining timing constraints, supporting verification, and serving as the key interface between the digital design and the analog and layout specialists who will carry the implementation through to silicon. On a test ASIC of this scope, the front-end digital work is the critical path, with contractor and layout support engaged as bandwidth demands warrant. What You’ll Own Digital Block Architecture & RTL Design Own the micro-architecture and RTL implement
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron's Global Supplier Quality organization is seeking a Controller Quality Principal Engineer to lead the quality strategy, qualification, and continuous improvement of storage and memory controller (ASIC/SoC) manufacturers supporting Micron's SSD, embedded, and storage solutions portfolio. This is a senior technical leadership role responsible for driving controller supplier quality performance from design qualification through mass production and field support The successful candidate will collaborate across functions with ASIC Development Team, Compose Engineering, Product Engineering, Dependability, Manufacturing, and Commodity Management, as well as directly with controller IC vendors, foundries, third party reliability labs and OSAT (outsourced assembly and test) partners, to ensure controller quality, reliability, and supply continuity meet Micron's standards. This role can be based in Taiwan, Hyderabad, or San Jose and will work extensively across time zones with global partners and suppliers! Develops, evaluates, revises, and applies technical quality assurance protocols/methods to inspect and test in-process raw materials, production equipment, and finished products. Ensures activities and items are in compliance with both company quality assurance standards and applicable government regulations. Performs analysis and identifies trends in the inspection of finished products, in-process materials and bulk raw materials, and recommends corrective actions when vital. Ensures that established manufacturing inspection, sampling and statisti
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. We are seeking a Principal Analog Design Engineer to lead the design, integration, and delivery of advanced analog and mixed-signal IPs for High Bandwidth Memory (HBM) products. This role is critical in developing and integrating high-performance analog subsystems within the HBM logic die, enabling industry-leading bandwidth, power efficiency, and reliability. As a principal engineer, you will provide deep technical leadership across analog design, IP integration, system alignment, and silicon execution, driving end-to-end success of HBM solutions. Responsibilities will include, but are not limited to: Lead the d esign and own critical HBM analog circuits, including: High ‑ speed transmitters and receivers Clock generation and distribution (PLLs, DLLs, CDRs) SerDes ‑ related analog blocks Biasing, reference, and calibration circuits
Job Details: Job Description: As the world's largest chip manufacturer, and a global leader in innovation and new technology, Intel strives to make every facet of semiconductor manufacturing state-of-the-art - from semiconductor process development and manufacturing, through yield improvement to packaging, final test and optimization, and world class Supply Chain and facilities support. Employees in the Technology Development and Manufacturing Group are part of a worldwide network of design, development, manufacturing, and assembly/test facilities, all focused on utilizing the power of Moore's Law to bring smart, connected devices to every person on Earth. We're constantly working on making a more connected and intelligent future, and we need your help. Change tomorrow. Start today. What we offer: We foster a collaborative, supportive, and exciting environment where the brightest minds in the world come together to achieve exceptional results. We give you opportunities to transform technology and create a better future, by delivering products that touch the lives of every person on earth. What we do: Join the Logic Technology Development (LTD) organization, a powerhouse of process technology innovation driving Intel's IDM 2.0 strategy to exceed the aspirations of our global customers. We are seeking a visionary Principal Engineer to lead the charge in Intel's foundry revolution, serving as a strategic technical architect and executive partner in guiding our most critical customers through the entire product development lifecycle. In this leadership capacity, you will orchestrate multidisciplinary strategies spanning device engineering, product integration, and yield analysis to steer customers from critical design tape-out and pre-silicon evaluation through to NPI silicon qualification. Furthermore, you will define t
What you’ll do Partner with medical image reconstruction scientists / engineers to build ML components that improve reconstruction quality, speed, robustness, or quantitative accuracy. Define training/evaluation pipelines, datasets, and metrics that map to user needs and design requirements. Productionize models: inference performance, reproducibility, monitoring for drift/regressions, and safe fallbacks. Collaborate on hybrid algorithms, incorporating physics and learned priors, denoisers, learned regularizers, and quality estimation. Help build tooling for rapid experimentation as well as rigorous verification of algorithm changes. What we’re looking for Strong applied ML experience plus comfort with signal processing / imaging or adjacent domains. Ability to move fluidly between research prototypes and production-quality systems. Strong evaluation discipline: metrics, ablations, data leakage avoidance, and reproducibility. A demonstrated track record of applying ML to physics-based or inverse problems (i.e., shipped projects, a portfolio, or publications.) Useful experience ML for imaging/inverse problems (or adjacent) with strong evaluation discipline and comfort with GPU performance constraints. Pragmatic production mindset: reproducible training/inference, regression testing, and safe deployment in high-stakes contexts. A background in computational physics or scientific computing. Leverage ML-based methods such as PiNNs and Neural Operators to solve partial differential equations arising in ultrasound simulation and imaging. Experience in Agentic-SciML is a plus. Hands-on experience with data curation for ML: building datasets from messy, real-world sources, defining ground truth, and managing labeling or simulation pipelines. Background in data assimilation: combining observations with physics-based models (Kalman filtering, variational methods, ensemble approaches, or learned variants).
NVIDIA DGX Cloud is an AI Factory designed to power the next generation of AI and industrial-scale breakthroughs. As a Principal Engineer for Security Architecture, within our Security Engineering organization, you will own a core security domain of the AI factory: the architecture, the paved road that delivers it, and much of the code underneath. You will hold the security design bar across DGX Cloud from inside the teams doing the building, and this is a founding seat on a new team. Security Engineering is a new organization at DGX Cloud, accountable for the security outcome of the platform, and Security Architecture is the function inside it that holds the design bar. Security here is fleet horizontal and stack vertical, so your work will cross every DGX Cloud engineering organization: you will embed with the teams building GPU clusters, control planes, and services, join their designs as a participant rather than an approver, and leave behind systems in which an entire class of risk is no longer possible. There is no architecture review board here and no approval queue. You are a senior IC with deep security domain knowledge, and the security bar holds because you helped set it and then helped ship it. What You Will Be Doing: Own a Security Domain End to End: Take architectural ownership of a core domain of DGX Cloud security, from the design through the system running in production. That could be tenant and GPU workload isolation, workload identity, infrastructure and network, supply-chain provenance, hardened baselines and patching, or deploy-time policy and admission control. Embed with the Teams Building It: Join the design early, write the code, and help land it. The posture is not "you did this wrong." It is "here are the considerations we need to meet, I will help, let's go to work." Build Paved Roads, Not
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta's Governance and Compliance platform is the operating layer enterprises trust to run their security programs. We are hiring an engineer who will own the architectural foundation that makes it work for their most complex organizational structures. Program Structure and Trust is a newly chartered group at Vanta with a focused mission: build the enterprise org model that lets customers bring their compliance and security structure — product lines, business units, isolated data environments, cross-cutting audits, scoped approvals, and data residency requirements — natively into the platform. The problems this team solves determine whether Vanta can serve the enterprise customers it's increasingly winning. This is the defining technical role of the group. The Principal Engineer owns the design, phased delivery, and long-term technical direction of Vanta's enterprise org model — a multi-quarter initiative that cuts across the platform and establishes the foundation for how enterprise customers structure, segment, and operate inside Vanta. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as a Principal Engineer at Vanta: Own the design and multi-quarter delivery of Vanta's enterprise org model, including hierarchical product lines and business units, isolated data access and ownership, cross-cutting audit workflows, scoped approvals, and EU and GovCloud data residency support Define and evolve the core abstractions that let the platform absorb structurally diverse, often conflicting enterprise requirements — solving for the general case rather than one-off customer accommodations R
What you’ll do Act as the technical lead for large parts of the scanner platform: system architecture, codebase structure, and long-term maintainability. Own core runtime foundations: distributed control, state management, fault handling, and reliability. Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. Build robust observability: logs, metrics, traces, and replayable diagnostics (with privacy constraints). Collaborate with hardware and recon/ML teams to define interfaces, data contracts, timing/synchronization, and failure modes. Lead complex refactors (e.g., message passing / RPC boundaries, modularization, concurrency model) without halting forward progress. What we’re looking for Deep software architecture experience for real-world systems: robotics, instrumentation, medical devices, or other complex distributed products. Strong Python and concurrency background (asyncio, multiprocessing, profiling, performance engineering). Track record of shipping systems that are observable, debuggable, and resilient. Strong technical leadership: clarity, pragmatic trade-offs, and mentoring. Useful experience Building but rock-solid systems: clear interfaces (gRPC/protobuf or equivalent), strong state modeling, and failure handling. High-leverage engineering habits on a lean team: good tests, CI, reproducible dev environments, and fast code review. Practical performance + concurrency work in Python (asyncio, profiling, multiprocessing) and comfort debugging distributed behavior. Security-minded device software: safe defaults, encrypted data paths, and disciplined handling of PII/PHI. Operational thinking: remote updates/management, excellent logging, and diagnostics that make real hardware debuggable.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the Team The Apps & Experiences (AppEx) organization powers the unified user experience (Snowsight) for Snowflake. Every major Snowflake customer interacts with our platform through these surfaces, making AppEx central to both customer satisfaction and revenue growth. Within Appex, our team sits at the intersection of AI, platform, and user experience. We are on a mission to create an AI‑first, scalable, maintainable, and extensible platform to build all Snowflake user experience features. The team owns: The serving layer that powers Snowsight, Snowflake CoWork, and UX features Developer tooling that enable our engineers to move fast and deliver Frontend infrastructure and shared services for product teams CI pipelines that ensure fast and reliable builds Testing and validation platform that maintains code and product quality The monitoring and logging stack that ensures observability and trust What You Will Do As a software engineer on the team, you will play a central role in delivering the next generation of tools and evolve our developer infrastructure and tooling to be scalable, highly performant, and AI-first. Design and build AI-first platforms, tools, and best practices to significantly enhance the developer experience for Snowsight Drive strategy and init
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a Principal Engineer II to architect the core data processing engine of the Snowflake Data & AI Cloud. At Snowflake, we believe that high-performance, unified compute fabrics are the indispensable building blocks for Agentic AI. Autonomous agents require more than just models; they require a high-fidelity, low-latency state layer to reason, act, and persist context. This role is not about building traditional data processing pipelines or legacy ETL/ELT workflows; it is about building the core distributed systems and atomic primitives that make those agentic workflows possible. In this role, you will be a lead architect of the Snowflake Data Transformation Engine. You will design and implement the fundamental transformations infrastructure—Stateful Stream Processing Engines, Incremental View Maintenance Engine, Materialization Internals, and the Distributed Orchestration Fabric. Our solid foundation supporting the seamless transition for enterprises between batch and streaming through Dynamic Tables, Streams & Tasks, and DBT Projects is the starting point. Your architectural work will extend the reach of the core engine to accelerate and support the massive scale of the Snowpark and Spark ecosystems. You are building the systems that allow both data eng
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are the Snowflake Metadata team. We own Snowflake’s metadata systems that make it easy for customers to query, modify and manage their petabyte-scale data. We develop distributed systems that store and maintain metadata, transaction frameworks that power Snowflake’s query and DML capabilities, asynchronous systems that provide time travel and lifecycle management capabilities and entity metadata supporting DDL capabilities. We also build foundational capabilities that deliver global features like cross-region replication (Snowgrid), data sharing, and data marketplace. AS A PRINCIPAL SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real business needs at large scale by applying your software engineering and analytical problem solving skills. Design, develop and support fault-tolerant scalable distributed systems for our Snowgrid and Data Sharing teams. Create architecture and design, influence our product roadmap, and take ownership and responsibility over new projects. Analyze fault-tolerance and high availability issues, performance and scale challenges, and solve them. Mentor and grow junior engineers. Understand trade-offs between consistency, performance and costs to build solutions which can meet the demands of rapidly growing services. Ensure operational readiness of
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron is seeking a highly motivated and experienced Technical Staff Member to join our Quality Engineering team in Boise, ID. Your role will be crucial in ensuring the reliability and quality of wafers shipped from our Boise manufacturing site! Summary The Product Quality and Reliability team ensures Micron delivers high-quality, reliable semiconductor solutions that meet customer expectations and business objectives. The team partners across product engineering, design, technology development, manufacturing, and reliability organizations to identify risks, solve complex technical challenges, and drive continuous improvements in product performance, yield, and customer satisfaction! Position Overview As a Senior Product Quality and Reliability Engineer, you will be a technical leader. You will drive product excellence in quality, dependability, and yield across advanced semiconductor technologies. You will lead complex technical investigations, influence product and technology decisions, and develop innovative approaches that strengthen quality systems and business outcomes. This role provides an opportunity to have broad interpersonal impact through technical leadership, multi-functional teamwork, and strategic problem-solving. Responsibilities Lead complex investigations involving yield excursions, product deviations, and quality issues, driving root cause identification, containment actions, and sustainable corrective solutions Define product quality and reliability moni
From $200K/yr
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. About the role You will work on core enterprise platform systems, focusing on services that ensure Synthesia is secure, reliable, and scalable for our largest customers. You will contribute to our new suite of APIs that will unlock our market leading Avatar technology to be utilised by external creative tools. We are an AI native company and use the most powerful assistant tools on a daily basis to increase our speed of development, automate repetitive tasks and widen the scope of what the team can worked on. This includes Claude and Cursor. You will have ownership of projects that span months and multiple teams, requiring you to break down complex, ambiguous problems into clear steps that can be delivered and validated iteratively. Engineers within Synthesia are empowered to contribute heavily to product discussions and build with both user and business considerations in mind. Impact is key to everything we do here. You will work closely with product, security, legal, and infrastructure partners, and will be expected to translate business and regulatory requirements into scalable technical solutions. You will evaluate your work through system health and reliability metrics, leveraging observability and monitori
Other cities to consider
More places hiring for this role
Get new principal ai engineer jobs in United States by email
Daily job updates · Unsubscribe anytime