Define the physical limits of what a chip can do, then break them! NVIDIA's Silicon Co-Design Group sits at the crossroads of architecture, silicon, systems, and manufacturing, where first principles thinking translates directly into product outcomes at scale. As GPU power density and thermal limits approach physical boundaries, the gap between system architecture needs and manufacturable, testable capabilities widens with every generation, and this role exists to close it. You wouldn't be maintaining existing solutions, you'd be identifying where physical constraints will erode NVIDIA's competitive edge before anyone else sees it coming, then building the cross-domain responses that ensure the next generation leapfrogs those limits entirely. The problems here don't have known answers yet: Every generation of NVIDIA silicon pushes closer to the limits of physics. The Co-Design Architect role isn't about optimizing within known bounds; it's about redrawing those bounds entirely. If you want your work to shape what's physically possible in the world's most demanding computing products, this is where that work happens. What you'll be doing: Power, Thermal & Packaging Limits Analysis: Gather and analyze the hard constraints of packaging, power delivery, and thermal dissipation against roadmap demands, predicting when those constraints will reduce competitiveness and acting before they do. Cross-Domain Solution Development: Develop solutions that span silicon features, packaging innovation, firmware and hardware co-design, and DFX updates, ensuring future products are not only performant but testable to the highest quality and reliability standards. Packaging Technology Evaluation: Assess emerging packaging technologies for their voltage/frequency, noise, reliability, and testability implications, determining which are worth adopting an
Jobs in United States
Ai Architect in United States
5,082 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
NVIDIA is seeking a world-class computer architect to contribute to the development of future high-performance computing systems, with a focus on enhancing the power-constrained performance of the hardware. Ideal candidates will have a strong track record of understanding and analyzing memory systems architecture to improve performance per watt (perf/W) and performance per millimeter (perf/mm). A broad perspective across the field of computer architecture and depth in the area of power, performance, and area (PPA) analysis is highly desirable. NVIDIA has pioneered programmable GPUs and the CUDA language and is a world leader in high-performance computing technology, with aggressive plans for future processors. This position offers the opportunity to have a real impact in a fast-moving, technology-focused company. What you will be doing: Develop innovative high-performance processor and system architectures, focusing on the memory system and energy efficiency. Develop architecture and micro-architecture features to improve the state-of-the-art in GPU memory systems, optimizing along the axes of perf/W, perf/mm, and perf/$. Develop and enhance architecture prototype models for power and noise analysis. Participate in performance and power simulation of features to analyze, define, and improve energy per byte. Analyze benchmarks, application workloads, and performance/power simulation and emulation results to identify areas for architecture optimizations. Debug power, performance, and functional issues with high-level models, RTL simulation and emulation, silicon, and systems. Collaborate with outside partners on system infrastructure. What we want to see: 10+ yrs of experience in CPU/GPU architecture, memory systems design with a focus on energy efficiency in the system. Bachelor
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. NVIDIA has a rapidly expanding ecosystem of data center platform & node designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each bringing together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We're searching for a highly motivated, technical leader to drive the engineering roadmap and innovation for our rack system software architecture. From firmware, kernel drivers, operating systems, networking, fabrics and associated user mode drivers + manageability software. You will work with component leads internally and engage with industry leading hyperscalar / cloud service providers on taking these products to market. What you’ll be doing: Drive the software end-to-end architecture for NVIDIA's rack-scale products Maintain deep understanding of the product portfolio and roadmap; translate forward-looking plans into clear, formal software requirements that anchor execution across the organization. Ensure high quality & reliable software; serving as a trusted architectural partner to teams requiring
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal is hiring a high-impact Solutions Architect to drive technical strategy across our most strategic enterprise accounts. You will operate as the executive technical counterpart to Enterprise Account Executives, leading complex evaluations, shaping infrastructure modernization roadmaps, and driving multi-product adoption across AI/ML workloads. This role is not demo support. It is a strategic, consultative position requiring strong architectural depth, executive presence, and the ability to influence 7–8 figure infrastructure decisions. You will work directly with CTOs, VPs of Engineering, and ML platform leaders to help them rethink how AI infrastructure should be built and operated. If you thrive in high-velocity technical sales environments and want to shape the infrastructure layer powering modern AI companies, this role is for you. What You’ll Do: Own the t
NVIDIA builds the silicon behind AI, accelerated computing, and graphics. Every watt of performance and every degree of thermal headroom traces back to decisions made in power, performance, and thermal architecture. We are the Silicon Co-Design Group (SCG). We identify, own, and drive system-level co-design ideas. We start with initial concepts and advance to product differentiation across NVIDIA's roadmap. We are hiring a Principal System Power Management and Performance Architect who operates at the ambiguous boundary where workload behavior, silicon capabilities, firmware policies, and platform constraints collide, and who turns that ambiguity into architecture that survives across multiple silicon generations. SCG scope spans architecture, design, software, operations, platforms, and productization. This role shapes system, platform, and data center features and behavior, and partners with teams across NVIDIA. What You'll Be Doing: The work here is rarely well-defined when it arrives. You will be given problems that appear to be performance gaps or power anomalies and encouraged to build a framework for solving them, not just tackle a single instance. Define the multi-generation roadmap for system-level power and performance features, grounded in prototyping, use-case analysis, and cost/benefit trade-offs across segments. You will decide what problems are worth solving and why. Own the architecture and integration strategy for HSIO power management, DVFS, P-states, and low-power features. Your decisions improve product performance, power, and reliability across product lines — not just the current program. Lead system-level boot and IST architecture defining how power and clock domains initialize, sequence, and recover across complex multi-IP systems where the interaction space is large and the failure modes matter. Drive power management strategy at data
£205K – £245K/yr
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? In this role as a Solutions Architect at Cohere, you will play a significant role in growing Cohere’s Public Sector business and have a great deal of autonomy when it comes to technical pre-sales and post-sales. In order to qualify for this exciting career opportunity, you must have a Security Clearance. Responsibilities include developing a deep understanding of customer problems, mapping them to Cohere solutions, and working closely with our partners as the trusted technical advisor who owns the technical relationship with our stakeholders. By leveraging your expertise, you will help to increase the adoption of Cohere products, both internally and externally, and will gather valuable insights and feedback from customers to help shape the future of our products. In this dynamic role, you will need to be both a strategic thinker and a hands-on doer. You will take a hands-on approach to building customer Proof of Concepts that showcase the business value of our platform. As the technical relationship owner, you will collaborate with stakeholders to understand their business objectives and translate those into techn
$110K – $165K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role Notion is reimagining how customers adopt, expand, and realize value from software in an AI-native world. As an Outcomes Architect, you’ll be the build-beside partner who helps customers turn AI ambition into working systems inside Notion. You’ll own workflow adoption, value delivery, and renewal outcomes for a book of named accounts, with a focus on customers that are flat, at-risk, or need a stronger path to measurable impact. You’ll partner closely with Solutions Consultants on account strategy, expansion qualification, and planning, while directly shaping outcomes that strengthen customer partnerships and contribute to predictable revenue. If you’re someone who loves helping customers build workflows that solve real business problems—and you want to work at the intersection of AI, productivity, customer value, and revenue impact—this is the role. What You'll Achieve Drive workflow adoption and value realization for a book of named accounts ensuring your customers know how to deliver outcomes with Notion Contribute directly to NRR with predictable renewal and expansion pipeline and proactive product adoption for you
$180K – $240K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role Notion is reimagining how customers adopt, expand, and realize value from software in an AI-native world. As an Outcomes Architect, you’ll be the build-beside partner who helps customers turn AI ambition into working systems inside Notion. You’ll own workflow adoption, value delivery, and renewal outcomes for a book of named accounts, with a focus on customers that are flat, at-risk, or need a stronger path to measurable impact. You’ll partner closely with Solutions Consultants on account strategy, expansion qualification, and planning, while directly shaping outcomes that strengthen customer partnerships and contribute to predictable revenue. If you’re someone who loves helping customers build workflows that solve real business problems—and you want to work at the intersection of AI, productivity, customer value, and revenue impact—this is the role. What You'll Achieve Drive workflow adoption and value realization for a book of named accounts ensuring your customers know how to deliver outcomes with Notion Contribute directly to NRR with predictable renewal and expansion pipeline and proactive product adoption for you
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. NVIDIA has a rapidly expanding ecosystem of data center platform & node designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each bringing together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We're searching for a highly technical, motivated manager to lead & manage the team responsible for rack-scale system software architecture. From firmware, kernel drivers, operating systems, networking, fabrics and associated user mode drivers + manageability software. You will work with component leads internally and engage with industry leading hyperscalar / cloud service providers on taking these products to market. What you’ll be doing: Drive the software end-to-end architecture for NVIDIA's rack-scale products Maintain deep understanding of the product portfolio and roadmap; translate forward-looking plans into clear, formal software requirements that anchor execution across the organization. Ensure high quality & reliable
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: As a Solution Architect at Baseten you will partner closely with Sales and customers to translate business needs into technical solutions, run technical discovery, and guide repeatable deployments and proofs of value for customers. This role is a great fit for entrepreneurial, customer-facing technical professionals who want a front-row view into how modern companies adopt AI at scale, and who enjoy working across technical discovery, solution design, demos, deployment scoping, and hands-on customer implementations, in close partnership with Sales and Engineering. RESPONSIBILITIES: Partner with Sales on customer discovery calls (most often second calls, occasionally first calls for large accounts). Lead demos and technical scoping to align on success criteria, architecture, and deployment approach. Own benchmarking and repeatable deployments , including: Handling standard deployment patterns and configurations across many modalities – LLMs, embeddings, image and video generation, VoiceAI, etc. Advising on tradeoffs like H100s vs B200s and latency-optimized vs throughput-optimized setups. Driving consistent “playbook” style deployments for common models and use cases. Become a power user of different runtimes such as vllm, sglang, and TRT-LLM and all the common configurations and tradeoffs between them Drive POC and project execution , including: Scoping POCs and keeping stakeholders aligned on timeline, deliv
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking an experienced SoC Architect to lead the definition and development of next-generation custom AI silicon for edge deployments. This role will be responsible for shaping the architecture of highly efficient, high-performance SoCs optimized for machine learning inference and on-device intelligence. You will work cross-functionally with internal engineering teams and external ecosystem partners to translate product requirements into scalable silicon solutions, driving execution from concept through delivery. In this role you will: Define the architecture and technical roadmap for custom SoCs targeted for edge applications. Drive system-level tradeoff analysis across compute, memory, interconnect, power, thermal, and cost constraints. Architect energy-efficient ML compute subsystems optimized for inference workloads and real-world deployment environments. Collaborate with internal hardware, software, systems, and product teams to align architecture with platform needs. Partner with external silicon vendors, IP providers, and manufacturing partners to execute development plans. Lead hardware/software co-design efforts to maximize performance per watt and end-to-end system efficiency. Guide implementation teams through microarchitecture, RTL development, validation, and bring-up phases. Operate effectively in agile development environments and help teams deliver against aggressive schedules and milestones. You might thrive in this role if: Proven exper
$136K – $160K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Own CX enablement knowledge strategy + systems globally. Notion ships quickly, and our customer experience depends on support teams being able to find accurate answers fast. This role exists to keep our internal knowledge base fresh as it scales—by designing durable information architecture and governance that keeps pace with frequent product change. What’s different at Notion: you’ll use Notion as customer zero, building a living system of record structured for both humans and AI—so the right people (and the right agents) can retrieve the right information at the right time. You’ll own the internal knowledge base end-to-end (content + page structure), partnering with CX Enablement, Product Ops, QA, and Vendor Ops to keep it reliable and effective. What You'll Achieve: Own end-to-end CX knowledge as a system: Make it dramatically faster for CX teams to retrieve trustworthy, up-to-date knowledge, improving support quality and efficiency as Notion grows. Design and evolve information architecture (IA): Define how knowledge is structured, governed, maintained, and measured across the full lifecycle (intake → draft → revi
Citi is evolving its current commerce offerings and branching into new avenues of growth by offering merchants access to high value cardmembers across Citi’s digital properties and partner placements. By combining Citi’s scale, trusted customer relationships, and real purchase insights, the business delivers advertising that is relevant to customers and drives measurable, incremental impact for advertisers. If you're early in your career and have been working within AdTech, MarTech, or data environments and are ready to deepen your technical expertise while helping build a platform from the ground up, this is the role. The Junior Solution Architect will support the design and implementation of Citi Commerce’s AdTech and data infrastructure by translating defined architecture into scalable, working technical components. This is a highly hands-on, detail-oriented role focused on execution—ensuring platforms are properly integrated, configured, and optimized to support campaign activation, ad serving, and measurement. This role is ideal for an early-career technologist with AdTech or MarTech exposure who enjoys working across systems, troubleshooting integrations, and enabling seamless data flow across platforms. Responsibilities: Support implementation of Citi Commerce’s end-to-end AdTech and data architecture across platforms Assist with integrations across ad servers, DSPs, CDPs, and measurement platforms (e.g., GAM, DV360, LiveRamp) Configure and maintain tagging, tracking, and pixels to ensure accurate campaign execution and reporting Help build and manage data pipelines, ensuring reliable data ingestion and flow across systems Troubleshoot technical issues related to integrations, tracking discrepancies, and data inconsistencies Assist in QA and validation of platform integrations, tagging, and measurement frameworks Document system c
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Every company grows differently, and at WRITER, our growth is directly tied to empowering our users to create better content, faster, and at an unprecedented scale. As a strategic solutions architect, you'll be at the forefront of this mission, working directly with our largest and most strategic prospects to identify, validate, and build innovative agentic AI solutions that unlock massive business value. This isn't just about selling a product; it's about deeply understanding complex enterprise needs and architecting a future where AI transforms how our customers operate. You'll be instrumental in shaping how the world's leading companies adopt and scale AI, making a tangible impact on both their success and WRITER’s continued leadership in the enterprise generative AI space. This is a full-time, hybrid role based out of our Chicago hub. You will report directly to the Director, solutions architecture. 🦸🏻♀️ What you'll do Drive strategic technical discovery
$100K – $500K/yr
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is building next-generation processors and systems, bringing together world-class expertise across silicon, systems, and software. We are looking for a CPU Performance Modeling Architect to help evaluate, shape, and optimize the performance of future CPU architectures. In this role, you’ll use performance modeling, workload analysis, and deep understanding of CPU architecture to answer complex questions about how a processor should be designed. You’ll work closely with CPU architects, RTL designers, software and compiler teams, and system engineers to identify performance opportunities, evaluate architectural tradeoffs, and turn modeling insights into actionable design decisions. This role is hybrid, based out of Santa Clara, CA or Austin, TX. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are A CPU architect, performance architect, or performance modeling engineer with experience influencing CPU architecture or microarchitecture decisions. You have a strong understanding of modern processor architecture and enjoy digging into why a CPU performs the way it does. You are comfortable combining hardware architecture, software, data, and
Other cities to consider
More places hiring for this role
Get new ai architect jobs in United States by email
Daily job updates · Unsubscribe anytime