Jobs in United States

Ai Architect in United States

5,082 active opportunities · Updated October 2026

Explore current ai architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing. NVIDIA is a “learning machine” that constantly evolves by adapting to new opportunities that are hard to solve, that only we can address, and that matter to the world. This is our life’s work, to amplify human creativity and intelligence. Make the choice to join us today. Design-for-Test Engineering at NVIDIA works on groundbreaking innovations involving crafting creative solutions for DFT architecture, verification and post-silicon validation on some of the industry's most complex semiconductor chips. What you'll be doing: As a senior member in our team, you will work with pre-silicon and post-silicon data analytics - visualization, insights and modeling. Design and uphold sturdy data pipelines and ETL processes for the ingestion and processing of DFX Engineering data from various origins Lead engineering efforts by collaborating with cross-functional teams (execution, analytics, data science, product) to define data requirements and ensure data quality and consistency You will work on hard-to-solve problems in the Design For Test space which will involve application of algorithm design, using statistical tools to analyze and interpret complex datasets and explorations using Applied AI methods. In addition, you will help develop and deploy DFT methodologies for our next generation products using Gen AI solutions. You will also help mentor junior engineers on test designs and trade-offs including cost and quality. What we need to see: BSEE (or equivalent experience) with 5+, MSEE with 3+, or PhD wi

PythonSQLAWSAzure
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's engineers want to work in an AI-first way. What's missing isn't enthusiasm — it's the platform underneath it. Today everyone assembles their own agent config, context files, and MCP servers, so the good patterns stay trapped in individual setups instead of becoming defaults everyone inherits. You'll build that platform: the agent configurations tuned to our monorepo, the context and tooling layer that makes agents competent in our codebase, the evals that tell us which approaches actually work, and the rollout mechanics that get a new engineer productive with agents in week one. You are not here to mandate how engineers use AI — you're here to make the good path the easy path. Success looks like teams adopting what you build because it beats what they'd cobble together themselves, not because a policy requires it. Platform engineer, not AI evangelist. Ship infrastructure, measure it, kill what doesn't work, let adoption be the referee. The playbook for AI-first SDLC doesn't exist at any company yet. You'll write ours. WHAT YOU'LL BUILD Agent substrate — Repo-level context infrastructure that makes agents competent in our codebase ( CLAUDE.md/AGENTS.md conventions, architecture and domain context, and the tooling to keep it accurate as code moves). Internal MCP servers giving agents scoped access to CI, observability, incident tooling, deployment state, and docs. Shared skills, subagents, and hooks th

PythonDockerKubernetesCI/CD
M
📍 New York City, New York, United States· Full-time
✓ Quality checkedCompany trend -100%

What we're building Mutiny is the self-improving AI infrastructure for GTM teams to execute faster and close more revenue. Our ambition is to do for revenue velocity what Cursor and Claude Code did for engineering velocity. With Mutiny, everyone in sales and marketing gets a bench of GTM athletes that handle any work across their revenue motion and learn from what's actually moved their deals. In April we re-launched the product as an agent-first platform. Anthropic showcased us as a leader in AI GTM. MRR is growing more than 70% month-over-month, with customers like Uber, Rippling, and Snowflake. We're backed by Sequoia, YC, and Insight, and we're building a generational company. The opportunity Most engineers spend their career making predictable systems faster. You'll spend yours making non-deterministic ones trustworthy. As a senior engineer on our AI product team, you'll architect the Campaign Builder and Agent experiences marketers and sellers open every day to go from idea to personalized assets in minutes. You'll partner directly with product, design, and the founders to define what an agent-first GTM platform should feel like, and your calls on architecture, evals, and guardrails compound across thousands of customer accounts. This role is in person in New York City, five days a week, and we ship weekly. What you'll own The core agent surfaces. Architect and ship the Campaign Builder and Agent experiences end-to-end. Frontend, backend, prompts, evals, the whole stack. Reliability on top of LLMs. Make non-deterministic models feel deterministic at the surface. Build the retries, fallbacks, and orchestration so the customer never sees the failure mode. Evals and guardrails. Define how we measure quality, catch regressions, and keep brand and tone consistent across thousands of customer accounts. Speed and feel. AI products live or die by latency and the loop between intent and output. You'll obsess over both, and use coding agents and agent networks to ship f

TypeScriptPythonAIKotlin
V
📍 United States· Full-time
✓ Quality checkedCompany trend -90%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. As the People Automation & AI Partner at Vanta, you'll be the hands-on builder embedded in our Recruiting and People teams, designing, shipping, and scaling the workflows, agents, and automations that remove manual work across recruiting and People operations so our team can focus on the work that moves the business. Recruiting is the front door to that talent, and the speed, quality, and consistency of our hiring engine is a top priority for this role. The People team at Vanta is building the infrastructure to hire, develop, and retain world-class talent at scale. We're not just running processes, we're rethinking them. That means bringing in the tooling, AI, and automation capabilities to match the speed and ambition of the company we're building. What you’ll do as a People Automation & AI Partner at Vanta: Design, build, and maintain AI-powered workflows, automations, apps, and agents (Claude, Claude Code, Dust, Zapier) that eliminate manual lift across Recruiting and People operations Own Dust agent architecture for the People team: build and maintain agents, skills, and connectors that support recruiters, coordinators, hiring managers, and People partners across pods Take team-generated concepts and one-off ideas and turn them into durable, documented, adopted production workflows Build and maintain automation across the recruiting funnel, sourcing outreach, interview scheduling and coordination, candidate communication, and offer-to-hire handoffs, as a first-class use case for this role Serve as a technical build partner for Ashby-driven workflows, treating the ATS as a primary automation surface alongside Dust an

RestAIRust
M
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

ABOUT THE TEAM The AI Foundations Team at Mural is pioneering how generative AI transforms visual collaboration and decision-making. We’re a remote-first group of engineers, designers, and product thinkers focused on helping teams work together more effectively. Our goal isn’t to replace human creativity. It’s to amplify it, building AI that enhances how people align, communicate, and make decisions visually. YOUR MISSION You will design and build the core AI systems and platforms that enable Mural’s next wave of agentic, AI-driven collaboration experiences. Rather than building isolated AI features, you’ll work on the core backend systems that power Mural’s agent platform, including agent orchestration, durable execution, contextual memory, tool integration, observability, and evaluation. Your work will enable intelligent agents to reason over product context, act on behalf of users, and operate reliably and safely at scale. Our stack at Mural includes Azure OpenAI, React, Node, MongoDB. WHAT YOU'LL DO Build the core backend systems that power Mural’s agent platform, including orchestration, durable execution, tool execution, memory, observability, and evaluation infrastructure Design scalable services and APIs that allow AI agents to retrieve context, coordinate multi-step workflows, interact with Mural data, and act reliably on behalf of users Develop the agent memory layer, including systems for conversation context, product context, retrieval, summarization, compaction, and long-term context management Create infrastructure to monitor, debug, and improve agent behavior through traces, metrics, feedback loops, and offline evaluation Translate complex, open-ended product needs into clear backend architectures, service boundaries, data models, and implementation plans that align technical capabilities with user value Help define the technical direction for agentic AI at Mural, contributing to long-term architecture and strategy Champion engineering excellence, men

ReactMongoDBAzureAI
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Cortex CoWork team is defining the future of AI for enterprise data. Our mission is to transform how the world’s largest enterprises interact with their data through flagship products like Snowflake (CoWork) Intelligence . As a Principal AI Engineer , you will be a technical North Star for our AI initiatives. You won't just execute on a roadmap; you will help define it. You will tackle the most complex, "frontier" problems in agentic reasoning, NL-to-SQL, and enterprise-scale RAG, ensuring our AI products are not only innovative but fundamentally reliable and scalable for the Fortune 500. What you will do in this role: Technical Strategy & Architecture: Define the long-term technical vision for Snowflake Intelligence. Lead the architectural design of multi-agent systems, complex tool-use frameworks, and self-correcting NL-to-SQL engines. Drive Industry-Leading Reliability: Move beyond simple evals to build world-class, automated "hill-climbing" infrastructure. You will establish the methodology for how Snowflake measures and guarantees LLM performance across diverse customer schemas. Cross-Functional Influence: Partner with Product and Engineering leadership to align AI capabilities with business goals. You will bridge the gap between Research (modeling) and Product

SQLAIGoRust
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $399.4K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. About the role: Roblox Studio is the creation engine behind millions of immersive 3D games built by creators around the world. We are entering a major platform transition, evolving Studio into an AI-Native IDE where intelligent systems plan, act, validate, and iterate alongside creators to amplify their output. We are seeking a Technical Director, Applied AI to own the technical direction of the agentic AI systems embedded deeply into Roblox Studio. This is a hands-on individual-contributor role operating at the intersection of platform engineering, AI systems, and developer experience, where you will set the technical direction, design the architecture, and write the code. You will: Design and build the agentic AI systems at the core of the AI-Native IDE (planning, execution, validation, evaluation) so they are production-grade and reliable. Decide how Studio uses coding models, agents, retrieval, tool calling, and context management across the Assistant and the broader creator workflow. Own the quality bar and evaluation strategy , standing up quantitative and qualitative eval pipelines, including human-in-the-loop, so we ship with confidence. Optimize these systems for latency, throughpu

AWSGitAIGo
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Cortex Apps team is building the future of AI for enterprise data. This role focuses on the backend infrastructure that powers our flagship products like Snowflake Intelligence , Cortex Agents and Search making agentic AI fast, reliable, scalable and secure at the enterprise level. You won’t just be using AI tools; you will be building the high-performance systems that orchestrate them. You’ll own and influence the architecture for agent execution environments, high-throughput context retrieval, or the ecosystem that allows our customers to iterate and launch agents in production. What you will do in this role: Architect Agentic Runtimes: Build and scale the orchestration engines that execute complex agentic workflows, ensuring low-latency tool execution and robust state management. Scale Context Engineering Infra: Design high-performance systems for RAG (Retrieval-Augmented Generation), including vector database integration, scalable and efficient search indexing, query processing, and result ranking, semantic caching, and automated metadata extraction. Build the "Evals Engine": Develop the automated infrastructure required to run massive-scale golden set simulations, error analysis pipelines, and "hillclimbing" experiments. Productionize AI Workflows: Collaborate with

PythonJavaSQLKubernetes
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Snowflake Cortex team is on a mission to bring the transformative power of generative AI and machine learning to enterprise customers, seamlessly and securely within the Snowflake Data Cloud. We are seeking an entrepreneurial and customer-obsessed Product Manager to lead the product vision and execution for our AI Solutions and Services. In this role, you will own the end-to-end product lifecycle for some of our most exciting AI services. You will be responsible for understanding customer needs, defining the product roadmap, and working with a world-class engineering team to deliver innovative solutions that make it easy for any user to leverage AI. You will sit at the critical intersection of customers, our forward-deployed engineering team, core engineering, and GTM strategy, driving the future of AI within the Data Cloud. If you are obsessed with building products that are both powerful and simple, and thrive on turning ambiguity into impact, this is the role for you. AS A SENIOR APPLIED AI PRODUCT MANAGER AT SNOWFLAKE, YOU WILL: Own the Product Vision & Roadmap: Define and articulate a clear, compelling strategy for large-scale AI solutions. You will architect and own the end-to-end product lifecycle, from deep discovery and detailed requirements to hands-on dev

Machine LearningAIGoRust
S
📍 Bellevue, Washington, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Data Clean Rooms team is Leading the market shift from traditional 2-party data sharing to multi-party collaboration hubs . Our vision is to provide a seamless, "safe-room" environment where enterprises can collaborate on shared datasets while maintaining absolute governance. We ensure that no party can exfiltrate another's underlying content, even while running complex joint workloads and getting high-value results. You will join a fast-paced, collaborative team of engineers on a journey to provide customers with an integrated set of innovative, AI-enabled capabilities to analyze data in a privacy-preserving way. You will have a real opportunity to impact and shape the future of secure data collaboration at Snowflake. AS A SOFTWARE ENGINEER IN DATA CLEAN ROOMS, YOU WILL: Architect and build highly scalable infrastructure that enables secure, multi-party collaboration. Design and implement core clean room features and services, intelligent agents, and robust developer APIs to expand platform capabilities and support custom AI/ML workflows. Partner closely with Product Management and cross-functional teams to drive complex projects from ideation and system design through to production deployment. Mentor peers and foster a warm, supportive culture of innovation, cross-tea

PythonJavaAIGo
S
📍 Bellevue, Washington, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a Staff Research Scientist, Physical AI for our AI Research team . You will build the next-generation training and learning platform for physical AI: models that perceive, reason about, and act within structured environments . This is a greenfield (0 to 1) effort at the intersection of representation learning, world models, and policy optimization. You will help define its technical direction from day one. AS A STAFF RESEARCH SCIENTIST YOU WILL: Design and build scalable training infrastructure for representation models (e.g., contrastive and self-supervised approaches like CLIP/SigLIP, DINO/MAE, and joint-embedding predictive architectures) Develop latent world models that learn environment dynamics through imagined rollouts, enabling model-based reasoning and planning (Dreamer-style, I-JEPA/V-JEPA families) Architect and implement action/policy model pipelines, including vision-language-action models and diffusion-based policy learning Build generative simulator frameworks that produce controllable, physically plausible future states (video world models in the spirit of Cosmos/Genie/Sora) Develop multimodal generative model capabilities that fuse visual, language, and structured inputs for downstream reasoning and decision-making Lead cross-team technical de

Machine LearningAIGoRust
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the AI Products team The AI Products team is part of the broader Marketplace & Collaboration organization and is focused on bringing AI products on top of Snowflake’s data and application platform to help customers discover, share, monetize, and act on data assets & applications more easily. This team is building the connective tissue of the agentic enterprise: the infrastructure and product surfaces that allow Snowflake customers to seamlessly share datasets, semantic views, and applications, and make them discoverable and executable through Cortex Code, CoWork, and other agentic harnesses. Our strategy is centered on evolving Snowflake Marketplace for the AI era, including packaging data and intelligence into ready-to-use agentic experiences, and enabling governed access patterns that let AI systems safely operate on enterprise data and applications. As a Staff Software Engineer on AI Products, you will Lead the design and delivery of large, complex initiatives spanning multiple teams, turning ambiguous product and platform opportunities into durable technical solutions. Shape the architecture for how datasets, applications, semantic assets, and agentic capabilities are shared, discovered, governed, and invoked across Snowflake surfaces and third-party agent

PythonJavaAIGo
H
📍 New York, NY, United States
✓ High-confidence listingCompany trend +310%
Quick readStrong listing-quality and freshness signals

Become a part of our caring community Help shape practical, responsible AI solutions that improve healthcare experiences and outcomes. Humana’s Enterprise AI organization develops safe, scalable AI solutions across our Insurance and CenterWell businesses. We bring together product managers, data scientists, engineers, policy experts, and business leaders to apply emerging technology to meaningful healthcare challenges. As Associate Director of Applied AI, you will lead teams that design, build, and deploy enterprise AI solutions, with a focus on generative AI and intelligent agents. You will connect technical strategy to business needs, guide responsible delivery in a regulated environment, and help teams turn promising ideas into measurable outcomes for members, patients, and associates. Key Responsibilities Lead and mentor teams developing production-ready AI solutions that improve healthcare delivery, member experiences, and business operations. Help define and execute the roadmap for applied AI initiatives in alignment with enterprise priorities and business needs. Guide the evaluation and adoption of machine learning, generative AI, large language models, multimodal models, and intelligent agent technologies. Oversee scalable APIs, frameworks, data pipelines, retrieval-augmented generation solutions, and agent orchestration capabilities. Partner with product, data science, engineering, architecture, security, and business teams to translate requirements into reliable solutions. Establish standards for AI evaluation, obse

PythonDockerKubernetesMachine Learning
C
📍 Tampa Florida United States, United States
✓ High-confidence listingCompany trend +800%
Quick readStrong listing-quality and freshness signals

At Citi Services - Global Trade and Working Capital Solutions (TWCS) Technology Organization, we are on a mission to harness the power of data to drive innovation, create exceptional customer experiences, and solve complex business challenges. Our data team is at the heart of this mission, building the scalable and resilient infrastructure that turns data into our most asset. We are a passionate, collaborative group dedicated to pushing the boundaries of what's possible. The Opportunity We are seeking an Engineering Director to join our development team. The ideal candidate is a seasoned technologist with extensive experience in building and delivering data & AI solutions for business functions. This individual will be directly and fully accountable for solution delivery and should have a history of creating strategic technology architecture roadmaps aligned with business outcomes. The successful candidate will define and execute the technology roadmap for the Data & AI portfolio of TWCS, providing strategic direction and critical input into technology decisions to ensure a scalable buildout. This role involves building strong relationships with senior business and technology partners, driving agile execution, and leading a team of expert engineers to deliver with velocity and quality. Applicants should demonstrate exceptional technical acumen, a strong data engineering background, and a proven ability to lead and provide technical direction to high-performing engineering teams. A proven expertise in Data Engineering, Data Analytics, and AI Engineering delivery at scale is essential. Responsibilities: Manage/develop multiple teams of professionals to accomplish established goals and conduct personnel duties for team (e.g. performance evaluations, hiring and disciplinary actions) as well as ensure team adheres to best practices and processes Develop vision for team aroun

PythonJavaAWSAzure
N
📍 Santa Clara, United States
✓ High-confidence listingCompany trend -8%
Quick readStrong listing-quality and freshness signals

We are seeking a mission-driven Developer Relations Manager focused on Foundational AI Research to engage leading academic labs advancing the next generation of AI models, systems, and methods. In this role, you will work directly with top researchers building frontier AI systems, including large language models, multimodal models, reasoning systems, training methods, inference systems, model serving, and scalable AI infrastructure. You will help researchers adopt NVIDIA’s AI and accelerated computing platforms to push the boundaries of model performance, efficiency, and scale. The ideal candidate brings deep technical credibility in foundational AI, strong research engagement experience, and hands-on expertise in either AI inference research or AI training research. What you'll be doing: Serve as a trusted technical advisor to leading academic AI labs working on foundation models, LLMs, multimodal AI, reasoning, training, inference, and AI systems. Identify high-impact research workloads where NVIDIA software, systems, and accelerated computing platforms can advance model performance, scale, and efficiency. Engage principal investigators, postdocs, graduate researchers, and lab leadership to understand research goals, technical blockers, infrastructure needs, and collaboration opportunities. Track frontier AI research across papers, benchmarks, open-source projects, and academic labs to identify emerging trends and future platform opportunities. Partner with Research Account Managers, Solution Architects, Product, Engineering, and Business Development teams to support researcher adoption and long-term engagement. Represent researcher needs internally by translating academic feedback into actionable insights for product roadmaps, developer programs, education, and platform strategy. Support NVIDIA participation in major AI, ML, and systems research venues through technical content,

Machine LearningAI
🔔

Get new ai architect jobs in United States by email

Daily job updates · Unsubscribe anytime