NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. At NVIDIA, as a Principal Rack Scale Systems Infrastructure Engineer, you will build and guide the development of software systems. These systems support our upcoming rack-scale infrastructure products and services. This exceptional role sits where software meets hardware. You will work on control planes, state machines, orchestration systems, firmware, OS lifecycle, and networking fabrics. Your task is to compose infrastructure-as-a-service control plane software that converts complex rack-scale hardware into dependable, manageable, and programmable infrastructure for NVIDIA, partners, and leading cloud and enterprise clients globally. What You Will Be Doing: Define the complete software architecture for rack-scale infrastructure products and services, covering control plane services, infrastructure management, firmware, operating systems, kernel drivers, networking fabrics, accelerator software, and user-mode manageability software. Use Kubernetes and cloud-native primitives as an infrastructure fabric when appropriate. This includes controllers, operators, reconciliation loops, and open source components. These components can operate safely at rack and fleet scale. Build open source infrastructure software that can b
Jobs in United States
Ai Architect in United States
5,348 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
As a Senior Backend Engineer on Coder's Enterprise Experience team, you'll build the systems that help large organizations run Coder in production with confidence. You'll improve how Coder scales, how it's upgraded, and how reliably it performs in regulated, air-gapped, and enterprise environments. You'll work on a genuinely cross-functional team of backend, platform, and QA engineers who own the end-to-end experience for Coder operators. From designing new features to evolving Coder's architecture, you'll partner across engineering and product to solve complex problems and ship software that operators trust. What you'll do here Design and build new features end to end, from technical design through production rollout. Design and implement backend architecture changes that support Coder's long-term scalability goals. Investigate and resolve scalability bottlenecks under production-like load, from database access patterns to concurrency handling in coderd. Improve database migration safety and upgrade reliability through schema compatibility, background migrations, and safe rollback strategies. Own the backend side of issues surfaced by Coder operators and administrators. Document the design, implementation, and operational tradeoffs of the systems you build. Participate in code reviews, RFC-style design discussions, and on-call rotations for the services you own. What we're looking for 5+ years of professional software engineering experience, including significant production experience with Go. Deep understanding of Go's concurrency model, including goroutines, channels, the sync package, and debugging race conditions under real-world load. Experience designing and operating relational databases in production, including schema design, migrations, and transactions. Strong verbal and written communication skills. Exceptional debugging and troubleshooting skills, with the persistence to drive complex problems to resolution. A self-motivated, analytical engineer who enj
About the Team Our mission at OpenAI is to discover and enact the path to safe, beneficial AGI. To do this, we believe that many technical breakthroughs are needed in generative modeling, reinforcement learning, large-scale optimization, active learning, and other areas. The team builds the performance-critical systems that allow OpenAI's models to run efficiently across a diverse set of AI accelerators. We work across the inference stack, from low-level kernels and compilers through model execution, to unlock the full capabilities of the underlying hardware. About the Role As a Software Engineer, Trainium, you will help bring OpenAI's inference workloads to AWS Trainium and build the software stack required to run cutting-edge frontier models efficiently on the platform. This is a deeply technical, cross-stack role spanning kernels, compilers, and model execution. You will work on the systems needed to support OpenAI's inference stack on Trainium, including developing and optimizing high-performance kernels, improving compiler support, and enabling efficient execution of the model forward pass. You'll work closely with engineers across inference, compilers, kernels, and ML systems to identify performance bottlenecks and build the software needed to take full advantage of Trainium. The work may range from low-level hardware-specific optimization to compiler and runtime improvements to integrating new model architectures into the inference stack. If you enjoy working at the intersection of ML systems, compilers, kernels, and accelerator hardware, this role is for you. We're looking for engineers who are self-directed, comfortable operating across abstraction layers, and excited to solve challenging performance problems for frontier-scale AI systems. In This Role, You Will Build and optimize OpenAI's inference stack for AWS Trainium. Develop high-performance kernels for critical model operations and workloads. Extend and improve compiler support to efficiently target
About the team The OpenAI for Government team is a dynamic, mission-driven group leveraging frontier AI to transform how governments achieve their missions. Our team works to empower public servants with secure, compliant AI tools (e.g., ChatGPT Enterprise in custom configurations) and mission-aligned deployments that meet government technical requirements with strong reliability and safety. As part of the OpenAI for Government team you will work across U.S. defense enterprises, helping to accelerate adoption through hands-on support, tailored training, and early insights—so civil servants can spend less time on red tape and more on meaningful work. If you're passionate about responsibly deploying frontier AI to uplift public institutions and transform how the government serves the American people, we want you on our team. About the Role The Strategic Delivery Lead (SDL) for DoW will identify and deliver on some of OpenAI’s most complex and high-impact deployments for the U.S. national security enterprise, with a focus on CDAO-led efforts including GenAI.mil , Advana (WDP), and other programs that accelerate DoW adoption of data, analytics, and artificial intelligence across the DoW enterprise. The role is highly cross-functional and customer-facing: You’ll partner with Research, Engineering, Go-to-Market, CDAO, the Military Departments and Services supported by CDAO, and other external stakeholders to align on scope, unblock delivery, and communicate progress at every level, from technical teams to C-suite. You will own and drive execution of technical delivery workstreams led by internal teams (Forward Deployed Engineers, Researchers, Solutions Architects, Enablement), ensuring clear problem-framing, crisp milestone definitions, proactive risk identification, rapid issue resolution, and effective translation between business objectives and technical solutions, while consistently communicating clear progress and demonstrating tangible value to national security cus
About Us: Blockworks is an information platform that sits at the center of the crypto industry. We transform raw, complex data and facts into actionable research, trusted alpha-driven insights, and world-class events. The result is transparency and confidence. Blockworks connects investors and businesses in onchain capital markets. We give businesses a platform to earn trust and provide investors with the information they need to underwrite the asset class. Who You Are: You have a keen focus on backend and API development and Software engineering is your passion. You are a player-coach and a natural leader who understands the technical and human elements that go into great software design. You have a results-oriented attitude and a passion for delivering flawless releases and developing digital product pipelines (CI/CD pipelines). You have a proven track record facilitating engineering teams to increase productivity and quality. You're excited at the possibility of being on the ground floor of the design and development of backend strategies. You bring a passion for designing and maintaining scalable API services that handle large amounts of data elegantly. You love moving quickly in a fast-paced start-up, but you also bring intentionality, sustainability and scalability to your approach as an engineer. What You’ll Do: As a Senior Backend Engineer at Blockworks, you’ll design, build, and maintain the systems that power our products end-to-end. From high-performance APIs to database architecture, you’ll own the backend layer that makes everything else possible. You won’t just be handed requirements, you’ll help define them, scope projects, and make the architectural decisions that shape our technical foundation. Your work will directly impact our research platform ( blockworksresearch.com ) and our media site ( blockworks.co ; 1M+ monthly active users). Every day will look a little different, but in general, you will do things like: Architect, build, and ship backend
$200K – $271.5K/yr
Why Join The Drata Team? The best way to understand the Driver’s Mindset is to see it in action. We’re an award-winning, mission-driven team of 600+ people worldwide , united by a culture that values trust, speed, and continuous growth. Hear the Voice of the Team : Explore our "Life at Drata" page for employee testimonials on our collaborative and the growth opportunities available. Connect with Us on Socials: LinkedIn - follow us for company updates, employee stories, and career news. Job Summary: The Staff Software Engineer, Monetization Platform serves as a technical leader with a primary focus on the systems that turn product usage, subscriptions, credits, commitments, and contracts into accurate customer outcomes and clean financial operations. This person will architect and evolve the core billing domain for Drata: usage event ingestion, billable metrics, pricing and packaging, invoice generation, entitlement enforcement, and the integrations that connect engineering to Accounting, Finance, RevOps, Salesforce, CPQ, and ERP systems like NetSuite. They should be equally comfortable designing event-driven systems, shaping commercial flexibility with business stakeholders, and building operationally trustworthy software where correctness, audit-ability, and scale matter. Modern billing is no longer just “charge a credit card once a month” — it is a real-time product control plane that must support evolving pricing models without turning the codebase into a fragile mess. What you’ll do: Partner with Product, Finance, Accounting, RevOps, and Engineering leadership to shape Drata’s long-term billing architecture and commercial flexibility across self-serve and enterprise motions. Design the primitives of a modern billing platform: usage events, billable metrics, products, rate cards, contracts, credits, subscriptions, overages, and invoice workflows. Architect a system that can accurately meter product usage at scale, transform raw events into billable quantities, an
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. We are looking for a Senior Software Engineer to join our Product Platform org - the group that powers product development across Vanta by building the shared foundations that every engineering team builds on top of. This hire will join one of two closely-aligned teams within the org, each working on foundational challenges with company-wide impact. Application Primitives is responsible for the shared, customer-facing platform components that product teams implement across the product: examples include comments, notifications platform, event log infrastructure, and agentic audit trails. This team sits at the intersection of platform and product - not pure infrastructure, not pure product - building the reusable building blocks that let other teams ship faster and more consistently. Libraries & Foundations is a brand-new team that owns Vanta's internal engineering libraries and golden-path patterns. Sitting between core platform and product platform in the stack, this team builds the TypeScript-first abstractions and shared foundations that all Vanta engineers build on top of - including event-driven architecture libraries, code quality tooling, and test frameworks. You'll be evaluated for both teams based on your background and interests, and the decision on team placement will reflect where your skills and passions are the strongest fit. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as a Senior Software Engineer, Product Platform at Vanta: Design and build high-quality, scalable systems and shared libraries that multiple engineering teams depend on - whether that's shared
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of researchers and engineers focused on building reliable ML systems and pushing the boundaries of LLM inference efficiency. We develop techniques that improve how models execute in production, driving lower latency, higher throughput, and consistent quality across diverse workloads. As an engineer on this team, you’ll work across the inference stack to improve core performance metrics by diving deep into model execution, identifying bottlenecks, and developing innovative optimizations. You’ll collaborate closely with modeling and systems teams to experiment, measure, and ship improvements that meaningfully accelerate inference. As the team evolves, you’ll have opportunities to build expertise in advanced performance techniques, including GPU/CUDA optimizations, kernel-level improvements, and model execution strategies for MoE and large-scale architectures. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, e
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Large Language Models (LLMs) continue to push the boundaries of what AI systems can do — but inference is still the bottleneck. The Model Efficiency team is responsible for pushing the limits of LLM inference efficiency across our foundation models. We explore and ship breakthroughs across the model execution stack, including: model architecture and MoE routing optimization decoding and inference-time algorithm improvements software/hardware co-design for GPU acceleration performance optimization without compromising model quality Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. As a Staff Research Engineer, you will develop, prototype, and deploy techniques that materially improve how fast and efficiently our models run in production. You may be a good fit
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 The Mission The Foundry is ClickUp's internal AI innovation lab — embedded inside GTM Systems and accountable for turning AI capabilities into production-grade, internally deployed products that make every GTM function faster and smarter. We build the infrastructure that powers AI-first work across Sales, Marketing, Post-Sales, and Revenue Operations. As the Senior Software Engineer on this team you will own the technical delivery of our MCP server platform, agent orchestration layer, and internal tooling — shipping production systems used daily by hundreds of ClickUp employees, and scaling your own throughput by treating AI tools as first-class engineering collaborators. What You'll Own MCP Server Platform Design, build, and operate Model Context Protocol servers that expose CRM, ticketing, analytics, and communication data to AI agents across the GTM stack Implement Okta PKCE authentication flows and RBAC policy enforcement so agents access only the data they're authorized to touch Maintain deployment infrastructure on AWS (Bedrock, Lambda, ECS, API Gateway) and contribute to GCP workloads where applicable Own observability: structured logging, distributed tracing, latency SLOs, and on-call runbooks for every production server Agent Orchestration & AI-Native Products Build and maintain multi-step autonomous agents that execute end-to-end GTM workflows — lead qualification, deal room assembly, onboarding automation, support triage, and more Architect prompt engineering frameworks, tool-call schemas, and agent evaluation harnesses that make AI behavior predictable and auditable Integrate with LLM p
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 As Senior Manager, CX Operations at ClickUp, you will own strategy, execution, and operational leadership for Technical Account Management (TAM) Operations , spanning Professional Services and Customer Success. Partnering closely with TAM leadership, you will serve as the day-to-day operational leader for the systems, automation, planning, quality, and insights that power our Services & Success organization. You will set the operating rhythm and performance standards for TAM Operations while owning the AI agent harness and automation layer within your domain. You will shape the cadence of business for TAM, unlock data insights, coach organizational performance, transform internal capabilities, and influence strategic decisions across Services, Success, and cross-functional leadership. You will architect, ship, operate, and iterate on AI-driven workflows that enhance and automate customer health inspection, engagement management, services delivery, quality measurement, and decision support across TAM and CX. About the Role Strategy & Operations Own the vision and strategy for delivering world-class customer experience through our Services & Success operating model Drive cross-functional alignment across Sales, Product & Engineering, Finance, and Support by synthesizing customer health, retention, and services delivery opportunities into operational priorities Lead and execute strategic initiatives to optimize and transform customer engagement, services delivery, and internal collaboration processes Extract key business insights from qualitative and quantitative data, identify risks and o
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Job Summary We are looking for a GTM DevOps Engineer to join our Business Systems team and own the reliability, automation, and delivery infrastructure behind our Go-To-Market (GTM) technology stack. This role sits at the intersection of platform reliability and CI/CD engineering, ensuring that our critical business systems — including Salesforce, NetSuite, MuleSoft, Workato, and an expanding portfolio of AI-powered workloads — are deployed consistently, operate resiliently, and scale with the business. You will partner closely with Business Systems developers, architects, and business stakeholders to build and maintain the pipelines, monitoring frameworks, and operational standards that keep our GTM systems healthy and our release cycles fast and predictable. As our team builds and deploys AI agents across GCP Cloud Run and AWS Bedrock AgentCore, you will serve as the infrastructure and deployment owner for these workloads — bringing engineering discipline to an environment where AI-generated code is increasingly entering production. This is a hands-on engineering role for someone who thrives in complexity, takes ownership of platform uptime, and brings a software engineering mindset to business application operations — directly supporting GTMSOE's broader mission of operational excellence across the GTM org. Key Responsibilities CI/CD & Release Engineering Design, build, and maintain CI/CD pipelines for Salesforce (SFDX/Salesforce CLI), NetSuite (SuiteScript/SuiteBundler), MuleSoft (Anypoint Platform), and Workato; establish branching strategies, environment promotion standards, and release gatin
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by leading the design of high-performance, scalable and reliable machine learning systems? Do you want to set technical direction and help shape the next generation of AI platforms powering advanced NLP applications? We are looking for a Lead Member of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team. You may be a good fit if you have: 8+ years of engineering experience running production infrastructure at a large scale, with a track record of technical leadership Demonstrated experience leading the architecture
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this team? The GPU Clusters team builds and operates the superclusters that train Cohere’s frontier models. We sit at the intersection of hardware, distributed systems, and AI research. We work with cloud providers, researchers, and other infrastructure teams on problems few companies get to take on. As an Engineering Manager, you’ll lead a team of engineers who care deeply about GPU infrastructure. You’ll set technical direction, grow people, and help the company scale a rapidly growing compute footprint. As an Engineering Manager, you will: Hire, mentor, and grow a team of GPU infrastructure engineers , including performance, career development, and technical guidance on hard infrastructure problems Own the technical roadmap for the fleet: how we deploy, operate, and scale Kubernetes clusters, including workload scheduling, hardware fault detection, and performance Partner with researchers and ML engineers so the training and inference stack works well on new GPU architectures Work with cross-functional stakeholders such as Capacity, Finance, Legal, Security, and other infrastructure teams on planning, cost, compliance, an
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview We’re looking for a Senior Frontend Engineer to help build and improve core product experiences at ClickUp. You’ll work closely with product, design, backend, and integrations teams to deliver high-impact features in a fast-moving environment. This role is ideal for someone who can move quickly, write high-quality code, and balance speed with long-term maintainability. What You’ll Do Build product features in Angular 2+ and React in partnership with designers, product managers, and engineers Architect efficient, reusable frontend code that powers the ClickUp user experience Collaborate closely with backend and integrations teams to deliver end-to-end features Work with QA to ensure strong test coverage and high product quality across edge cases Identify and resolve performance, scalability, and usability issues Build and maintain unit and integration tests Fix bugs quickly and deliver durable solutions to complex product challenges Follow established frontend architecture and state management patterns across the app Manage project priorities, deadlines, and deliverables in a fast-paced environment Qualifications 5+ years of experience with JavaScript and modern frontend frameworks, with Angular 2+ or React required Strong experience with TypeScript, RxJS, and NgRx or similar Redux-style state management Solid understanding of reusable component architecture and frontend performance Strong HTML/CSS fundamentals, including layout, accessibility, cross-browser compatibility, and maintainable styling practices Proven ability to move quickly, take ownership, and execute with urgency Strong com
Other cities to consider
More places hiring for this role
Get new ai architect jobs in United States by email
Daily job updates · Unsubscribe anytime