About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team is responsible for building the next generation of AI-native silicon while working closely with software and research partners to co-design hardware tightly integrated with AI models. In addition to delivering production-grade silicon for OpenAI’s supercomputing infrastructure, the team also creates custom design tools and methodologies that accelerate innovation and enable hardware optimized specifically for AI. About the Role As an Engineer on our hardware optimization and co-design team, you will co-design future hardware from different vendors for programmability and performance. You will work with our kernel, compiler and machine learning engineers to understand their unique needs related to ML techniques, algorithms, numerical approximations, programming expressivity, and compiler optimizations. You will evangelize these constraints with various vendors to develop and influence future hardware architectures towards efficient training and inference on our models. If you are excited about efficiently distributing a large language model across devices, dealing with and optimizing system-wide/rack-wide networking bottlenecks and eventually tailoring the compute pipe and memory hierarchy of the hardware platform, simulating workloads at different abstractions and working closely with our partners, this is the perfect opportunity! This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. Key Responsibilities Co-design future hardware for programmability and performance with our hardware vendors Assist hardware vendors in developing optimal kernels and add support for it in our compiler Develop performance estimates for critical kernels for different hardware configurations and drive decisions on compute core and memory h
Jobs in United States
Ai Architect in United States
5,203 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is the world's leading API platform, and APIs and AI agents are increasingly the backbone of modern software. We're looking for a Technical Community Manager who will build and lead a thriving global developer community. One that goes beyond Postman users to become the destination for any developer working with APIs and AI agents. This is not a typical community role. You'll be architecting programs, designing engagement loops, and building something genuinely new: a community where developers come to learn, build, share, and grow — and where Postman is recognized as an indispensable part of their toolchain. The intersection of APIs and AI agents is one of the most exciting spaces in software right now, and Postman sits right in the middle of it. We have millions of developers already on the platform — the opportunity is to turn that user base into a community that educates, inspires, and advocates for one another. The person who builds this will leave a real mark on how developers collaborate and grow for years to come. What You'll Do Discord Community Growth & Engagement Own the end-to-end strategy for
$135K – $225K/yr
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Who you are We are seeking an experienced DevOps Engineer to join our growing team and play a pivotal role in designing and building our platform and infrastructure as we continue to scale our product and user base. As a part of our team, you will be working in a dynamic, fast-paced environment to ensure the reliability, scalability, and performance of our systems, while focusing on service architecture and deployment, query optimization, distributed systems, data and machine learning infrastructure, and security and authentication. Most importantly, you are excited to be part of a mission-oriented, fast-paced, high-growth startup that can create a lasting impact. You will: Partner with product teams to architect, design, and build the foundational infrastructure for our products. Design, develop, and deploy highly available and scalable Multi-tenant SaaS solutions on any one of the public cloud networks like AWS, Azure and GCP. Leverage technologies such as Kubernetes, Helm, Terraform, and Istio to achieve infrastructure resilience. Drive the automation of infrastructure tasks, from provisioning to configuration management and deployment, utilizing tools like Terraform, Ansible, a
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a talented Tech Lead Manager (TLM) to lead Snowflake’s System Under Test (SUT) team, responsible for evolving how Snowflake engineers test the Snowflake product locally and at scale in CI. The SUT empowers Snowflake engineers by delivering a reliable, low-latency and cost-efficient developer experience across a high-growth, high-demand surface area. As the TLM for SUT, you will lead a small and highly technical team at the intersection of CI, developer infrastructure, and product engineering. You will set direction, drive execution, and partner broadly across Engineering Systems and product teams to deliver a more reliable, faster, and more maintainable test platform for Snowflake’s engineers. In this role, you will: Lead, coach, and grow the SUT team while creating a high-energy, cohesive environment with strong planning, ownership, and career development. Own the roadmap and execution for SUT rollout across development environments, CI and AI workflows. Drive measurable improvements in startup reliability, latency, and cost, using clear SLOs, dashboards, and operational metrics to guide decisions and raise the bar on execution. Serve as the technical anchor for the SUT domain, shaping architecture and guiding the evolution from legacy systems to a composable
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Build the future of the AI Data Cloud. Join the Snowflake team. Do you want to own one of the most consequential narratives in enterprise data? As AI forces organizations to rethink their architectures, the ability to act on governed data from any engine, any cloud, any format is becoming a foundational requirement. We're looking for a technically fluent, driven PMM to lead Snowflake's Interoperable Lakehouse story across the Apache Iceberg™ and Apache Polaris™ ecosystems, turning a fast-moving technical landscape into sharp positioning that wins in the market. This is a hybrid role requiring 3 days per week in Snowflake’s Menlo Park, CA, OR Bellevue, WA office. WHAT YOU'LL DO: Build and execute a go-to-market strategy and innovative programs that position the Snowflake platform as the leading Lakehouse, with interoperability at the core of our offerings. Create crisp and compelling messaging, content, sales enablement, and more to be used by Snowflake marketing and sales teams, as well as partner teams Drive Snowflake's go-to-market narrative across the Apache Iceberg™ and Polaris™ open source ecosystems alongside our community team, including developer and community audiences Collaborate cross-functionally with other PMM teams, demand generation, content marketing, sales,
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Our team builds and operates Snowflake's Streaming Platform — the core infrastructure responsible for bringing data into Snowflake continuously and at scale. This includes Snowpipe Streaming and Datastream, services used by large set of enterprise customers to move mission-critical data in real time. Data is the fuel powering the new enterprise AI world, and we are the team responsible for getting it there reliably, continuously, and fast. Responsibilities Design, build, and maintain core components of Snowflake's streaming ingestion platform Improve service reliability, scalability, and latency under high-throughput production workloads Debug and resolve incidents in a distributed, multi-tenant cloud service Contribute to the design and evolution of streaming APIs, SDKs, and server-side protocols Write comprehensive tests including unit, integration, and chaos/fault-injection scenarios Collaborate with partner teams (storage, query, compute) on cross-cutting platform concerns Participate in code reviews, on-call rotations, and architecture discussions Required Qualifications 3–5 years of software engineering experience on large-scale distributed systems or cloud services Strong proficiency in Java or C++ Deep understanding of distributed systems concepts: consistency, faul
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are the Snowflake Metadata team. We own Snowflake’s metadata systems that make it easy for customers to query, modify and manage their petabyte-scale data. We develop distributed systems that store and maintain metadata, transaction frameworks that power Snowflake’s query and DML capabilities, asynchronous systems that provide time travel and lifecycle management capabilities and entity metadata supporting DDL capabilities. We also build foundational capabilities that deliver global features like cross-region replication (Snowgrid), data sharing, and data marketplace. AS A PRINCIPAL SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real business needs at large scale by applying your software engineering and analytical problem solving skills. Design, develop and support fault-tolerant scalable distributed systems for our Snowgrid and Data Sharing teams. Create architecture and design, influence our product roadmap, and take ownership and responsibility over new projects. Analyze fault-tolerance and high availability issues, performance and scale challenges, and solve them. Mentor and grow junior engineers. Understand trade-offs between consistency, performance and costs to build solutions which can meet the demands of rapidly growing services. Ensure operational readiness of
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary We are seeking a Software Development Engineer In Test, to apply advanced software engineering skills to improve quality across the development lifecycle through scalable test frameworks, developer tooling, and reusable automation solutions. This is a senior individual contributor role for a developer who specializes in testing, testability, and quality engineering. This role combines hands-on engineering with technical leadership across teams. The individual will influence architecture, strengthen automation strategy, improve developer feedback loops, and help establish consistent quality engineering practices that scale across products and platforms. The role carries strong Software Development Engineer in Test expectations, with an emphasis on building engineering solutions that improve product quality, platform reliability, and development velocity. Primary Responsibilities Design, develop, and evolve scalable test frameworks, automation libraries, and developer-facing quality tools Understand how the broader software ecosystem works together and define quality engineering approaches that align to platform strategy Partner with Product, Architecture, and Engineering teams to define test strategy, coverage goals, and acceptance criteria Build reusable utilities, harnesses, mocks, stubs, and data solutions that improve system
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Staff Hardware Engineer to provide advanced operational, diagnostic, and engineering support for Graphcore’s Arm-based hardware platforms across lab and data center environments. This role focuses on supporting hardware bring-up, validation, and troubleshooting of complex AI compute platforms, including server blades, racks, and rack-scale infrastructure. The successful candidate will collaborate closely with engineering, platform, and data center teams to ensure the reliability and performance of next-generation AI systems. The Team The Systems Engineering and Hardware Engineering teams are responsible for enabling the bring-up, validation, and operational reliability of Graphcore’s AI infrastructure platforms. The team works closely with server engineering, firmware teams, platform architects, and data center operations to support the development, testing, and deployment of next-generation AI compute systems. This collaborative environment enables rapid problem-solving and continuous improvement of Graphcore’s hardware platforms from early development through production deployment.
Job Details: Job Description: Intel is shaping the future of technology to help create a better future for the entire world. Our work in pushing forward fields like AI, analytics, and cloud-to-edge technology is at the heart of countless innovations. With a career at Intel, you'll have the opportunity to use technology to power major breakthroughs and create enhancements that improve our everyday quality of life. Join us and help make the future more wonderful for everyone. Want to learn more? Visit our YouTube Channel or the link below. Life at Intel The Role and Impact As a System Software Architect, you will play a pivotal role in designing and developing innovative software solutions that drive efficiency and scalability across Intel's manufacturing processes. You will be responsible for architecting robust systems and frameworks that enhance operational performance while ensuring seamless integration with existing technologies. Your contributions will directly impact Intel's ability to deliver world-class semiconductor products efficiently and reliably. Business Group Intel Foundry Services is at the forefront of manufacturing excellence, delivering cutting-edge solutions to meet the complex demands of the semiconductor industry. The group focuses on pioneering advancements in manufacturing technologies, infrastructure optimization, and secure operations, enabling Intel to stay ahead in a fast-evolving landscape. By joining this team, you will contribute to Intel's broader mission of solving technological challenges and empowering innovation worldwide. Key Responsibilities: - Architect scalable and efficient system sof
Job Details: Job Description: Do Something Wonderful At Intel, we put silicon at the center of the world’s most meaningful innovations. From personal computing and AI to cloud-scale data centers and edge platforms, our processors power technologies that touch billions of lives. If you’re passionate about building world-class silicon at massive scale, we want to build the future with you. The Team You’ll join Intel’s All Cores Engineering (ACE) organization, the team responsible for designing and delivering Intel’s industry‑leading CPU cores. Within ACE, the Atom group focuses on high‑efficiency, performance‑per‑watt optimized cores that enable everything from client platforms to emerging edge and scalable solutions. This is a hands‑on senior engineering role with direct impact on next‑generation CPU silicon. What You’ll Do As a Senior CPU Design Verification Engineer you will play a critical role in ensuring architectural correctness, functional robustness, and power‑efficient performance of Intel’s Atom CPU designs before first silicon. Key responsibilities include but not limited to the following: Lead pre‑silicon functional verification of complex CPU and SoC logic to ensure full compliance with architectural and microarchitectural specifications. Develop and execute verification plans, testbenches, and SystemVerilog/UVM environments, ensuring comprehensive functional and coverage closure. Define, run, and analyze system‑level simulations to uncover functional, performance, power, and corner‑case issues. Debug and root‑cause complex RTL, microarchitecture, and integration issues; drive issues to resolution with design teams. Collaborate cross‑functionally with CPU architects, RTL de
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Job Summary We are seeking a motivated engineer to join the DRAM Systems Engineering team, focusing on the development, evaluation, and optimization of next-generation memory systems for AI accelerators. This role emphasizes research and development across hardware architecture, operating systems, and performance analysis to support Agentic AI inference workloads. If you are ambitious and eager to make an impact in the exciting world of AI and memory systems, this is the perfect opportunity for you! Responsibilities Characterize AI inference workloads and examine memory behavior Build and evaluate tiered memory hierarchies for AI accelerators Study KV cache lifecycles, MoE models, and data placement strategies Compare and optimize explicit versus hardware-assisted data movement Develop, test, debug, and detail system-level and OS components Prototype and evaluate agentic AI systems by building agents and multi-agent workflows using modern frameworks and orchestration patterns (planning, tool use, memory, and context management). Apply these technologies both as workloads under study and as accelerators for internal engineering workflows <h2 style="color:!importan
NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and establish teams with the most thoughtful people in the world. NVIDIA GH200 superchip provides performance and productivity required for strong scaling for HPC and generative AI workload. Scale out is inherent to design of this massive superchip. We are looking for expert engineers to come and help design rack level solutions for next generation scaling AI supercomputing platforms. We are looking for a strong technical architect to own end to end manageability architecture for these products in data centers. You will work with various component leads internally and externally, drive customer use cases, align architecture with customer requirements and release best products to market. Join us at the forefront of technological advancement. What you’ll be doing: Drive server management for large clusters and data centers deploying GPUs and Grace solution from Nvidia. Work with data center architects and cloud customers to narrow down on requirements for implementation to ensure speed of light product development. Work with internal teams to make sure requirements are designed and implemented in right way with each firmware and software module Collaborate with other leads to design & build data center health management workflow. Drive reliability and optimization in firmware architecture from a data center view point. Work closely with cluster bring up team and resolve is
Yext (NYSE: YEXT) is the enterprise agentic marketing platform. Built on the world's most comprehensive structured data platform for local businesses, Yext gives brands and their partners the visibility intelligence to win every moment of discovery – across AI and traditional search. Yext's API-first architecture connects structured data to APIs, MCP servers, and generative interfaces, so partners and developers can build purpose-built experiences on the same infrastructure powering Yext's own products. Thousands of brands and digital marketing partners in financial services, healthcare, retail, hospitality, and food rely on Yext to manage, measure, and optimize visibility at scale. For more information, visit yext.com . At Yext, Product Engineering builds and evolves the technology behind our products and services. We’re looking for software engineers who want to solve meaningful technical problems, contribute to systems at scale, and help shape what we build next. We work in an agile environment with two-week sprints and regular demos that keep teams aligned and give engineers clear visibility into the impact of their work. From day one, you’ll contribute directly to the codebase and collaborate with experienced engineers from a wide range of leading universities and technology companies. We are looking for an engineer to join Team Watson , which owns and develops the systems that power Yext Search and Yext Chat. The team builds the indexing, retrieval, and serving technology that enables brands to deliver fast, relevant answers across their websites and digital experiences. Yext Search handles more than 50 million requests each month, serving users around the world in over a dozen languages. Watson also brings these search and retrieval capabilities to Yext Chat, helping conversational experiences generate useful answers grounded in trusted customer content. Because Watson’s systems serve real-time, customer-facing experiences at a global scale, engineers o
Job Details: Job Description: The Role and Impact As a Physical Design Engineer, you will play a critical role in driving the development of cutting-edge technologies at Intel. You will work hands-on to deliver high-performance physical designs, ensuring the successful integration of advanced semiconductor technologies. Your work will directly impact Intel's product innovation, addressing complex challenges in power, performance, and area optimization to shape the future of computing architectures. From synthesis to sign-off, your contributions will enable Intel to meet ambitious design goals and deliver transformative solutions to the market. Business Group This role is part of Intel's Corporate Technology Office (CTO), a dynamic group dedicated to advancing the company's technological leadership and innovation. The group focuses on developing foundational technologies, architectures, and methodologies that drive Intel's product roadmap and enable breakthrough computing solutions. Collaborating with cross-disciplinary teams, the CTO plays a pivotal role in delivering impactful designs that support Intel's broader mission of creating world-changing technology. Key Responsibilities - Drive RTL-to-GDS design convergence using advanced synthesis and place-and-route tools targeting performance, power, and area (PPA) goals. - Deliver block-level physical design, including closure of backend flows, electrical requirements, and enhancing silicon yield. - Collaborate with CAD and physical design methodology teams to adopt industry-leading tools and optimizations for custom designs. - Execute physical implementation tasks such as floor planning, bus/pin placement, power/clock distribution, congestion analysis, timing closure, IR drop analysis, and physical verification. - Debug timing, EM/IR, LVS, and DRC violations, and driv
Other cities to consider
More places hiring for this role
Get new ai architect jobs in United States by email
Daily job updates · Unsubscribe anytime