ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a Delivery Director, Capacity programs for our on-premises data center builds and neo cloud (GPU cloud) delivery programs. This is a high-visibility, execution-critical role sitting at the intersection of infrastructure engineering, capacity planning, vendor/partner management, and customer delivery. You will own the end-to-end delivery lifecycle for large-scale compute infrastructure — from initial site/capacity commitments through power, networking, and hardware bring-up, to production-ready GPU/compute capacity landing in the hands of internal teams or customers. You'll be the person who turns ambitious infrastructure roadmaps into predictable, on-time, delivery. RESPONSIBILITIES Own delivery of on-prem infrastructure builds — colocation expansions, power/cooling readiness, rack-and-stack, network fabric bring-up, and hardware acceptance testing — coordinating across colo providers and partners, network engineering, hardware ops, and vendor teams. Drive neo cloud delivery programs — manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps, SLAs, and go-live readiness. Build and maintain master delivery schedules across concurrent, multi-site, multi-vendor programs, integrating power/shell timelines, hardware lead times, logistics, and software/platform readiness into a single critical path.
Jobs in United States
Ai Architect in United States
5,246 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role At WRITER, our mission to expand human capacity with superintelligence relies on a foundational truth: our platform must be available, performant, and reliable, 24/7. As an Infrastructure engineer, you'll be at the heart of making this a reality, impacting every enterprise customer who trusts us with their AI-powered workflows. This isn't just about keeping the lights on; it's about pushing the boundaries of what's possible, proactively identifying and solving complex systemic challenges, and laying the groundwork for our rapid growth and the evolving demands of enterprise generative AI. You'll build resilient systems, automate across the stack, and champion reliability best practices, directly enabling our ambitious product roadmap and ensuring our customers always have access to the powerful tools they need. This is a hybrid position, based out of our New York City, San Francisco, Seattle, or London hubs. You'll report to our director of engineering. 🦸🏻♀️
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Join WRITER's security team as a staff detection and response engineer and help protect the AI infrastructure that's transforming how the world works. You'll build sophisticated detection systems that identify attacks targeting our AI platform, training data, and model deployments while creating automated response capabilities that scale with our explosive growth. This isn't just traditional security work – you're defending cutting-edge AI/AGI systems against adversaries who are evolving their tactics as fast as AI itself advances. This role combines hands-on security engineering with strategic thinking to stay ahead of novel threats that don't exist in textbooks yet. You'll be the operational arm of our security function, translating threat intelligence into real-time detections, coordinating incident response across multiple teams, and hunting for sophisticated attacks across GPU clusters and distributed training environments. If you're excited by the challen
About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight: Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful autonomy (for example future versions of auto-review ). About the Role We’re looking for strong executors with excellent judgment, comfort with ambiguity, and an understanding of frontier model research. You don’t need prior safety or alignment experience, we also welcome people that recently realized that alignment and safety is a critical area to contribute to. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Train and evaluate frontier models to reduce harmful or misaligned agent actions, forming clear hypotheses and executing independently through ambiguity. Mine incidents and build scalable measurement, data-processing, and evaluation systems that turn real failures into repeatable safety signals. Collaborate closely with post-training, capabilities, oversight, and pre-training partners to ship research-backed mitigations into large-scale training and agent systems. You might thrive in this role if you: Have demonstrated strength in research engineering, ML en
About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight : Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful agent autonomy (for example future versions of auto-review ). About the Role This role focuses on oversight and system-level mitigations that enable increasingly capable agents to operate safely and autonomously in real environments. We prioritize building oversight systems that are used in practice today, both internally and externally (see our recent work on action monitoring for codex and former code review ). We also study longer-term questions about how increasingly capable agentis systems can be supervised, constrained, and corrected. We’re looking for a safety&security minded researcher or engineer who can reason rigorously about security boundaries and agent behavior, then build and test practical mitigations. A background in AI control or security is welcome but not required. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and evaluate system-level controls for agent actions like agent-based review. Plan how they fit in a broader syste
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE The largest, most demanding enterprises run on Baseten, and they bring exacting requirements for how people, services, and agents access the platform. This is the founding role for our identity and authorization team within enterprise engineering. You'll own the identity and access layer of the Baseten platform: the authorization model, credential systems, and admin experiences that enterprise IT teams use to govern access for organizations like Harvey, HubSpot, and Notion. You'll design and build Baseten's fine-grained authorization system from the ground up to support the workflows customers depend on today while giving them cleaner, more precise ways to manage access as the platform grows. Authorization at Baseten requires low-latency permission checks at high request volume, consistent contracts and behaviors across the product suite, and strong security guarantees for mission-critical, highly regulated workloads. EXAMPLE INITIATIVES Recent and upcoming work in this area: Fine-grained authorization for users, service accounts, and agentic workloads: per-resource permissions at the organization, team, and workload scope to support both common workflows and complex enterprise access policies Programmatic authentication allowing high-compliance customers to connect service principles securely via short-lived, workload-based credentials Agent credentials that grant an agent exactly the access it needs for the gi
About the Team OpenAI’s User Operations team shepherds our customer’s adoption of AI and ensures that our customers' product experience is nothing short of exceptional. We are building the very first post-AGI support team. We resolve complex issues, provide technical guidance, and support customers in maximizing value and adoption from deploying our products. We work closely with Sales, Technical Success, Product, Engineering and others to deliver the best possible experience to our customers at scale. OpenAI's customers represent a range of diverse backgrounds and maturity, from early-stage startups to established global enterprises. About the Role We are looking for a Support Program Manager to join our Support Delivery team. This role is an exciting opportunity to help define and implement foundational support practices that will scale with OpenAI’s growth. You will lead efforts to establish new operational frameworks, driving process alignment with various internal teams, and leading tooling and automation projects. This position offers the chance to make a significant impact in shaping customer experience while collaborating across multiple teams. We’re looking for people who thrive at the intersection of project management, systems building, data science/data engineering/software engineering, team enablement, and customer advocacy – and enjoy working cross-functionally in a fast-paced, evolving environment. This role is based in San Francisco, CA, and follows a hybrid work model of 3 days in-office per week. Relocation assistance is available. In this role, you will: Lead support delivery programs for Tier 3 frontline delivery for our most strategic customers, including productivity and workflow improvements, team operations, and knowledge management for the Support Delivery team Partner with Support Delivery and cross-functional leadership to define the experience, stand up the program, manage pilots and rollout, and ensure premium operations and playbooks st
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're hiring a Product Data Scientist to establish how product decisions at Baseten are made with data. You'll work directly with Product and Engineering, alongside GTM to determine measurement, strategy, experimentation and implementation. This is a foundational, hands-on role. You'll define what success looks like across a technical, usage-based platform and turn ambiguous questions into analyses, forecasts, and experiments that shape product strategy. You'll work from clickstream and product events through inference telemetry and observability data, helping Baseten make faster decisions about reliability, performance, adoption and developer experience. RESPONSIBILITIES Partner directly with Product and Engineering: frame the questions that matter, define success criteria, and turn analysis into roadmap, launch, and prioritization decisions. Define how product success is measured: establish metrics across activation, adoption, retention, expansion, reliability and user experience. Support experimentation and launches: design measurement plans, analyze A/B experiments and controlled rollouts, and translate results into product decisions. Diagnose reliability and scaling behavior: join customer signals with request, replica, deployment, and cluster telemetry to find patterns in release bottlenecks, unhealthy replicas, and models without traffic. Define the enterprise customer journey and measure feature adoption
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We’re hiring a Data Scientist to help build and scale our internal analytics capabilities. This is a foundational role where you’ll create dashboards, data models and insights to power business and product teams alike. You’ll collect requirements, define key metrics, and deliver insights directly to stakeholders. You'll define what success looks like across a technical, usage-based platform and turn ambiguous questions into analyses, forecasts, and experiments that shape Baseten’s product and strategy. RESPONSIBILITIES Build and maintain production-grade dbt models and dashboards across multiple functions with a focus on accuracy, simplicity and user experience. Define and instrument core metrics around ROI, product adoption, customer lifecycle, capacity, availability, revenue and costs. Ingest and transform raw data using tools like dbt, Airbyte, and BigQuery. Partner with Engineering, Finance, Marketing, and Sales teams to understand goals and translate them into data solutions REQUIREMENTS 5+ years of experience in analytics engineering, data analysis, analytics, data science or a related role Advanced SQL and dbt skills, with a record of building models, tests, semantic layers and lineage in a cloud data warehouse. Prior experience supporting complex cross-functional projects across GTM, Finance and Engineering across various stages of the customer journey. Experience building dashboards and self-serve analy
About the Team The Search research team focuses on building the systems that help AI systems find, retrieve, and use information from the world. We aim to make answers more useful and grounded for more than a billion ChatGPT users. About the Role We’re looking for a Technical Program Manager to lead a broad portfolio of research and engineering programs that power search. You’ll partner closely with researchers, engineers, and product leaders to turn ambitious goals into clear plans, resolve dependencies, and move complex technical work forward. This role combines technical depth, product judgment, and hands-on execution. You’ll work across retrieval, indexing, and model improvements, while collaborating with policy, legal, and external data partners. You’ll help teams make informed tradeoffs and build practical ways of working that support a fast-moving research environment. This role is based in San Francisco, CA. In this role, you will: Lead programs across model training, retrieval, large-scale indexing, and search infrastructure. Translate evolving goals into prioritized workstreams with clear owners, milestones, dependencies, and resource needs. Partner with research, engineering, and product leads to define requirements and make tradeoffs across scope, quality, performance, timelines, and cost. Establish program success metrics and use them to guide priorities and track improvements in coverage, answer quality, responsiveness, and trust. Identify technical and cross-functional risks early, drive blockers to resolution, and communicate progress and decisions clearly to teams and leadership. Coordinate with product, policy, legal, and external partners on data access, use, and presentation, helping teams resolve decisions that span technical and non-technical domains. Manage dependencies with data providers and build repeatable processes that help research and engineering teams execute effectively as the search effort grows. You might thrive in this role if you
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building the infrastructure layer for AI — and we're now building the team that will scale how we take it to market. This is a foundational hire on our GTM Strategy & Revenue Operations team, sitting at the intersection of strategic planning and field execution. You'll work directly with the CRO, Head of Revenue Operations and partner closely with Finance, field leadership, and our Central Ops team. On any given week, you might be refining our pipeline generation model, building a territory coverage analysis, designing a new GTM motion, or partnering with a regional leader to understand what's driving a trend in their pipeline. This role requires someone who can think rigorously, build things from scratch, and operate with speed and judgment in an environment where the playbook is still being written. This is a rare opportunity to be an early GTM strategy and operations hire at one of the fastest-growing companies in AI infrastructure — and to help define how we scale. WHAT YOU'LL DO: GTM Planning, Target Setting & Market Intelligence Contribute to the annual and quarterly GTM planning process in partnership with Finance, including headcount modeling, ramp assumptions, and revenue target-setting Build and maintain coverage models aligned to 2-year growth projections, incorporating territory design, account segmentation, and capacity planning Support quota framework development, helping trans
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building the infrastructure layer for AI — and we're now building the team that will scale how we take it to market. This is a foundational hire on our GTM Strategy & Revenue Operations team, sitting at the intersection of strategic planning and field execution. You'll work directly with the CRO, Head of Revenue Operations and partner closely with Finance, field leadership, and our Central Ops team. On any given week, you might be refining our pipeline generation model, building a territory coverage analysis, designing a new GTM motion, or partnering with a regional leader to understand what's driving a trend in their pipeline. This role requires someone who can think rigorously, build things from scratch, and operate with speed and judgment in an environment where the playbook is still being written. This is a rare opportunity to be an early GTM strategy and operations hire at one of the fastest-growing companies in AI infrastructure — and to help define how we scale. WHAT YOU'LL DO: GTM Planning, Target Setting & Market Intelligence Contribute to the annual and quarterly GTM planning process in partnership with Finance, including headcount modeling, ramp assumptions, and revenue target-setting Build and maintain coverage models aligned to 2-year growth projections, incorporating territory design, account segmentation, and capacity planning Support quota framework development, helping trans
About the Team OpenAI's Research Team is at the forefront of AI research, pushing the limits of what AI can achieve. Our team is dedicated to developing advanced AI systems that are powerful, safe, and beneficial for everyone. About the Role The Research IP Partnerships team is in need of Technical Program Managers (TPMs) to streamline the integration of our applied research with external strategic partners. This role is critical for synthesizing research from cross functional teams, enabling model deployment, and ensuring new technologies are effectively adopted. You will act as the connection that enables our partners to deploy the most advanced AI models. Your primary focus will be to increase our research velocity and ensure that our deployments are successful and collaborative with our partners. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Build and share a deep understanding of frontier AI model development Partner with internal and external teams to drive deployment of the latest OpenAI technologies Manage critical inquiries from both technical and non-technical partners Design and implement simple, scalable processes that solve complex problems Deliver high-profile pipeline and tooling projects on tight deadlines Work across research and engineering to align goals, streamline communication, and support business priorities You might thrive in this role if you: Have experience in a strategic partnerships and technical program management role Can right-size process to align stakeholders while ensuring speed of delivery (action-oriented) Are fantastic at building cross-functional relationships and having empathy for the many roles involved in deploying research Can design and build tools (via code / no-code / AI) to facilitate internal processes Are a great communicator across written, presentation, and visual forms. Are engaged and c
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Web Designer / Design Engineer to own baseten.co day to day, from concept through production. You'll design and build landing pages, bring product launches to life, run experiments, improve conversion, and keep raising the quality bar on the site. This is a high-ownership role on the brand design team, sitting between brand design and frontend engineering. We want the website to be one of the best expressions of the Baseten brand and one of the best developer-focused sites on the internet, which means treating it as a product that's constantly evolving rather than something we redesign every few years. You'll have a lot of autonomy to ship, while partnering closely with brand, product marketing, engineering, and leadership on bigger launches. RESPONSIBILITIES Own the design and frontend implementation of the Baseten marketing site. Ship landing pages, product pages, launches, and campaigns from concept through production. Improve the site continuously, finding and acting on opportunities to make it better rather than waiting for a redesign cycle. Partner with growth and product marketing to test messaging, layouts, and conversion improvements without sacrificing craft. Build polished interactions, motion, and details that make the site feel distinctly Baseten. Build and evolve the components, patterns, and design systems that let the team move quickly without lowering the bar. Turn complex
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is looking for a Product Designer to help shape the next generation of our platform. You'll join a small design team with a lot of ownership and a high bar for craft, working across the entire Baseten product to solve hard problems for technical users. This role covers the full arc of the work: early exploration through production. You'll establish patterns that show up across the product, evolve our design system, and raise the quality bar as we scale. RESPONSIBILITIES Lead product design end to end, including research, product definition, prototypes, design reviews, specs, and final implementation. Work directly with engineering and product to frame problems, explore solutions, and ship quickly. Collaborate with customers to understand their workflows, validate ideas, and test prototypes. Push the visual and interaction quality of the Baseten product across surfaces. Create interactions and details that make complex technical workflows feel clear, fast, and polished. Evolve our design system and establish patterns that scale across a growing product. Look across the product for opportunities to improve consistency, usability, and overall quality. Get into the code with engineers to polish UI and make sure the shipped experience matches the design. Help shape how the design team works and raise the bar for product quality across the company. REQUIREMENTS A portfolio that demonstrates strong visual and
Other cities to consider
More places hiring for this role
Get new ai architect jobs in United States by email
Daily job updates · Unsubscribe anytime