Jobs in United States

Applied Ai Architect in San Francisco

288 active opportunities · Updated October 2026

Explore current applied ai architect jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

Hiring demand

23/100

cooling · 8 related jobs

Hiring trend

-66.7%

Job postings compared with the previous 30 days

Remote options

37.5%

Share of matching jobs listed as remote

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Codex Core Agent team builds the kernel of Codex. We own making the agent better, accelerating research, and making those improvements real in production for our users. That means working across the systems that make Codex actually function as an agent in the real world: the production performance envelope around tokens, latency, reliability, cost, and capacity; the core execution loop and interfaces that turn models into useful behavior; the shared infrastructure that enables other teams to build on Codex; and the feedback loops that turn real-world usage into better models and better agent behavior over time. About the Role We’re looking for engineers to build the infrastructure that powers Codex agents in production. This role focuses on the systems that let models safely execute code, interact with tools, complete long-running tasks, and operate reliably and efficiently at scale. You’ll design and operate the infrastructure behind sandboxed execution, orchestration, stateful workflows, app-server and SDK boundaries, and model rollouts. You’ll work at the intersection of distributed systems, developer tooling, and AI, building primitives that make Codex faster, safer, more reliable, and easier for the rest of the organization to build on. What You’ll Do Design and build execution environments for AI agents, including sandboxing, isolation, and reproducibility. Develop systems for agent orchestration across multi-step, tool-using workflows. Build infrastructure for running, testing, and debugging code generated by models. Create state and memory systems that allow agents to persist context across long-running tasks. Optimize tokens, latency, reliability, and cost across Codex’s production fleet. Support model rollouts, capacity planning, and the core tradeoffs between quality, speed, and economics to manage a fleet of frontier agents at scale. Build shared platform capabilities that unblock product teams, partner teams, and open source Codex. Yo

AWSCI/CDRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Education team is building products and experiences that help learners, educators, and institutions benefit from AI in ways that are rigorous, useful, and grounded in real learning outcomes. The work spans both consumer and B2B education, with close collaboration across engineering, learning science, design, data, and research. This team sits in a highly strategic investment area for OpenAI, with strong opportunities to shape how product ideas flow across consumer and institution-facing experiences. Some of our recent work: New Education Plugins for ChatGPT Work and Codex New tools for understanding AI and learning outcomes Education for countries Advancements in higher education Early product work - Introducing Study Mode About the role We’re looking for a product-minded Full Stack Engineer to help build OpenAI’s education products from the ground up. You’ll own end-to-end development across the stack, from early concepting and prototyping through production launch and iteration. This is an opportunity to work on a highly strategic, early-stage product area where engineering judgment, product sense, and customer empathy all matter. You’ll partner closely with leaders across the education org, including learning scientists, researchers, designers, and cross-functional partners, to turn emerging ideas into durable product experiences for schools, universities, and other education stakeholders. In this role, you will: Build and ship product experiences across the full stack for OpenAI’s education offerings Own projects end-to-end, from ideation and technical design through implementation, launch, and iteration Work closely with learning scientists and researchers to translate learning goals and evidence into product decisions Collaborate with design, data, and cross-functional partners to build thoughtful, high-quality user experiences Help define the engineering foundation for a growing education pod, including patterns, systems, and technical

TypeScriptPythonReactSQL
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Online Data team builds and operates the core online database and indexing services for OpenAI’s production AI applications, including supporting the explosive growth of ChatGPT, the #1 AI app in the world, and Codex, the fastest growing agentic development toolset in the world. Our mission is to ensure the reliability, correctness, and scalability of our online data stack and to curate a comprehensive portfolio of services that matches the relentless ambition of OpenAI, enabling our product and research teams to build 0-100 without getting bogged down in the minutiae of multi-region, multi-cloud, exabyte-scale data infrastructure. About the Role We are seeking an Engineering Manager to lead our Online Data Systems team, responsible for our in-house database and indexing technology. This role is about shepherding a team of world-class engineers tasked with building and operating hyperscale data storage and retrieval technology. You’ll be overseeing the delivery of extremely challenging engineering work in areas like distributed query execution, multi-region federation, self-orchestrating and self-healing services, low-level performance optimization, and more. There are few companies in the world building this kind of technology in-house at this scale where you’ll still be getting in on the ground floor. Instead of being a cog in the machine spending months chasing small optimizations, you’ll play a major part of shaping our future. In this role, you will: Build, lead, and grow high-performing infrastructure engineering teams. Drive the evolution of OpenAI’s in-house online data technologies, our core, hyper-scale database systems, indexing technologies, and vector search. Anchor delivery around measurable reliability goals (SLOs, etc) to ensure system performance and resiliency is above reproach. Champion pragmatic use of agent technology to amplify execution velocity. Reduce operational toil and incident frequency through better abstractions, gua

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -80.2%

From $230K/yr

Quick readStrong listing-quality and freshness signals

About the Role The Engineering Acceleration Delivery / Continuous Deployment team builds and operates the systems that safely ship OpenAI’s infrastructure and product code to production. We own the deployment platform, release pipelines, and rollout safety mechanisms that allow engineers across OpenAI to deploy changes rapidly while minimizing operational risk. Our mission is to make production deployments fast, safe, and increasingly autonomous. This role sits at the intersection of developer productivity, distributed systems reliability, and large-scale infrastructure orchestration. In This Role, You Will Design and build continuous deployment infrastructure that safely rolls out changes across dozens of Kubernetes clusters and global regions. Develop systems for progressive delivery, including canary releases, staged rollouts, and automated rollback. Improve engineering velocity by reducing friction in the release pipeline and automating manual operational workflows. Work with product and infrastructure teams to ensure their services are deployable, observable, and resilient at scale. Implement and evolve deployment methodologies such as GitOps, infrastructure-as-code, and progressive delivery patterns. Build systems that automatically evaluate deployment health using metrics, logs, traces, and alerts to detect regressions and trigger safe rollbacks. Build systems that support agent-assisted or autonomous deployment workflows using modern AI tooling. Technologies commonly used in this environment include: Kubernetes for large-scale container orchestration and runtime infrastructure Python and FastAPI for internal services Terraform for infrastructure as code GitOps-based deployment workflows (e.g., ArgoCD, Flux, or similar systems) Buildkite for CI orchestration You may be a strong fit if you: Have worked with Kubernetes-based deployment systems at scale Have experience building or operating continuous deployment platforms Are familiar with GitOps tooling such as

PythonAWSKubernetesGit
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Premium team owns some of the highest-leverage customer-facing levers in ChatGPT’s consumer revenue business, spanning the paid customer journey: helping users understand the value of paid plans, convert with confidence, and continue finding lasting value in their subscription. Our work is highly cross-functional, partnering with Product, Data Science, Design, FinEng, Finance, Legal, Support, and Marketing to improve free-to-paid conversion, renewal, customer lifetime value, and revenue while keeping the experience trustworthy, scalable, and low-friction. In This Role, You Will: Lead and scale an engineering team responsible for some of ChatGPT’s most important subscription and monetization experiences. Own the technical execution for Premium customer experiences across plan merchandising, paywalls, upgrade flows, checkout UX, plan management, renewals, downgrades, and cancellation. Partner with Product and Data Science to run high-quality experiments across upgrade, trial, renewal, downgrade, and cancellation flows. Improve key subscription metrics including conversion, renewal, churn, ARPU, and lifetime value. Build reliable customer-facing Premium experiences for purchase, plan management, renewal, downgrade, cancellation, and access-related states at scale. Partner closely with FinEng and other platform teams to evolve the billing, payments, and entitlement capabilities that power Premium experiences. Collaborate closely with Product, Design, Data Science, Finance, Legal, Support, and Marketing on monetization strategy and execution. You Might Thrive in This Role If You: Have 5+ years of engineering management experience, Have strong technical expertise in backend, frontend, or full-stack development, with experience building growth-oriented features. Have a track record of improving conversion, retention, or monetization through experimentation and data-driven product engineering. Are experienced with subscription products, plan merchandising

SQLAWSRestAI
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As a senior member of Baseten's Platform Team, you will own the systems that let every engineer at Baseten prove their code works before it reaches production. Our product runs mission-critical AI inference for customers who measure downtime in dollars per second, which means our internal bar for correctness, performance, and failure tolerance has to be exceptional. Your focus is the full testing stack: fast and reliable unit test tooling, integration harnesses that spin up realistic environments on demand, load and performance testing for GPU-backed inference workloads, and resilience testing that deliberately breaks things so our customers never have to find out what happens when a node dies mid-request. This is a builder role with org-wide leverage. You won't be writing tests for other teams — you'll be building the frameworks, harnesses, and feedback loops that make writing good tests the path of least resistance, and you'll set the standards for what "well-tested" means at Baseten. RESPONSIBILITIES Own Baseten's testing strategy end to end — define the standards, the tiers, and the tooling that engineering teams build against. Build and maintain unit, integration, load and performance testing frameworks Design end to end test infrastructure that provisions realistic dependencies

PythonDockerKubernetesCI/CD
S
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -87.9%

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Support Experience group develops and applies technology at all points of the user support journey to solve customer problems at scale, keeping millions of businesses running and unlocking growth across Stripe’s product suite. What you’ll do The rise of AI agents has fundamentally changed how users interact with Stripe, with a supermajority of conversations with Stripe now occurring via AI across various surfaces including third-party platforms. These conversations now blend planning, building, configuration, and troubleshooting. As a Product Manager for Support Experience, you will be responsible for building the platform that enables Stripe’s user-facing AI agent to solve problems across the Stripe product suite. You’ll develop rapid AI-powered feedback loops that enable product teams across Stripe to continuously improve the conversational experience of their users and build the infrastructure to enable powerful and flexible conversational experiences to run in diverse properties such as Stripe’s merchant dashboard, consumer apps, and experiences mediated by third-party agents. You’ll be at the forefront of applied AI, solving problems for businesses and consumers, and redefining what great can look like. Responsibilities Create the home for product teams to safely build out, understand and improve their conversational experiences. Build feedback loops from conversations through improvement recommendations to generating ev

O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Team The Cooperative AI team is scaling OpenAI with OpenAI. We are building a model-powered scaled automated workforce and knowledge system that evolves and learns alongside a human workforce. By leveraging OpenAI’s state-of-the-art models and technologies, some already in production, others still in the lab, we develop systems that reason and work autonomously for a wide variety of operational work. We leverage real workloads for critical systems across finance, sales, customer support, integrity, product insights, internal operations, and more in order to drive insights into product and industry. We partner closely with internal teams and external customers globally, operating in a hyper-fast feedback loop where many of our users are just a few steps away. This proximity allows us to iterate quickly, validate impact in real time, and accelerate industry impacting learnings and systems builds. We are a highly multidisciplinary, self-contained team focused on transforming the workplace via smart systems, knowledge, scalable and reliable primitives that apply world-class AI capabilities across domains. Our mission is to learn fast and transform how humans collaborate with AI at scale. About the Role We are looking for a hands-on Engineering Manager to lead a small, fast-moving team building AI-powered automation systems that redefine how work gets done across OpenAI. This role sits at the intersection of applied AI, research, and product engineering. You’ll lead a team that builds systems that know how to learn from humans, and carry real workloads across, sales, support, finance, IT, and more, while staying deeply involved in the technical work. You will operate in a highly iterative environment, deploying systems directly to internal users, gathering rapid feedback, and evolving solutions in real time. This is a high-ownership role for someone excited about building 0→1 systems, working closely with customers, and shaping how AI transforms operational wor

AWSRestAIRust
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We are hiring at a rapid rate, and every one of our new hires deserves a great new hire experience from the moment they sign their offer letter through their entire onboarding journey. As our People Operations Coordinator, you'll own the tactical side of onboarding: collecting and tracking the completion of required documentation, greeting new hires, assisting in running sessions, and keeping the operations of onboarding on track as we grow. The role is heavily onboarding-focused today and will broaden across the employee lifecycle as our new people systems take over more of the routine work. Things change quickly here, and this role will too. This role is based in our San Francisco office. RESPONSIBILITIES Own the full onboarding experience for every new hire, including: new hire communication from offer signature through Day 1; ensuring completion of onboarding tasks like background checks, Form I-9s, setting up HRIS profiles, etc; coordinating travel logistics; partnering with the Global Mobility Manager on immigration matters that may affect start dates Support day one and San Francisco week one onboarding program, including coordinating with managers and buddies, room booking and coordination of start location Partner with IT and Workplace so new hires have the right access, equipment, and desk waiting on day one Manage the logistics behind our onboarding platform and manage the day-to-day vendor relationsh

Machine LearningAILogistics
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We’re looking for a Corporate Accounting Manager to support the general ledger, the close, and our accounting processes and controls as Baseten scales. This is a hands-on, individual-contributor role for someone who wants full ownership of the core accounting function at a company where the entity structure, transaction volume, and reporting requirements are growing quickly. You’ll join the corporate accounting team and take on the general ledger, vendor contract review, the close calendar, and our internal control environment as the business grows. We are focused on tightening close procedures, strengthening controls, and building reporting rigor to support the scale ahead. You’ll partner closely with FP&A, Data, and cross-functional partners to make sure the books close on time and accurately every cycle. Baseten is building the infrastructure layer for AI-native companies, and we're scaling quickly - in headcount, entity structure, transaction volume, and contract complexity and volume and vendor size. If you want real ownership over a growing function on a lean team, this role offers real scope. RESPONSIBILITIES General Ledger and Financial Reporting Own day-to-day execution of general ledger journal entries, account reconciliations, and monthly financial statement preparation as the entity structure grows Own recurring close areas, including accruals, prepaids, fixed assets, and equity compensation Lead

Machine LearningAIAccountingFinance
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We are bringing our people systems in-house on Workday, and we are hiring our first dedicated Workday Analyst to make it excellent. You will join while the implementation is underway, ramp alongside our deployment partner, and own the tenant from go-live onward. This is a hands-on configuration role: you will build business processes, manage security, load data, and test releases yourself, and you will teach others on the People team to do the same as we grow. If you want to shape a Workday environment from its first day in production instead of inheriting years of someone else's decisions, this is that rare opening. RESPONSIBILITIES Own day-to-day Workday configuration: business processes, security groups and roles, custom reports, calculated fields, and tenant settings Build and run EIB loads for data changes, mass updates, and audits Own the twice-yearly Workday release cycle: evaluate new features, regression-test, and roll out changes safely Shadow the implementation build, then take over tenant ownership at go-live Monitor integrations and triage issues with our IT team and vendors Support payroll configuration in partnership with our Accounting team and payroll services provider Field and resolve system requests from employees, managers, and the People team Mentor teammates so Workday administration becomes a team capability, not a single point of failure REQUIREMENTS 4+ years administering a live Workday

Machine LearningAIAccountingPayroll
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE Our Sales and Solutions teams navigate hard technical conversations spanning inference performance, GPU economics, latency budgets, deployment shape. As Baseten’s platform matures, we need a dedicated owner to translate launch velocity into field readiness. As our first Product Enablement Lead, you'll sit between Product, Marketing, and Sales GTM and own how Baseten's products, features, campaigns, and market moments like the launch of GLM-5.2 or Kimi K3 or the sudden evolution of Tokenomics as a discipline get translated into field execution. You will own how these launches land with the field, how AEs and SAs stay credible on a highly dynamic technical ecosystem, and how what the field hears from customers makes it back to Product. This is a hands-on individual contributor role. You are the bridge between product, marketing, and sales. You'll build the system and run it, which includes cross-functional program leadership, direct training and enablement of in-seat reps, and content and curriculum development for managers, sellers, and new hires. Success here will depend on your ability to build repeatable systems and rhythms and to partner across the business and with your enablement colleagues to ensure alignment and speed of execution. RESPONSIBILITIES Own launch readiness: partner with Product and Marketing on positioning, write internal launch comms, and run readiness sessions so AEs and SAs can sell new pr

Machine LearningAIMarketingHR
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. ABOUT THE ROLE This role owns Baseten's relationships and market intelligence across the hardware and chip layer of the compute stack: NVIDIA directly and key OEM partners such as Dell, Lenovo, Pegatron, and Supermicro. As Baseten's compute strategy increasingly depends on hardware access and terms, this role is central to keeping Baseten ahead of the market. WHAT YOU'LL DO Build and maintain relationships across NVIDIA and key OEM partners (e.g. Dell, Supermicro) Track market intelligence on hardware availability, roadmaps, and terms to keep Baseten informed and strategically well-positioned Support deal structuring and negotiation in partnership with Baseten's deal-making function Work closely with Infrastructure and Hardware Platform engineering teams to ensure consistent, high-quality provider relationships and engineering partnerships Represent Baseten credibly across senior relationships in the hardware ecosystem, escalating to company leadership when strategically valuable WHAT WE'RE LOOKING FOR Existing relationships and credibility within the NVIDIA, OEM, and HPC ecosystem Strong relationship-management instincts, with the judgment to know when to bring in senior leadership for maximum impact Comfort operating in a fast-moving, high-stakes market where hardware access can be a major competitive differentiator Collaborative style — this role depends on close coordination with engineering counterparts, not just ex

Machine LearningAIGoHR
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. ABOUT THE TEAM Supply is responsible for knowing everything happening in the compute market: who's building, who's buying, and on what terms. This role owns a specific and fast-moving slice of that map — emerging clouds and international markets — and owns the full relationship lifecycle in that space, from first outreach through to closed terms. RESPONSIBILITIES Build and maintain a real-time picture of the emerging cloud and international compute landscape — who's active, what they're building, and what terms are available Own the full partnership lifecycle in this space — from identifying and sourcing new providers, to negotiating terms, to ongoing relationship management Develop and manage relationships across a broad set of emerging and international providers, from account reps up through leadership Identify, structure, and help close opportunities where Baseten can move quickly to secure favorable capacity terms Define compelling value propositions tailored to different types of providers, rather than a one-size-fits-all pitch Partner closely with others in the team already covering this space to build out a durable, well-organized intelligence and relationship function Collaborate with the broader Supply and Deals functions to bring opportunities to the table and support negotiation when it's time to close WHAT WE’RE LOOKING FOR Equal parts relationship-builder and operator — you can open a door and also drive it

Machine LearningAIGoHR
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE This role owns Baseten's relationships and market intelligence across hyperscalers and strategic neoclouds, including NVIDIA cloud partners. This is a technical and commercial role in equal measure: you'll evaluate capacity from the GPU to the data center, negotiate cost and terms with suppliers, and stay close enough to the market to develop and defend a real point of view on where it's heading. Given current market conditions, Baseten needs a much stronger pulse on this part of the market so we can track pricing, stay close to the right relationships, and move fast the moment more capacity is needed. This is a senior, experienced hire who will also help pair with and develop 1-2 junior to mid-level teammates covering the same space. WHAT YOU'LL DO Build and maintain deep relationships across hyperscalers and strategic neoclouds (including NVIDIA cloud partners), working each organization from top to bottom rather than a single point of contact Maintain a consistent, "top of mind" presence with key accounts so Baseten is positioned to move quickly when capacity needs arise Evaluate capacity from the GPU to the data center — hardware generation, rack and node configuration, interconnect, power density, and cooling — so you know what a configuration will actually deliver, not just what the spec sheet claims Live in compute pricing daily: track rates by GPU generation, region, and contract term to keep Baseten inf

Machine LearningAIGoHR
🔔

Get new applied ai architect jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime