Jobs in United States

Engineer 3 in United States

3,770 active opportunities · Updated October 2026

Explore current engineer 3 jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -84.1%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI is building the next generation of advertising for an AI-powered world. As AI changes how people discover information, explore possibilities, and make decisions, we believe this presents a huge opportunity for small and medium-sized businesses to connect with customers in new ways. The SMB Ads Marketing team is building this opportunity from the ground up. We define how businesses discover OpenAI’s advertising platform, become successful advertisers, and grow their investment over time. Working closely with Product, Ads Engineering, Data Science, and Sales, we combine customer insight, rigorous experimentation, and fast execution to turn early opportunities into scalable growth. About the Role We are seeking a Lifecycle Marketing Manager to build, own, and lead lifecycle programs for SMB advertisers at OpenAI. This is a zero-to-one opportunity to establish the strategy and programs that guide businesses from initial interest and their first campaign to sustained success and growth. You will set the lifecycle strategy, own the program roadmap and performance goals, and lead delivery from customer insight and journey design through launch, measurement, and iteration. Combining behavioral insights, experimentation, and personalized engagement across email, product experiences, education, and other channels, you will build programs that scale with the business. Success requires independent judgment, analytical depth, cross-functional leadership, and hands-on execution. In This Role, You Will Own the lifecycle program portfolio across prospect conversion, activation, retention, repeat spend, reactivation, and expansion. Map advertiser journeys using cohort analysis, product behavior, campaign performance, and customer feedback to identify high-impact growth opportunities. Build segmented, triggered journeys that connect advertiser needs, intent, readiness, and behavior to relevant messages and next steps. Lead prospect nurture and email acquisition

ReactAWSRestAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -84.1%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI is building the next generation of advertising for an AI-powered world. As AI changes how people discover information, explore possibilities, and make decisions, we believe this presents a huge opportunity for small and medium-sized businesses to connect with customers in new ways. The SMB Ads Marketing team is building this opportunity from the ground up. We define how businesses discover OpenAI’s advertising platform, become successful advertisers, and grow their investment over time. Working closely with Product, Ads Engineering, Data Science, and Sales, we combine customer insight, rigorous experimentation, and fast execution to turn early opportunities into scalable growth. About the Role We are looking for a Scaled Programs Lead to build and scale global programs that engage SMB-focused agencies and other partners helping small and medium-sized businesses grow. This is a zero-to-one opportunity to shape how these audiences discover, understand and adopt OpenAI’s advertising solutions. You will develop programs that build awareness, recruit participants, strengthen their capabilities and help them bring the advertising opportunity to their SMB customers. Your remit will span audience segmentation, recruitment campaigns, education, certification, incentives and co-marketing. You will design and launch new programs, validate their impact and scale the approaches that deliver meaningful results. Working across Marketing, Partnerships, Sales, Product and Operations, you will combine strategic judgment with hands-on execution. As the programs grow, you will build and lead a team to expand them globally. In This Role, You Will Set the global program strategy for SMB-focused agencies and scaled partners, defining priority audiences, growth goals, and investment choices. Build awareness and recruit qualified participants through campaigns and outreach that demonstrate the value of OpenAI’s advertising solutions. Segment agencies and partners by ca

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -84.1%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI is building the next generation of advertising for an AI-powered world. As AI changes how people discover information, explore possibilities, and make decisions, we believe this presents a huge opportunity for small and medium-sized businesses to connect with customers in new ways. The SMB Ads Marketing team is building this opportunity from the ground up. We define how businesses discover OpenAI’s advertising platform, become successful advertisers, and grow their investment over time. Working closely with Product, Ads Engineering, Data Science, and Sales, we combine customer insight, rigorous experimentation, and fast execution to turn early opportunities into scalable growth. About the Role We are hiring a Lead, Ads Prospecting & Growth Intelligence to build the intelligence and orchestration platform powering OpenAI’s ads business. This person will help us identify the right companies and people, understand why they are relevant, determine the right next action, and learn from what happens afterward. Own the prospecting function for the ads business, serving SMB Ads while building a foundation reusable across segments. Set the vision, operating model, roadmap, platform choices, business rules, adoption, and outcomes; and partner with Ads Engineering, Data Science, and internal platform teams. Decide what to build in-house, enable through internal platforms, or deliver through selected vendors and partners. In This Role, You Will Lead the function’s strategy, roadmap, operating model, and adoption. Connect platform investments to measurable advertiser and business outcomes. Define a unified prospect, account, contact, and advertiser data model. Establish standards for identity, data quality, provenance, freshness, confidence, and consent. Combine market, intent, behavioral, product, CRM, campaign, and customer-success signals to improve targeting and customer engagement. Design enrichment and verification workflows. Evaluate in-house cap

SQLAWSRestAI
P
📍 New York, New York, United States· Full-time· Remote
✓ High-confidence listingCompany trend -73.5%
Quick readStrong listing-quality and freshness signals

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. Fraud Data is the data science and machine learning team within Plaid’s Fraud organization, responsible for using data and ML to improve and scale Plaid’s fraud products. Within Fraud Data, the Customer & Product Intelligence team focuses on understanding product performance, uncovering customer insights, and enabling go-to-market teams with data-driven solutions. The team partners closely with customers and GTM teams on fraud analyses and proofs of concept, turning customer learnings into scalable, reusable product capabilities. We also build the metrics, analytics, and data foundations that measure product health, identify opportunities for improvement, and guide product decisions across Plaid’s Fraud portfolio. As a Data Science Manager, you will lead a team responsible for customer-facing data science and Fraud product analytics. You will set the team's roadmap, develop its data scientists, and remain involved in analytical methods, technical reviews, and customer investigations. You will: Set a 6–12-month roadmap with Product, Engineering, and GTM, and assign priorities and responsibilities across the team. Define product metrics, their underlying data, and reporting and

PythonSQLAWSMachine Learning
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. ABOUT THE ROLE This role owns Baseten's relationships and market intelligence across the hardware and chip layer of the compute stack: NVIDIA directly and key OEM partners such as Dell, Lenovo, Pegatron, and Supermicro. As Baseten's compute strategy increasingly depends on hardware access and terms, this role is central to keeping Baseten ahead of the market. WHAT YOU'LL DO Build and maintain relationships across NVIDIA and key OEM partners (e.g. Dell, Supermicro) Track market intelligence on hardware availability, roadmaps, and terms to keep Baseten informed and strategically well-positioned Support deal structuring and negotiation in partnership with Baseten's deal-making function Work closely with Infrastructure and Hardware Platform engineering teams to ensure consistent, high-quality provider relationships and engineering partnerships Represent Baseten credibly across senior relationships in the hardware ecosystem, escalating to company leadership when strategically valuable WHAT WE'RE LOOKING FOR Existing relationships and credibility within the NVIDIA, OEM, and HPC ecosystem Strong relationship-management instincts, with the judgment to know when to bring in senior leadership for maximum impact Comfort operating in a fast-moving, high-stakes market where hardware access can be a major competitive differentiator Collaborative style — this role depends on close coordination with engineering counterparts, not just ex

Machine LearningAIGoHR
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. ABOUT THE TEAM Supply is responsible for knowing everything happening in the compute market: who's building, who's buying, and on what terms. This role owns a specific and fast-moving slice of that map — emerging clouds and international markets — and owns the full relationship lifecycle in that space, from first outreach through to closed terms. RESPONSIBILITIES Build and maintain a real-time picture of the emerging cloud and international compute landscape — who's active, what they're building, and what terms are available Own the full partnership lifecycle in this space — from identifying and sourcing new providers, to negotiating terms, to ongoing relationship management Develop and manage relationships across a broad set of emerging and international providers, from account reps up through leadership Identify, structure, and help close opportunities where Baseten can move quickly to secure favorable capacity terms Define compelling value propositions tailored to different types of providers, rather than a one-size-fits-all pitch Partner closely with others in the team already covering this space to build out a durable, well-organized intelligence and relationship function Collaborate with the broader Supply and Deals functions to bring opportunities to the table and support negotiation when it's time to close WHAT WE’RE LOOKING FOR Equal parts relationship-builder and operator — you can open a door and also drive it

Machine LearningAIGoHR
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE This role owns Baseten's relationships and market intelligence across hyperscalers and strategic neoclouds, including NVIDIA cloud partners. This is a technical and commercial role in equal measure: you'll evaluate capacity from the GPU to the data center, negotiate cost and terms with suppliers, and stay close enough to the market to develop and defend a real point of view on where it's heading. Given current market conditions, Baseten needs a much stronger pulse on this part of the market so we can track pricing, stay close to the right relationships, and move fast the moment more capacity is needed. This is a senior, experienced hire who will also help pair with and develop 1-2 junior to mid-level teammates covering the same space. WHAT YOU'LL DO Build and maintain deep relationships across hyperscalers and strategic neoclouds (including NVIDIA cloud partners), working each organization from top to bottom rather than a single point of contact Maintain a consistent, "top of mind" presence with key accounts so Baseten is positioned to move quickly when capacity needs arise Evaluate capacity from the GPU to the data center — hardware generation, rack and node configuration, interconnect, power density, and cooling — so you know what a configuration will actually deliver, not just what the spec sheet claims Live in compute pricing daily: track rates by GPU generation, region, and contract term to keep Baseten inf

Machine LearningAIGoHR
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE This is a sourcing-first role, not a deal-closing role. Baseten needs someone who can build and maintain deep relationships across the long tail of data center and powered land providers, well beyond the handful of large, well-known players that everyone in the market is already competing for. This coverage area is a key differentiator for Baseten's broader compute strategy, so we're looking for the best possible person in this specific lane rather than a generalist. You'll own the full lifecycle of a sourcing relationship — from first outreach to ongoing management — not just the introduction. WHAT YOU'LL DO Build and maintain a comprehensive map of data center and powered land opportunities, with a particular focus on the long tail rather than the handful of major, oversubscribed players Own the full sourcing lifecycle for each relationship — from identifying and reaching out to new providers, through negotiation support, to ongoing relationship management — not just the initial introduction Develop and manage sourcing relationships across neoclouds, hyperscalers, brokers, and independent operators Quickly and independently evaluate new sites and spaces to determine fit and priority Prepare business cases and cost analysis to support new data center and powered land opportunities, partnering with Finance where needed Maintain accurate records of suppliers, contracts, and commercial terms so the team has a reli

Machine LearningAIGoProject Management
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Manager focused on TPUs at Baseten, you will lead the "engine room" for our non-NVIDIA accelerator fleet, architecting, securing, and optimizing the Google Cloud TPU (and broader emerging accelerator) capacity that powers our customers' AI workloads. You'll own the end-to-end journey of capacity management for this fleet, from securing large-scale TPU pod allocations to building the automation that ensures reliable uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering, with a specific focus on the TPU ecosystem. You will act as the fleet orchestrator for Google's TPU architecture, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics as we diversify beyond NVIDIA. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the latest generation of TPU hardware, like Google's Trillium (v6e) architecture, and partnering closely with the Model Performance (MP) team to ensure workloads are tuned for TPU-specific execution. EXAMPLE INITIATIVES The TPU Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's TPU clusters, including pod slicing and topology planning Global Workload Orchestration: Bui

PythonAWSAzureGCP
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -84.1%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI's Industrial Compute organization builds and operates the infrastructure required to train and serve frontier AI models. The Capacity Planning team connects rapidly changing research and product demand with the compute, networking, storage, power, data center, hardware, and operational resources required to make that demand executable. About the Role We are seeking a Technical Program Manager to build and lead capacity planning across OpenAI's large-scale AI infrastructure. You will translate uncertain workload demand into clear infrastructure requirements, allocation decisions, supply commitments, activation priorities, and long-range capacity strategies. This role sits at the intersection of research, engineering, infrastructure, finance, sourcing, deployment, and operations. You will create the planning models, operating cadences, governance mechanisms, and source-of-truth systems that allow teams to understand what capacity is required, what is available, what is at risk, and what decisions must be made. This is not a finance-only forecasting or reporting role. Success requires technical fluency across the infrastructure stack, strong analytical judgment, and the ability to move consequential decisions forward when requirements, timelines, and supply conditions change quickly. Key Responsibilities Own capacity-planning processes across near-term workload allocation, quarterly execution, and longer-range infrastructure horizons. Translate research, training, inference, and product demand into compute, accelerator, cluster, networking, storage, rack, power, and site requirements. Develop scenarios that make assumptions, confidence levels, constraints, sensitivities, and decision points explicit. Reconcile requested demand against contracted, delivered, installed, activated, and workload-usable capacity. Partner with research and engineering teams to understand workload priorities, technical dependencies, utilization patterns, and changing req

PythonSQLAWSRest
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a member of the Capacity Strategy & Operations team, you will sit at the intersection of supply intelligence, demand forecasting, and cross-functional execution, turning a complex, fast-moving hardware market into a predictable, reliable foundation for our customers and internal engineering teams. This is not a purely analytical role. You will own the end-to-end capacity planning process: from translating customer commitments and growth forecasts into concrete supply requirements, to coordinating fulfillment across vendors, finance, and the infrastructure team, to building the systems that make all of this repeatable and scalable. When supply is constrained and tradeoffs are unavoidable, you are the person in the room who can model the options, make a clear recommendation, and drive alignment fast. You are a strong fit if you have operated at the intersection of strategy and execution before — someone who is equally comfortable building a capacity model in a spreadsheet and running a cross-functional war room when a customer deployment is at risk. EXAMPLE INITIATIVES Demand-Supply Alignment Framework: Build and own the process that translates customer pipeline, signed commitments, and growth projections into a forward-looking GPU demand signal — so the team is never caught flat-footed when a customer scales faster than expected. Constrained Allocation Playbook: Define the decision framework for how Basete

Machine LearningAIGoExcel
B
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We’re looking for a Recruiting Coordinator to help create a seamless, welcoming, and well-organized interview experience for every candidate who engages with our team. You’ll work closely with our recruiters to coordinate both virtual and in-person interviews, support executive involvement when needed, and ensure candidates have everything they need during throughout their interview process. This role is ideal for someone who thrives on operational excellence, loves solving logistics problems on the fly, and brings both warmth and precision to every interaction. RESPONSIBILITIES Work closely with recruiters and hiring managers to coordinate interview loops and debriefs for candidates and the internal team members conducting interviews Ensure every candidate has a smooth, well-communicated, and positive experience Manage logistics for onsite interviews, including candidate arrival and workspace setup Proactively identify and solve day-of issues, including last-minute changes or scheduling conflicts Communicate clearly and promptly with candidates and internal teams about interview logistics and updates REQUIREMENTS 1+ year of recruiting or HR experience Detail-oriented and operationally strong—you know how to keep things moving Clear and professional written and verbal communication skills Personable and warm—you're great at making candidates feel welcome and supported Ability to think on your feet and respond to

Machine LearningAIExcelLogistics
R
📍 Foster City, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -87.5%
Quick readStrong listing-quality and freshness signals

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: We're hiring a Product Partnerships Manager to own Replit's most important strategic partnerships end-to-end: from identifying the opportunity to shipping the outcome and measuring the impact. This is a product-focused role where you will work with our consumer technology partners, such as Stripe, Shopify, Google, and the broader connector ecosystem. You'll work at the intersection of product strategy, ecosystem thinking, and deal execution. You'll need to be a serious Replit power user. You'll speak credibly with engineers. You'll negotiate with senior partner stakeholders. And you'll be the person accountable for turning ambiguous ecosystem opportunities into shipped product and business outcomes. What You'll Do Develop a clear point of view on Replit's partner landscape and build a prioritized pipeline of high-leverage opportunities across consumer tooling, AI tooling, cloud and infrastructure, and payments technology partners. Lead partner conversations from early exploration through joint product thesis, business case, commercial terms, launch plan, and post-launch iteration. Partner with Product, Engineering, Partner Engineering, Legal, Finance, Marketing, and Sales to turn agreements into shipped outcomes. Work directly with Partner Engineers to scope integrations, demos, prototypes, reference apps, and partner enablement assets. Define success metrics before every launch activation: retained usage, apps created, deployments, revenue, partner-sourced users, and use them to decide when to scale, iterate, or sunset a partnership. Build lightweight operating systems: partner scorecards, launch checklists, partner roadmap tracking, and repeatable frameworks for evaluating new opportunities. Represent

AIGoRustMarketing
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -84.1%
Quick readStrong listing-quality and freshness signals

About the Team: We are a small and fast-moving partnerships team that shapes and executes OpenAI’s most important collaborations. Your mission is to identify, structure, and scale partnerships that expand how businesses adopt and benefit from OpenAI. You will work across priority sectors, including legal, professional services, information services, and other B2B categories, while maintaining the flexibility to pursue the highest-impact opportunities as the market evolves. About the Role: You are a senior business development and partnerships leader with strong judgment, high ownership, and a track record of originating and closing complex partnerships. You can move from market strategy and executive engagement through deal development, launch, and growth. You combine strategic thinking, commercial discipline, and hands-on execution, and you can create clarity and momentum across multiple internal and external stakeholders. Key Responsibilities: Develop a portfolio strategy for high-impact B2B partnerships. Identify and prioritize partners based on customer reach, differentiated data or workflows, distribution, strategic value, and growth potential. Originate, structure, negotiate, and launch complex product and commercial partnerships. Build joint value propositions and business plans spanning integration, distribution, go-to-market, and customer adoption. Lead executive relationships and establish durable cross-functional partnership governance. Coordinate product, engineering, sales, marketing, legal, finance, security, policy, and operations to deliver partnership outcomes. Translate partner and market insight into product strategy and new partnership models. Track adoption, pipeline, revenue impact, strategic milestones, and partner performance, and communicate progress and risks clearly. Qualifications: 10+ years of experience in strategic partnerships, business development, enterprise technology, or platform ecosystems. Proven experience personally originatin

AWSRestAIGo
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Lead at Baseten, you will lead the "engine room" of the company, architecting, securing, and optimizing the global GPU fleet that powers our customers' AI workloads. You’ll own the end-to-end journey of capacity management, from securing multi-million dollar GPU clusters to building the automation that ensures 99.9% uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering. You will act as the fleet orchestrator for the world's most advanced chips, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the next generation of hardware, like NVIDIA’s Blackwell (B200) architecture. EXAMPLE INITIATIVES The B200 Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's first Blackwell GPU clusters. Global Workload Orchestration: Building "Multi-cloud Capacity Management" systems to move customer workloads seamlessly across regions to optimize cost and latency. Precision GPU Triage: Developing automated Go-based operators to identify, cordon, and repair unhealthy H100 nodes in under an hour. The Supply Chain of Intelligence: Partnering with lead

PythonAWSAzureGCP
🔔

Get new engineer 3 jobs in United States by email

Daily job updates · Unsubscribe anytime