ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Manager focused on TPUs at Baseten, you will lead the "engine room" for our non-NVIDIA accelerator fleet, architecting, securing, and optimizing the Google Cloud TPU (and broader emerging accelerator) capacity that powers our customers' AI workloads. You'll own the end-to-end journey of capacity management for this fleet, from securing large-scale TPU pod allocations to building the automation that ensures reliable uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering, with a specific focus on the TPU ecosystem. You will act as the fleet orchestrator for Google's TPU architecture, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics as we diversify beyond NVIDIA. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the latest generation of TPU hardware, like Google's Trillium (v6e) architecture, and partnering closely with the Model Performance (MP) team to ensure workloads are tuned for TPU-specific execution. EXAMPLE INITIATIVES The TPU Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's TPU clusters, including pod slicing and topology planning Global Workload Orchestration: Bui
Jobs in United States
Scaled Partnerships Manager in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current scaled partnerships manager jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a member of the Capacity Strategy & Operations team, you will sit at the intersection of supply intelligence, demand forecasting, and cross-functional execution, turning a complex, fast-moving hardware market into a predictable, reliable foundation for our customers and internal engineering teams. This is not a purely analytical role. You will own the end-to-end capacity planning process: from translating customer commitments and growth forecasts into concrete supply requirements, to coordinating fulfillment across vendors, finance, and the infrastructure team, to building the systems that make all of this repeatable and scalable. When supply is constrained and tradeoffs are unavoidable, you are the person in the room who can model the options, make a clear recommendation, and drive alignment fast. You are a strong fit if you have operated at the intersection of strategy and execution before — someone who is equally comfortable building a capacity model in a spreadsheet and running a cross-functional war room when a customer deployment is at risk. EXAMPLE INITIATIVES Demand-Supply Alignment Framework: Build and own the process that translates customer pipeline, signed commitments, and growth projections into a forward-looking GPU demand signal — so the team is never caught flat-footed when a customer scales faster than expected. Constrained Allocation Playbook: Define the decision framework for how Basete
About the Team The Emerging Products team is a lean, high-output product lab group that builds products at the forefront of model capabilities. We collaborate across all teams within the company, from research and infrastructure to consumer products. The team is responsible for identifying new product opportunities, building them quickly, dogfooding them internally, and then launching the successful products to users. We use data, user research, and analytics to inform our ideas, and make decisions on what experiments are worth iterating, stopping, or scaling. About the Role We’re looking for a senior, product-minded software engineer to own ambiguous 0-to-1 work from idea through prototype, validation, and handoff. This is a full-stack role with a strong frontend and product emphasis: you will build the interfaces and supporting backend systems needed to test new experiences quickly, while making sound architectural choices that enable successful concepts to scale. This role is based in our Mission Bay office in San Francisco. In this role, you will: Build and ship high-quality, product experiments across the full stack. Turn ambiguous user needs and emerging technical capabilities into testable product concepts, using research and metrics to guide iteration. Own technical direction for 0-to-1 projects, balancing speed, reliability, and a clear path from prototype to scalable product. Partner closely with design, product, research, and engineering teams to dogfood, evaluate, launch, and transition successful experiments. You might thrive in this role if you: Have a track record of building and shipping end-to-end products in fast-moving, startup, founder-led, growth, or other high-ownership environments. Bring strong frontend engineering skills and enough backend and systems depth to make sound full-stack architectural decisions. Pair product intuition with evidence, using user research and product data to identify opportunities and make pragmatic tradeoffs. Operat
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? The Data Infrastructure team at Cohere is responsible for the storage and data movement layer underlying every model training run. We're building the unified storage layer that feeds our training workloads. It needs to serve petabytes of training data and model checkpoints fast enough to keep thousands of GPUs busy across several training clusters. In this role, you’d have an opportunity to build this system from the ground up. You’d be a key contributor, working on a problem few teams have had to solve at this scale. In this role, you will: Design, build, and operate the distributed storage system that feeds model training and evaluation. Run this system multiple on Kubernetes clusters at petabyte scale. Work with researchers and training-infra teams on how jobs actually read and write data, and turn that into throughput, latency, and durability requirements Work through the networking, I/O, and consistency problems of moving large datasets and checkpoints across regions and backends, with GPU idle time and time-to-insight as the measures of success You may be a good fit if you have: Strong storage fundamentals,
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring a Senior Software Engineer for our Anti-Abuse team within Security Foundations. This team is responsible for protecting Snowflake and our customers from abuse on the Snowflake platform — building foundational security primitives across a complex multi-cloud environment to combat fraud, account takeovers, data exfiltration, and AI security threats at the scale of Snowflake’s hyper-growth. AS A SENIOR SOFTWARE ENGINEER, SECURITY FOUNDATIONS AT SNOWFLAKE, YOU WILL: Design, build, and scale security-by-default capabilities that protect Snowflake products across a complex multi-cloud environment Develop intelligent security platforms and services that leverage AI/ML and risk scoring to identify, prioritize, and respond to security threats and policy violations Build real-time detection and prevention systems to defend against account compromise, AI abuse, and data exfiltration through adaptive enforcement and granular policy controls Create security intelligence and monitoring capabilities that detect suspicious account & user activity, generate actionable insights, and enable timely response Partner across Security, Engineering, and Product teams to define security strategy, influence product design, and drive adoption of secure engineering practices Raise the
About the Team OpenAI's Environmental, Health & Safety (EHS) team partners across the company to enable safe, responsible growth. This role will be a senior EHS partner to our robotics operations as our operating footprint expands. You will work closely with Robotics, Engineering, Operations, Facilities, Workplace, Construction, Security, People, Legal, and external partners to integrate safety into how our facilities are designed, built, staffed, and operated. About the Role We are seeking a Senior Manager, EHS - Robotics to provide senior, hands-on safety leadership across multiple sites. This role is for an experienced EHS leader who can operate independently in a fast-moving technical environment, anticipate risk before it becomes an incident, and build practical safety systems that scale with the business. Our robotics operations are scaling quickly, with a growing mix of construction, commissioning, workforce expansion, and around-the-clock operations. This creates a complex and evolving risk environment and requires strong preventive planning, consistent field presence, and sustained incident-management oversight. This leader will own site-level EHS strategy and execution for the robotics portfolio while partnering with the broader EHS team on company-wide standards and programs. The successful candidate will look around corners: planning for future ramps, identifying requirements early, influencing design and operating decisions, and creating durable systems that enable teams to move quickly without compromising safety. In this role you will: Serve as a senior EHS partner across robotics R&D operations, owning site-level safety strategy, priorities, and execution as the footprint scales. Anticipate EHS requirements for new operations, equipment, processes, construction phases, staffing ramps, and 24/7 operations, and translate them into practical plans before work begins. Lead proactive risk identification and control across robotics activities, incl
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We’re looking for a high-performing strategic finance professional to join our growing GTM Finance team. Our business grows with our customers' usage, which makes the finance function highly strategic at Baseten: growth, pricing, margin, and capacity decisions are business model decisions. You'll sit at the center of them, partnering directly with GTM leadership and reporting into a finance team with a seat at the table for the calls that shape the company's trajectory. This role is ideal for someone with 3 to 7 years of experience across strategic finance, investing, and/or investment banking who wants broad exposure to company-building inside a fast-scaling AI infrastructure company. Experience at a usage-based software company is a plus. RESPONSIBILITIES Own financial planning, forecasting, and budgeting processes for the GTM org Build and maintain financial models across revenue, S&M spend, headcount, and strategic bets Analyze the metrics that define a usage-based business – ARR, gross margin, consumption trends, retention, and GTM efficiency Partner with GTM leaders to set targets, evaluate growth initiatives, shape pricing, and design sales compensation Help prepare board materials, investor updates, and fundraising analyses Improve financial reporting, dashboards, and operational rigor so our infrastructure scales as fast as our revenue Work cross-functionally to turn ambiguous business questions int
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. As a Premium Support Engineer at Replit, you’ll be the front line for our highest-value customers — delivering fast, expert, and reliable technical support when it matters most. You’ll handle complex product issues, guide customers through critical incidents, and ensure every interaction meets the highest standard of quality and speed. Replit is at the forefront of AI-driven software development, and how we support customers is constantly evolving. You’ll play a critical role in shaping how Premium Support adapts to new products, new customer expectations, and AI-assisted workflows, operating effectively in ambiguity and driving clarity for your team. You’ll combine deep technical troubleshooting with calm, confident communication to keep builders moving — whether it’s an enterprise team deploying at scale or a top-tier developer relying on Replit to power their business. In this role you will: Provide swift, high-priority support to Premium customers, responding within strict SLAs. Diagnose, reproduce, and resolve complex technical issues across the Replit platform. Escalate and track high-impact issues with Product and Engineering, ensuring timely fixes and transparent communication. Lead customer-facing communications during outages or incidents. Identify recurring issues and collaborate internally to reduce time-to-resolution. Contribute to internal tooling, automation, and documentation that improves team efficiency. Partner with Engineering, Product, Sales and other internal teams to ensure Premium customers receive a consistent, high-quality experience. Help onboard and mentor other support engineers, raising the team’s overall bar for responsiveness and quality. Required skills and experience: 3+ years in techn
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. As a Premium Support Engineer at Replit, you’ll be the front line for our highest-value customers — delivering fast, expert, and reliable technical support when it matters most. You’ll handle complex product issues, guide customers through critical incidents, and ensure every interaction meets the highest standard of quality and speed. Replit is at the forefront of AI-driven software development, and how we support customers is constantly evolving. You’ll play a critical role in shaping how Premium Support adapts to new products, new customer expectations, and AI-assisted workflows, operating effectively in ambiguity and driving clarity for your team. You’ll combine deep technical troubleshooting with calm, confident communication to keep builders moving — whether it’s an enterprise team deploying at scale or a top-tier developer relying on Replit to power their business. In this role you will: Provide swift, high-priority support to Premium customers, responding within strict SLAs. Diagnose, reproduce, and resolve complex technical issues across the Replit platform. Escalate and track high-impact issues with Product and Engineering, ensuring timely fixes and transparent communication. Lead customer-facing communications during outages or incidents. Identify recurring issues and collaborate internally to reduce time-to-resolution. Contribute to internal tooling, automation, and documentation that improves team efficiency. Partner with Engineering, Product, Sales and other internal teams to ensure Premium customers receive a consistent, high-quality experience. Help onboard and mentor other support engineers, raising the team’s overall bar for responsiveness and quality. Required skills and experience: 3+ years in techn
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake empowers more than 10,000 organizations worldwide to unlock the full potential of AI-driven data, shaping the future of intelligent business. As companies embrace transformative AI and data workloads, controlling who can access what — and proving it — is not just important: it's essential. Join the Security Platform team at the heart of this revolution, where YOU will lead the charge in building the identity and access foundation that customers trust to protect their most sensitive data and AI workloads. Identity and access management is both the foundation and the catalyst for Snowflake's bold mission — and your work will directly shape how thousands of organizations govern access at global scale. As a Senior Product Manager in our fast-moving, hands-on, customer-focused team, you'll drive innovative IAM and authorization solutions that give customers precise control over who can see, do, and build on their data. You'll have ownership of delivering cutting-edge capabilities including role-based access control (RBAC), privilege management, identity governance, authentication policy enforcement, and AI-powered access intelligence experiences. You will team up with talented engineers, collaborate across functions, and shape the roadmap for IAM products that not only
About the Team OpenAI is building the world’s most advanced AI infrastructure ecosystem. The Site Readiness & Development team owns the upstream diligence and development work required to convert powered-land opportunities into executable infrastructure options. About the Role The Site Selection Lead sets the site-selection strategy for powered-land and greenfield/brownfield opportunities. This role defines, with cross-functional partners, what makes a site attractive, executable, and scalable, then turns that shared rubric into portfolio choices and deployment plans. Key Responsibilities Own the site-selection strategy across priority geographies, including market maps, pipeline segmentation, source refreshes, live-location tracking, and portfolio prioritization. Convene Power, Land, Development, Engineering, Construction, Environmental, Community, Commercial, Legal, and Finance partners to define the shared rubric for a high-quality site. Translate that rubric into clear screening criteria across power readiness, land control, timing, scale, cost, community risk, environmental constraints, AHJ path, budget, schedule, and counterparty credibility. Orchestrate Commercial, Legal, Finance, Power, and Development teams to shape exclusivity, initial terms, site-control strategy, funding assumptions, and the path to deployment before deeper commitment. Build a deployment plan for top candidates that identifies capacity, sequencing, capital implications, critical-path approvals, risks, owners, and stage-gate decisions. Lead conceptual site planning and test-fit screening using boundaries, map layers, buildings, substations, roads, parking, laydown, stormwater, utility corridors, acreage, and estimated load. Maintain a single decision-support view of pipeline stage, site health, power readiness, finance review, maps, project evidence, outstanding source data, owners, deadlines, and escalations. Present executive-ready advance, defer, or stop recommendations with clear
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a Delivery Director, Capacity programs for our on-premises data center builds and neo cloud (GPU cloud) delivery programs. This is a high-visibility, execution-critical role sitting at the intersection of infrastructure engineering, capacity planning, vendor/partner management, and customer delivery. You will own the end-to-end delivery lifecycle for large-scale compute infrastructure — from initial site/capacity commitments through power, networking, and hardware bring-up, to production-ready GPU/compute capacity landing in the hands of internal teams or customers. You'll be the person who turns ambitious infrastructure roadmaps into predictable, on-time, delivery. RESPONSIBILITIES Own delivery of on-prem infrastructure builds — colocation expansions, power/cooling readiness, rack-and-stack, network fabric bring-up, and hardware acceptance testing — coordinating across colo providers and partners, network engineering, hardware ops, and vendor teams. Drive neo cloud delivery programs — manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps, SLAs, and go-live readiness. Build and maintain master delivery schedules across concurrent, multi-site, multi-vendor programs, integrating power/shell timelines, hardware lead times, logistics, and software/platform readiness into a single critical path.
About the Team The Industrial Security team enables OpenAI’s classified and national security work by building secure, compliant programs that support government customers and protect sensitive information. We partner across Security, Legal, People, Facilities, Engineering, IT, and program leadership to translate complex government requirements into practical operating models that allow important work to move forward responsibly. Our work spans personnel security, classified facilities, special programs, government customer engagement, and the safeguards required to operate in highly regulated environments. About the Role As a Senior Special Programs Security Manager, you will lead security operations for OpenAI’s Sensitive Compartmented Information (SCI) and Special Access Programs (SAP), serving in customer-designated roles such as Contractor Special Security Officer (CSSO) or Contractor Program Security Officer (CPSO). You will help shape how OpenAI operates across the Intelligence Community and other national security environments, building the relationships, processes, and security foundations required to support classified work at scale. We’re looking for someone who brings deep special programs expertise, sound operational judgment, and the credibility to work directly with government customers and senior internal stakeholders. This is a hands-on leadership role for someone comfortable navigating ambiguity, helping build the function, mentoring future team members, and finding compliant paths forward when mission needs and security requirements intersect. This role is based in San Francisco, CA, or Washington, DC. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. Additional onsite presence and occasional travel may be required based on classified program, facility, and government customer needs. This position requires active Top Secret eligibility, current SCI access, and a current counterintelli
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Join WRITER's security team as a staff detection and response engineer and help protect the AI infrastructure that's transforming how the world works. You'll build sophisticated detection systems that identify attacks targeting our AI platform, training data, and model deployments while creating automated response capabilities that scale with our explosive growth. This isn't just traditional security work – you're defending cutting-edge AI/AGI systems against adversaries who are evolving their tactics as fast as AI itself advances. This role combines hands-on security engineering with strategic thinking to stay ahead of novel threats that don't exist in textbooks yet. You'll be the operational arm of our security function, translating threat intelligence into real-time detections, coordinating incident response across multiple teams, and hunting for sophisticated attacks across GPU clusters and distributed training environments. If you're excited by the challen
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. Build the operational and technical skill set that runs global trade At Flexport, we believe global trade can move the human race forward. We're shaping the future of a $10T industry with technology and people who understand it end to end. The Rotational Development Program (RDP) exists because that combination — real supply chain fluency plus the ability to build and automate — doesn't exist at scale in the market today. So we're building it ourselves. This is an 18-month program, not a rotation for rotation's sake. Every assignment is live, revenue-generating work. Every participant leaves with a certification in AI and low-code tooling running alongside their operational and commercial training. And every graduate places into a role with real scope: primarily Account Management, with paths into Automation Engineering, Forward Deployed Engineering/Consulting, Sales Engineering, or a traditional Ops/Sales/Customs track. What you'll do You'll rotate through three functions, six months each, with AI/low-code curriculum running continuously across all three: Gateway Ops (Air or Ocean). Own the shipment lifecycle. Execute bookings, manage exceptions, and hold data int
Other cities to consider
More places hiring for this role
Get new scaled partnerships manager jobs in United States by email
Daily job updates · Unsubscribe anytime