ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. At Baseten, we are building the global operating system for distributed, heterogeneous AI hardware. We believe that as LLM and multi-modal workloads scale, the network is the computer. We are looking for foundational engineers to lead our GPU Networking efforts, making RDMA a first-class building block in our infrastructure and unlocking the next generation of distributed inference optimizations. THE OPPORTUNITY Networking and compute are no longer separate disciplines; they are converging. The massive throughput of H100, B200, and NVL72 architectures enables and demands a new approach where communication is co-optimized alongside computation. We are entering an era where the network is an active accelerator, leveraging smart hardware offloads and direct interconnects to ensure that data movement operates at wire-speed. In this role, you will go beyond network configuration to architect the software fabric that unifies thousands of GPUs into a cohesive operating system. While you will leverage the best of the open-source ecosystem, you won't be limited by it. Where off-the-shelf solutions stop, you will build from scratch, engineering the primitives required to co-optimize communication and compute for Disaggregated Serving, Wide Expert Parallelism (WideEP), and lightening cold starts. WHAT YOU'LL DO Make RDMA First-Class: You will work on integrating RDMA/RoCE/InfiniBand capabilities directly into our inference stack,
Jobs in United States
Ai Deployment Manager in United States
5,246 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai deployment manager jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Software Engineer at on the Training Infrastructure team, you'll architect and lead development of our training platform, supporting top tier research engineers and model developers. You'll make key technical decisions for the infrastructure enabling developers to deploy, scale, and monitor their workloads with high performance and reliability. You’ll own scheduling, storage, networking, reliability, and observability of technical systems in the training stack EXAMPLE INITIATIVES Take a look at what we’ve built so far: Overview of the product so far Training docs overview Story of the Training product Research we've done RESPONSIBILITIES Design and architect scalable infrastructure systems for our ML training platform (e.g. scheduling, storage, and networking) Partner closely with developers and research engineers to translate complex training requirements into technical solutions Design and architect a global training scheduler Design and architect reinforcement learning systems and continuous learning pipelines Drive long-term improvements to improve reliability of systems and velocity of development Partner closely with SRE and Capacity teams to unlock state of the art training infrastructure Make critical architectural decisions balancing performance with system reliability Lead technical discussions and mentor junior engineers on infrastructure best practices Contribute to long-term technical strateg
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Are you passionate about ensuring the highest quality for cutting-edge generative AI applications? As a software quality engineer at WRITER, you'll play a critical role in shaping the reliability, performance, and trustworthiness of our AI-powered work orchestration platform. You’ll be at the forefront of defining and implementing rigorous quality strategies for our enterprise-grade LLMs and AI agents, directly impacting how hundreds of global companies unlock transformational value through AI. This is a unique chance to dive deep into the unique challenges of AI quality assurance and make a tangible difference in a rapidly evolving field. This is a hybrid role based out of our London, San Francisco, Seattle, and New York City hubs. You will report directly to the director of engineering. 🦸🏻♀️ What you'll do Define and implement comprehensive quality assurance strategies and test plans for our AI agents and LLM-powered applications, ensuring exceptional prod
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role We’re seeking a highly skilled fullstack software engineer to join our engineering team building advanced AI-driven agent systems that execute autonomous workflows, orchestrate multi-step tasks, and extend human capacity across enterprise applications. In this role, you’ll play a key part in designing, building, and scaling next-generation AI agents that integrate with enterprise data and services to solve real-world problems. You will collaborate with cross-functional teams to turn complex agent concepts into production-ready systems. 🦸🏻♀️ What you'll do Design, implement, and maintain scalable, secure agent-driven services and systems that autonomously accomplish tasks using modern AI frameworks. Develop and enhance robust infrastructure and high-throughput APIs, focusing on core agent capabilities such as memory, communication channels, skills, intelligent decision logic, security and workflow management. Integrate agent capabilities with backend services
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role At WRITER, our enterprise customers are building mission-critical AI applications that transform how their businesses operate. As a senior support engineer, your top priority is helping our developer personas and technical users succeed with our platform. You'll act as a leading voice for our enterprise support function, handling intricate customer issues directly while working in a complex and ambiguous environment to drive impactful value across our customers' businesses. You'll collaborate closely with Customer Success, Education, Product, Engineering, and Sales to create great experiences and help customers get the most out of the platform. By leveraging an automation-first mindset, you'll scale our support operations and use our own AI tools to solve problems faster and more efficiently. This role is available for hybrid work in our New York City and London hubs. and you will report directly to the manager of support engineering. 🦸🏻♀️ What you'll do Own
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role WRITER is looking for a senior paid marketing manager to lead our paid media programs. In this role, you'll use AI to personalize and scale paid marketing across channels, with a focus on account-based marketing (ABM) and supporting enterprise sales cycles. You'll work cross-functionally with the broader marketing team to keep paid programs aligned with our overall strategy. The ideal candidate will have a strategic, data-driven mindset, with deep expertise in full-funnel performance marketing. You’ll bring a validated POV on digital planning, media execution, ad operations, and stakeholder management, paired with the passion of a marketer, the insight of a strategist, and the drive of a doer. This position reports to the director of lifecycle marketing. 🦸🏻♀️ What you’ll do Build and manage paid marketing programs across social, search, and display channels Use AI tools to personalize campaigns and scale them efficiently Develop and execute account-based mar
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role We're looking for a hands-on, detail-obsessed global payroll manager to join our team. This is a high-ownership individual contributor role that sits at the intersection of finance and human resources. You will own end-to-end payroll for a multi-entity, multi-currency, multi-regulatory environment spanning the US, UK, Canada, Singapore, the UAE, and a growing set of Employer of Record (EOR) countries across Europe. You'll also manage contractor payroll and oversee 401(k) and pension administration. This is a role for someone who is energized by complexity — who sees a fragmented global payroll landscape and wants to bring order, accuracy, and scale to it. 🦸🏻♀️ What you’ll do End-to-end payroll execution across all of WRITER’s active entities and pay schedules, including multi-state US payroll with varying pay frequencies, and commission cycles; UK PAYE with HMRC filings, National Insurance contributions, pension auto-enrolment, and statutory pay calculations
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Every company grows differently, and at WRITER, our growth is directly tied to empowering our users to create better content, faster, and at an unprecedented scale. As a strategic solutions architect, you'll be at the forefront of this mission, working directly with our largest and most strategic prospects to identify, validate, and build innovative agentic AI solutions that unlock massive business value. This isn't just about selling a product; it's about deeply understanding complex enterprise needs and architecting a future where AI transforms how our customers operate. You'll be instrumental in shaping how the world's leading companies adopt and scale AI, making a tangible impact on both their success and WRITER’s continued leadership in the enterprise generative AI space. This is a full-time, hybrid role based out of our Chicago hub. You will report directly to the Director, solutions architecture. 🦸🏻♀️ What you'll do Drive strategic technical discovery
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role At WRITER, our mission to expand human capacity with superintelligence relies on a foundational truth: our platform must be available, performant, and reliable, 24/7. As an Infrastructure engineer, you'll be at the heart of making this a reality, impacting every enterprise customer who trusts us with their AI-powered workflows. This isn't just about keeping the lights on; it's about pushing the boundaries of what's possible, proactively identifying and solving complex systemic challenges, and laying the groundwork for our rapid growth and the evolving demands of enterprise generative AI. You'll build resilient systems, automate across the stack, and champion reliability best practices, directly enabling our ambitious product roadmap and ensuring our customers always have access to the powerful tools they need. This is a hybrid position, based out of our New York City, San Francisco, Seattle, or London hubs. You'll report to our director of engineering. 🦸🏻♀️
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role We're looking for an exceptional software engineer to join our rapidly evolving team at WRITER. In this pivotal role, you'll be at the forefront of expanding human capacity by building the next generation of AI-powered solutions that transform how leading enterprises operate. You'll dive deep into developing a state-of-the-art platform that leverages cutting-edge generative AI technologies, from large language models to sophisticated agentic workflows, delivering seamless, scalable, and secure applications that redefine enterprise productivity. This is an unparalleled opportunity to make a tangible impact, shaping the future of AI and contributing to a product that’s changing how the world works. This role is hybrid, based out of our San Francisco, New York City, or Seattle hubs. You'll report to our senior director, engineering . 🦸🏻♀️ What you’ll do Design and deliver secure, scalable AI integration platforms that connect enterprise systems and power missio
About Pinecone: Pinecone is the leading vector database for building accurate and performant AI applications at scale in production. Pinecone’s mission is to make AI knowledgeable. More than 109,000 customers across various industries have shipped AI applications faster and more confidently with Pinecone’s developer-friendly technology. Pinecone is based in New York and has raised $138M in funding from Andreessen Horowitz, ICONIQ, Menlo Ventures, and Wing Venture Capital. About the Role/Team: As a Technical PMM , you will help developers, software engineers, and machine learning scientists understand how Pinecone enables knowledgeable AI applications. You will accomplish this through messaging and positioning, GTM strategy, launch planning, and enablement for key product features. You will work closely with Product, Sales, Developer Relations, and Marketing teams to ensure our customers understand the value of our platform. Responsibilities: Product Launches & GTM Strategy: Develop launch plans for product releases, align with product teams, and manage GTM execution. Lead core launch teams with cross-functional members (documentation , communications, developer relations, product, growth marketing, etc). Measure customer impact and adoption post-launch, iterating on strategies as needed. Positioning & Messaging: Craft compelling, technically accurate messaging for AI-powered search and retrieval solutions. Build product narratives for enterprise and developer audiences that differentiate Pinecone from competitors. Develop and maintain positioning for new product capabilities Customer & Market Insights: Partner with customers to develop case studies, gathering business impact metrics and architectural insights. Engage in customer interviews and research to refine messaging and identify key value drivers. Content Development & Enablement: Develop internal and external enablement materials, including sales training, pitch decks, and GTM enablement sessi
From $244K/yr
We're on a mission to build the best platform in the world for engineers to understand and scale their systems, applications, and teams. We operate at high scale—trillions of data points per day—allowing for seamless collaboration and problem-solving among Dev, Ops and Security teams globally for tens of thousands of companies. Our engineering culture values pragmatism, honesty, and simplicity to solve hard problems the right way. The Team: As organizations rapidly invest in AI applications and build out AI labs, telemetry volumes are growing exponentially and costs are becoming unpredictable. From LLM interactions to agentic workflows, AI systems generate unpredictable streams of logs, driving up costs and making it harder to maintain efficient observability. These challenges are critical for organizations in regulated industries with strict data residency requirements, where data must remain within controlled environments. Datadog’s Bring Your Own Cloud (BYOC) team is reimagining what observability and security look like at petabyte scale in the AI era. The Opportunity: The Group Product Manager - Bring Your Own Cloud (BYOC) role is responsible for defining and bringing to market the next generation of telemetry analytics and insights capabilities in an AI-first environment. This role is highly technical and creative in nature as you will envision novel ways to enable customers to cost-effectively explore, analyze and report over petabytes of data through a welcoming and easy-to-use interface. You will partner with various teams to take advantage of BitsAI capabilities and surface critical insights on volume usage and retention for popular use cases. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a team
About the Team OpenAI’s Agentic Data Science team helps shape how AI agents are built, deployed, and improved across our products. We partner with product, engineering, research, and security teams to define meaningful measures of success, understand how our systems behave in the real world, and translate evidence into better decisions. As AI agents become more capable, they can write and execute code, access sensitive systems, and complete increasingly complex tasks with greater autonomy. These capabilities create powerful opportunities to improve cybersecurity, but they also introduce risks that traditional security tools and processes were not designed to address. Meeting this moment requires new ways to measure security, evaluate defenses, and distinguish genuine risk reduction from friction that slows users down. About the Role We are looking for a senior data scientist to help define what effective cybersecurity looks like in the age of AI agents. You will work across OpenAI’s Security organization and cybersecurity product teams to measure emerging risks, improve internal security controls, and shape AI-powered security products. The problems are foundational: How do we know whether an agent’s security controls are effective? Which safeguards meaningfully reduce risk, and which create unnecessary friction? When an AI system identifies a potential vulnerability, how do we determine whether the finding is accurate, actionable, and ultimately resolved? How do we detect anomalous behavior or risky access when the systems themselves are changing rapidly? You will report into Data Science while partnering closely with Security, Cyber Product, Engineering, and Research. This is a high-ownership role for someone who can establish a new analytical discipline, operate across organizational boundaries, and turn ambiguous security challenges into measurable improvements. In This Role You Will Define how we measure AI-agent security. Establish metrics and evaluation frame
About the Team The B2B Content team helps business audiences understand what AI makes possible and how to put it to work. We develop narrative, editorial, demo, and adoption programs that connect OpenAI’s products to concrete business value. We partner closely with Product Marketing, Demand Generation, Customer Education, Sales, and other teams to create content that is useful on its own and supports the broader customer journey. We value people who are eager to learn, have a bias for action, and look beyond the immediate deliverable. We build the tools, systems, and shared practices that help the whole team produce better work over time. About the Role We’re looking for a Content Marketing Manager, Demand Generation to turn foundational B2B messaging, flagship narratives, research, launches, events, and product workflows into the right content for different industries, lines of business, audiences, and stages of the customer journey. You’ll create content across webinars, guides, ebooks, blogs, customer proof, and decision-enablement materials. You’ll also build the modular frameworks that allow strong source material to be adapted and reused without recreating the work from scratch. You’ll work closely with Product Marketing, which owns product positioning and messaging, and Demand Generation, which owns audience and channel strategy. You’ll bring strong content judgment to that partnership: helping determine what should be created, how it should be structured, and how it can work across different contexts while preserving a coherent central narrative. This is a senior individual-contributor role for a hands-on marketer, writer, and editor. It does not include people management or ownership of the broader content portfolio. This is a hybrid role and will be based in San Francisco, Monday through Wednesday. Thursday and Friday are work-from-home days, with occasional in-office attendance as needed for live webinars. In this role, you will Partner with Product Marke
About the Team Our economics team is continuously working to improve our understanding of an AI-driven economy. About the Role We are seeking a highly technical Economist to join the OpenAI Economic Research team studying the real-world economic impacts of AI. This role is designed for economists with up to 5 years of professional experience post-Ph.D. who are interested in using novel, large-scale datasets to study how AI is reshaping economic systems. We are looking for candidates with deep expertise in at least one core domain relevant to AI’s economic impact, and an interest in contributing to a broader research agenda spanning labor markets, firm behavior, market dynamics, and macroeconomic change. This is an individual contributor role where the candidate will organize and execute on their own data-oriented projects. You will work at the intersection of economic research, data science, and public policy to produce rigorous empirical work that informs decision-makers across the public, industry, and government. Research Areas of Interest We are particularly interested in candidates with demonstrated expertise in one or more of the following areas: Economic Measurement of AI Impact (e.g., adoption trajectories, labor market transitions, productivity growth, and forecasting/scenario modeling for AI-driven economic change) Macroeconomic Implications of AI (e.g., productivity, technology diffusion, economic growth) AI and the Labor Market (e.g., employment, wages, job search, task-level impacts, skill acquisition) Applicants are not expected to have experience across all domains. We aim to build a team with complementary strengths across these areas. In this role, you will: Design and execute empirical research using large-scale observational or experimental data. Apply causal inference and/or structural modeling techniques to study AI-driven economic change. Collaborate with cross-functional teams to translate research questions into testable frameworks and applic
Other cities to consider
More places hiring for this role
Get new ai deployment manager jobs in United States by email
Daily job updates · Unsubscribe anytime