Distinguished Engineer, Storage – AI Cloud — US, CA, Santa Clara. Apply via Workday.
Jobs in United States
Ai Compiler Engineer in United States
5,082 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai compiler engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Are you the person on your team who builds the agent everyone else ends up using? We're looking for an AI Engineer to join our Training Product team and do that at Baseten. You'll build AI-driven product features for the customers training and post-training frontier models on our platform, and you'll raise the ceiling on how Baseten itself uses AI internally, turning manual workflows into agentic ones that make every other team faster. You'll work directly with our research engineers to scope and build products, taking ideas from a research loop that already works internally to something customers can run themselves. This is a hands-on role with real autonomy. You'll pick the problems worth solving, build the harnesses, execution flows, and guardrails that make AI systems reliable, and own the results. If you've been shipping agents and want that to be the job, let's talk. EXAMPLE INITIATIVES: Take a look at these blog posts written by members of our team: Baseten Training: an autoresearch substrate Introducing Baseten Loops Harnesses are everything. Here's how to optimize yours. Building with NVIDIA Nemotron 3 Ultra and LangChain Deep Agents Code on Baseten RESPONSIBILITIES: Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to GA. Design the harnesses, execution flows, and guardrails that make AI systems reliable in production. Build internal autom
AI Research Engineer/Scientist — US, California, Santa Clara. Apply via Workday.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
$145K – $196.4K/yr
Drata is building the trust layer between great companies - automating compliance, managing risk, and helping organizations prove trust continuously as they scale. We're Dratanauts: a global crew of 600+ professionals united by a culture that rewards integrity, ownership, and raising the bar, no matter where in the world we're working from. Why Join the Drata Team? At Drata, you're not maintaining legacy compliance software - you're building the agentic AI platform defining what trust looks like for the next generation of companies. Here's what makes the work itself worth showing up for: Problems without a playbook: You'll work at the edge of AI and security, building agentic governance, continuous compliance, and real-time trust verification to solve problems that don't have an established answer yet. You're writing it as you go. Real ownership, not just process: Our values center on owning outcomes and raising the bar, not checking boxes. You're expected to have opinions and back them. A seat at the table: Your perspective is unique and valued. Open debate and diverse viewpoints are built into how decisions actually get made here, at every level. Growth at rocketship speed: Drata is scaling fast, which means scope grows fast too. High performers get more ownership, visibility, and experience. A crew, not just coworkers: Dratanauts consistently describe a "come as you are" culture with sharp, curious people—the kind of team that makes hard problems genuinely fun to solve. See what they say here and follow us on LinkedIn for company news, employee stories, and career updates. Job Summary: Drata is reimagining compliance as an intelligent, always-on experience — and AI is at the center of that vision. We are seeking a Senior AI Product Engineer to own the full-stack development of customer-facing AI features, embedded directly within our product teams. This is not a platform or infrastructure role. You'll translate the capabilities of LLMs, agents, and RAG pipelines
We’re looking for a Staff Software Engineer to help shape AI governance for developer tooling at Coder. This role sits on our AI Governance team, which builds and maintains two enterprise-grade components of Coder's AI governance stack. AI Gateway is a centralized LLM gateway that sits between coding agents and providers such as OpenAI or Anthropic, providing organizations with audit trails, token tracking, cost control, and centralized authentication. Agent Firewall wraps those agents with default-deny network policies, controlling which domains and methods they can reach inside workspaces. This team works across the full stack - from Go backend and React frontend to integrating with LLM provider APIs. Day to day, you'll be shipping features, hardening security boundaries, collaborating with enterprise customers on real-world policy needs, and contributing to Coder's open-source ecosystem. What you'll do here Design and build product features that push the standard for remote development in self-hosted environments Create and improve upon popular open source projects that integrate with VS Code, JetBrains, and other developer tools Champion best practices to both internal team members and external contributors Collaborate with Product and Design teams at Coder, as well as with partners like JetBrains, to execute key product integrations Document the design, implementation, and operations of systems for knowledge sharing within the team Work alongside Customer Success teams to support Coder’s enterprise user base Work with cutting-edge AI technologies to create seamless, painless developer experiences Rapidly iterate from prototype to implementation in a highly adaptive, reactive team environment What we're looking for 8+ years of full-stack experience writing code in a professional setting, with 1+ year(s) writing Go (ideally in current or most recent position) Proficiency in building distributed systems in Go Excellent verbal and written communication skills Excep
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As a deployment engineer at WRITER, you'll be at the forefront of expanding human capacity through superintelligence. This isn't just a technical role; it's a deeply influential one where you'll partner directly with our leading enterprise customers. Your expertise will be crucial in uncovering their unique business challenges and architecting AI-powered solutions that leverage our powerful platform and enterprise-grade LLMs. You'll transform complex needs into tangible, high-impact applications, creating champions and driving tangible business results. Your builder's mentality and passion for bringing cutting-edge AI into the hands of real users will directly shape the future of work for some of the world's largest companies. This is a critical role that directly impacts our customers' success and product evolution. You'll contribute significantly to WRITER's mission, working with a dynamic team to push the boundaries of what's possible with generative AI. Thi
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role AI research at WRITER isn't just about publishing papers — it's about building the scientific foundation that powers some of the most ambitious enterprise AI deployments in the world. As an AI research scientist, you'll be at the center of that work. You'll drive a high-impact research agenda focused on large language models, agentic reasoning, and the system-level capabilities that make AI genuinely useful at enterprise scale. This is a rare opportunity to do research that matters twice over — advancing the field and shipping directly into products used by hundreds of thousands of people every day. We're at an inflection point. Enterprises are moving from experimenting with AI to deeply embedding it across their operations, and WRITER's models are the engine making that possible. The work you do here — on post-training, planning, multi-step reasoning, and agentic workflows — will directly shape how the next generation of enterprise AI behaves, performs, and scales. You
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As a deployment engineer at WRITER, you'll be at the forefront of expanding human capacity through superintelligence. This isn't just a technical role; it's a deeply influential one where you'll partner directly with our leading enterprise customers. Your expertise will be crucial in uncovering their unique business challenges and architecting AI-powered solutions that leverage our powerful platform and enterprise-grade LLMs. You'll transform complex needs into tangible, high-impact applications, creating champions and driving tangible business results. Your builder's mentality and passion for bringing cutting-edge AI into the hands of real users will directly shape the future of work for some of the world's largest companies. This is a critical role that directly impacts our customers' success and product evolution. You'll contribute significantly to WRITER's mission, working with a dynamic team to push the boundaries of what's possible with generative AI. Thi
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As an AI engineer at WRITER, you'll be at the forefront of shaping how enterprises harness superintelligence. This isn't just about theory; you'll be building tangible AI solutions that power the future of work for hundreds of the world's leading companies. Your work will directly impact the performance, scalability, and ethical alignment of our cutting-edge LLMs and AI agents, enabling businesses to unlock unprecedented levels of productivity and innovation with AI that is truly grounded in their data. This role can be hybrid in our San Francisco, New York City, or Seattle hubs. You'll report to the head of AI engineering. 🦸🏻♀️ What you'll do Architect, develop, and deploy high-performance, scalable AI applications into production environments, ensuring robust integrations with our end-to-end platform. Drive the development of intelligent agents and AI-powered features, translating complex research into practical, impactful solutions for our customers. Coll
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As a deployment engineer at WRITER, you'll be at the forefront of expanding human capacity through superintelligence. This isn't just a technical role; it's a deeply influential one where you'll partner directly with our leading enterprise customers. Your expertise will be crucial in uncovering their unique business challenges and architecting AI-powered solutions that leverage our powerful platform and enterprise-grade LLMs. You'll transform complex needs into tangible, high-impact applications, creating champions and driving tangible business results. Your builder's mentality and passion for bringing cutting-edge AI into the hands of real users will directly shape the future of work for some of the world's largest companies. This is a critical role that directly impacts our customers' success and product evolution. You'll contribute significantly to WRITER's mission, working with a dynamic team to push the boundaries of what's possible with generative AI. Thi
What we're building Mutiny is the self-improving AI infrastructure for GTM teams to execute faster and close more revenue. Our ambition is to do for revenue velocity what Cursor and Claude Code did for engineering velocity. With Mutiny, everyone in sales and marketing gets a bench of GTM athletes that handle any work across their revenue motion and learn from what's actually moved their deals. In April we re-launched the product as an agent-first platform. Anthropic showcased us as a leader in AI GTM. MRR is growing more than 70% month-over-month, with customers like Uber, Rippling, and Snowflake. We're backed by Sequoia, YC, and Insight, and we're building a generational company. The opportunity Most CS roles are graded on how few problems the customer has. You'll be graded on how much they grow. As a growth strategist at Mutiny, you'll own a book of managed accounts and become the reason teams at customers like Rippling, Uber, and Snowflake get more out of AI than their peers. CS isn't a support function here. It's how we turn early traction into the customer base that defines a category. You'll build the playbooks as you go and shape what customer success looks like at an AI-native company from the ground up. This role is in person in New York City, five days a week. What you'll own Your book. Onboarding, adoption, retention, and expansion across your managed accounts. Every customer outcome runs through you. Customer expertise. Become the expert on how each customer's business works and where Mutiny fits in. Translate that into "aha" moments that move their roadmap. The AI motion. Teach customers how to use our agent-first platform to create assets that actually move the needle for their sales and marketing teams. Their team uses Mutiny because you made it obvious how. EBRs that move things. Run business reviews that tell a real story and drive action. Follow-ups that are crisp and lead somewhere. Renewals and expansion. Build the case six months before it comes
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. As a AI Optimization Specialist, Support, you will empower both our customers and Support team by building and maintaining the AI-powered knowledge that fuels our customer-facing chatbot and internal AI Copilot. You’ll collaborate closely with Support, Customer Education, Product, and Engineering teams to ensure our AI tools deliver accurate, helpful responses while enhancing customer experience and support efficiency at scale. This is a unique opportunity to shape the future of AI in support operations. You’ll own the strategy, design, optimization, and performance tracking of our chatbot and internal Copilot, ensuring both are trusted, high-performing tools. Our goal is to provide best-in-class Technical Support to our customers. Our team’s mission: “We make complex solutions seem simple, and are leaders in customer education”. We build strong partnerships with our customers through trust and transparency, for this reason our Technical Support Metrics are available publicly. What you’ll do as a AI Optimization Specialist, Support at Vanta: Leverage and Test AI-Driven Knowledge Collaborate with Customer Education to ensure AI tools (chatbot and Copilot) are drawing from the most effective and accurate content. Identify performance gaps, broken experiences, or edge cases that indicate content or structural improvements are needed. Provide feedback loops and insights to inform content updates based on how users interact with AI systems. Enhance the Internal AI Copilot Build and maintain internal knowledge resources (e.g., Guru cards, internal snippets) to help the support team resolve customer issues quickly and confidently. Ens
From $215.4K/yr
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. About the team The AI GTM team works with a cross-functional pod to accelerate Plaid’s footprint in AI. Our goal is to sell into and partner with AI companies of all sizes, from the largest model companies to small startups. This team works cross functionally, internally and externally, to ensure Plaid is the financial and identity provider of choice for AI companies; while also working closely with product and engineering to future proof our product set in an AI-driven ecosystem. In this role, you will serve as a critical link between our GTM teams (across sales, account management and partnerships), our product team, and the external AI market. You will be responsible for operationalizing our AI strategy, helping us win key deals, and ensuring our roadmap reflects the needs of our current customers and partners. What you’ll do You’ll work at the forefront of AI and fintech, driving deals and partnerships with some of the most innovative companies in the ecosystem. You’ll identify new opportunities, build relationships with AI-native partners, and translate emerging use cases into real business outcomes for Plaid. You’ll collaborate across teams to bring deals to life and help ensure Plaid shows up
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Marketing Operations (Ops) Team at Plaid builds the essential foundation that enables the Marketing function to operate efficiently, scale sustainably, and align with the company’s strategic goals. We focus on people, technology and process to drive operational excellence and optimize marketing performance. Our team is focused on creating and executing impactful marketing strategies that resonate with our target audiences, drive sustainable growth through high-quality pipeline generation, and enhance marketing efficiency. By integrating innovative technologies, we strive for seamless workflows and data-driven decision-making. We also continuously strengthen the foundation of a high-performing marketing organization, ensuring that processes and systems are in place to support long-term success. Responsibilities: Own the strategy and execution for integrating AI into marketing workflows, identifying high-impact opportunities to improve efficiency, scalability, and performance. Design, build, and deploy AI-powered agents and automations that support marketing workflows. Evaluate emerging AI tools and technologies, staying current on industry trends and translating new capabilities into practical ma
Other cities to consider
More places hiring for this role
Get new ai compiler engineer jobs in United States by email
Daily job updates · Unsubscribe anytime