Jobs in United States

Machine Learning Intern in San Francisco

238 active opportunities · Updated October 2026

Explore current machine learning intern jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building for talent density. We believe attracting and retaining exceptional people, and ensuring they feel recognized and valued for their impact, is core to becoming the best place to work. Compensation is a critical lever in that mission. As our Compensation Manager, you will own compensation programs company-wide. You’ll be a trusted advisor to senior leaders, shaping our compensation philosophy, leveling framework, and equity programs to ensure we remain competitive, principled, and performance-oriented as we scale. This role blends strategy and execution: designing clear, fair systems while moving quickly in a high-growth environment. RESPONSIBILITIES Own and evolve Baseten’s company-wide compensation strategy, philosophy, and programs. Collaborate with leadership and HRBP to create and evolve job architecture and leveling frameworks. Build and maintain compensation bands. Conduct regular market benchmarking to ensure comp bands and strategy remain competitive in a fast moving industry. Partner closely with Talent to design and approve competitive new hire offers, advising on negotiation strategy within our compensation principles. Lead bi-annual leveling and compensation review cycles to ensure market competitiveness and reward high performance across teams. Manage new hire equity grants in partnership with Finance and Legal. Design and administer a thoughtful equity refresher program for ten

Machine LearningAIGoRust
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE At Baseten, we’re looking for a Technical Program Manager to drive our most complex, cross-cutting infrastructure programs. This role will operate across all domains of AI infrastructure, from the GPUs up to the multi-cluster orchestration layer. This is an execution-first role. The work is less about owning a single system and more about imposing order on ambiguity: standing up the right structures, driving decisions to closure, and making sure nothing falls through the cracks across dozens of stakeholders. If you take satisfaction in turning a chaotic, half-defined initiative into a predictable, well-governed program, this role is for you. RESPONSIBILITIES Own complex migrations end to end. Lead large-scale infrastructure migrations across teams and domains. This will involve scoping the work, sequencing dependencies, managing risk, and driving them to completion without surprises. Drive process across infrastructure. Establish and run the operating rhythms that keep programs healthy: planning cadences, status reporting, decision logs, risk reviews, and escalation paths. Make the process light enough that teams adopt it and rigorous enough that it actually works. Help managers build the right structures. Partner with engineering managers and leads to design the team structures, ownership boundaries, and working models a program needs to succeed. Spot gaps in accountability before they become problems. Own fo

Machine LearningAIGoRust
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE This role sits at the frontier of our research agenda. You will pursue open problems at the intersection of post-training methodology and performant inference, and then collaborate with research engineering to translate findings into production systems. A meaningful portion of your time will be dedicated to research that deepens our understanding of how models learn, alignment, and architectural efficiency — questions that may not have immediate product application. The remainder will be directed toward research that solves concrete problems for Baseten's platform and customers, who are the fastest growing AI companies in the world like Cursor, Lovable, and Notion. We are looking for someone with sharp research taste and genuine creative instinct for problem selection. Someone who can identify questions that matter, design clean experiments to answer them, and push the state of the art. The environment here is not theoretical, but rather research that can be validated with eager customers who are serving billions of tokens a second. RECENT RESEARCH Towards infinite context windows: neural KV cache compaction Dense, on-policy or both? Repeated kv cache for long-running agents Distillation without the dark – replicating black-box on-policy distillation on Baseten RESPONSIBILITIES Define and pursue a research agenda spanning both foundational and applied work, with the applied component connected to Baseten's pla

Machine LearningAIGo
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're looking for a senior social media manager to own how Baseten shows up on social. Our audience is ML engineers, infrastructure teams, and technical founders, and most of them meet us first on X or LinkedIn, around a model launch, a benchmark, an open source release, or a customer result. This role decides what that first impression is. This is a senior individual contributor role. You set the strategy and you write the posts. Day to day you'll work with product marketing, comms, design, our engineers, and our founders. This role relies on technical credibility. You don't need an engineering background, but you do need to understand what we're claiming and why it matters. We post about latency, throughput, and GPU cost, and we hold ourselves to getting those details right. RESPONSIBILITIES This role is the face of the Baseten brand on our social channels and builds our direct line of communication with the community across X, LinkedIn, YouTube, and the communities where our audience already spends time. Track the conversation across AI and open source, and move quickly when we have something useful to add. Set the social strategy: what we post where, how each account grows, and how we measure it, with a clear point of view on which channels deserve investment and which don't. Translate technical work into posts worth sharing: model launches, benchmark results, open source projects, engineering deep dives, an

ReactMachine LearningAI
W
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend +8.1%
Quick readStrong listing-quality and freshness signals

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role AI research at WRITER isn't just about publishing papers — it's about building the scientific foundation that powers some of the most ambitious enterprise AI deployments in the world. As an AI research scientist, you'll be at the center of that work. You'll drive a high-impact research agenda focused on large language models, agentic reasoning, and the system-level capabilities that make AI genuinely useful at enterprise scale. This is a rare opportunity to do research that matters twice over — advancing the field and shipping directly into products used by hundreds of thousands of people every day. We're at an inflection point. Enterprises are moving from experimenting with AI to deeply embedding it across their operations, and WRITER's models are the engine making that possible. The work you do here — on post-training, planning, multi-step reasoning, and agentic workflows — will directly shape how the next generation of enterprise AI behaves, performs, and scales. You

PythonMachine LearningAI
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE We're hiring a Marketing Analytics Manager to establish how Marketing at Baseten makes decisions with data. Marketing at Baseten is scaling fast: more spend, more campaigns, more model launches and more inbound. This is a foundational, hands-on role and the first dedicated Marketing Analytics hire. You'll work directly with Demand Gen, Field Marketing, MOps and Product Marketing alongside GTM, Finance, Product and Engineering to stitch together activities and outcomes across the funnel. You'll build the data models that connect acquisition, engagement, activation, product usage and revenue. You’ll design dashboards, tools, semantic layers and plugins that enable teams to answer questions. Along the way, you’ll develop an understanding of how customers use Baseten and with this, define how our systems and product can improve. RESPONSIBILITIES Define how Marketing success is measured: establish metrics across audience growth, acquisition, activation, engagement, usage, pipeline and revenue. Build the marketing data foundation: ingest and model data across Salesforce, HubSpot, Google Analytics, advertising platforms, web, email, product and third-party sources. Create a medallion architecture that connects the prospect and customer journeys across systems, from first touch through signup, onboarding, activation and expansion. Understand our audiences: develop audience and segmentation frameworks based on customer a

SQLMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building the infrastructure layer for AI — and we're now building the team that will scale how we take it to market. This is a foundational hire on our GTM Strategy & Revenue Operations team, sitting at the intersection of strategic planning and field execution. You'll work directly with the CRO, Head of Revenue Operations and partner closely with Finance, field leadership, and our Central Ops team. On any given week, you might be refining our pipeline generation model, building a territory coverage analysis, designing a new GTM motion, or partnering with a regional leader to understand what's driving a trend in their pipeline. This role requires someone who can think rigorously, build things from scratch, and operate with speed and judgment in an environment where the playbook is still being written. This is a rare opportunity to be an early GTM strategy and operations hire at one of the fastest-growing companies in AI infrastructure — and to help define how we scale. WHAT YOU'LL DO: GTM Planning, Target Setting & Market Intelligence Contribute to the annual and quarterly GTM planning process in partnership with Finance, including headcount modeling, ramp assumptions, and revenue target-setting Build and maintain coverage models aligned to 2-year growth projections, incorporating territory design, account segmentation, and capacity planning Support quota framework development, helping trans

SQLMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building the infrastructure layer for AI — and we're now building the team that will scale how we take it to market. This is a foundational hire on our GTM Strategy & Revenue Operations team, sitting at the intersection of strategic planning and field execution. You'll work directly with the CRO, Head of Revenue Operations and partner closely with Finance, field leadership, and our Central Ops team. On any given week, you might be refining our pipeline generation model, building a territory coverage analysis, designing a new GTM motion, or partnering with a regional leader to understand what's driving a trend in their pipeline. This role requires someone who can think rigorously, build things from scratch, and operate with speed and judgment in an environment where the playbook is still being written. This is a rare opportunity to be an early GTM strategy and operations hire at one of the fastest-growing companies in AI infrastructure — and to help define how we scale. WHAT YOU'LL DO: GTM Planning, Target Setting & Market Intelligence Contribute to the annual and quarterly GTM planning process in partnership with Finance, including headcount modeling, ramp assumptions, and revenue target-setting Build and maintain coverage models aligned to 2-year growth projections, incorporating territory design, account segmentation, and capacity planning Support quota framework development, helping trans

SQLMachine LearningAIGo
D
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -93.7%

$135K – $207K/yr

Quick readStrong listing-quality and freshness signals

Drata is building the trust layer between great companies - automating compliance, managing risk, and helping organizations prove trust continuously as they scale. We're Dratanauts: a global crew of 600+ professionals united by a culture that rewards integrity, ownership, and raising the bar, no matter where in the world we're working from. Why Join the Drata Team? At Drata, you're not maintaining legacy compliance software - you're building the agentic AI platform defining what trust looks like for the next generation of companies. Here's what makes the work itself worth showing up for: Problems without a playbook: You'll work at the edge of AI and security, building agentic governance, continuous compliance, and real-time trust verification to solve problems that don't have an established answer yet. You're writing it as you go. Real ownership, not just process: Our values center on owning outcomes and raising the bar, not checking boxes. You're expected to have opinions and back them. A seat at the table: Your perspective is unique and valued. Open debate and diverse viewpoints are built into how decisions actually get made here, at every level. Growth at rocketship speed: Drata is scaling fast, which means scope grows fast too. High performers get more ownership, visibility, and experience. A crew, not just coworkers: Dratanauts consistently describe a "come as you are" culture with sharp, curious people—the kind of team that makes hard problems genuinely fun to solve. See what they say here and follow us on LinkedIn for company news, employee stories, and career updates. Job Summary: As Senior Platform Operations Manager, you will lead the strategy, execution, and optimization of our marketing technology ecosystem. You will architect, implement, and manage the systems and integrations that power our go-to-market (GTM) engine, with a sharp focus on scalability, data integrity, automation, and lead orchestration. This role is pivotal in ensuring that marketing, s

RestMachine LearningAIGo
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? We’re seeking an experienced Engineering Program Manager to join Cohere’s customer-facing Engineering Program Management team. We need someone with curiosity, drive, independence, and leadership, who has hands-on experience working directly with customers and managing projects for enterprise-grade software or enterprise-focused machine learning solutions. In return, you’ll have the unique opportunity to shape Cohere’s operations, collaborate with leading minds in the LLM space, work directly with our Strategic Customers as well as Applied ML (AML) Engineering, Forward Deployed Engineering (FDE), Platform , Product and Go-to-Market teams, and be the “technical” voice of Cohere for the customer. You will get a chance to create extremely high-impact contributions to our fast-growing company, product and culture. As an Engineering Program Manager/ Technical Program Manager, you will: Communicate: Provide clear, timely, and objective communication across the tech organization, Cohere teams, leadership, and most importantly - our strategic customers and external partners. Optimize: Break down complex issues into strateg

GitMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Lead large-scale brand campaigns across digital, events, and out-of-home. Partner with engineering and product marketing on major product launches. Turn complex technical ideas into clear, compelling visual communication. Evolve the Baseten visual identity into a brand system that scales. Own projects end-to-end, from concept through launch, collaborating directly with marketing, product, engineering, and leadership. Shape our employer brand and help define how new hires experience Baseten. Raise the bar for craft across every customer touchpoint. ABOUT YOU You have a portfolio of exceptional work with outstanding visual craft and attention to detail. You care deeply about quality and sweat the details. You communicate ideas clearly and thrive in collaborative environments. You know when to build systems, not just execute on assets. You default to ownership and are comfortable leading highly cross-functional projects from concept through launch. You're excited by technical products and know how to make complex ideas feel consumable without oversimplifying them. You thrive in a fast-moving environment with a high bar for quality. You have strong opinions about design and can articulate why something works - and why it doesn't. BONUS Motion design and animation experience Experience designing for developers or highly technical audiences. BENEFITS Competitive compensation, including meaningful equity. 100% covera

GitMachine LearningAI
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As Manager (Player & Coach) of the Solution Architect team, you will lead and mentor a team of Solution Architects who will partner closely with Sales and customers to translate business needs into technical solutions, run technical discovery, and guide repeatable deployments and proofs of value for customers. Applying both hands-on technical ownership and managerial leadership, you will guide your team through the processes of owning discovery calls, demos, technical scoping and driving POC’s through to execution as well as designing, deploying, and managing high performance, low latency AI applications on Baseten’s platform. You will also partner with product, infrastructure, and other customer engineering teams to ensure that large language models (LLMs) and other generative AI systems deliver best-in-class performance, reliability, and cost efficiency in production environments. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Leadership Lead, mentor, and grow a team of Solution Architects, providing guidance on technical direction, project execution, and professional development. Set clear goals and ensure timely, high-quality d

DockerMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. ABOUT THE ROLE Baseten is building the infrastructure layer for AI — and we're now building the team that will scale how we take it to market. This is a foundational hire on our GTM Strategy & Revenue Operations team, sitting at the intersection of strategic planning and field execution. You'll work directly with the CRO, Head of Revenue Operations and partner closely with Finance, field leadership, and our Central Ops team. On any given week, you might be refining our pipeline generation model, building a territory coverage analysis, designing a new GTM motion, or partnering with a regional leader to understand what's driving a trend in their pipeline. This role requires someone who can think rigorously, build things from scratch, and operate with speed and judgment in an environment where the playbook is still being written. This is a rare opportunity to be an early GTM strategy and operations hire at one of the fastest-growing companies in AI infrastructure — and to help define how we scale. WHAT YOU'LL DO: GTM Planning, Target Setting & Market Intelligence Contribute to the annual and quarterly GTM planning process in partnership with Finance, including headcount modeling, ramp assumptions, and revenue target-setting Build and maintain coverage models aligned to 2-year growth projections, incorporating territory design, account segmentation, and capacity planning Support quota framework development, helping

SQLMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Container runtimes were designed for general-purpose software workloads. AI inference is not a general-purpose workload. Running large models at production scale exposes cracks in every layer of the container stack: runtimes unaware of GPU memory constraints, images that take minutes to pull when a model needs to scale to thousands of replicas, and isolation mechanisms that weren't designed for the multi-tenant serving environments that production AI requires. The tools the industry has relied on for a decade weren't built for this, and patching around those limitations at higher layers only goes so far. Baseten owns the entire pipeline, from the moment a developer pushes a model to the moment a request gets a response. That vertical ownership means we can fix these problems at the root. The Runtime Fabrics team is doing exactly that: purpose-building the container runtime and storage layers for AI inference workloads, led by some of the world's top containerd maintainers. As Engineering Manager of the Runtime Fabrics team, you will lead this work, setting technical direction, growing a world-class team of systems engineers, and ensuring the team's output shapes not just Baseten's infrastructure but the open-source container ecosystem at large. If you've contributed to containerd, runc, or related OCI projects and are ready to lead a team solving some of the hardest problems in infrastructure today, we'd love

LinuxMachine LearningAIC++
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re seeking a GPU Kernel Engineer to join our team at the cutting edge of AI acceleration, where your code directly impacts the performance of state-of-the-art machine learning models. As a GPU Kernel Engineer, you'll craft the foundation that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You'll work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Core Engineering Responsibilities Design and implement high-performance GPU kernels for key ML operations, including matrix multiplications, attention mechanisms, and mixture-of-experts routing Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core acceleration, and compute/memory overlap Performance & Innovation Impl

AWSMachine LearningAIC++
🔔

Get new machine learning intern jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime