ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
Jobs in United States
Ai Solution Engineer in San Francisco
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current ai solution engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
From $192.5K/yr
WHAT IS BOX? Box (NYSE:BOX) is the leader in Intelligent Content Management. Our platform enables organizations to fuel collaboration, manage the entire content lifecycle, secure critical content, and transform business workflows with enterprise AI. We help companies thrive in the new AI-first era of business. Founded in 2005, Box simplifies work for leading global organizations, including JLL, Morgan Stanley, and Nationwide. Box is headquartered in Redwood City, CA, with offices across the United States, Europe, and Asia. By joining Box, you will have the unique opportunity to continue driving our platform forward. Content powers how we work. It’s the billions of files and information flowing across teams, departments, and key business processes every single day: contracts, invoices, employee records, financials, product specs, marketing assets, and more. Our mission is to bring intelligence to the world of content management and empower our customers to completely transform workflows across their organizations. With the combination of AI and enterprise content, the opportunity has never been greater to transform how the world works together and at Box you will be on the front lines of this massive shift. WHY BOX NEEDS YOU The Solutions Engineering Team at Box includes solutions engineers, value engineering, platform solution engineering, enterprise architects, and demo engineering. As a Solutions Engineer, you are empowered to sell to business and IT leaders in every space and vertical, and take ownership in crafting customer-centric solutions. You will work alongside the account team to define and expand revenue opportunities, and ensure the solution is ready for cross company deployments. You also act as a critical liaison between Sales and Product; sharing customer feedback with the Product Management, Operations and Engineering functions at Box. Our highest performers have a natural curiosity, develop deep knowledge of the Box platform, have
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is the world’s leading API platform, used by millions of developers and thousands of organizations to design, build, test, and scale APIs faster. As Postman continues to grow, we are investing in enabling our technical customer-facing teams to deliver exceptional experiences for developers and enterprises. We’re looking for a Technical Enablement Program Manager to build and scale enablement programs that empower our Solution Engineers (SEs) and Customer Solution Engineers (CSEs) to succeed with customers. In this role, you will ensure our technical teams have the deep platform knowledge, technical engagement skills, and practical tools they need to help customers evaluate, adopt, and expand their use of Postman. You’ll partner closely with Sales Engineering, Product, Product Marketing, and Revenue Operations to create structured technical enablement programs, hands-on learning experiences, and technical accreditation paths that raise the bar for technical excellence across the organization. This role is ideal for someone who thrives at the intersection of technical depth, program management, and field enabl
About the Team Like every team at OpenAI, the Marketing team contributes to our broader mission of ensuring responsible and widespread adoption of artificial intelligence. With that aim in mind, we are responsible for developing and executing strategies that drive awareness, engagement, and usage for OpenAI’s products and platform amongst our core audiences. Our focus extends beyond just promoting product features; we aim to provide valuable insights and resources that help our users make the most out of AI technologies. About the Role As an Industry PMM, you will help define how OpenAI brings frontier AI to priority industries. This role sits at the center of OpenAI’s industry marketing motion, helping customers understand where AI can create practical value, improve workflows, and support responsible adoption. You will shape the market narrative, build the field-facing operating system, and coordinate cross-functional execution across product launches, customer proof, partner moments, events, and account-based campaigns. You will help senior industry leaders understand how OpenAI models, products, and workflows can support meaningful work in either Life Sciences, Banking, or Healthcare. We’re looking for a product marketer who can combine strategic narrative, technical curiosity, enterprise GTM judgment, and program leadership. The right person can turn fast-moving capability into clear positioning, credible assets, and practical field execution. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own positioning and messaging for OpenAI’s industry offering for either Life Sciences, Banking, or Healthcare. Build field-ready assets for account directors, solution engineers, industry leaders, and executive audiences, including first-call decks, one-pagers, use-case libraries, customer stories, and proof-point packages Translate frontier model ca
About the Team OpenAI’s Forward Deployed Engineering team partners with leading semiconductor companies to deploy production-grade AI systems across the entire chip design lifecycle: design, verification, and physical design. We operate at the intersection of customer delivery and core platform development, embedding deeply with customers to translate frontier model capabilities into systems that materially improve engineering workflows and accelerate innovation. Our work turns early, high-touch deployments into repeatable solution patterns, reference architectures, and evaluation practices that scale across the semiconductor ecosystem. About the Role We are seeking a highly skilled Physical Design Engineer to join our semiconductor-focused Forward Deployed Engineering team. This is a senior IC role that will begin with a strong emphasis on physical design expertise, technical judgment, advisory leverage, and customer credibility, with the expectation that the person will grow into a broader Forward Deployed Engineering role over time. In the near term, you will serve as the team’s physical design SME across semiconductor deployments: helping FDEs, Product, and Research understand backend implementation workflows, pressure-test AI-assisted solution ideas against real physical design constraints, and raise the quality of our customer-facing technical work. You will help the broader team build fluency in implementation flows, EDA tooling, signoff methodology, and the trade-offs that shape physical design decisions in practice. Over time, we expect this role to expand beyond SME support into broader FDE ownership: partnering directly with customers, shaping deployment strategy, building and iterating production-grade AI systems, driving technical workstreams, and helping turn high-touch semiconductor deployments into repeatable solutions. This is a strong fit for someone who brings deep physical design expertise today and is excited to grow into a customer-facing, syst
About the team The AI Deployment Engineering team is responsible for helping developers and enterprises safely and effectively deploy OpenAI technologies in production. We act as trusted technical advisors and thought partners for customers, working side by side with their teams to identify high-value use cases, design practical architectures, and move from prototype to durable deployment. Cybersecurity is one of the most urgent domains where AI can help. Security teams are under pressure to reason across code, logs, infrastructure, tickets, alerts, and vulnerability data faster than ever. As frontier models become more capable, organizations need deep technical guidance on how to evaluate, validate, and safely deploy AI systems in security-critical workflows. About the role We are looking for a Cyber AI Deployment Engineer to partner with customers and help them apply OpenAI models, APIs, Codex, and agentic workflows to real cybersecurity use cases. You will work with CISOs, security executives, application security leaders, SOC teams, security engineering teams, and hands-on practitioners to identify where AI can create measurable security outcomes. This is a customer-facing technical role for someone who can move fluidly between executive strategy, practitioner-level cyber depth, and hands-on solution design. You will help customers evaluate and deploy workflows such as secure code review, vulnerability triage, threat modeling, remediation, SOC and incident response workflows, detection engineering, cloud security, GRC automation, and security validation. You will collaborate closely with Sales, Solutions Engineering, Product, Engineering, Research, and Security to turn customer needs into safe deployment patterns, reusable field assets, and product feedback. This role is based in our San Francisco HQ. We offer relocation support to new employees. In this role, you will: Deeply embed with strategic customers as the technical lead for AI-enabled cybersecurity work
About the Team The Applied AI Engineering team is responsible for helping customers turn frontier AI capabilities into real products, workflows, and business impact. We act as trusted technical partners across solution design, architecture, implementation, evaluation, and adoption, working alongside customers to build and scale effective AI applications with OpenAI’s technologies. The Codex Applied AI Engineering team focuses on helping organizations transform how software is built with AI. We partner directly with engineering teams and technical leaders to integrate Codex into their software development lifecycle — from identifying high-impact use cases and designing AI-enabled workflows to implementation, evaluation, and scaled adoption. Our work helps ensure AI-powered software development is effective, reliable, secure, and deeply integrated into how engineering organizations operate. About the Role We are seeking an experienced technical leader to join as Manager, Applied AI Engineering (Codex) , leading a team of Applied AI Engineers responsible for driving successful Codex adoption across strategic customers. Your team will work hands-on with customer engineering organizations to design and build AI-enabled development workflows, solve complex implementation challenges, and establish scalable patterns for AI-powered software development. As a manager, you will shape how these technical engagements operate at scale — setting strategy, coaching engineers, determining where the team can have the greatest impact, and ensuring consistently strong execution across customers. You will serve as both a people leader and senior technical advisor, partnering closely with Sales, Product, Research, and Engineering to translate customer needs and real-world usage into better technical approaches, reusable patterns, and product insights. Success in this role will be measured by meaningful and sustained Codex adoption, successful customer outcomes, and the creation of repeat
About the Team The Ona team at OpenAI is helping build the software factory for the enterprise. We build infrastructure that enables AI agents to work in secure, customer-controlled cloud environments, with the context, tools, and controls they need to make progress across the software lifecycle—beyond a single developer’s laptop or active session. Our focus is helping enterprises move from experimenting with agents to using them reliably in production. That means solving challenging problems in cloud environments, orchestration, security, and collaboration, while making the experience straightforward for the people directing and reviewing the work. We’re a team that values initiative, close relationships with customers, and exceptional engineering craft. We take ownership, learn quickly, and communicate directly and kindly. About the Role We’re hiring backend-focused Product Engineers across our platform and security product teams. You’ll build infrastructure and customer-facing workflows that let developers and AI agents work reliably in parallel. You’ll work primarily in Go on APIs, complex networking, development environments, and orchestration for long-running tasks. You’ll own outcomes from understanding a user’s problem and choosing an approach through shipping, operating, and improving the solution, working closely with frontend, infrastructure, and security engineers. In this role, you will: Work directly with customers to build developer and security workflows, from getting a project running to investigating findings, reviewing agent-generated changes, and verifying fixes. Build Go services and APIs for provisioning cloud environments, running agents in customer infrastructure, and integrating with source control, CI, and other developer tools. Design reliable orchestration for long-running, parallel work, including durable state, retries, cancellation, and recovery. Build security into execution workflows through clear permissions, credential handling, is
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Senior Software Engineer (Federal) Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Okta Identity Governance Team Okta Identity Governance (OIG) is Okta's Identity Governance and Administration solution, responsible for some of the most critical and visible workflows in enterprise identity: how people request access, how approvers grant it, how entitlements are enforced, and how organizations prove to auditors that only the right people have the right access. OIG is the leading contributor to Okta's new product bookings, and it runs inside some of the largest enterprises in the world — the kind of accounts where a single tenant carries hundreds of thousands of identities across thousands of applications. OIG currently supports high-compliance public sector workloads and plans to expand our footprint to support the most secure, highly classified environments in government. The Senior Software Engineer (Federal) Opportunity As a Senior Fullstack Engineer on the OIG (Federal) team, your mission
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Sales Engineering unlocks and empowers prospects’ full potential on the Plaid Network. We bring product expertise and drive Revenue outcomes by being customer-obsessed technical advisors in the pre-sales process. You’ll work closely with our Enterprise Sales team as the technical lead during the pre-sales process. You will be responsible for partnering with Account Executives, proactively pushing deals forward by overcoming technical objections and building solutions around Plaid products. You will ensure a successful transition to the post-sale stage of the customer lifecycle and partner with account teams to drive adoption of new products by existing customers. You will be expected to have significant knowledge of the Fintech ecosystem and a deep understanding of how clients can best utilize Plaid. Responsibilities Serve as the lead technical advisor for Enterprise accounts from discovery through solution design, evaluation, and post-sale handoff. Partner with Account Executives to align Plaid capabilities to customer priorities across payments, wallets, gaming, travel, and lending. Build alignment with technical, product, risk, and operations stakeholders on complex enterprise initiatives. Lead d
About Supabase Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the Role We’re looking for a Developer Relations Engineer based in San Francisco to join our team and to help more developers discover, learn, and build with Supabase. You’ll create high-impact content, build real-world projects, and represent Supabase across communities and events. If you’re equally energized by writing code and teaching others, this is the role for you. Why this role matters Supabase is growing fast with 350,000+ developers , 1,000+ OSS contributors , and a thriving open source ecosystem. Our users are builders, startup founders, weekend hackers, and engineers scaling to millions of users. DevRel is how we meet them where they are: through content, community, and code. We’re building a community of communities that brings together developers from many backgrounds, including first-time open source contributors. As a DevRel Engineer, you’ll be a bridge between Supabase and the broader developer ecosystem, helping people get started, go deep, and feel connected. What you'll do Make content Publish compelling technical content, especially video, to help developers learn Supabase quickly Build demos and tutorials Ship real-world apps using Supabase and tools like Next.js, React, and Stripe. Write guides that others can follow and remix Represent Supabase Speak at meetups, livestream builds, and engage with the developer ecosystem. You’ll be a visible and trusted voice of the platform Support and grow the community Celebrate contributors, answer questions, highlight cool projects, and bring developer feedback to the team Collaborate across the company Work with engineering, product, and growth to amplify launches, prioritize content, and reduce friction for new users You migh
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Are you passionate about advancing the application of artificial intelligence? We are looking for a Software Engineer focused on ML performance to join our dynamic team. This role is ideal for someone who thrives in a fast-paced startup environment and is eager to make significant contributions to the exciting field of LLM Inference. If you are a backend engineer who thrives on making things faster and is excited about open-source ML models, we look forward to your application. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and other libraries to debug ML performance issues. Apply and scale optimization techniques across a wide range of ML models, particularly large language models. Collaborate with a diverse team to design and implement innovative solutions. Own projects from idea to production. REQUIREMENTS Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or related field. Experience with one
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re seeking a GPU Kernel Engineer to join our team at the cutting edge of AI acceleration, where your code directly impacts the performance of state-of-the-art machine learning models. As a GPU Kernel Engineer, you'll craft the foundation that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You'll work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Core Engineering Responsibilities Design and implement high-performance GPU kernels for key ML operations, including matrix multiplications, attention mechanisms, and mixture-of-experts routing Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core acceleration, and compute/memory overlap Performance & Innovation Impl
From $151K/yr
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are seeking an experienced and vision-driven Lead Enterprise Systems Engineer to join our engineering team. In this role, you will bridge the gap between business objectives, solution architecture, and hands-on execution. The ideal candidate remains actively involved in coding (roughly 70–80% of the time) while serving as the primary technical point of contact for project stakeholders. What you'll do Technical Vision & Solution Architecture Lead the architectural design, development, and deployment of resilient, scalable solutions in our Salesforce Platform for both Sales & CPQ. Translate business and product requirements into clear, technical roadmaps and system specifications. Establish engineering best practices, design patterns, coding standards, and testing strategies. Hands-On Execution & Quality Assurance Write clean, maintainable, and highly efficient APEX code alongside the Salesforce development team. Conduct thorough code reviews to ensure quality, security, and performance. Manage technical debt, proactively balancing speed of delivery with long-term system health. Team Leadership & Mentorship Provide technical guidance, direct support, and actionable feedback to Salesforce engineers. Mentor team members to foster technical growth and career advancement. Lead agile ceremonies (sprint planning, daily stand-ups, technical grooming, post-mortems). Cross-Functional Collaboration Partner closely with Technical Managers, Enterpri
Other cities to consider
More places hiring for this role
Get new ai solution engineer jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime