Jobs in United States

Performance And Systems Engineer in San Francisco

364 active opportunities · Updated October 2026

Explore current performance and systems engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The B2B Marketing team is responsible for helping businesses understand, adopt, and get value from OpenAI’s products. B2B marketing is a major and growing priority for OpenAI as we scale our work with companies, developers, and institutions around the world. About the Role Within B2B Marketing, Demand Generation builds the integrated, full-funnel engine that connects audience insights, content, field and digital experiences, paid media, lifecycle, and sales follow-through to qualified pipeline. We partner closely with Sales, Partnerships, Product Marketing, Communications, Creative, Web, RevOps, Analytics, and regional teams to create a cohesive customer experience and scale what works. We’re looking for a Senior Marketing Strategist to lead paid and emerging demand-channel strategy for OpenAI’s B2B business. You’ll own the B2B audience, channel, offer, partner, and measurement strategy, translating pipeline goals into clear briefs and investment recommendations. You’ll partner closely with Performance Marketing, which owns hands-on media execution, while informing audience strategy, campaign priorities, experimentation, and downstream performance decisions. You’ll also build our approach to content syndication and third-party demand media partnerships. In this role, you will: Owning the B2B paid and emerging-channel demand strategy across paid search, paid social, display and retargeting, content syndication, third-party demand media partnerships, sponsorships, review platforms, and selected emerging channels. Translating pipeline goals into audience and channel strategies, including ICP and segment priorities, buying groups, intent signals, channel roles, forecast assumptions, budget recommendations, campaign architecture, and a rolling experimentation roadmap. Partnering closely with Performance Marketing to define B2B audience, campaign, offer, creative, landing-page, conversion, and measurement requirements, while Performance Marketing owns media

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s mission is to build safe artificial general intelligence (AGI) which benefits all of humanity. This long-term undertaking brings the world’s best scientists, engineers, and business professionals into one lab together to accomplish this. In pursuit of this mission, our Go To Market (GTM) team is responsible for helping customers learn how to leverage and deploy our highly capable AI products across their business. The team is made of Sales, Solutions, Support, Marketing, and Partnership professionals that work together to create valuable solutions that will help bring AI to as many users as possible. About the Role OpenAI is building a dedicated Digital Natives vertical focused on high-growth technology and internet companies shaping the future of software. As a Sales Manager, you will recruit, develop, and lead a team of senior Account Directors responsible for landing and expanding OpenAI’s largest and most strategic Digital Native customers. You will define how we win in this segment — establishing the operating cadence, deal strategy, and vertical GTM playbook for engaging sophisticated, technically fluent buyers building AI into their core products. You will operate as a hands-on front-line leader: deeply involved in complex opportunities, elevating sales discipline, and shaping how OpenAI scales enterprise revenue in the most dynamic segment of the market. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. Key Responsibilities Recruit and develop a high-performing team of Account Directors. Establish a coaching culture centered on rigorous deal reviews, structured 1:1s, and clear performance standards. Raise the talent bar continuously through thoughtful hiring, development, and accountability. Architect and refine the GTM strategy for high-growth tech and internet companies. Shape territory design, account prioritization, and coverage mo

AWSGitRestAI
D
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -93.7%

$207.7K – $281K/yr

Quick readStrong listing-quality and freshness signals

Drata is building the trust layer between great companies - automating compliance, managing risk, and helping organizations prove trust continuously as they scale. We're Dratanauts: a global crew of 600+ professionals united by a culture that rewards integrity, ownership, and raising the bar, no matter where in the world we're working from. Why Join the Drata Team? At Drata, you're not maintaining legacy compliance software - you're building the agentic AI platform defining what trust looks like for the next generation of companies. Here's what makes the work itself worth showing up for: Problems without a playbook: You'll work at the edge of AI and security, building agentic governance, continuous compliance, and real-time trust verification to solve problems that don't have an established answer yet. You're writing it as you go. Real ownership, not just process: Our values center on owning outcomes and raising the bar, not checking boxes. You're expected to have opinions and back them. A seat at the table: Your perspective is unique and valued. Open debate and diverse viewpoints are built into how decisions actually get made here, at every level. Growth at rocketship speed: Drata is scaling fast, which means scope grows fast too. High performers get more ownership, visibility, and experience. A crew, not just coworkers: Dratanauts consistently describe a "come as you are" culture with sharp, curious people—the kind of team that makes hard problems genuinely fun to solve. See what they say here and follow us on LinkedIn for company news, employee stories, and career updates. Job Summary: Drata is looking for a Head of Product, Assurance to lead product strategy, execution, and team development for significant product lines at Drata – Trust Center & AI Questionnaire Assistance. This is a high-impact leadership role with end-to-end ownership of product vision, roadmap, customer outcomes, and business performance for a core area of the company. You will partner cl

RestAIGoRust
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The User Operations team (Support) is central to ensuring that our customers' experience with our products is nothing short of exceptional. We resolve complex issues, provide technical guidance, and support customers in maximizing value and adoption from deploying our products. We work closely with Sales, Technical Success, Product, Engineering and others to deliver the best possible experience to our customers at scale. OpenAI's customers represent a range of diverse backgrounds and maturity, from early-stage startups to established global enterprises. About the Role We’re seeking a Program Manager to lead support readiness and operational programs for OpenAI’s cloud and strategic partnerships. You’ll work closely with Engineering, Product, Support Delivery, and external partners to translate complex technical and business requirements into scalable customer support experiences. This role will help define how we support customers across partner ecosystems, from launch planning and issue-routing workflows to escalation management and ongoing operational improvements. In this role, you will: Lead support readiness for new and existing cloud and strategic partnerships, including launch planning, operational design, and ongoing program execution. Partner closely with Engineering, Product, go-to-market teams, and external partners to develop support models for technically complex products and integrations. Define partner-specific customer journeys, support workflows, escalation paths, ownership models, and cross-company handoffs. Identify operational and technical risks, align stakeholders on solutions, and drive improvements that strengthen the customer experience. Establish clear success metrics and operating rhythms to monitor partnership health, launch readiness, and support performance. Use AI and automation to improve partner-related support workflows and scale operations effectively. You might thrive in this role if you: Have 8+ years of experience

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the Role As Head of Finance - Data Centers, you will be the finance leader for OpenAI’s self-built data center efforts. You will partner with teams across infrastructure, real estate, energy, construction, procurement, and finance to turn proposed sites into sound investment decisions and funded projects. This role spans the full development lifecycle: evaluating opportunities, building investment cases, forecasting capital needs, managing construction budgets, and helping determine how projects should be financed. You will give leadership a clear view of project economics, funding requirements, and risks as we build data center capacity at scale. In this role, you will: Lead financial evaluation of proposed data center and related power infrastructure projects, including site economics, development costs, capacity phasing, lifecycle costs, and key risks. Build and own project-level models and capital expenditure forecasts that connect construction schedules, power delivery, equipment procurement, contingencies, and funding needs. Establish capital budgets and financial controls for active builds. Track commitments, actual spending, change orders, and forecasts to completion; identify cost or schedule risks early. Partner with development, engineering, energy, construction, and procurement leaders on decisions that affect cost, timing, and long-term performance. Work with Treasury, Corporate Finance, Tax, and Legal to evaluate financing options, including project or construction debt, leases, joint ventures, and other partnership structures where appropriate. Prepare investment recommendations and capital approval materials for senior leadership, translating complex project details into clear choices and tradeoffs. Build a consistent portfolio view of project costs, cash requirements, milestones, and financial performance. Partner with Accounting and operations teams through project completion and handoff. You might thrive in this role if you have: Prior exper

AWSRestAIRust
M
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -67.9%
Quick readStrong listing-quality and freshness signals

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role Modal's LLM inference platform delivers frontier performance for open-source models with best-in-class elasticity and developer experience, made in part possible by our custom runtime with GPU memory snapshots and multi-cloud substrate . We're looking for a leader to own the direction and execution of this platform to continue to establish us as the clear market leader, working closely with customers like Cognition, Doordash, Ramp, and many more. You'll be leading a group of highly talented engineers working on our market-leading LLM inference offering, spanning the serving stack, routing infrastructure, internal agentic optimization platform, and the user-facing product surface area. This is a hands-on leadership role — expect to split your time between technical contribution, product shaping and people management depending on what the team needs. You'll set direct

LinuxAIGoExcel
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Are you passionate about advancing the application of artificial intelligence? We are looking for a Software Engineer focused on ML performance to join our dynamic team. This role is ideal for someone who thrives in a fast-paced startup environment and is eager to make significant contributions to the exciting field of LLM Inference. If you are a backend engineer who thrives on making things faster and is excited about open-source ML models, we look forward to your application. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and other libraries to debug ML performance issues. Apply and scale optimization techniques across a wide range of ML models, particularly large language models. Collaborate with a diverse team to design and implement innovative solutions. Own projects from idea to production. REQUIREMENTS Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or related field. Experience with one

PythonDockerKubernetesRest
B
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE The Model Performance organization at Baseten is looking to hire our first Technical Program Manager. This is a zero-to-one role in a team that is responsible for building the core algorithms and methods that power Baseten’s high performance inference stack. You won't inherit an existing program framework, you'll build one from the ground up: the planning structure, execution processes, metrics and the cross-functional alignment that a fast-growing organization needs. Your contributions will directly impact how fast our performance R&D gets productized. If you can drive turning a set of ambitious but loosely defined initiatives into a predictable, well-governed program, this role is for you. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Model Performance team: How to build a day-0 API for Kimi K3 How we built the new fastest API for GLM-5.2 Inference engineering for DeepSeek V4 Pro 0813 RESPONSIBILITIES Own execution across Model Performance's active project portfolio, freeing the team's technical leads to focus on technical direction rather than tracking. Design and stand up the planning structures, operating cadences, and status reporting mechanisms that best fits the team’s DNA. Coordinate model release and optimization programs end to end, including day-zero launches, sequencing the work across performance engineering, infra, and release stakeholders. Drive cross-team al

Machine LearningAIGoExcel
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team : We build core first-party app experiences in ChatGPT and Codex, define the primitives for high-quality app third-party app experiences, and collaborate with best-in-class partners across consumer and enterprise categories to bring delightful experiences to our customers. About the Role: We are hiring a Product Manager to shape and scale the app ecosystem across ChatGPT and Codex. This person will own both 1P product experiences and partner-led launches, translating user needs, model capabilities, platform constraints, developer and partner requirements, and enterprise controls into products that feel delightful, reliable, and safe. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Develop the strategy and roadmap for ChatGPT and Codex app ecosystem experiences across consumer and enterprise use cases. Build and ship high-quality first-party app experiences that demonstrate the best of what apps can do inside ChatGPT and Codex. Collaborate with best-in-class partners to create app experiences that solve real user and business workflows. Define the product foundations and quality standards needed for a trusted app ecosystem. Lead cross-functional execution across engineering, design, research, partnerships, GTM, legal, privacy, security, support, and data/evals. Use customer, user, partner, and model-behavior insights to prioritize the roadmap and improve post-launch performance. You might thrive in this role if you: Have built and scaled consumer or enterprise product experiences with strong product taste and measurable user impact. Have built app platforms, marketplaces, partner ecosystems. Are technically fluent enough to reason about MCPs, APIs, SDKs, and model/product constraints. Can move fluidly between strategy, product, partner judgment, and operational execution. Communicate crisply in writing and bring clarity to ambiguous, fast-moving product areas. Care deeply about user trust, safe

AWSRestAIGo
P
📍 San Francisco, CA, United States· Full-time· Remote
✓ High-confidence listingCompany trend -85.6%

From $92K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . About tvScientific We are the first CTV advertising platform purpose-built for performance marketers. For game developers and publishers, we bridge the gap between massive TV reach and granular User Acquisition (UA) metrics. Built by ad-tech veterans, our platform combines media buying, optimization, and MMP attribution to help gaming brands automate CTV campaigns, drive app installs, and maximize Return on Ad Spend (ROAS). Join the tvScientific team as an Account Manager (Gaming), where you'll lead strategic client relationships for gaming and app clients, drive revenue growth, and ensure client success on our cutting-edge platform. As an Account Manager on our team, you'll be responsible for managing a portfolio of key client accounts, developing and executing strategic account plans, and driving revenue growth through upsell, cross-sell, and

AWSGitRestAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the Team The Growth team drives user and revenue growth across ChatGPT’s consumer and business segments as well as other OpenAI products worldwide. We operate across the full funnel - from awareness and acquisition through activation, retention, and expansion - using a combination of global performance marketing, AI-powered workflows, in-product optimization, insights, experimentation, and creative ops engineering. About the Role We are hiring a Growth Marketing Manager to turn product-led growth priorities into clear audience strategies, compelling cross-channel messaging, coordinated go-to-market programs, and drive measurable, incremental growth. You will partner closely with Growth Product, Consumer Product Marketing, Lifecycle, Performance Marketing, Brand, Creative, Research, Data Science, and International Marketing to ensure that full-funnel campaigns and initiatives are supported and connected to Growth capabilities and maximize the performance of these campaigns. In this role, you will: Build the marketing layer across the Growth roadmap: acquisition, access, activation and resurrection, monetization, and the shared Growth Platform. Translate product hypotheses into audience insights, positioning, messaging, launch briefs, lifecycle journeys, landing-page narratives, performance creative, and in-product education. Maintain an integrated plan that connects the Growth product roadmap to the consumer product marketing calendar across lifecycle, performance, social, creators, partnerships, and brand moments. Serve as the subject matter expert across all Growth levers, advising teams across the organization on the most effective ways to integrate Growth into their initiatives. Partner with cross-functional teams to design and interrogate the user journey so we can align our external promises with in-product and owned-channel landing, access, onboarding, continuation, and conversion experiences. Create audience and market playbooks for new users, high-valu

AWSRestAIGo
P
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -72.3%
Quick readStrong listing-quality and freshness signals

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. Our Payments Risk team builds and operates decisioning capabilities across Signal and Guarantee. We help customers tune thresholds and rules so they can approve more good transactions while reducing returns. For Guarantee, this work also protects the economics of the risk Plaid underwrites. You will be the technical, customer-facing Payment Risk Consulting Lead for Signal and Guarantee. You will interpret model outputs, diagnose customer performance, and turn that analysis into concrete threshold and rules recommendations. You will own proofs of concept and retros, improve existing integrations, and help the team identify patterns that can be scaled through better processes and product capabilities. Responsibilities Own customer proofs of concept and retros from analysis through recommendations and follow-through. Proactively optimize Signal customers, prioritizing accounts with high return rates. Read model outputs and diagnose the drivers of authorization and return-rate performance. Recommend threshold and rules changes that align with each customer's risk and authorization goals. Help Guarantee customers tune thresholds to achieve target authorization rates while protecting lo

SQLAWSAIGo
B
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -79.1%
Quick readStrong listing-quality and freshness signals

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Manager focused on TPUs at Baseten, you will lead the "engine room" for our non-NVIDIA accelerator fleet, architecting, securing, and optimizing the Google Cloud TPU (and broader emerging accelerator) capacity that powers our customers' AI workloads. You'll own the end-to-end journey of capacity management for this fleet, from securing large-scale TPU pod allocations to building the automation that ensures reliable uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering, with a specific focus on the TPU ecosystem. You will act as the fleet orchestrator for Google's TPU architecture, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics as we diversify beyond NVIDIA. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the latest generation of TPU hardware, like Google's Trillium (v6e) architecture, and partnering closely with the Model Performance (MP) team to ensure workloads are tuned for TPU-specific execution. EXAMPLE INITIATIVES The TPU Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's TPU clusters, including pod slicing and topology planning Global Workload Orchestration: Bui

PythonAWSAzureGCP
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -80.2%

About the Team We are a small and fast-moving partnerships team that shapes and executes OpenAI’s most important collaborations. Your mission is to define and lead the strategy, roadmap, operating model, and growth of OpenAI’s B2B marketplace. You will build the scalable foundation through which customers discover, evaluate, adopt, and manage trusted third-party solutions built on and with OpenAI’s models. About the Role You are a senior marketplace and ecosystem leader with strong judgment, high ownership, and a track record of building scaled B2B platforms. You are comfortable moving between product design, commercial strategy, ecosystem development, and operations, and you can create clarity where there is no playbook. You bring the credibility to align senior internal and external stakeholders while moving quickly and executing at a high bar. Key Responsibilities Define the vision, strategy, roadmap, and success metrics for OpenAI’s marketplace. Own marketplace business performance across partner supply, customer demand, adoption, transaction volume, partner success, solution quality, and customer trust. Partner closely with product and engineering to shape discovery, listing, evaluation, procurement, billing, deployment, identity, governance, security, and measurement capabilities. Build a high-quality ecosystem of ISVs, developers, service providers, and other third-party solution partners. Design marketplace commercial models, policies, incentives, and partner economics with finance, legal, security, policy, and operations. Create scalable go-to-market motions across sales, co-sell, private offers, marketing, and customer success. Establish standards for solution quality, safety, compliance, performance, and lifecycle management. Build and lead the marketplace team and operating cadence, aligning cross-functional owners against clear priorities and outcomes. Qualifications 15+ years of experience in marketplaces, platforms, ecosystems, product, partnerships,

Artificial IntelligenceAIFinanceProcurement
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for a customer-obsessed software engineer to come ship with us. You’ll own features like multi-node training and products like serverless reinforcement learning (RL) from conception to MVP (and from MVP to GA!). You’ll work through the stack, architecting solutions from API and UI down to our infrastructure layer. You’ll fine tune models yourself to develop an understanding of user workflows. You’ll work closely with research engineers leveraging state-of-the-art training techniques to build experiences that accelerate model development and solve for real pain points. If you’re excited to dive deep into the training, let’s talk! THE PRODUCT Take a look at what we’ve built so far: Overview of the product so far Training docs overview Story of the Training product Research we've done EXAMPLE INITIATIVES Checkpointing Pipeline: Our checkpointing pipeline starts with automated checkpointing, a feature that ensures that versions of models created during training are automatically backed up to the cloud. Users are able to then deploy checkpoints seamlessly into inference servers, providing point-and-click integrations into inference frameworks like vLLM and Baseten’s Inference Stack. This enables customers to quickly evaluate the performance of their checkpoints with real traffic. Multinode training: Multinode training enables customers to easily run training jobs across multiple compute nodes, enablin

KubernetesRestMachine LearningAI
🔔

Get new performance and systems engineer jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime