ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Voice is becoming the internet’s next interface, but a production-grade Voice AI system is "hard to build" . You’ll join a small founding team of Baseten Voice AI, focused on bringing state-of-the-art open source models into production for Voice AI customers across productivity, customer service, clinical conversation, creator tools, education, and more. You’ll make a meaningful impact on people’s daily lives and help reshape these industries. This is a high-impact, high-ownership role. You will be the primary owner of Baseten Voice AI - our in-house inference stack to power Voice AI models - from product roadmap through engineering implementation. You’ll partner closely with Forward Deployed Engineers, Model Performance Engineers, and sister engineering teams to push the boundaries of Voice AI. EXAMPLE INITIATIVES: Develop world-class model serving stack for state-of-the-art open-source voice models - reduce end-to-end and tail latency (p95/p99), increase throughput, and improve GPU efficiency via profiling, runtime tuning, and server-level optimizations. Build large-scale, real-time infrastructure for multi-model voice agents - orchestrate STT, TTS, and agent components with streaming I/O to meet customer SLOs. Design tight training and inference iteration loops for voice model customization - enable fast evaluation, safe rollout, and rapid experimentation for custom voice model development. Past projects:
Jobs in United States
Performance And Systems Engineer in San Francisco
364 active opportunities · Updated October 2026
Showing
15 jobs
Explore current performance and systems engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
Overview: The Data Acquisition team within the Foundations organization at OpenAI is responsible for all aspects of data collection to support our model training operations. Our team manages web crawling and GPTBot services and works closely with Data Processing, Architecture, and Scaling teams. We are looking for a skilled Full-Stack Engineer to join our Data Acquisition team to build and optimize the interfaces and tools that power our data infrastructure. Responsibilities: Develop and maintain full-stack applications that support data acquisition, including internal tools and dashboards. Collaborate closely with cross-functional teams, including Data Processing, Architecture, and Scaling, to ensure seamless data ingestion and workflow management. Design and implement APIs to facilitate data interactions between internal services and external data sources. Enhance user experience by developing intuitive web-based interfaces for managing and monitoring data pipelines. Optimize backend services for performance, scalability, and security in a distributed computing environment. Work with legal and compliance teams to ensure our data acquisition processes adhere to privacy regulations and best practices. Deploy and maintain infrastructure using Kubernetes and Infrastructure-as-Code (IaC) methodologies. Analyze system performance, conduct experiments, and improve data workflows to maximize efficiency. Qualifications: BS/MS/PhD in Computer Science or a related field. 4+ years of industry experience in full-stack development. Proficiency in frontend frameworks (React, Vue, or similar) and backend technologies such as Python, Node.js, or Go. Strong expertise in RESTful APIs, GraphQL, and database design (SQL and NoSQL). Experience building data-intensive applications that handle large-scale datasets. Familiarity with cloud platforms (AWS, GCP, or Azure) and container orchestration (Kubernetes, Docker). Prior experience with web crawling and large-scale data processing is a
The ChatGPT Finances team builds experiences that help people connect their financial accounts, understand their financial picture, and ask useful questions about their finances through ChatGPT. Our work spans account connectivity, data ingestion, dashboards, personalized insights, and conversational experiences. We collaborate across product, design, research, infrastructure, security, and data integrations to make complex financial information understandable and actionable. This is an early and ambitious product area with a substantial roadmap. We are looking for engineers who want to shape both the first user experiences and the durable systems required to earn and keep users’ trust. About the role We’re looking for full-stack product engineers to build and scale ChatGPT Finances. You will own features across the stack—from polished frontend experiences to the APIs, services, and data models that power them. This role is well suited to engineers who combine strong product judgment with broad technical depth. You should care about how quickly users can understand their financial lives, how reliably data moves through the system, and how AI can answer financial questions in a grounded, transparent, and useful way. You will work closely with product, design, research, infrastructure, security, and data integration teams to take ideas from early prototypes to reliable production experiences. In this role, you will Own full-stack product features from user experience and frontend implementation through backend services, data models, deployment, and observability. Build polished, accessible, and performant interfaces for account connection, dashboards, insights, and conversational workflows. Design APIs and backend systems that safely ingest, normalize, and serve financial data. Build resilient integrations that handle synchronization, data freshness, partial failures, permissions, and user consent. Bring new AI capabilities into production while prioritizing grounding
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Making data driven decisions is key to Plaid's culture. To support that, we need to scale our data systems while maintaining correct and complete data. We provide tooling and guidance to teams across engineering, product, and business and help them explore our data quickly and safely to get the data insights they need, which ultimately helps Plaid serve our customers more effectively. Engineers on Data Infrastructure are domain experts in Data Warehouse, Data Lakehouse, Spark, Workflow Orchestration, and Streaming technologies. We scale our existing data pipelines in a performant and cost efficient way while creating the necessary abstractions to make developing on top of this platform extremely simple for other engineers at Plaid. Responsibilities Contribute towards the long-term technical roadmap for data-driven and machine learning iteration at Plaid Leading key data infrastructure projects such as improving ML development golden paths, implementing offline streaming solutions for data freshness, building net new ETL pipeline infrastructure, and evolving data warehouse or data lakehouse capabilities. Working with stakeholders in other teams and functions to define technical roadmaps for key backe
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE This role sits at the frontier of our research agenda. You will pursue open problems at the intersection of post-training methodology and performant inference, and then collaborate with research engineering to translate findings into production systems. A meaningful portion of your time will be dedicated to research that deepens our understanding of how models learn, alignment, and architectural efficiency — questions that may not have immediate product application. The remainder will be directed toward research that solves concrete problems for Baseten's platform and customers, who are the fastest growing AI companies in the world like Cursor, Lovable, and Notion. We are looking for someone with sharp research taste and genuine creative instinct for problem selection. Someone who can identify questions that matter, design clean experiments to answer them, and push the state of the art. The environment here is not theoretical, but rather research that can be validated with eager customers who are serving billions of tokens a second. RECENT RESEARCH Towards infinite context windows: neural KV cache compaction Dense, on-policy or both? Repeated kv cache for long-running agents Distillation without the dark – replicating black-box on-policy distillation on Baseten RESPONSIBILITIES Define and pursue a research agenda spanning both foundational and applied work, with the applied component connected to Baseten's pla
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As a Member of Technical Staff and AI Agent Development Lead, you will lead the design, development, and deployment of next-generation AI agents that interact with users and complex environments. You will drive the architecture and implementation of scalable, reliable AI systems, working closely with research, product and engineering teams to build safe, interpretable, and performant AI technology. What You’ll Do Lead a cross-functional engineering team focused on AI agent development, from conceptual design to production deployment. Design and implement AI agent architectures leveraging state-of-the-art language models and associated technologies. Collaborate with research scientists on scalable experiments and productize research innovations. Drive the development of agent capabilities including dialogue management, decision making, and autonomy. Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. Mentor and grow technical staff, fostering an environment of collaboration and innovation. Evaluate new tools, frameworks, and methodologies to enhance AI agent capabilities. Partner
About the Team The ChatGPT organization at OpenAI supports our mission by bringing advanced AI capabilities to hundreds of millions of users worldwide. The Image Generation team is responsible for one of the fastest-growing experiences in ChatGPT, enabling users to create, edit, and transform images through natural language. Recent advances in our multimodal image models have dramatically improved image quality, instruction following, editing precision, consistency, and text rendering, unlocking entirely new creative and professional workflows. We work at the intersection of research, infrastructure, and product to build the systems that power image generation at global scale. Our team partners closely with researchers, product engineers, designers, and platform teams to bring state-of-the-art image capabilities to millions of users while continuously pushing the boundaries of what AI-powered creation can do. About the Role We are looking for an experienced Backend Engineer to join the Image Generation team and help build the systems that power image creation and editing across ChatGPT. You'll work on the core backend infrastructure that enables users to generate, edit, and iterate on visual content using cutting-edge multimodal AI models. This includes building highly scalable services, orchestration systems, APIs, storage platforms, and distributed infrastructure that support billions of image generations and editing workflows. You'll partner closely with product, research, and mobile teams to transform breakthrough AI capabilities into reliable, performant experiences used by millions around the world. In this role, you will: Design, build, and operate backend systems that power image generation and image editing experiences in ChatGPT. Develop scalable APIs, services, and infrastructure that support multimodal AI workflows. Optimize reliability, latency, throughput, and cost across large-scale distributed systems. Partner with researchers to productionize new im
Join the engineering teams that bring OpenAI’s ideas safely to the world!! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role We’re building the observability product for OpenAI—from scalable infrastructure to a rich, AI-powered UI. Our systems ingest over petabytes of logs and billions of time series metrics across our fleet. We're now layering intelligence on top—think agents that summarize SEVs, auto-generate dashboards, or help engineers debug through notebook-like UIs. We’re hiring software engineers across the stack—infra, backend, and product. You’ll join a small, gritty team building both foundational infra and novel internal tools to make OpenAI's production systems reliable, performant, and observable. What You’ll Do Own core observability infrastructure, including distributed logging, time series, and trace storage Build AI-native tools that help engineers detect, understand, and resolve issues autonomously. Contribute to UI experiences like dashboards, notebooking, or interactive debugging Collaborate closely with engineers, researchers, user ops, and other teams across the company to build the next generation observability product You Might Be a Fit If You: Have operated large-scale distributed systems in production. ( especially logging systems or some other time series databases) Thrive in ambiguous environments and roll up your sleeves to solve unscoped problems. Have full-stack chops or product sensibilities—you're excited to build real tools people use. Have strong fundamentals in systems, networking, and cloud infra (Kubernetes, AWS, etc). Bonus : built or contributed to observability systems (e.g. Prometheus, OpenTelemetry, etc). Why This Team We’re b
About the Team ChatGPT is evolving from answering questions to becoming a deeply personalized assistant that helps people discover, create, and make decisions across everyday life. We're building new multimodal product experiences that combine language, images, personalization, and interactive interfaces to help millions of users accomplish tasks in entirely new ways. This team sits at the intersection of AI research, product engineering, design, and consumer experiences. We move quickly, ship frequently, and work on products that define how people interact with AI every day. About the Role We're looking for exceptional full stack product engineers who love building polished consumer experiences from the ground up. You'll work across frontend, backend, AI-powered workflows, and rich interactive interfaces to create new product experiences that blend conversation, visual understanding, personalization, and commerce. You'll collaborate closely with designers, researchers, product managers, and model teams to rapidly prototype, launch, and iterate on experiences used by millions of people. This is an opportunity to help invent entirely new interaction paradigms—not just build traditional web applications. In This Role, You Will Design and build end-to-end product experiences across web services, APIs, and modern frontend applications. Partner closely with product, design, and research to rapidly prototype and launch new AI-native experiences. Build intuitive, performant interfaces that make advanced AI capabilities feel simple and delightful. Develop scalable backend systems that power personalized, real-time product experiences. Work with multimodal capabilities including text, images, and interactive UI components. Iterate quickly using user feedback, experimentation, and product metrics. Help define engineering standards, architecture, and technical direction for a fast-growing product area. You Might Thrive If You Have significant experience building consumer-facin
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role Sentry's Revenue Operations team is built around three pillars — Analytics, GTM Technology, and Deal Desk — and we're hiring a Revenue Operations Manager who will work across all three. This isn't a role where you sit at the end of the sales process waiting for deals to land in your queue. We want someone who is curious about why deals are structured the way they are, who spots inefficiencies before anyone else does, and who sees our CPQ and Salesforce environment as a system to be improved, not just operated. Roughly half your time will be deal desk ownership while the other half is the broader rev ops surface area — process design, systems and tooling, reporting and insights, and the cross-functional projects that keep sales, finance, and marketing running off the same playbook. Across both, your mandate is the same: reduce friction, automate the routine, and make more of the work self-service for reps. This role is an opportunity to grow into a deep operator and systems thinker. If you're someone who thrives on bringing order to chaos, who gets excited about eliminating the manual step that shouldn't exist in the first place, this is the job. In This Role You Will Own Deal Desk - end-to-end processing of deals — including quote creation, discount review, order form generation, and approval routing — with minimal escalation and fast turnaround. Become an expert in our CPQ system and Salesforce integration– not just as a user, but as a critical voice for what should change. Document pain points, propose solutions, and work alongside our GTM Tech team to drive improvements. Help design and implement self-servi
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Marketing Operations (Ops) Team at Plaid builds the essential foundation that enables the Marketing function to operate efficiently, scale sustainably, and align with the company’s strategic goals. We focus on people, technology and process to drive operational excellence and optimize marketing performance. Our team is focused on creating and executing impactful marketing strategies that resonate with our target audiences, drive sustainable growth through high-quality pipeline generation, and enhance marketing efficiency. By integrating innovative technologies, we strive for seamless workflows and data-driven decision-making. We also continuously strengthen the foundation of a high-performing marketing organization, ensuring that processes and systems are in place to support long-term success. Responsibilities: Own the strategy and execution for integrating AI into marketing workflows, identifying high-impact opportunities to improve efficiency, scalability, and performance. Design, build, and deploy AI-powered agents and automations that support marketing workflows. Evaluate emerging AI tools and technologies, staying current on industry trends and translating new capabilities into practical ma
$1.2M – $1.4M/yr
At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. About the People team The People team helps Affirmers cultivate a high-performing, inclusive organization. We focus on three core areas: attracting talent aligned to our mission, enabling growth through people and team development, and retaining employees through thoughtful experiences and rewards. We develop Affirm’s people programs and systems to create a workplace where all employees can thrive and contribute meaningfully. By investing in our people, we create a long-term advantage for the business. About the Role Affirm’s Workplace team is seeking an experienced Workplace Specialist II to oversee the day-to-day operations of our San Francisco headquarters. This role will focus on delivering exceptional workplace experiences that support connection, productivity, and flexibility in a remote-first environment. We are looking for someone with a hospitality mindset, great operational skills, an upbeat and approachable demeanor, and a willingness to cultivate the best employee experience for all Affirmers. This role is based out of our San Francisco HQ office and reports to the Workplace Manager. What You'll Do Own day-to-day workplace operations for our San Francisco headquarters, including janitorial services, HVAC coordination, repairs and maintenance, office supplies, the snack and beverage program, and in-office events. Serve as the primary on-site contact for building management, reporting and tracking service requests, coordinating access and maintenance activity, and driving issues through timely resolution. Manage relationships with local workplace vendors, monitor service quality and performance, address gaps, and help ensure services are delivered efficiently and within scope. Respond to and resolve local Workplace Support Portal requests, keeping employees informed and identifyin
About the Team OpenAI’s Hardware organization develops system and infrastructure solutions optimized for advanced AI workloads. We collaborate across research, software, and external hardware partners to design and deploy next-generation AI systems at scale. Our team works closely with silicon vendors and system partners to evaluate emerging technologies, validate performance characteristics, and ensure that hardware capabilities translate effectively to real-world AI workloads. About the Role We are seeking a 3P Hardware Architecture Expert with deep expertise in GPU and accelerator architectures to engage directly with silicon vendors and guide hardware decisions for AI infrastructure. In this role, you will evaluate architectural tradeoffs across compute, memory, and interconnect systems, translating vendor specifications into real-world workload impact. You will play a critical role in early silicon evaluation, benchmarking, and performance validation, helping ensure that next-generation hardware meets the needs of our workloads. This role is highly hands-on and requires both deep technical understanding and the ability to engage at a high level with partners such as NVIDIA and AMD on architectural direction and design tradeoffs. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Engage deeply with silicon vendors (e.g NVIDIA & AMD) on GPU and accelerator architecture tradeoffs. Analyze and interpret performance, power, and efficiency characteristics of next-generation hardware. Translate vendor specifications into expected real-world performance for AI workloads. Evaluate architectural aspects including: compute throughput and utilization memory systems (HBM, cache hierarchies, bandwidth constraints) data types and precision tradeoffs (FP16, BF16, FP8, etc.) interconnect and scaling behavior. Run benchmarks and profiling to validate hardware performance a
About the Team The Strategic Finance team provides financial insights and guidance to support OpenAI’s long-term goals and strategies. We partner across the business to allocate and deploy our resources for the highest-impact outcomes.  Within Strategic Finance, the B2B team focuses on the financial performance of our products and GTM functions, ensuring tight alignment between financial objectives and company strategy. We partner with leaders across Product, GTM, Research, Partnerships, and Operations to: Drive operational planning, financial forecasting, and performance management. Provide decision-quality insights on product and financial performance to inform strategic resource allocation. Build the “0→1” financial foundations required to scale and accelerate growth. About the Role We are hiring a senior leader in B2B Strategic Finance to build and scale a new pillar within our finance organization. This is a highly visible role that reports into the Head of B2B Strategic Finance and supports some of our most critical executive stakeholders, including our COO, CFO, and CRO, among others on the B2B Leadership Team. This role is ideally based in our San Francisco HQ, but we are open to NYC and Seattle. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Drive B2B finance scale and rigor Build and scale core financial infrastructure across the B2B business, including forecasting methodology, variance management, and performance narratives that drive accountability and decision-making velocity. Lead consolidated planning across revenue, gross margin (including compute), and opex for annual budget, forecasts, regular business reviews, and long-range planning. Establish durable management reporting: KPI definitions, dashboards, month-end/quarter-end deliverables, and exec-ready readouts. Partner with Corporate FP&A, Accounting, and Finance Systems/Data to evolve processes and contro
ABOUT THE TEAM Critical Harm Operations sits within User Safety & Risk Operations and builds enforcement systems for Frontier Risk and Material Harm that are accurate, fast, defensible, and built to scale. The Cyber vertical turns policy into reviewer standards, calibrated judgment, quality systems, escalation paths, and automation guardrails. ABOUT THE ROLE We are looking for a senior cybersecurity practitioner and operations strategist to raise the quality, scalability, and technical rigor of our Cyber Operations. You will combine hands-on cyber judgment with systems-level operating design: resolve the hardest dual-use questions, evolve SOPs, uplift reviewers and vendors, and build practical tools and automations. This is a senior IC role. Success is not primarily cases closed; it is durable improvement in the operating model and the reviewers who run it. IN THIS ROLE, YOU WILL: Drive the Cyber Operations operating model across domain priorities, SOPs, escalation paths, quality health, vendor capability, roadmap inputs and help inform trusted access strategies. Serve as the senior cyber expert for complex or high-risk decisions across ChatGPT, API, Codex, agents, and emerging product surfaces. Translate policy ambiguity, quality misses, appeals, and reviewer disagreement into clear decision rules, calibration examples, training, and tooling requirements. Build durable operating systems and quality loops: golden sets, holdouts, double-labeling, adjudication, error taxonomies, reviewer calibration, and automation evaluations. Raise FTE and BPO capability through onboarding, certification, coaching, recurring calibration, and vendor-performance partnership. Use quality, appeals, SLA, backlog, and disagreement signals to diagnose root causes and prioritize high-leverage fixes. Build hands-on solutions—SQL analyses, scripts, dashboards, LLM eval workflows, evidence enrichment, routing logic, and lightweight automations—that improve decision quality and reduce manua
Other cities to consider
More places hiring for this role
Get new performance and systems engineer jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime