Jobs in United States

Performance Modeling Engineer in San Francisco

388 active opportunities · Updated October 2026

Explore current performance modeling engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -88.6%

From $180K/yr

Quick readStrong listing-quality and freshness signals

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Millions of people use Notion — and this number is increasing every day. That means millions of people trust us to deliver a fast, reliable, and secure experience, and we value this more than anything. We want to keep earning trust, while also continuing to amaze our users with the tools they can build in Notion. The AI Platform team is responsible for building the shared foundations that let Notion ship AI products quickly and operate them safely at scale. You’ll join a team of talented engineers focused on making speed and quality compatible: reliability and availability through provider changes, quality and correctness systems like evals and release gates, observability that makes failures explainable, and shared primitives for model integrations, context management, long-running actions, and cost/performance tradeoffs. Notion’s AI platform is vital to helping product teams move faster with production-grade guardrails as models, providers, and AI capabilities rapidly evolve. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days)

Node.jsRestAIGo
N
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -88.6%

From $245K/yr

Quick readStrong listing-quality and freshness signals

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role Notion is looking for a Head of Demand Engine to build the engine that turns our best product stories, customer proof, and market opportunities into scaled pipeline. You’ll lead demand programs across campaigns, digital, paid strategy, webinars, email, web, and emerging channels. You’ll partner closely with Product Marketing on audience and offer strategy, with Performance Marketing on paid execution, and with Field Marketing & Events on the moments that benefit from in-person activation. This is a senior leadership role with both AMER and global scope. You’ll help set the global standard for how Notion plans, runs, tests, and measures demand programs, while learning from strong regional teams and making the best ideas repeatable across markets. You’ll lead a growing team spanning demand programs and digital demand, with room to expand into areas like partner marketing over time. This is a build job as much as a run job. We’re looking for someone with real taste who can spot a strong idea, turn it into a sharp program, and build a repeatable engine around what works. You should care about quality and craft, but als

GitRestAIGo
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As an Applied Scientist specializing in Small Language Models and AI Training, you will lead research and development efforts focused on building efficient, high-performance language models tailored for practical applications. You will work closely with research, engineering, and product teams to advance model training techniques, optimize architectures, and scale AI solutions. Your work will directly contribute to AI systems that are safe, interpretable, and impactful across diverse usage scenarios. What You’ll Do Lead research and development of novel training methodologies and architectures for small and efficient language models. Design, implement, and evaluate model training experiments to improve performance, robustness, and generalization of language models. Collaborate closely with research scientists and engineers on scalable training pipelines and model deployment strategies. Develop techniques for model compression, fine-tuning, and domain adaptation to optimize models for real-world applications. Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluat

PythonMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.2%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE This role sits at the frontier of our research agenda. You will pursue open problems at the intersection of post-training methodology and performant inference, and then collaborate with research engineering to translate findings into production systems. A meaningful portion of your time will be dedicated to research that deepens our understanding of how models learn, alignment, and architectural efficiency — questions that may not have immediate product application. The remainder will be directed toward research that solves concrete problems for Baseten's platform and customers, who are the fastest growing AI companies in the world like Cursor, Lovable, and Notion. We are looking for someone with sharp research taste and genuine creative instinct for problem selection. Someone who can identify questions that matter, design clean experiments to answer them, and push the state of the art. The environment here is not theoretical, but rather research that can be validated with eager customers who are serving billions of tokens a second. RECENT RESEARCH Towards infinite context windows: neural KV cache compaction Dense, on-policy or both? Repeated kv cache for long-running agents Distillation without the dark – replicating black-box on-policy distillation on Baseten RESPONSIBILITIES Define and pursue a research agenda spanning both foundational and applied work, with the applied component connected to Baseten's pla

Machine LearningAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -86.4%
Quick readStrong listing-quality and freshness signals

About the Team The ChatGPT organization at OpenAI supports our mission by bringing advanced AI capabilities to hundreds of millions of users worldwide. The Image Generation team is responsible for one of the fastest-growing experiences in ChatGPT, enabling users to create, edit, and transform images through natural language. Recent advances in our multimodal image models have dramatically improved image quality, instruction following, editing precision, consistency, and text rendering, unlocking entirely new creative and professional workflows. We work at the intersection of research, infrastructure, and product to build the systems that power image generation at global scale. Our team partners closely with researchers, product engineers, designers, and platform teams to bring state-of-the-art image capabilities to millions of users while continuously pushing the boundaries of what AI-powered creation can do. About the Role We are looking for an experienced Backend Engineer to join the Image Generation team and help build the systems that power image creation and editing across ChatGPT. You'll work on the core backend infrastructure that enables users to generate, edit, and iterate on visual content using cutting-edge multimodal AI models. This includes building highly scalable services, orchestration systems, APIs, storage platforms, and distributed infrastructure that support billions of image generations and editing workflows. You'll partner closely with product, research, and mobile teams to transform breakthrough AI capabilities into reliable, performant experiences used by millions around the world. In this role, you will: Design, build, and operate backend systems that power image generation and image editing experiences in ChatGPT. Develop scalable APIs, services, and infrastructure that support multimodal AI workflows. Optimize reliability, latency, throughput, and cost across large-scale distributed systems. Partner with researchers to productionize new im

AWSRestAIRust
P
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -74.6%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Making data driven decisions is key to Plaid's culture. To support that, we need to scale our data systems while maintaining correct and complete data. We provide tooling and guidance to teams across engineering, product, and business and help them explore our data quickly and safely to get the data insights they need, which ultimately helps Plaid serve our customers more effectively. Engineers on Data Infrastructure are domain experts in Data Warehouse, Data Lakehouse, Spark, Workflow Orchestration, and Streaming technologies. We scale our existing data pipelines in a performant and cost efficient way while creating the necessary abstractions to make developing on top of this platform extremely simple for other engineers at Plaid. Responsibilities Contribute towards the long-term technical roadmap for data-driven and machine learning iteration at Plaid Leading key data infrastructure projects such as improving ML development golden paths, implementing offline streaming solutions for data freshness, building net new ETL pipeline infrastructure, and evolving data warehouse or data lakehouse capabilities. Working with stakeholders in other teams and functions to define technical roadmaps for key backe

PythonAWSMachine LearningAI
D
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -89.8%

From $214K/yr

Quick readStrong listing-quality and freshness signals

Datadog is seeking a strategic, visionary, and results-oriented Senior Director, Growth Marketing – Organic Growth (SEO/GEO/PLG) to lead our organic acquisition strategy across traditional search engines and emerging AI/LLM platforms. This leader will own the vision, strategy, and operating model responsible for driving measurable growth in organic traffic and inbound pipeline through content programs, off-page authority, and AI/LLM discoverability initiatives. In this role, you will lead a team of organic growth specialists, while partnering closely with Website Experience, Product Marketing, and Engineering teams. You will define Datadog's long-term organic growth strategy, establish investment priorities, and ensure the organization is positioned to win across an increasingly complex discovery ecosystem. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Own Datadog’s content-led SEO/GEO/PLG strategy, defining the topics, formats, and ecosystems that drive measurable growth in traffic and pipeline Lead, coach, and develop a high-performing team Identify and prioritize high-impact content opportunities across product areas and influence cross-functional teams to bring that content to life Map and optimize Datadog’s presence across the full ecosystem of LLM-ingested content (e.g., YouTube, Reddit, review sites) to improve AI-driven discoverability Design and execute a comprehensive off-page strategy, including link acquisition, digital PR, and authority-building initiatives Partner closely with the Website Experience team, who owns technical SEO, to ensure content is effectively surfaced, indexed, and performant Create and contribute to high-impact content (e.g., flagship pieces, new formats, or experimental channels), setting the standard for qualit

SQLGitAIGo
P
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, United States· Full-time· Remote
✓ High-confidence listingCompany trend -87%

From $177.2K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Intro : The Service Communications team manages the foundational high-layer networking systems ensuring reliable, secure, and performant service-to-service interactions at Pinterest. Our Envoy-based mesh and multi-language frameworks provide critical primitives for hundreds of internal services, and we are looking for a Staff Engineer to lead technical strategy across identity, traffic optimization, and platform scaling. What you’ll do: Architect and deploy advanced service mesh features, focusing on service discovery, traffic shaping, and deep observability using Envoy proxy. Lead the organization-wide adoption of service identity and mTLS to satisfy critical AAA security requirements for service-to-service paths. Design traffic optimization primitives like locality-aware routing to materially reduce data transfer costs for high-vo

PythonJavaAWSRest
P
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, United States· Full-time
✓ High-confidence listingCompany trend -87%

From $285.5K/yr

Quick readStrong listing-quality and freshness signals

About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Millions of people across the world come to Pinterest to find new ideas every day. It’s where they get inspiration, dream about new possibilities and plan for what matters most. Our mission is to help those people find their inspiration and create a life they love. We are looking for a Director of Engineering to lead the strategic vision and execution for Pinterest’s core and ads serving platforms, overseeing the infrastructure that enables high-scale, low-latency delivery of content and ads across all Pinterest surfaces. They will define the multi-year architecture to ensure our serving systems are reliable, performant, and cost-efficient while enabling rapid innovation for product and ML teams. What you’ll do: Set the long-term technical and operational vision for the Core and Ads Serving Platform, aligning infrastructure investme

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -86.4%

About the Team ChatGPT is evolving from answering questions to becoming a deeply personalized assistant that helps people discover, create, and make decisions across everyday life. We're building new multimodal product experiences that combine language, images, personalization, and interactive interfaces to help millions of users accomplish tasks in entirely new ways. This team sits at the intersection of AI research, product engineering, design, and consumer experiences. We move quickly, ship frequently, and work on products that define how people interact with AI every day. About the Role We're looking for exceptional full stack product engineers who love building polished consumer experiences from the ground up. You'll work across frontend, backend, AI-powered workflows, and rich interactive interfaces to create new product experiences that blend conversation, visual understanding, personalization, and commerce. You'll collaborate closely with designers, researchers, product managers, and model teams to rapidly prototype, launch, and iterate on experiences used by millions of people. This is an opportunity to help invent entirely new interaction paradigms—not just build traditional web applications. In This Role, You Will Design and build end-to-end product experiences across web services, APIs, and modern frontend applications. Partner closely with product, design, and research to rapidly prototype and launch new AI-native experiences. Build intuitive, performant interfaces that make advanced AI capabilities feel simple and delightful. Develop scalable backend systems that power personalized, real-time product experiences. Work with multimodal capabilities including text, images, and interactive UI components. Iterate quickly using user feedback, experimentation, and product metrics. Help define engineering standards, architecture, and technical direction for a fast-growing product area. You Might Thrive If You Have significant experience building consumer-facin

TypeScriptPythonReactNode.js
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -86.4%

The ChatGPT Finances team builds experiences that help people connect their financial accounts, understand their financial picture, and ask useful questions about their finances through ChatGPT. Our work spans account connectivity, data ingestion, dashboards, personalized insights, and conversational experiences. We collaborate across product, design, research, infrastructure, security, and data integrations to make complex financial information understandable and actionable. This is an early and ambitious product area with a substantial roadmap. We are looking for engineers who want to shape both the first user experiences and the durable systems required to earn and keep users’ trust. About the role We’re looking for full-stack product engineers to build and scale ChatGPT Finances. You will own features across the stack—from polished frontend experiences to the APIs, services, and data models that power them. This role is well suited to engineers who combine strong product judgment with broad technical depth. You should care about how quickly users can understand their financial lives, how reliably data moves through the system, and how AI can answer financial questions in a grounded, transparent, and useful way. You will work closely with product, design, research, infrastructure, security, and data integration teams to take ideas from early prototypes to reliable production experiences. In this role, you will Own full-stack product features from user experience and frontend implementation through backend services, data models, deployment, and observability. Build polished, accessible, and performant interfaces for account connection, dashboards, insights, and conversational workflows. Design APIs and backend systems that safely ingest, normalize, and serve financial data. Build resilient integrations that handle synchronization, data freshness, partial failures, permissions, and user consent. Bring new AI capabilities into production while prioritizing grounding

TypeScriptPythonReactNode.js
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -86.4%

Join the engineering teams that bring OpenAI’s ideas safely to the world!! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role We’re building the observability product for OpenAI—from scalable infrastructure to a rich, AI-powered UI. Our systems ingest over petabytes of logs and billions of time series metrics across our fleet. We're now layering intelligence on top—think agents that summarize SEVs, auto-generate dashboards, or help engineers debug through notebook-like UIs. We’re hiring software engineers across the stack—infra, backend, and product. You’ll join a small, gritty team building both foundational infra and novel internal tools to make OpenAI's production systems reliable, performant, and observable. What You’ll Do Own core observability infrastructure, including distributed logging, time series, and trace storage Build AI-native tools that help engineers detect, understand, and resolve issues autonomously. Contribute to UI experiences like dashboards, notebooking, or interactive debugging Collaborate closely with engineers, researchers, user ops, and other teams across the company to build the next generation observability product You Might Be a Fit If You: Have operated large-scale distributed systems in production. ( especially logging systems or some other time series databases) Thrive in ambiguous environments and roll up your sleeves to solve unscoped problems. Have full-stack chops or product sensibilities—you're excited to build real tools people use. Have strong fundamentals in systems, networking, and cloud infra (Kubernetes, AWS, etc). Bonus : built or contributed to observability systems (e.g. Prometheus, OpenTelemetry, etc). Why This Team We’re b

AWSKubernetesRestAI
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As a Member of Technical Staff and AI Agent Development Lead, you will lead the design, development, and deployment of next-generation AI agents that interact with users and complex environments. You will drive the architecture and implementation of scalable, reliable AI systems, working closely with research, product and engineering teams to build safe, interpretable, and performant AI technology. What You’ll Do Lead a cross-functional engineering team focused on AI agent development, from conceptual design to production deployment. Design and implement AI agent architectures leveraging state-of-the-art language models and associated technologies. Collaborate with research scientists on scalable experiments and productize research innovations. Drive the development of agent capabilities including dialogue management, decision making, and autonomy. Ensure AI safety and alignment principles are integrated throughout the agent lifecycle. Mentor and grow technical staff, fostering an environment of collaboration and innovation. Evaluate new tools, frameworks, and methodologies to enhance AI agent capabilities. Partner

PythonMachine LearningAIGo
🔔

Get new performance modeling engineer jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime