Jobs in United States

Performance Modeling Engineer 2 in San Francisco

364 active opportunities · Updated October 2026

Explore current performance modeling engineer 2 jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.

B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -73.6%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As a senior member of Baseten's Platform Team, you will own the systems that let every engineer at Baseten prove their code works before it reaches production. Our product runs mission-critical AI inference for customers who measure downtime in dollars per second, which means our internal bar for correctness, performance, and failure tolerance has to be exceptional. Your focus is the full testing stack: fast and reliable unit test tooling, integration harnesses that spin up realistic environments on demand, load and performance testing for GPU-backed inference workloads, and resilience testing that deliberately breaks things so our customers never have to find out what happens when a node dies mid-request. This is a builder role with org-wide leverage. You won't be writing tests for other teams — you'll be building the frameworks, harnesses, and feedback loops that make writing good tests the path of least resistance, and you'll set the standards for what "well-tested" means at Baseten. RESPONSIBILITIES Own Baseten's testing strategy end to end — define the standards, the tiers, and the tooling that engineering teams build against. Build and maintain unit, integration, load and performance testing frameworks Design end to end test infrastructure that provisions realistic dependencies

PythonDockerKubernetesCI/CD
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -73.6%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten’s Inference Stack team builds the distributed runtime that powers large-scale LLM inference across our platform. We operate at the intersection of distributed systems, model performance, infrastructure, and developer experience. We enable customers to deploy and operate cutting-edge LLM models with industry-leading performance, scalability, reliability, and ease of use. As a Software Engineer on the Inference Stack team, you’ll work across the stack - from the developer experience customers use to deploy models, the libraries used for features like tool calling and reasoning, all the way down to the systems we use to orchestrate deployments in Kubernetes and route traffic efficiently. This is an ideal role for engineers who enjoy owning systems in production, solving hard integration problems, and making complex infrastructure simple and reliable for users. EXAMPLE INITIATIVES Blog Posts https://www.baseten.co/blog/nvidia-dynamo-day-baseten-inference-stack/ https://www.baseten.co/blog/how-baseten-achieved-2x-faster-inference-with-nvidia-dynamo/ https://www.baseten.co/blog/how-baseten-multi-cloud-capacity-management-mcm-powers-cloud-self-hosted-and-hybr/#comparing-deployment-options-cloud-vs-self-hosted-vs-hybrid RESPONSIBILITIES Develop infrastructure and orchestration systems for deploying and managing large-scale distributed LLM inference Work across the stack, from customer-facing features to low-le

KubernetesCI/CDRestMachine Learning
W
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend +8.1%

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Are you passionate about ensuring the highest quality for cutting-edge generative AI applications? As a software quality engineer at WRITER, you'll play a critical role in shaping the reliability, performance, and trustworthiness of our AI-powered work orchestration platform. You’ll be at the forefront of defining and implementing rigorous quality strategies for our enterprise-grade LLMs and AI agents, directly impacting how hundreds of global companies unlock transformational value through AI. This is a unique chance to dive deep into the unique challenges of AI quality assurance and make a tangible difference in a rapidly evolving field. This is a hybrid role based out of our London, San Francisco, Seattle, and New York City hubs. You will report directly to the director of engineering. 🦸🏻‍♀️ What you'll do Define and implement comprehensive quality assurance strategies and test plans for our AI agents and LLM-powered applications, ensuring exceptional prod

TypeScriptPythonAWSAzure
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Events Analytics Platform (EAP) team is responsible for the infrastructure that powers all of Sentry's time-series data and searching capabilities across billions of events with sub-second latency. We started this initiative by building Snuba, the primary storage and query service for Sentry's event data powered by ClickHouse, and we are now focused on unlocking deeper visibility and reporting across the terabytes of event data our users generate. As a Senior Software Engineer, you will lead efforts to push the boundaries of data visibility at Sentry. You will do this by expanding the capabilities of our search infrastructure, building new capabilities on top of our state-of-the-art storage layer and increasing the performance and integrity of Sentry’s core data services. You will also help shape Infrastructure's technical direction at Sentry and collaborate with Product and other Engineering teams to turn that vision into a reality. If you want to solve the hard problems that come with scaling event data into the petabyte range, this could be the job for you. In this role you will: Expand EAP's ability to deliver data at world-class speed and reliability. Architect and automate services and systems to scale reliably under growing demand. Make architectural trade-offs that balance product requirements with engineering constraints. Maintain and grow the team's code quality initiatives by regularly reviewing code and contributing to design decisions. Lead design and discussions around deliverables the team is working towards. Improve the maintainability and developer experience of the codebases EAP owns. Exa

PythonSQLPostgreSQLRedis
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About this role Are you ready to redefine the future of JavaScript development? Are you a seasoned JavaScript expert who gets a thrill from tackling complex challenges across the entire JavaScript ecosystem? Do you believe that AI can be a powerful partner in crafting elegant, high-impact code? If you're ready to leave the mundane behind and join a team that's shaping the tools used by millions of developers globally, we've got an opportunity for you at Sentry. This isn't your typical Senior Software Engineer position. As a key member of our growing JavaScript SDK team, you'll be at the forefront of innovation, working on everything from our cutting-edge SDKs for Node.js, Bun, Deno, Cloudflare Workers, and other modern server runtimes. You won't just be maintaining code; you'll be pushing the boundaries of what's possible in developer tooling across the rapidly evolving server-side JavaScript landscape. In this role you will Join our JavaScript SDK team and get ready to build the future. You'll be at the forefront, working on: A Universe of JavaScript Challenges: Dive deep into our extensive suite of JavaScript SDKs, with a sharp focus on server-side and edge runtimes — from the battle-tested Node.js ecosystem to cutting-edge alternatives like Bun and Deno, and distributed edge environments like Cloudflare Workers. Your work will directly empower millions of developers to build better, more reliable software, no matter which runtime powers their stack End-to-End Ownership: We believe in giving our engineers the autonomy to see their vision through. You'll have the freedom to plan, implement, and ship your code, from writing

JavaScriptTypeScriptJavaNode.js
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About this role Are you ready to redefine the future of JavaScript development? Are you a seasoned JavaScript expert who gets a thrill from tackling complex challenges across the entire JavaScript ecosystem? Do you believe that AI can be a powerful partner in crafting elegant, high-impact code? If you're ready to leave the mundane behind and join a team that's shaping the tools used by millions of developers globally, we've got an opportunity for you at Sentry. This isn't your typical Senior Software Engineer position. As a key member of our growing JavaScript SDK team, you'll be at the forefront of innovation, working on everything from our cutting-edge SDKs for the React, Next.js, Vue, Nuxt, Hono, NestJS, and beyond. You won't just be maintaining code; you'll be pushing the boundaries of what's possible in developer tooling across the full spectrum of the modern JavaScript framework landscape. In this role you will Join our JavaScript SDK team and get ready to build the future. You'll be at the forefront, working on: A Universe of JavaScript Challenges: Dive deep into our extensive suite of JavaScript SDKs, with a broad focus on framework support spanning the modern JS ecosystem — from frontend frameworks like React, Vue, and their meta-frameworks Next.js and Nuxt, to server-side and edge runtimes like NestJS and Hono. Your work will directly empower millions of developers to build better, more reliable software, regardless of their framework of choice End-to-End Ownership: We believe in giving our engineers the autonomy to see their vision through. You'll have the freedom to plan, implement, and ship your code, from writi

JavaScriptTypeScriptJavaReact
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$150K – $190K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Developer Experience team is growing in San Francisco! This is a hybrid role out of our main headquarters and must be based there. We’re looking for a hands-on builder with strong opinions on great developer docs. Are you someone who loves tinkering with the newest features in your favorite products? Do you enjoy taking the next random JavaScript framework for a spin and figuring out how to make them tick? Does it especially irk you when the code snippets are wrong, or the only docs for a feature are on X? Developer Experience at Sentry lives in the intersection of shipping really cool products and getting developers set up to use them. If you’re someone who is confident in partnering with Product and Engineering to test and ship the latest features, not afraid to jump in and go hands-on to solve problems for our biggest customers, and a good eye for what good docs looks like - this is a dream role. Developer Experience at Sentry is a team of builders who are constantly looking for ways to make it easier for every developer to use Sentry. We engage with the challenges facing technical communities, to support developers, gather feedback, and help our product and engineering teams ship new capabilities. An engineer in this role should be confident to “come with an answer," propose the product solutions, identify the content and/or execution plans and people to partner with, and go execute. In this role you will Attend and even host events and meetups in the developer community Have a point of view and aren’t shy about expressing it. You get your energy from both knowing the latest trends and having a POV on

JavaScriptJavaRestAI
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$220K – $450K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role AI and machine learning are reshaping how developers debug, monitor, and ship software, and Sentry is uniquely positioned to lead that shift. We sit on a novel and massive dataset of real production errors, spans, and logs from tens of thousands of engineering organizations — the kind of signal that makes ML genuinely useful, whether it's a clustering model that groups related issues, a ranking system that surfaces the right alert at the right time, or an agent that proposes a fix. We're looking for an Engineering Manager to lead and grow our Machine Learning Engineering team. This team owns the full spectrum of ML at Sentry: classical techniques like clustering, ranking, anomaly detection, and embeddings that quietly power core product surfaces today, alongside the LLM-based and agentic systems shaping where the product is headed. You'll partner closely with product, design, and engineering leaders to decide where ML belongs in our products, what kind of ML actually fits the problem, and how we translate that work into experiences millions of developers rely on every day. In this role you will Set technical direction across the team's full ML surface area — from classical models for clustering, ranking, and anomaly detection to LLM-based and agentic systems — and make sharp calls about which approach fits each problem Define how the team evaluates and monitors ML systems in production, from offline metrics to online experimentation to model and agent observability Stay hands-on enough to review code and model designs, contribute to architecture discussions, and unblock engineers on complex ML problems Define

RestMachine LearningAIGo
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$140K – $190K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role: The Solutions Engineering team at Sentry is responsible for helping our largest customers successfully embed Sentry into their applications and workflows, from integrating our SDK to making sure alerts and issues make it to the right place at the right time. Solutions Engineers will be interfacing with customers to ensure optimal implementation and utilization ensuring that: Developers can effectively identify and resolve slowdowns and errors in their applications. Development teams increase developer productivity and customer satisfaction. In this role you will Understand customers' pain points and objectives and translate those into implementation and onboarding plans to meet those objectives and demonstrate business value. Develop best practices, educational offerings, and content around error and performance management, open-source, or the Sentry product (ex: “how-to” articles and/or videos on specific functionality of the product). Act as a trusted advisor and subject-matter expert for customers to assist with instrumenting software engineering and monitoring best practices Provide workshop-level interaction to customers in order to drive the maturity of use and broader adoption. Work cross-functionally (internally) to provide feedback to the product and development teams and align Sentry's product roadmap with improving time to value-delivered. You’ll love this job if you Enjoy talking about technology and interfacing with engineers and engineering leaders. Appreciate working in a dynamic environment, on a variety of projects with customers from lots of different industries. Love flexing your creative b

JavaScriptPythonJavaReact
S
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -72.4%

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role Sentry already has a startup program: credits, partner deals with YC and a16z, a landing page with a terminal theme that developers actually think is cool. What we don't have is someone whose full-time job is making this program a real growth engine. That's this role. You'll own startup marketing end-to-end: the partner ecosystem, the community presence, the acquisition programs, the brand in rooms full of founders and early engineers. You'll take what exists today and turn it into something that makes Sentry the default monitoring choice for every new company shipping code. This role reports to the Head of Enterprise Marketing (yes, we get how this can be confusing, but she’s pretty cool) and works across Growth, Product Marketing, and DevEx. You'll have real autonomy to shape the strategy and build programs from scratch – but you'll also need to execute quickly, (kind of – we can explain more) measure what works, and kill what doesn't. In this role you will Build and scale startup partnerships. Expand relationships with accelerators, incubators, VC platforms, and startup deal aggregators. Make Sentry's credits program a no-brainer inclusion in every startup perks stack. Own the startup acquisition funnel. Design programs that move early-stage companies from free plan → credits program → paid customers → expansion. Instrument everything so we know where the pipeline is and where it leaks. Be Sentry's face in the startup ecosystem. Show up at demo days, founder events, and SF tech week. Build relationships with the people building the next generation of companies. Not because it's a line item, but because you'

S
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -72.4%

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role: Sentry is expanding our Technical Customer Success team to support our growing customer base and drive deeper adoption of the Sentry Platform worldwide. As a Technical Customer Success Manager (TCSM), you’ll be a Sentry product expert responsible for ensuring customers are successfully onboarded, achieve maximum value from our platform, and uncover new opportunities for growth through additional use cases and products. You’ll play a key role in helping customers realize measurable outcomes with Sentry. In this highly cross-functional role, you’ll collaborate closely with Account Executives (AEs), Sales Engineers (SEs), and Engineering teams to ensure customers’ technical and business goals are met. This role requires strong technical expertise and a deep understanding of the Software Development Life Cycle (SDLC) and related technologies. If you’re a technologist with experience supporting technical products in customer-facing roles—and you’re eager to join a fast-growing team delivering real value through an exceptional product—we’d love to meet you. In this role you will: Become a Sentry product expert and support customers by understanding their needs and helping them achieve their goals using the Sentry Platform. Drive customer success and health through effective onboarding, adoption, value realization, and retention. Collaborate with customer key stakeholders to define business and technical objectives, and work with the customer team to achieve them. Act as a trusted, strategic advisor to each assigned customer—driving best practices, innovation, and long-term success. Partner closely with the sales te

AIGoRustDevOps
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$220K – $300K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Developer experience at Sentry spans four areas: the integration and developer platform, docs, community, and DevRel. The DevEx team is responsible for how developers first encounter Sentry, how they learn to use it, and how they build on top of it — from SDK configuration docs to the Discord community to the integrations ecosystem. This role owns all of it. We recently had our Head of DevEx depart, and we're looking for a senior leader to step in with real ownership and real scope. You'll sit in Marketing, partner closely with EPD, and have direct influence over both the product surface (the integration platform) and the team responsible for developer education, docs, and community. In this role you will Own strategy and execution for Sentry's integration ecosystem and developer platform APIs — working closely with EPD to define the developer-facing surface and ensure building on Sentry is a genuinely good experience Lead the documentation function across SDK configuration docs, API reference, and tutorials — the places where developer activation actually happens or doesn't Connect docs to the product experience: logged-in state, contextual guidance, health checks, and surfacing adoption gaps in-context, in partnership with EPD and the Docs platform team Own Sentry's developer community across Discord, social, and events, and build the team and processes that turn engagement into measurable outcomes: signups, adoption, support deflection Build and deliver the content and education programs — workshops, deep dives, fireside chats — that teach developers tracing, debugging, and how to actually use Sentry's full

S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Senior Software Engineer on Sentry’s AI team, you’ll be directly responsible for developing the platform used by our debugging agents. This role is crucial; you will be at the forefront of integrating AI and machine learning into our core products, from issue triage and resolution to predictive analytics for application performance monitoring. Your work will help companies around the globe gain actionable insights into their software, enabling them to build better products, faster. In this role you will Build state-of-the-art agentic AI platforms to triage, debug, and solve real production issues Leverage Sentry’s novel (and massive) dataset of errors, spans, and profiles Own the development of major initiatives in the AI/ML space You'll love this job if you Are driven by impact and enjoy working on high-stakes, high-visibility projects Enjoy building things. You will have the opportunity to join the AI/ML team as one of its foundational members Thrive in cross-functional teams and enjoy building features alongside developers and product teams Qualifications Minimum 5+ years of professional experience with Bachelor’s degree in computer science, machine learning, or a related field Demonstrated expertise building production-grade agentic systems and tools You are comfortable writing production quality code (we use Python and Typescript) Familiarity with deep learning frameworks (we use PyTorch) Familiarity in deploying machine learning models at scale in production environments The base salary range (or hourly wage range, if applicable) that Sentry reasonably expects to pay for this position is $155,000 to

TypeScriptPythonMachine LearningAI
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Staff Machine Learning Engineer on Sentry’s AI/ML team, you’ll be directly responsible for developing the models and agents used to make our product smarter and more capable. This role is crucial; you will be at the forefront of integrating AI and machine learning into our core products, from issue triage and resolution to predictive analytics for application performance monitoring. Your work will help companies around the globe gain actionable insights into their software, enabling them to build better products, faster. In this role you will Build state-of-the-art agentic AI systems to triage, debug, and solve real production issues Leverage Sentry’s novel (and massive) dataset of errors, spans, and profiles Own the development of major initiatives in the AI/ML space You'll love this job if you Are driven by impact and enjoy working on high-stakes, high-visibility projects Enjoy building things. You will have the opportunity to join the AI/ML team as one of its foundational members Thrive in cross-functional teams and enjoy building features alongside developers and product teams Qualifications Minimum 4+ years of professional experience with a MS/PhD degree in computer science, machine learning, or a related field Minimum 6+ years of professional experience with Bachelor’s degree in computer science, machine learning, or a related field Demonstrated expertise building production-grade agentic systems and tools You are comfortable writing production quality code (we use Python) Expertise with deep learning frameworks (we use PyTorch) Familiarity in deploying machine learning models at scale in production

PythonMachine LearningAIRust
S
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -72.4%

$155K – $400K/yr

Quick readStrong listing-quality and freshness signals

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Senior Software Engineer on Sentry’s AI/ML team, you’ll be responsible for building the evaluation infrastructure that measures the accuracy, reliability, and real-world performance of our AI systems. This role is critical to ensuring that our debugging agents and AI-powered features behave correctly, safely, and predictably as they scale. You’ll design datasets, benchmarks, and test harnesses that turn ambiguous AI behavior into measurable signals, helping the team ship AI with confidence. In this role you will Design and build robust evaluation frameworks to measure accuracy, reliability, regressions, and edge cases in AI systems Create and curate high-quality datasets, golden test cases, and benchmarks grounded in real production data Build automated test harnesses and metrics pipelines to continuously evaluate models, prompts, and agentic workflows Partner closely with applied AI engineers and product leaders to define what “good” looks like and translate it into measurable criteria Own the evaluation lifecycle for major AI initiatives, from early experimentation through production monitoring You’ll love this job if you Care deeply about correctness, rigor, and measurement in AI systems Enjoy turning fuzzy product goals and model behavior into concrete tests and metrics Like building foundational infrastructure that unlocks faster iteration and higher confidence for the entire AI team Thrive in cross-functional environments and enjoy influencing model design through better evaluation Qualifications Minimum 5+ years of professional experience with a Bachelor’s degree in computer science, machine learni

TypeScriptPythonMachine LearningAI
🔔

Get new performance modeling engineer 2 jobs in San Francisco, United States by email

Daily job updates · Unsubscribe anytime