About the Team We’re hiring Software Engineers to join our broader Infrastructure organization, which supports multiple high-impact teams. Depending on your interests and experience, you could work on one of several focus areas—including Core Distributed Systems, Reliability Engineering, Observability, Developer Productivity or Cloud Infrastructure. About the Role All teams are deeply collaborative, work on mission-critical services, and are responsible for building distributed, scalable infrastructure to bring OpenAI’s technology to the world through products like ChatGPT and the OpenAI API. You’ll work closely with stakeholders to understand infrastructure, data and compute needs, setting the technical strategy that supports cutting-edge research and product development. This is a critical role for someone who is passionate about solving complex engineering problems at scale, ensuring their performance, scalability and reliability Team Focus Areas Distributed Systems: Owning and building important, highly scalable, available, performant, and reliable distributed systems (and their building blocks) to power the entire stack at OpenAI Systems Engineering: Work across layers of the stack—debugging system bottlenecks, evolving core infrastructure, and solving novel problems in performance and scalability. Reliability Engineering: Build scalable, fault-tolerant systems and lead efforts around service health, incident response, and resilience. Observability: Design and maintain observability tooling (metrics, logs, tracing) to give teams visibility into production systems at scale. Developer Productivity: Create tools, environments, and workflows that help engineers ship high-quality software faster and more safely. Cloud Infrastructure: Own the cloud-native infrastructure (compute, networking, storage) that underpins all services and research workloads. Databases: Building high performance, distributed database systems that power all of OpenAI's product stack. In this
Jobs in United States
Performance Modeling Engineer 2 in San Francisco
364 active opportunities · Updated October 2026
Showing
15 jobs
Explore current performance modeling engineer 2 jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role: Cohere seeks a Strategic Sourcing Specialist to manage our Professional Services Procurement function. In this role you will be responsible for executing Strategic Sourcing activities across these categories as we continue to grow rapidly. You will collaborate with FP&A, Accounting, Legal, and multiple business stakeholders to support our Procurement operations. This is an opportunity for a hands-on Procurement professional who thrives in a high-growth environment, is comfortable navigating ambiguity, and can help implement efficient processes. You’ll be instrumental in building Procurement categories for Professional Services spend and owning their development and optimization. You’ll have the opportunity to establish best practices, drive significant cost savings, and create strategic partnerships with key service providers that support our business growth. Contract Terms: This is a 12-month fixed-term contract position with the potential for conversion to a permanent role based on business needs and performance. Please note that conversion to a permanent position is not automatic and will be subject to a se
About the Team The Consumer Devices team at OpenAI builds end-to-end hardware and software systems that bring AI into the physical world. We work at the intersection of custom silicon, embedded systems, operating systems, and cloud services to deliver reliable, production-ready devices at scale. About the role We are looking for an Operating Systems Engineer to build and harden the OS foundations for OpenAI products. We are especially interested in experienced, passionate, and innovative operating systems developers who thrive on building foundational platform software and solving hard problems in security, privacy, performance, power, and reliability. You will work across the OS kernel, core OS services, security and privacy primitives, performance and power, and the frameworks that connect applications and UI to the system. This role emphasizes deep debugging and systems ownership from development through production. You will collaborate closely with embedded, firmware, hardware, application, and product engineering teams. Experience with hardware bring-up is a plus, but not required. What you will do Work on end-to-end OS capabilities spanning the OS kernel, userspace services, application frameworks, UI toolkits, and application-facing APIs. Develop, integrate, and maintain OS components, both kernel-bound and in userspace, including scheduling, memory management, filesystems, drivers, IPC/RPC mechanisms, and security-relevant subsystems. Build and maintain core OS services and daemons (init, service management, device discovery, networking primitives, time, logging, update hooks, crash handling, and so on). Design and implement security and privacy mechanisms: Secure boot and measured boot integration points (where applicable). Mandatory access control and sandboxing. Secrets management, secure storage, key handling, and least-privilege service design. Privacy-preserving telemetry, data minimization, and user-consent oriented system behaviors. Establish a perfo
About the Team OpenAI's People team helps hire, develop, and support the people building safe and beneficial AGI. Within that team, People Systems builds the technical foundation that enables our HR, recruiting, payroll, benefits, and performance operations to scale with quality, speed, and rigor. We work at the intersection of HR systems, software engineering, and internal tooling. Our goal is not just to keep core systems running, but to build durable technical leverage for the company. About the Role We're hiring a Workday Engineer to help design, build, and operate the systems that power critical people workflows at OpenAI. This is a highly technical role for someone who combines strong Workday expertise with real engineering fluency. You'll build reliable integrations, improve system architecture, automate complex workflows, and help connect Workday to internal tools, external platforms, and emerging AI-driven systems. You should be comfortable going beyond configuration work. We're looking for someone who can reason through ambiguous systems problems, write and debug technical solutions, work effectively in Git-based environments, and use modern developer workflows, including CLI-driven tooling, to build and operate with speed and discipline. You'll partner closely with cross-functional teams across People, Finance, Security, and Engineering, including our People Innovations team, to build systems that are secure, scalable, and practical. Some work will involve improving mature production infrastructure; some will involve building entirely new workflows and capabilities from scratch. In this role, you will Design, build, and maintain Workday integrations, applications, and workflow automations across domains such as payroll, benefits, recruiting, performance, and case management Improve the reliability, quality, and scalability of People systems through strong engineering, testing, and operational practices Build technical solutions that connect Workday with i
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will work on the systems software strategy and execution that brings new AI silicon from first power-on to a fully integrated system running production-representative models at expected functionality and performance. You will define how software exercises and validates compute, memory, interconnect, and I/O subsystems, then build the diagnostics, automation, and observability needed to find issues quickly. This role sits at the center of silicon, firmware, platform, systems, and workload teams. You will turn hardware specifications and performance targets into an end-to-end bringup plan, drive cross-functional debug, and establish the stress and regression infrastructure that makes each new platform reliable across operating environments. In this role, you will: Contribute to the end-to-end software bringup and validation strategy for new silicon and first-party systems. Define software-driven test coverage across compute, memory, interconnect, I/O, and their system-level interactions. Build diagnostics, test automation, telemetry, and regression infrastructure that accelerate first-silicon learning and issue isolation. Lead bringup from initial silicon arrival through board and system integration, docking, runtime enablement, and model execution. Design stress tests that characterize reliability, performance, and stability across workloads and operating conditions. Translate architecture specifications and performance models into measurable acceptance crit
About the Team OpenAI, in partnership with our capital and technology partners, is building a global network of advanced datacenters to support the most demanding AI workloads. The Infrastructure Quality team ensures that all datacenter systems are manufactured, delivered, and commissioned to the highest standards of quality, reliability, and performance. We work closely with manufacturing partners, general contractors, engineering teams, and operations staff to ensure that every component is delivered ready for installation, startup, and long-term service. Our work spans from vendor qualification through commissioning, ensuring operational readiness across our global portfolio. About the Role We are seeking an experienced Manufacturing Quality Engineer (MQE) to establish, implement, and manage a manufacturing-focused quality program for datacenter infrastructure. This role will be responsible for vendor oversight, quality assurance, process improvement, and issue resolution for all critical systems. You will lead vendor audits, monitor performance metrics, and coordinate corrective actions to ensure predictable delivery schedules, reduced risks, and operational reliability. By partnering with vendors, construction teams, and internal stakeholders, you will help ensure OpenAI’s datacenters are delivered on time and built to the highest operational standards. Travel Domestic and international travel as needed (estimated 40–60%) to manufacturing sites, datacenter locations, and partner facilities. Key Responsibilities Vendor Oversight & Performance Management Conduct manufacturing evaluation, audits, and improve vendor performance across production, inspection, testing, and delivery phases. Develop and track quality metrics to assess manufacturing performance and identify trends. Partner with vendors to refine processes, training, and quality controls to mitigate risks before shipment. Program Development & Execution Develop and maintain a datacenter-focused m
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking an Electrical Engineer to build and own the electrical backbone of our robotic actuator dynamometer and test infrastructure. You will design, integrate, and operate the load motor drives, power distribution, instrumentation wiring, DAQ interfaces, and safety systems that make high-performance robotic actuator testing repeatable, safe, and scalable. This role spans hands-on lab execution and system architecture: selecting and commissioning power electronics, designing robust test-cell electrical systems, bringing up sensors and DAQ, and partnering with mechanical and software engineers to turn robotic actuator hardware into trustworthy data. In this role, you will Own the electrical architecture of dynamometer and actuator test cells, from mains distribution and protection through load motor drives, braking, and auxiliary power. Specify, integrate, commission, and tune motor drives and load machines for robotic actuator torque, speed, efficiency, thermal, and durability testing. Design power distribution, grounding, shielding, cable routing, and connectorization for high-current, high-voltage, and low-level measurement systems. Integrate torque, position, speed, temperature, voltage, current, vibration, and other instrumentation from robotic actuators into DAQ and control systems. Develop electrical schematics, wiring diagrams, panel layouts, harness documentation, and test-cell interface definitions. Build, debug, and maintain test-cell electrical hardware, rapidly diagnosing noise, EMI, grounding, drive, sensor, and power-q
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re seeking a GPU Kernel Engineer to join our team at the cutting edge of AI acceleration, where your code directly impacts the performance of state-of-the-art machine learning models. As a GPU Kernel Engineer, you'll craft the foundation that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You'll work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Core Engineering Responsibilities Design and implement high-performance GPU kernels for key ML operations, including matrix multiplications, attention mechanisms, and mixture-of-experts routing Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core acceleration, and compute/memory overlap Performance & Innovation Impl
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten’s Model Performance (MP) team is responsible for ensuring the models running on our platform are fast, reliable, and cost‑efficient. As part of this team, you’ll focus on Model APIs — the infrastructure powering our hosted API endpoints for the latest open‑source models. This work spans distributed systems, model serving, and developer experience. You’ll join a small, high‑impact team operating at the intersection of product, model performance, and infra, helping to define how developers interact with AI models at scale. RESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups Productionize performance improvements across runtimes with deep understanding of their internals: speculative decoding implementations, guided generation for structured outputs, custom scheduling and routing algorithms for high-performance serving Build comprehensive benchmarking frameworks that measure real-world performance across different model architectures, batch sizes, sequence lengths, and hardware configurations Productionize performa
What you’ll do Partner with medical image reconstruction scientists / engineers to build ML components that improve reconstruction quality, speed, robustness, or quantitative accuracy. Define training/evaluation pipelines, datasets, and metrics that map to user needs and design requirements. Productionize models: inference performance, reproducibility, monitoring for drift/regressions, and safe fallbacks. Collaborate on hybrid algorithms, incorporating physics and learned priors, denoisers, learned regularizers, and quality estimation. Help build tooling for rapid experimentation as well as rigorous verification of algorithm changes. What we’re looking for Strong applied ML experience plus comfort with signal processing / imaging or adjacent domains. Ability to move fluidly between research prototypes and production-quality systems. Strong evaluation discipline: metrics, ablations, data leakage avoidance, and reproducibility. A demonstrated track record of applying ML to physics-based or inverse problems (i.e., shipped projects, a portfolio, or publications.) Useful experience ML for imaging/inverse problems (or adjacent) with strong evaluation discipline and comfort with GPU performance constraints. Pragmatic production mindset: reproducible training/inference, regression testing, and safe deployment in high-stakes contexts. A background in computational physics or scientific computing. Leverage ML-based methods such as PiNNs and Neural Operators to solve partial differential equations arising in ultrasound simulation and imaging. Experience in Agentic-SciML is a plus. Hands-on experience with data curation for ML: building datasets from messy, real-world sources, defining ground truth, and managing labeling or simulation pipelines. Background in data assimilation: combining observations with physics-based models (Kalman filtering, variational methods, ensemble approaches, or learned variants).
About the Team We are a small and fast-moving partnerships team that shapes and executes OpenAI’s most important collaborations. Your mission is to design and operate the partner program for technology partners, ISVs, and marketplace participants. You will turn partner strategy into a simple, scalable experience from recruitment and onboarding through enablement, benefits, performance management, and growth. About the Role You are an operational builder with strong judgment, high ownership, and a track record of turning ambiguous goals into clear programs and measurable outcomes. You are comfortable working across product, engineering, GTM, marketing, operations, legal, finance, security, and support. You balance rigorous systems and governance with a thoughtful partner experience, and you can move quickly as the ecosystem evolves. In this role, you will: Design and run the end-to-end ISV and marketplace partner program. Define partner eligibility, segmentation, tiers, benefits, requirements, and lifecycle stages. Build repeatable onboarding, technical enablement, validation, listing, launch, and growth journeys. Establish program policies, documentation, service levels, governance, and escalation paths. Create partner communications, education, office hours, playbooks, and self-service resources. Develop program metrics and dashboards covering recruitment, activation, solution quality, adoption, revenue impact, and partner health. Coordinate launches and program changes across product, engineering, GTM, marketing, legal, security, finance, support, and operations. Capture partner feedback and translate it into program, product, and operating improvements. You might thrive in this role if you have: 8+ years of experience in partner programs, marketplaces, ecosystem operations, business operations, or program management. Proven experience building or scaling a technology partner, ISV, or marketplace program. Exceptional ability to convert ambiguous goals into clear p
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking an Actuator Gear Design Engineer to lead the development of custom gears and gear stages for advanced robotic systems. You will own actuator development from early architecture and concept generation through prototype validation and system integration, partnering closely with mechanical, electrical, controls, firmware, and reliability teams. You will partner with external suppliers and internal manufacturing to create full gearbox assemblies. This role focuses on the design, integration, and validation of precision gearing, including broader knowledge around motor electromagnetics, transmission types, sensing, structural components, and thermal architectures. You will help drive actuator development across the full engineering lifecycle while establishing scalable design, test, and integration practices for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will: Lead the architecture, design, and integration of custom robotic actuator gearing. Define actuator requirements and system-level trade studies around torque density, bandwidth, efficiency, thermal performance, back drivability, inertia, reliability, manufacturability, and cost. Design precision electromechanical assemblies with strong attention to tolerances, alignment, load paths, thermal expansion, sealing, wear, and serviceability. Drive actuator integration into robotic systems, partnering closely with controls, firmware, electrical, and robotics software teams to optimize closed-lo
About the Team The B2B Marketing team is responsible for helping businesses understand, adopt, and get value from OpenAI’s products. B2B marketing is a major and growing priority for OpenAI as we scale our work with companies, developers, and institutions around the world. About the Role Within B2B Marketing, Demand Generation builds the integrated, full-funnel engine that connects audience insights, content, field and digital experiences, paid media, lifecycle, and sales follow-through to qualified pipeline. We partner closely with Sales, Partnerships, Product Marketing, Communications, Creative, Web, RevOps, Analytics, and regional teams to create a cohesive customer experience and scale what works. We’re looking for a Senior Marketing Strategist to lead paid and emerging demand-channel strategy for OpenAI’s B2B business. You’ll own the B2B audience, channel, offer, partner, and measurement strategy, translating pipeline goals into clear briefs and investment recommendations. You’ll partner closely with Performance Marketing, which owns hands-on media execution, while informing audience strategy, campaign priorities, experimentation, and downstream performance decisions. You’ll also build our approach to content syndication and third-party demand media partnerships. In this role, you will: Owning the B2B paid and emerging-channel demand strategy across paid search, paid social, display and retargeting, content syndication, third-party demand media partnerships, sponsorships, review platforms, and selected emerging channels. Translating pipeline goals into audience and channel strategies, including ICP and segment priorities, buying groups, intent signals, channel roles, forecast assumptions, budget recommendations, campaign architecture, and a rolling experimentation roadmap. Partnering closely with Performance Marketing to define B2B audience, campaign, offer, creative, landing-page, conversion, and measurement requirements, while Performance Marketing owns media
About the Team The Growth team drives user and revenue growth across ChatGPT’s consumer and business segments as well as other OpenAI products worldwide. We operate across the full funnel - from awareness and acquisition through activation, retention, and expansion - using a combination of global performance marketing, AI-powered workflows, in-product optimization, insights, experimentation, and creative ops engineering. About the Role We are hiring a Growth Marketing Manager to turn product-led growth priorities into clear audience strategies, compelling cross-channel messaging, coordinated go-to-market programs, and drive measurable, incremental growth. You will partner closely with Growth Product, Consumer Product Marketing, Lifecycle, Performance Marketing, Brand, Creative, Research, Data Science, and International Marketing to ensure that full-funnel campaigns and initiatives are supported and connected to Growth capabilities and maximize the performance of these campaigns. In this role, you will: Build the marketing layer across the Growth roadmap: acquisition, access, activation and resurrection, monetization, and the shared Growth Platform. Translate product hypotheses into audience insights, positioning, messaging, launch briefs, lifecycle journeys, landing-page narratives, performance creative, and in-product education. Maintain an integrated plan that connects the Growth product roadmap to the consumer product marketing calendar across lifecycle, performance, social, creators, partnerships, and brand moments. Serve as the subject matter expert across all Growth levers, advising teams across the organization on the most effective ways to integrate Growth into their initiatives. Partner with cross-functional teams to design and interrogate the user journey so we can align our external promises with in-product and owned-channel landing, access, onboarding, continuation, and conversion experiences. Create audience and market playbooks for new users, high-valu
About the Team The International Strategy & Operations team supports the growth and management of OpenAI’s business outside the United States. We work across products, markets, and functions to ensure our technology reaches and benefits users and customers around the world. Our team is responsible for being the glue between global and regional teams—bringing together market performance, local context, and cross-functional priorities to create a clear plan for winning in international markets. We advocate for the regional leaders closest to our users and customers, while giving company leadership the ground truth and structured recommendations needed to make the right decisions and tradeoffs. About the Role We are looking for a world-class, scrappy, and dynamic strategy and operations leader who spikes in analytical thinking, business judgment, executive communication, and operational execution. You will work closely with global leadership, regional general managers, and cross-functional partners to identify growth opportunities, build operating plans, and drive high-priority initiatives from conception through execution. You will be expected to operate at all altitudes—from translating complex market dynamics into clear recommendations for senior executives to working directly with product, growth, data science, and go-to-market teams to unblock execution. Success in this role requires the ability to bring structure to ambiguity, turn data into actionable insights, influence without authority, and drive meaningful business outcomes across a fast-moving and increasingly global organization. In this role, you will: Build the operating plan. Translate strategic priorities into clear goals, workstreams, owners, milestones, and decisions. Identify cross-functional dependencies early and ensure teams remain aligned on execution. Drive market growth initiatives. Work with product, growth, marketing, data science, and regional teams to identify and execute opportunities
Other cities to consider
More places hiring for this role
Get new performance modeling engineer 2 jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime