About the Team OpenAI's Training team is responsible for producing the large language models that power our research, our products, and ultimately bring us closer to AGI. Achieving this goal requires combining deep research into improving our current architecture, datasets and optimization techniques, alongside long-term bets aimed at improving the efficiency and capability of future generations of models. We are responsible for integrating these techniques and producing model artifacts used by the rest of the company, and ensuring that these models are world-class in every respect. Recent examples of artifacts with major contributions from our team include GPT4-Turbo, GPT-4o and o1-mini. About the Role As a member of the architecture team, you will push the frontier of architecture development for OpenAI's flagship models, enhancing intelligence, efficiency, and adding new capabilities. Ideal candidates have a deep understanding of LLM architectures, a sophisticated understanding of model inference, and a hands-on empirical approach. A good fit for this role will be equally happy coming up with a creative breakthrough, investing in strengthening a baseline, designing an eval, debugging a thorny regression, or tracking down a bottleneck. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, prototype and scale up new architectures to improve model intelligence Execute and analyze experiments autonomously and collaboratively Study, debug, and optimize both model performance and computational performance Contribute to training and inference infrastructure You might thrive in this role if you: Have experience landing contributions to major LLM training runs Can thoroughly evaluate and improve deep learning architectures in a self-directed fashion Are motivated by safely deploying LLMs in the real world Are well-versed in the state of the art tran
Jobs in United States
Model Behavior Engineer in United States
2,174 active opportunities · Updated October 2026
Showing
15 jobs
Explore current model behavior engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team Training Runtime designs the core distributed machine-learning training runtime that powers everything from early research experiments to frontier-scale model runs. With a dual mandate to accelerate researchers and enable frontier scale, we’re building a unified, modular runtime that meets researchers where they are and moves with them up the scaling curve. Our work focuses on three pillars: high-performance, asynchronous, zero-copy tensor and optimizer-state-aware data movement; performant, high-uptime, fault-tolerant training frameworks (training loop, state management, resilient checkpointing, deterministic orchestration, and observability); and distributed process management for long-lived, job-specific and user-provided processes. We integrate proven large-scale capabilities into a composable, developer-facing runtime so teams can iterate quickly and run reliably at any scale, partnering closely with model-stack, research, and platform teams. Success for us is measured by raising both training throughput (how fast models train) and researcher throughput (how fast ideas become experiments and products). About the Role As a Training Performance Engineer, you’ll drive efficiency improvements across our distributed training stack. You’ll analyze large-scale training runs, identify utilization gaps, and design optimizations that push the boundaries of throughput and uptime. This role blends deep systems understanding with practical performance engineering — analyzing GPU kernel performance, collective communication throughput, investigating I/O bottlenecks, and sharding our models so we can train them at massive scale. You’ll help ensure that our clusters are running at peak performance, enabling OpenAI to train larger, more capable models with the same compute budget. This role is based in San Francisco, CA. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you will: Profil
About the Team The Cooperative AI team is scaling OpenAI with OpenAI. We are building a model powered knowledge system that evolves and learns as our products, systems and customers evolve. We leverage our state of the art models, technologies, and products (some external, some still in the lab) to assist or completely automate robust operations supporting both internal and external customers. We support OpenAI customers and internal partners globally, powering systems from customer support to integrity to product insights. We are a self-contained multi-disciplinary team, who enjoy a lightning fast feedback loop with customers at scale, some of whom sit just a few pods away. We iterate fast, and engineer for reliable long-term impact. We're constantly looking for the similarities and patterns in different types of work, and focus on building simple primitives, to apply world class knowledge to many domains. The work of this team exemplifies use of OpenAI technologies. We build systems so everyone can see the leverage that is possible with well designed AI-based implementations. We do this by working through internal use cases focused on Customers (specifically knowledge systems, automation systems, and automated agent systems) to prove impact, then we scale. About the Role We’re looking for a Backend Software Engineer to help architect and scale the infrastructure that powers our knowledge systems. This is a deeply technical and highly cross-functional role where you’ll build robust systems and backend services that serve as the foundation for how knowledge is created, accessed, and applied across OpenAI. In this role, you will: Design, build, and maintain backend services and APIs to support intelligent automation and knowledge systems Integrate and structure data across internal platforms, transforming it into formats optimized for use by downstream systems and AI workflows. Collaborate closely with product, research, and engineering teams to integrate OpenAI mode
About the Team The Cooperative AI team is scaling OpenAI with OpenAI. We are building a model-powered scaled automated workforce and knowledge system that evolves and learns alongside a human workforce. By leveraging OpenAI’s state-of-the-art models and technologies, some already in production, others still in the lab, we develop systems that reason and work autonomously for a wide variety of operational work. We leverage real workloads for critical systems across finance, sales, customer support, integrity, product insights, internal operations, and more in order to drive insights into product and industry. We partner closely with internal teams and external customers globally, operating in a hyper-fast feedback loop where many of our users are just a few steps away. This proximity allows us to iterate quickly, validate impact in real time, and accelerate industry impacting learnings and systems builds. We are a highly multidisciplinary, self-contained team focused on transforming the workplace via smart systems, knowledge, scalable and reliable primitives that apply world-class AI capabilities across domains. Our mission is to learn fast and transform how humans collaborate with AI at scale. About the Role We are looking for a hands-on Engineering Manager to lead a small, fast-moving team building AI-powered automation systems that redefine how work gets done across OpenAI. This role sits at the intersection of applied AI, research, and product engineering. You’ll lead a team that builds systems that know how to learn from humans, and carry real workloads across, sales, support, finance, IT, and more, while staying deeply involved in the technical work. You will operate in a highly iterative environment, deploying systems directly to internal users, gathering rapid feedback, and evolving solutions in real time. This is a high-ownership role for someone excited about building 0→1 systems, working closely with customers, and shaping how AI transforms operational wor
Join a pioneering team at the forefront of financial technology innovation. We are building a cutting-edge, agentic AI platform designed to revolutionize the end-to-end credit risk model development lifecycle. By leveraging the power of Large Language Models (LLMs) and intelligent workflows, we aim to augment our quantitative modelers, dramatically increasing their productivity, enhancing model governance, and accelerating the delivery of critical risk models. We are seeking a senior, hands-on technology leader to drive this transformation. In this role, you will partner directly with quantitative analysts and key stakeholders to shape the most impactful use cases and then lead a team of talented developers to design, build, and deploy the generative AI platform that brings this vision to life. Key Responsibilities Strategic Vision & Stakeholder Partnership: Collaborate closely with quantitative model developers, Model Risk Management (MRM), and business leaders to deeply understand their pain points and translate them into a technical vision and product roadmap. Identify, scope, and prioritize use cases for the agentic workflow platform, ensuring they deliver measurable value and align with strategic objectives. Serve as the primary technical liaison between the engineering team and its users, ensuring a tight feedback loop and continuous alignment. Platform Design & Hands-On Development: Lead the architectural design of a scalable, robust, and secure agentic AI platform, including core components for orchestration, knowledge retrieval (e.g. RAG), and tool integration. Engage in hands-on software development to build foundational components, create proofs-of-concept, and tackle the most complex technical challenges. Design and implement secure integrations with internal data sources, external APIs, and various LLMs, ensuring compliance
About ElevenLabs ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in January 2023 with the first human-like AI voice model. Today, we serve millions of users and thousands of businesses - from fast-growing startups to large enterprises like Deutsche Telekom and Meta. Our investors are some of the world's most prominent, including Andreessen Horowitz, ICONIQ Growth and Sequoia. We've raised $781M in funding and our last valuation was $11B - multiples of 11, always. We have expanded from voice into three main platforms: ElevenAgents enables businesses to deliver seamless and intelligent customer experiences, with the integrations, testing, monitoring, and reliability necessary to deploy voice and chat agents at scale. ElevenCreative empowers creators and marketers to generate and edit speech, music, image, and video across 70+ languages. ElevenAPI gives developers access to our leading AI audio foundational models. Everything we do is the result of the creativity and commitment of our team - builders doing the best work of their lives. We are researchers, engineers, and operators. IOI medalists and ex-founders. If you want to work hard and create lasting positive impact, we want to hear from you. How we work High-velocity: Rapid experimentation, lean autonomous teams, and minimal bureaucracy. Impact not job titles: We don’t have job titles. Instead, it’s about the impact you have. No task is above or beneath you. AI first: We use AI to move faster with higher-quality results. We do this across the whole company—from engineering to growth to operations. Excellence everywhere: Everything we do should match the quality of our AI models. Global team: We prioritize your talent, not your location. What we offer Innovative culture: You’ll be part of a generational opportunity to define the trajectory of AI, surrounded by a team pushing the boundaries of what’s possible. Growth paths: Joining ElevenLabs means joining a
About the Team Life sciences is one of the clearest areas where advances in intelligence can meaningfully benefit the world at large. The OpenAI Life Sciences team works at the intersection of advancing frontier life sciences model capabilities and building products to help scientists accelerate their research and leverage the full potential of AI for scientific work. Rosalind Workbench brings together the scientific tools, data sources, interactive biology file viewers and core life sciences workflows into a central environment to help scientists investigate questions, design experiments, analyze results, and advance discovery. Rosalind Workbench can be used with any OpenAI model, including GPT-Rosalind--our dedicated life sciences model, which combines frontier reasoning with specialized tool orchestration across medicinal chemistry, genomics, wet-lab assistance, and other scientific applications. Our long term vision is for teams of agents to work together across these domains, giving researchers access to broader expertise and the ability to pursue more ambitious scientific questions. About the Role We’re looking for a Product Manager to shape Rosalind Workbench for life sciences. You will own product strategy and execution, working directly with researchers in academic labs, biotech, and pharma to understand where AI can meaningfully improve their work. You’ll partner with engineering, research, design, and customer-facing teams to turn emerging capabilities into intuitive products that scientists return to. This role calls for strong product judgment, depth in life sciences, and the ability to move from an ambiguous research problem to a focused, shippable experience. In this role, you will have the opportunity to define the future of AI guided scientific discovery and build the capabilities and tools that help advance the scientific frontier for researchers and help increase the accessibility of model intelligence for scientific use. This is a chance to shape
$99.2K – $165.4K/yr
Role Summary The CISO, Infrastructure and Cloud Services organization enables Pfizer’s mission by delivering secure, resilient, and scalable technology platforms that support operations and future growth. Through a unified operating model, it strengthens accountability, accelerates decision‑making, and improves availability, security, and cost efficiency. The CISO Marketing Manager will support the Senior Manager, Marketing and Enterprise Awareness to develop meaningful marketing materials that tell the CISO story to external clients by executing strategic, creative, and visually compelling communications including presentations, graphics, pictures, charts and tables that support cybersecurity awareness, organizational priorities, and business outcomes. The manager will lead the creation of a marketing strategy for the Chief Information Security Office Team. Role Responsibilities Lead the development and execution of integrated marketing and communications strategies in support of CISO priorities, strategic initiatives, and portfolio management objectives. Produce high-quality written, visual, and digital content that strengthens awareness, engagement, and understanding of cybersecurity programs across the organization. Create and manage internal communications, including newsletters, leadership messages, campaign materials, SharePoint content, presentations, infographics, and other digital assets. Simplify complex cybersecurity and technical information into clear, engaging, and actionable content for diverse audiences, including executives, business partners, and internal colleagues. Partner closely with cross-functional teams across Cyber Defense, Infrastructure, Cloud, and related Digital and Technology groups to ensure consistent, coordinated, and impactful messaging. Develop executive-ready presentations and visual storytelling materials
$1.2M – $1.3M/yr
About us EVERY™ is a leading VC-backed food tech ingredient company and market leader using precision fermentation to create animal proteins without the animal for the global food and beverage industry. EVERY™ is a team of passionate change-makers who are reimagining the factory farm model with a kinder, more sustainable alternative. Leveraging precision fermentation to produce hyper-functional and one-to-one replacement proteins from microorganisms, EVERY™ is on a mission to decouple the world’s proteins from the animals that make them. We are a passionate, determined (and fun!) team with a vital objective, and we're on the lookout for like-minded people to join our mission. For more information, visit www.every.com The Downstream Process Engineer II will be an integral member of our downstream process development team. You will use experimentation to optimize Every’s production process and then see the results of your changes in action at pilot and commercial-scale biomanufacturing sites. This is an excellent opportunity for someone with laboratory and tech transfer experience, who wants to make an impact at scale. What you'll Accomplish Optimize the Every downstream process via an iterative cycle. Improvements are developed in the laboratory, scaled up to an external pilot plant, and learnings are taken back to the lab for further troubleshooting and improvement. Perform various unit operations at the Every HQ such as TFF, depth filtration, chromatography, spray drying and other purification/separation processes. Develop and review tech transfer documentation to ensure successful scale up trials. Travel to external pilot plants to review the scale up of novel processes. Coordinate with third parties such as pilot scale equipment vendors to arrange internal and external trials. Partner with Every scientists and engineers to bring their bench ideas to pilot scale. Analyze results and report data to enable appropriate interpretation and
$900K – $1M/yr
About us EVERY™ is a leading VC-backed food tech ingredient company and market leader using precision fermentation to create animal proteins without the animal for the global food and beverage industry. EVERY™ is a team of passionate change-makers who are reimagining the factory farm model with a kinder, more sustainable alternative. Leveraging precision fermentation to produce hyper-functional and one-to-one replacement proteins from microorganisms, EVERY™ is on a mission to decouple the world’s proteins from the animals that make them. We are a passionate, determined (and fun!) team with a vital objective, and we're on the lookout for like-minded people to join our mission. For more information, visit www.every.com The Role: This unique entry-level Research Associate I position offers the rare chance to work across two core teams, Analytics and Protein Science. You’ll gain hands-on experience supporting protein development, purification, and analysis, while learning how these disciplines work together to drive innovation in precision fermentation. This is a great opportunity for someone early in their career who thrives in the lab, loves variety, and wants to learn fast in a collaborative, mission-driven environment. What you'll accomplish Generate data through basic biochemistry/molecular biology techniques (including but not limited to BCA, SDS-PAGE) Operate and maintain analytical equipment (e.g., HPLC-UV/RI and Combustion Analyzer, dynamic light scatterer, FPLC-UV, fluorescence/UV plate reader) Support protein characterization workflows through lab-scale protein powder generation involving bench-scale downstream processing unit operations (microfiltration, ultrafiltration, diafiltration) Prepare samples, reagents, and buffers to support cross-functional experiments Collaborate with scientists and engineers across teams to troubleshoot and iterate quickly as part of our Design, Build, Test, and Learn pipeline Present results
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. Our Payments Risk team builds and operates decisioning capabilities across Signal and Guarantee. We help customers tune thresholds and rules so they can approve more good transactions while reducing returns. For Guarantee, this work also protects the economics of the risk Plaid underwrites. You will be the technical, customer-facing Payment Risk Consulting Lead for Signal and Guarantee. You will interpret model outputs, diagnose customer performance, and turn that analysis into concrete threshold and rules recommendations. You will own proofs of concept and retros, improve existing integrations, and help the team identify patterns that can be scaled through better processes and product capabilities. Responsibilities Own customer proofs of concept and retros from analysis through recommendations and follow-through. Proactively optimize Signal customers, prioritizing accounts with high return rates. Read model outputs and diagnose the drivers of authorization and return-rate performance. Recommend threshold and rules changes that align with each customer's risk and authorization goals. Help Guarantee customers tune thresholds to achieve target authorization rates while protecting lo
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking Manufacturing Engineers to lead the development of processes, tooling, and prototype builds for custom motors and actuators. You will own a primary area as a Stator Process / Manufacturing Engineer, Tooling / Fixture Engineer, or Prototype Manufacturing Engineer, taking that work from early development through validation and repeatable execution in close partnership with mechanical, electromagnetic, electrical, test, quality, and supplier teams. These roles focus on the development, integration, and validation of electromechanical manufacturing capabilities, including stator winding processes, assembly and inspection tooling, and actuator prototype builds. You will help translate engineering designs into reliable hardware while establishing scalable processes, equipment, and build practices for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 5 days a week. In this role, you will Each opening focuses on one of the three specialties below, with shared responsibility for reliable processes and hardware. Stator Process / Manufacturing: develop and validate stacking, winding, termination, and potting processes. Establish process parameters, work instructions, and defect controls that produce consistent results across trained operators. Tooling / Fixture: design and deliver winding tools, assembly fixtures, inspection gauges, and bench equipment, from CAD and drawings through fabrication, commissioning, and troubleshooting. Improve setup time, labor, and repeatability. Prototype M
About the Team The Emerging Products team is a lean, high-output product lab group that builds products at the forefront of model capabilities. We collaborate across all teams within the company, from research and infrastructure to consumer products. The team is responsible for identifying new product opportunities, building them quickly, dogfooding them internally, and then launching the successful products to users. We use data, user research, and analytics to inform our ideas, and make decisions on what experiments are worth iterating, stopping, or scaling. About the Role We’re looking for a senior, product-minded software engineer to own ambiguous 0-to-1 work from idea through prototype, validation, and handoff. This is a full-stack role with a strong frontend and product emphasis: you will build the interfaces and supporting backend systems needed to test new experiences quickly, while making sound architectural choices that enable successful concepts to scale. This role is based in our Mission Bay office in San Francisco. In this role, you will: Build and ship high-quality, product experiments across the full stack. Turn ambiguous user needs and emerging technical capabilities into testable product concepts, using research and metrics to guide iteration. Own technical direction for 0-to-1 projects, balancing speed, reliability, and a clear path from prototype to scalable product. Partner closely with design, product, research, and engineering teams to dogfood, evaluate, launch, and transition successful experiments. You might thrive in this role if you: Have a track record of building and shipping end-to-end products in fast-moving, startup, founder-led, growth, or other high-ownership environments. Bring strong frontend engineering skills and enough backend and systems depth to make sound full-stack architectural decisions. Pair product intuition with evidence, using user research and product data to identify opportunities and make pragmatic tradeoffs. Operat
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking a Mechanical Design Engineer to lead the development of motor mechanical designs and prototype hardware for advanced robotic systems. You will own stator and rotor mechanical development from early concepts and detailed design through hands-on builds and prototype validation, partnering closely with electromagnetic, electrical, test, and manufacturing teams. This role focuses on the design, integration, and validation of motor components and prototype processes, including laminations, stack assemblies, bobbins, winding interfaces, and fixtures. You will help drive motor development from initial design through repeatable low-volume builds while establishing the tooling, work instructions, and validation practices needed for future robotic platforms. This role is based in San Francisco, CA, and requires in-person presence 4 days a week. In this role, you will Own stator and rotor mechanical designs and released CAD and drawings, including geometry, interfaces, fits, tolerances, retention, assembly access, and mechanical validation. Develop laminations, stack assembly methods, bobbins, and winding interfaces that control alignment, insulation, conductor placement, and end-turn packaging. Design, fabricate, and debug fixtures for winding, stacking, assembly, alignment, and inspection. Build and troubleshoot prototypes. Use measurements, defects, rework, and assembly effort to improve designs and processes. Establish low-volume prototype production with equipment, build sequences, work instructions, revision control, traceability
About the Team Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role We are seeking a Mechanical Engineer to design, build, and own the mechanical side of our robotic actuator dynamometer and test infrastructure. You will create the test stands, couplings, fixtures, load paths, guarding, and serviceable lab hardware that enable rigorous characterization of robotic actuators. This role combines precision mechanical design with hands-on lab work. You will take robotic actuator test infrastructure from requirements and analysis through CAD, fabrication, assembly, commissioning, and iteration, partnering closely with electrical and software engineers to deliver safe, flexible, high-uptime test cells. In this role, you will Own the mechanical architecture of dynamometer and actuator test cells, including frames, bases, load paths, alignment, guarding, and serviceability. Design dynamometer structures, robotic actuator fixtures, load-motor mounts, couplings, shafts, bearings, adapters, and torque-reaction hardware. Translate robotic actuator test requirements into robust mechanical systems for torque, speed, thermal, durability, backdrive, efficiency, and failure testing. Perform first-principles analysis and simulation for stiffness, strength, fatigue, vibration, thermal growth, critical speed, and safety factors. Create precise, repeatable alignment strategies that protect test articles, load machines, sensors, and couplings. Design modular fixturing that supports rapid changeover across actuator and motor variants without compromising measurement quality. Work closely with electrical engineers on cable routing
Other cities to consider
More places hiring for this role
Get new model behavior engineer jobs in United States by email
Daily job updates · Unsubscribe anytime