Jobs in United States

Model Behavior Engineer in United States

2,174 active opportunities · Updated October 2026

Explore current model behavior engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

C
📍 Redwood City, United States
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary: At CVS Health, we are focused on transforming health care for our customers, and on making our company a great place to work. We help people navigate the health care system – and their personal health care – by improving access, lowering costs and being a trusted partner for every meaningful moment of health. Within our Retail locations, we bring this promise to life with heart, every day. Our retail Store Team Leaders play a critical role in building and leading teams to consistently deliver on our brand promise. As a CVS Health Store Team Leader (STL), you will be a single-unit retail leader responsible for leading your team to achieve operational and service excellence. The STL will achieve success through managing, inspiring and coaching their direct reports, consisting of Front Store (FS) crew and FS management positions. The STL will lead the store team through effective communication, consistent application of SOPs and best practices, provisioning of appropriate support to all departments, and scheduling to the needs of the business. A model for all, the STL will demonstrate CVS Health’s Heart At Work behaviors, HelpingWithHeart Actions, and will lead their team to do the same. Primary Roles and Responsibilities: To successfully operate their retail location, the Store Team Leader is specifically responsible to: <p

NR
📍 Atlanta, Georgia, United States· Full-time
✓ High-confidence listingCompany trend -73.9%
Quick readStrong listing-quality and freshness signals

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Opportunity New Relic is looking for a Senior Revenue Operations Manager – GTM Business Planning to own the operating system behind how we plan, pay, and scale our Go-To-Market (GTM) organization. Sitting at the intersection of Sales, Finance, and Systems, you will architect the annual GTM plan, design incentive compensation programs, and govern processes across quota, territory, and comp data. What You'll Do GTM Planning & Architecture: Own the annual planning cycle end-to-end, building capacity models and territory frameworks to ensure optimal market coverage and alignment with targets. Incentive Compensation Design: Partner with FP&A to architect and govern financially sound sales incentive plans and compensation policies that motivate seller behavior, drive strategic priorities, and eliminate payout leakage. Quota Management & Strategy: Establish fair quota methodologies, manage ongoing adjustments for transfers/hires, and evaluate rep productivity to handle market shifts. Operations & Governance: Govern monthly sales compensation data within Salesforce and downstream systems to maintain audit-ready accuracy across territories and quotas. Systems & Automation: Own the deployment of annual territories and quotas in Salesforce, drive the roadmap for GTM planning systems, and lead automation initiatives to eliminate manual effort and scale operations. This Role Requires B2B SaaS Experience: 6–8+ years in Revenue/Sales Operations, Sales Comp, FP&a

SQLAWSGitRest
C
📍 Minneapolis, United States
✓ Quality checkedCompany trend +340.2%

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary: At CVS Health, we are focused on transforming health care for our customers, and on making our company a great place to work. We help people navigate the health care system – and their personal health care – by improving access, lowering costs and being a trusted partner for every meaningful moment of health. Within our Retail locations, we bring this promise to life with heart, every day. Our Retail Store Managers play a critical role in building and leading teams to consistently deliver on our brand promise. As a CVS Health Store Manager (SM), you will be a single-unit retail leader responsible for leading your team to achieve operational and service excellence. The SM will achieve success through managing, inspiring and coaching their direct reports, consisting of Front Store (FS) crew and FS management positions. The SM will lead the store team through effective communication, consistent application of SOPs and best practices, provisioning of appropriate support to all departments, and scheduling to the needs of the business. A model for all, the SM will demonstrate CVS Health’s Heart At Work behaviors, HelpingWithHeart Actions, and will lead their team to do the same. Primary Roles and Responsibilities: To successfully operate their retail location, the Store Manager is specifically responsible to: Cultivate an

O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Pricing team sits at the center of product, go-to-market, finance, and strategy. We define how OpenAI packages, prices, and scales access to our products across consumer, SMB, and enterprise customers, turning deeply technical product usage and market signal into company-level decisions. We’re looking for a senior Data Scientist to be the first dedicated data science hire on the Pricing team. This is a rare zero-to-one role with direct exposure to OpenAI’s CFO, Head of Pricing, and senior leaders across Product and GTM. You will help build the analytical foundation for pricing at OpenAI, shape executive decisions, and define what excellent pricing data science looks like. About the Role As a founding Data Scientist for Pricing, you will design the analyses, models, experiments, and decision frameworks that guide pricing strategy across OpenAI’s business. You’ll work side-by-side with the CFO, Head of Pricing, and senior leaders across Product and GTM on ambiguous, high-leverage questions where simple reporting is not enough, translating customer behavior, product usage, revenue outcomes, and market dynamics into clear recommendations. This role combines hands-on technical depth with executive-ready storytelling. You should be excited to build from first principles, operate with high independence, and influence decisions that shape how OpenAI grows and serves customers around the world. In This Role, You Will Serve as a senior analytical partner to the CFO, Head of Pricing, Product, and GTM leaders on pricing and monetization decisions. Build the analytical foundation for pricing across consumer, SMB, and enterprise segments, from exploratory analysis to repeatable decision systems. Design and execute analyses that connect customer behavior, product usage, conversion, retention, revenue outcomes, and pricing strategy. Develop models, algorithms, experiments, and decision frameworks for complex pricing, packaging, discounting, and willingness-t

AWSRestMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -82%

About the Team OpenAI’s Hardware organization develops system and infrastructure solutions tailored to the demands of advanced AI workloads. We work across the full stack—from silicon to system integration—partnering closely with internal teams and external vendors to define and deliver next-generation AI infrastructure. Our team focuses on defining scalable, high-performance system architectures and reference designs that balance performance, cost, and operational efficiency across rapidly evolving technologies. About the Role We are seeking a 3P Architect to define and drive rack- and cluster-level reference designs in collaboration with external partners. This role is responsible for translating workload requirements and system-level goals into concrete architectures, aligning partners on critical design attributes, and ensuring vendor roadmaps meet our infrastructure needs. You will work closely with performance modeling and internal architecture teams to evaluate tradeoffs, while owning the end-to-end definition and execution of third-party system designs. This includes identifying gaps in current technologies, driving vendor development, and shaping future infrastructure capabilities. This role requires strong system intuition, cross-functional leadership, and the ability to operate effectively across internal teams and external ecosystems. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance. Key Responsibilities Define rack- and cluster-level reference architectures for AI infrastructure deployments. Translate workload requirements into clear system design specifications and partner deliverables. Collaborate with performance modeling teams to evaluate architectural tradeoffs and system behaviors. Align internal stakeholders and external partners on critical system attributes (performance, cost, power, reliability, scalability). Identify gaps in current technology offerings and dr

AWSRestAIGo
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. About the Team The Finance Data Science team owns the forecasting systems that power Snowflake’s financial planning, operating cadence, and long-term strategy. Our work informs executive decision-making, corporate planning, investor reporting, and cross-functional decisions across Finance, Sales, and Product. We build and operate production forecasting systems for Snowflake’s core money-in metrics, with a particular focus on revenue and bookings in a consumption-based business. Our forecasts are highly visible, widely used, and foundational to how the company plans and operates. This is a high-trust team operating at the intersection of statistical modeling, production systems, and financial decision-making. The Role We are hiring a Senior Applied Scientist to own and advance mission-critical forecasting systems used across the company. This role is not just about building models. It is about developing reliable, explainable, production-grade forecasting systems that leaders can trust to make decisions. You will work on high-impact, open-ended problems involving revenue forecasting, customer consumption behavior, workload ramps, renewals, and other leading indicators that feed Snowflake’s broader financial planning processes. You will partner closely with Finance, Sales, Pr

PythonSQLMachine LearningAI
C
📍 United States· Remote
✓ High-confidence listingCompany trend +340.2%
Quick readStrong listing-quality and freshness signals

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary The Manager, Care Model is responsible for owning the day‑to‑day execution of a defined care model workstream, ensuring planned activities are delivered on time and with high quality. This role translates care model strategy and direction into detailed task‑level plans, coordinates functional contributors, and drives daily follow‑through on actions, milestones, and deliverables. The Manager focuses on tactical execution, detailed tracking, and issue management, proactively identifying risks and escalating impacts that may affect the broader care model timeline. This role serves as a key operational partner to the Workstream PM and Senior Care Model leadership by providing accurate status, maintaining workstream artifacts, and ensuring reliable inputs into integrated reporting. Own day‑to‑day execution of a defined care model workstream, delivering tasks and milestones as planned Develop and maintain task‑level work plans; track actions, owners, dependencies, and due dates Coordinate functional contributors to support timely completion of workstream deliverables Identify and manage workstream‑level risks and issues; propose mitigation options and timing impacts Escalate risks, issues, or slippage impacting the broader care model critical path Provide accurate, timely status updates and inputs to integrated workstream reporti

C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Large Language Models (LLMs) continue to push the boundaries of what AI systems can do — but inference is still the bottleneck. The Model Efficiency team is responsible for pushing the limits of LLM inference efficiency across our foundation models. We explore and ship breakthroughs across the model execution stack, including: model architecture and MoE routing optimization decoding and inference-time algorithm improvements software/hardware co-design for GPU acceleration performance optimization without compromising model quality Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. As a Staff Research Engineer, you will develop, prototype, and deploy techniques that materially improve how fast and efficiently our models run in production. You may be a good fit

GitRestMachine LearningAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $295.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. ML Platform @ Roblox today supports hundreds of ML use cases and billions of inferences per day across Discovery, Safety, Engine, and much more. As a Model Optimization engineer on ML Platform, you will be responsible for digging deep into model internals to optimize performance, for both training and inference. We are looking for accomplished engineers to help us maximize performance of our platform. You Will: Optimize machine learning models for performance on GPU architectures, focusing on both training and inference workflows. Conduct low-level performance profiling analysis to identify bottlenecks in existing machine learning pipelines and propose actionable improvements. Contribute to the development of best practices and tooling for model optimization and deployment. Collaborate with cross-functional teams, including data scientists and software engineers, to integrate and deploy optimized models into production environments. Partner across organizations to build tooling, interfaces, and visualizations that make the ML@Roblox a delight to use. You Have: 6+ years of professional experience and a tool chest of system design experience upon which to draw to build performant system

AWSGitMachine LearningAI
F
📍 Mclean, Virginia, United States
✓ High-confidence listingCompany trend -26.7%
Quick readStrong listing-quality and freshness signals

At Freddie Mac, our mission of Making Home Possible is what motivates us, and it’s at the core of everything we do. Since our charter in 1970, we have made home possible for more than 90 million families across the country. Join an organization where your work contributes to a greater purpose. Position Overview: We need a highly innovative Technical Lead! How confident are you that you can build sophisticated analytic systems? If you believe you could contribute to the development of innovative principles and ideas in a matrixed environment, please keep reading as we are seeking an individual contributor who has experience with Java and Python and can lead and nurture an inspiring environment in our Virginia office. Our Impact: The Investments and Capital Markets (I&CM) division is looking for a capable technology lead for its trading and analytics development team. This could be you! To thrive in this division, you must have a comprehensive understanding of system implementation and design, experience working in capital markets, and be enthusiastic about leading development of new paradigms in software system architecture. Your Impact: As a Trading Analytics Development Tech Lead, you will develop and maintain software using Java and Python tech stack that adheres to software engineering best practices. You will influence technical decisions, mentor developers, resolve engineering blockers, and partner with engineering managers to help teams deliver secure, reliable, and maintainable solutions. You will provide hands-on directions for full-stack applications, APIs, microservices, and integration services while reinforcing engineering discipline across design, development, testing, deployment, observability, and production readiness. Partner closely with Product Owners, engineering managers, architecture, business stakeholders, and cross-functional tec

PythonJavaReactAngular
B
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -80.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE The Model Performance organization at Baseten is looking to hire our first Technical Program Manager. This is a zero-to-one role in a team that is responsible for building the core algorithms and methods that power Baseten’s high performance inference stack. You won't inherit an existing program framework, you'll build one from the ground up: the planning structure, execution processes, metrics and the cross-functional alignment that a fast-growing organization needs. Your contributions will directly impact how fast our performance R&D gets productized. If you can drive turning a set of ambitious but loosely defined initiatives into a predictable, well-governed program, this role is for you. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Model Performance team: How to build a day-0 API for Kimi K3 How we built the new fastest API for GLM-5.2 Inference engineering for DeepSeek V4 Pro 0813 RESPONSIBILITIES Own execution across Model Performance's active project portfolio, freeing the team's technical leads to focus on technical direction rather than tracking. Design and stand up the planning structures, operating cadences, and status reporting mechanisms that best fits the team’s DNA. Coordinate model release and optimization programs end to end, including day-zero launches, sequencing the work across performance engineering, infra, and release stakeholders. Drive cross-team al

Machine LearningAIGoExcel
P
📍 Pittsburgh, United States
✓ Quality checked

$55K – $157.3K/yr

Position Overview At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. We work together each day to foster an inclusive workplace culture where all of our employees feel respected, valued and have an opportunity to contribute to the company’s success. As a Quantitative Analytics and Modeling Analyst Senior within PNC's Model Risk Management organization, you will be based in Pittsburgh, PA, Boston, MA or Tysons Corner, VA. We are seeking an experienced model validator to be part of our Model Risk Management team at PNC. The position reports to a validation manager in Commercial Credit and Financial Valuation Models and is part of the Independent Risk Management organization. This role involves performing rigorous independent reviews on some of PNC’s most important models including Commercial & Industrial, Commercial Real Estate and retail commercial loss forecasting models, risk rating models, as well as financial valuation and investment models. This role also participates at and provide individual and aggregate model risk assessment in various model working groups and forums. PNC is an in-office company that fosters a supportive culture

B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Are you passionate about advancing the application of artificial intelligence? We are looking for a Software Engineer focused on ML performance to join our dynamic team. This role is ideal for someone who thrives in a fast-paced startup environment and is eager to make significant contributions to the exciting field of LLM Inference. If you are a backend engineer who thrives on making things faster and is excited about open-source ML models, we look forward to your application. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and other libraries to debug ML performance issues. Apply and scale optimization techniques across a wide range of ML models, particularly large language models. Collaborate with a diverse team to design and implement innovative solutions. Own projects from idea to production. REQUIREMENTS Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or related field. Experience with one

PythonDockerKubernetesRest
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten’s Model Performance (MP) team is responsible for ensuring the models running on our platform are fast, reliable, and cost‑efficient. As part of this team, you’ll focus on Model APIs — the infrastructure powering our hosted API endpoints for the latest open‑source models. This work spans distributed systems, model serving, and developer experience. You’ll join a small, high‑impact team operating at the intersection of product, model performance, and infra, helping to define how developers interact with AI models at scale. RESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups Productionize performance improvements across runtimes with deep understanding of their internals: speculative decoding implementations, guided generation for structured outputs, custom scheduling and routing algorithms for high-performance serving Build comprehensive benchmarking frameworks that measure real-world performance across different model architectures, batch sizes, sequence lengths, and hardware configurations Productionize performa

KubernetesMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -80.4%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE OPPORTUNITY We are looking for Senior Software Engineers to join our team. This is a specialized, high-impact role sitting at the intersection of high-performance computing (HPC) and Large Language Model (LLM) engineering. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work. RESPONSIBILITIES Benchmarking : Evaluate, run and automate standard LLM quality benchmarks (GSM8K, MMLU) alongside custom performance suites for specific workloads (e.g., long-context window, KV cache reuse, disaggregated serving). DevEx Improvement : Develop and maintain internal GPU-enabled development environments (similar to GitHub Codespaces). You will ensure the team has seamless, high-performance "dev machines" optimized for model experimentation. Tool Development : Build and contribute to open-source tools such as InferenceMAX and genai-bench to automate model evaluation, benchmarking and analysis. System Profiling : Use profilers like PyTorch Profiler, NVIDIA Nsight Systems and py-spy to collect performance profiles, identify bottlenecks, and debug the compute/networking stack. Monitoring & Observability : Develop real-time dashboards and alerts to monitor system health, model startup times, and runtime performance. Continuous Integration : Auto

PythonCI/CDGitRest
🔔

Get new model behavior engineer jobs in United States by email

Daily job updates · Unsubscribe anytime