Jobs in United Kingdom

Model Behavior Engineer in London

66 active opportunities · Updated October 2026

Explore current model behavior engineer jobs in London. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 London, Greater London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

About the Team OpenAI’s Applied AI Engineering team helps organizations turn frontier AI capabilities into safe, reliable, and high-impact production systems. We work with customer executives, product and engineering teams, security leaders, and transformation teams to identify valuable opportunities, accelerate technical implementation, and scale what works. Enterprise deployments are defined by complexity rather than any one industry: existing architectures, diverse data environments, security and governance requirements, multiple stakeholder groups, and organization-wide change. We turn lessons from these deployments into better products and reusable patterns for customers everywhere. About the Role As an Applied AI Engineer you will partner directly with leading organizations to design, build, and deploy AI systems that deliver measurable business outcomes. You will combine deep technical judgment, hands-on engineering, and customer leadership to take ambitious ideas from use-case selection and architecture through prototyping, evaluation, production launch, and scale. You will write and debug code, build evaluation systems, resolve complex integrations, and guide decisions involving model behavior, reliability, latency, cost, safety, security, governance, and operational readiness. Success is measured by production systems, sustained adoption, and meaningful customer impact—not simply activity or successful demonstrations. This is a rare opportunity to work on consequential real-world deployments at the frontier of AI while directly influencing how OpenAI’s products evolve. This role is based in London. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees In this role, you will: Partner directly with enterprise customers to identify high-value opportunities and translate them into technical architectures, implementation plans, evaluation strategies, and measurable success criteria. Design, build, and deplo

JavaScriptTypeScriptPythonJava
O
📍 London, Greater London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Threat Intelligence team protects OpenAI’s technology, people, research, and infrastructure by proactively identifying and disrupting adversaries who seek to compromise our systems or misuse our models. We investigate sophisticated threats, build tooling to scale and augment analysis, and deliver intelligence that shapes security strategy and equips leadership with timely, risk-aware insights. We combine technical depth, investigative rigor, and strong cross-functional partnerships to uncover threats and drive impact across OpenAI’s security and research organizations. About the Role As a Technical Threat Investigator at OpenAI, you will help protect the company from sophisticated adversaries targeting OpenAI and the broader ecosystem, as well as those attempting to misuse our models in support of cyber operations. This is a deeply investigative role. You will independently conduct complex, end-to-end investigations into capable threat actors to understand their behavior, infrastructure, emerging techniques, and how AI is integrated into their workflows. You’ll use these insights to proactively identify malicious activity and drive detection, disruption, enforcement, and safety improvements across the company. You’ll translate your investigative findings into durable solutions that scale impact. You’ll build and own lightweight tooling, automate where it matters, and create AI-assisted workflows to make investigations faster, more repeatable, and more effective over time. In this role, you will: Conduct deep, end-to-end investigations into sophisticated threat actors interacting with OpenAI’s models, products, and broader ecosystem. Think like an adversary — model attacker behavior, anticipate misuse patterns, and proactively hunt for, identify, and disrupt malicious activity. Leverage internal telemetry, OSINT, vendor data, a

AWSRestAIGo
O
📍 London, Greater London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

About the Team The Platform Systems team at OpenAI operates at the intersection of cutting-edge AI and large-scale distributed systems. We build the engineering and research infrastructure required to train OpenAI’s flagship models on some of the world’s largest, custom-built supercomputers. Our team develops core model training software and works deep in the stack - spanning collective communication, compute efficiency, parallelism strategies, fault tolerance, failure detection, and observability. The systems we build are foundational to OpenAI’s research velocity, enabling reliable, efficient training at frontier scale. We collaborate closely with researchers across the organization, continuously incorporating learnings from across OpenAI into the evolution of our training platform. About the Role As a Software Engineer, Platform Systems, you will design and build distributed systems that provide visibility into large-scale training workloads and help operate them reliably at scale. You’ll work on failure detection, tracing, and observability systems that identify slow or faulty nodes, surface performance bottlenecks, and help engineers understand and optimize massive distributed training jobs. This infrastructure is critical to operating OpenAI’s training stack and is actively evolving to support new use cases and increasingly complex workloads. This role sits at the core of our training infrastructure, blending systems engineering, performance analysis, and large-scale debugging. In This Role, You Will Design and build distributed failure detection, tracing, and profiling systems for large-scale AI training jobs Develop tooling to identify slow, faulty, or misbehaving nodes and provide actionable visibility into system behavior Improve observability, reliability, and performance across OpenAI’s training platform Debug and resolve issues in complex, high-throughput distributed systems Collaborate with systems, infrastructure, and research teams to evolve platform

AWSRestAIRust
C
📍 London, London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About the Role We're seeking a Senior/Staff Engineer to build and maintain the automation infrastructure that powers the development cycles of our North platform. This engineer will design and implement robust automation systems that enable engineers to efficiently test and validate changes across diverse environments and configurations. This role sits at the intersection of infrastructure and standards. You'll build the systems, frameworks, and culture that allow the rest of engineering to own quality themselves; improving and extending our testing platform by creating the infrastructure that allows engineers to write and execute tests, and enable every engineering team to ship with more confidence. Key Responsibilities Design and implement automation pipelines that support comprehensive testing across multiple environments with varying feature flags and realistic customer data profiles Create intelligent testing agents that simulate real user behavior to validate different configuration combinations Develop and maintain GitHub workflows and actions to automate testing, deployment, and validation processes Manage and optimize H

TypeScriptPythonAWSAzure
DC
📍 London, England, United Kingdom· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Here's a summary of the role Data is only powerful when everyone agrees what it means. This is your chance to build a world class product behavioural data system and process from the ground up. We're rebuilding the internal product our teams use to measure product success and end-user value across all our 24+ products in four business units. You'll design the single source of truth: the metric definitions, the telemetry standards, the data contract engineering instruments against, the change governance processes and tooling, and the models that turn behaviour into portfolio insight. You'll educate and support the teams using this internal behavioural data product to set expectations and and plot a rational roadmap for the maturing of this product, helping our user to use the data honestly and wisely. Usage tells you what happened, not why, or whether it mattered. We want someone who sets quantitative behaviour against qualitative evidence and treats the disagreement as the interesting part. AWS/Snowflake is our spine; capture and BI are open decisions you'll help make. Your models feed the business reviews our executive team, CTO and investors use to run the portfolio, and engineering is committed to instrument against your contract. Here's what your first twelve months should look like First 90 days. Learn the estate, meet the engineers who instrument it, and agree the measurement model for one product end to end — with a defensible adoption number and the qualitative read on what it means. Three to six months. Data contract published, instrumentation spec landed with engineering, first business unit building against it. Definitions, naming and versioned change control governed. Six to twelve months. Rolling out across the remaining business units through the BU-aligned analysts, portfolio reporting running on your numbers. Here's a breakdown of what you'll do (not all

SQLAWSGitRest
FB
📍 London, England, United Kingdom· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Location: London At Freenow by Lyft, we are on a mission to empower smarter mobility decisions, helping people to move freely and cities to thrive. Reporting directly to the Senior VP of Central Operations, we are seeking a visionary Director of Premium Mobility Segments (Europe). This is a high-octane leadership mandate designed to spearhead Freenow by Lyft evolution into the definitive premium mobility leader across the continent. You will own the strategy to rapidly scale our European Premium segment, driving a monumental step-change in market share and setting new industry benchmarks. This role is a cornerstone of Freenow by Lyft global growth engine. You will be the architect of our premium sector expansion, elevating brand equity and capturing structurally superior margins. By harmonising sophisticated European product alignment with aggressive local growth and operational excellence, you will ensure Freenow by Lyft remains the unrivaled choice for premium mobility, delivering high-impact results that resonate globally. YOUR DAILY ADVENTURES WILL INCLUDE: European Premium Strategy: Define and execute long-term growth roadmaps, using London as an anchor for broader European expansion and scaling. Business Development: Drive premium supply growth by building profitable partnerships and serving as the primary face of Lyft Premium in Europe. Demand Generation: Collaborate with marketing teams to build segment demand through effective partnerships and campaigns. P&L & Financial Forecasting: Maintain full P&L ownership, ensuring target margins and accurate growth forecasting. Operational Framework: Create a scalable operating model for the Premium segment while ensuring compliance and localized deployment. Service Standards: Enforce "best-in-class" luxury experience standards, covering vehicle quality, safety, and driver behavior to ensure premium brand differentiation. Cross-Functional Alignment: Lead product and technical requirements for p

AIGoExcelMarketing
O
📍 London, Greater London, United Kingdom· Full-time· Remote
✓ Quality checkedCompany trend -100%

About the Team ChatGPT relies on a large and growing GPU fleet to serve inference workloads reliably and efficiently. We develop the systems and tools that make it possible to introduce new models, manage production deployments, respond to operational issues, and use infrastructure effectively at scale. Our work spans distributed systems, platform engineering, infrastructure automation, and developer experience. We partner closely with research, infrastructure, and product teams to make model deployment more reliable, more efficient, and easier to manage. About the Role We are looking for a software engineer with experience building or operating large-scale production systems. You will design and develop systems that support the model lifecycle in production, including deployment orchestration, configuration management, operational automation, reliability, and capacity management. You will help transform complex operational processes into scalable platform capabilities that enable teams across OpenAI to deploy and manage models with greater confidence and less manual effort. This role is a good fit for engineers who enjoy solving complex operational problems and building software that makes production infrastructure easier to run at scale. In This Role, You Will Build and evolve the platform used to deploy, configure, and manage models across ChatGPT. Develop systems for deployment orchestration, model rollouts, operational visibility, and production readiness. Create abstractions and tooling that simplify complex infrastructure and improve the developer experience. Automate operational workflows, including incident detection, diagnosis, mitigation, and recovery. Improve the reliability, scalability, and efficiency of model deployments and the infrastructure that supports them. Build systems that support capacity planning, resource allocation, and infrastructure utilization. Partner with research, infrastructure, and product engineering teams to identify common chal

PythonAWSRestAI
O
📍 London, Greater London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

About the Team Training Runtime designs the core distributed runtime that powers everything from early research experiments to frontier-scale model runs. We work on building robust, scalable, high performance components to support our distributed training workloads. Our priorities are to maximize the productivity of our researchers and our hardware, with the goal of accelerating progress towards AGI. Within Training Runtime, the Process Management team develops the distributed OS responsible for launching, coordinating, and supervising the large numbers of processes that make up modern training workloads. Our runtime sits beneath training frameworks and on top of research infrastructure, ensuring jobs run reliably across massive clusters while maintaining performance, stability, and observability. Success for us is measured by both system reliability and researcher velocity - enabling ideas to scale from experiments to production training runs. About the Role As a Training Runtime: Process Management Engineer , you will work on the software that ties thousands of computers together and exposes them as a unified system. This system has to serve individual researchers running multiple parallel experiments, as well as our largest training runs spanning 100’s of thousands and even millions of machines and accelerators. This requires easy to use, introspectable systems that can promote a fast debugging and development cycle, as well as relentless optimization for scale while maintaining stability and performance throughout. You will work primarily in Rust , building high-performance asynchronous systems with a strong emphasis on performance, correctness, and scalability. Working at this scale and at the frontier of AI development poses novel challenges. Out-of-the-box approaches often don’t work. The problems you will be working on are highly ambiguous and require strong design judgment as well as proficient execution to advance the state of our infrastructure. We’re loo

PythonAWSLinuxRest
F
📍 London, England, United Kingdom· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before it touches production. That founding instinct still defines how we work. We're accurate and evidence-driven, relentless about clarity, and we'd rather be certain than comfortable, building a groundbreaking platform that transforms how teams run and secure networks across every major cloud and vendor environment. Global leaders like Goldman Sachs, PayPal, S&P Global, IBM, and Dell trust Forward, alongside fast-growing enterprises and government agencies, realizing an average of $14.2 million in annual benefits, according to IDC. Backed by top-tier investors, including A. Capital, Andreessen Horowitz, Goldman Sachs, MSD Partners, Omega Venture Partners, Section 32, and Threshold Ventures, and headquartered in Santa Clara, we're most proud of our team: curious people who'd rather build what doesn't exist than accept how things have always been done. Forward Networks is looking for a Systems (Sales) Engineer Do want to create a category and help build a special company? Do you want to help sell a platform that solves real networking problems? Join a company that has been in market 5+ years and has some of the top Federal agencies and F500/Global 2000 already buying and referenceable. If you have 5-10 years of wildly successful experience as a Sales Engineer selling to large enterprise accounts..you may be the one! We are building a special team and hope you consider us if you want to have the experience of changing the networking world as we know it. Responsibilities: Serve as the primary technical resource for the sales organization working with large enterprise accounts. Work with the sales team as the product advocate and key technical adviser Drive and manage the te

PythonGitGraphqlAI
S
📍 London, England, United Kingdom· Full-time
✓ High-confidence listing

£107K – £262K/yr

Quick readStrong listing-quality and freshness signals

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: We’re seeking an exceptional iOS engineer to join a small, high-impact team dedicated to creating the world’s most valuable AI application, judged by the impact it delivers to users. You’ll shape cutting-edge mobile experiences, blending technical mastery with product brilliance. RESPONSIBILITIES: Dream up and craft delightful, intuitive user experiences that redefine how people interact with AI on mobile. Obsess over every pixel, animation, and interaction, ensuring the experience feels magical and seamless. Build blazing-fast, highly performant systems where every millisecond counts. Drive technical and product decisions, collaborating with AI researchers and engineers to deploy state-of-the-art models globally. BASIC QUALIFICATIONS: Expert in Swift, with deep knowledge of SwiftUI and UIKit. Proficient in performance optimization—memory, CPU, GPU—using low-level tools to deliver blazing-fast systems. Experienced in designing and shipping intuitive, pixel-perfect mobile UIs that redefine how users interact with AI. Skilled in concurrency and reactive programming (Combine) for responsive, real-time apps. Built high-throughput integrations with APIs (REST, gRPC) or AI model outputs, ensuring seamless data

ReactRestAISwift
DC
📍 London, England, United Kingdom· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Position Overview: We are excited to introduce a newly created Senior HR Business Partner role, partnering directly with the Senior Leadership Team (SLT) to drive organizational effectiveness, leadership capability, and strategic workforce planning across the business. Central to the role is holding senior leaders accountable for their own growth as leaders: coaching them with the same rigor they are expected to apply to their teams, so that as leaders level up, their team’s level up alongside them. You will help drive effective people practices by balancing business priorities with employee experience and will play a key role in evolving Diligent’s HR operating model toward more strategic, AI-enabled ways of working. The ideal candidate is collaborative, business-minded, and solutions-oriented, with strong judgment, communication skills, and the ability to build trusted relationships with senior stakeholders – while being direct enough to hold them to account. Key Responsibilities Serve as a senior strategic HR partner to the Senior Leadership Team (SLT) for an assigned business unit or regional pod, leading org design, talent planning, succession planning, and coaching. Develop the people strategy for the business unit or function you support, in partnership with senior leadership, translating business strategy and objectives into a people plan that delivers maximum impact. Lead strategic workforce planning for the business unit — including how work is organised across people and AI agents — and navigating the balance between building AI capability and developing the human skills needed for the future in an AI-first organisation. Partner with the HR Director on implementation and functional roll-out of the broader people strategy. Lead the business unit’s talent and performance cycles, including goal-setting/OKR cadence, SLT-level calibrations, compensation review for the SLT-and-below population, and 9-box grid review. Drive turning workforce, engageme

AWSGitAIGo
C
📍 London, London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role? North is Cohere’s AI workspace platform for enterprises: a secure, customizable environment where companies can use AI across their real workflows while maintaining control over sensitive data. North connects AI agents with workplace tools, applications, and business context, helping users delegate complex work, build automations, inspect outputs, and collaborate with AI in production environments. As North becomes more capable, one of the most important questions is also one of the hardest: how do we know whether the model is actually getting better for the workflows customers care about? This role is about being the voice of North inside modelling. You will build the evaluation systems, feedback loops, and applied modelling workflows that make sure model progress translates into better product outcomes for North users. You will work closely with North product teams, customer-facing teams, and modelling teams to define what “good” means across the product surface, turn real usage and product direction into high-quality evals, and use those evals to guide model selection, patches, and regular model updates. This i

O
📍 London, Greater London, United Kingdom· Full-time
✓ Quality checkedCompany trend -100%

About the Team OpenAI's Training team is responsible for producing the large language models that power our research, our products, and ultimately bring us closer to AGI. Achieving this goal requires combining deep research into improving our current architecture and optimization techniques, alongside long-term bets aimed at improving the efficiency and capability of future generations of models. We are responsible for integrating these techniques and producing model artifacts used by the rest of the company, and ensuring that these models are world-class in every respect. About the Role As a member of the training team, you will push the frontier of LLM development for OpenAI's flagship models, enhancing intelligence, efficiency, and adding new capabilities. Relevant interests may include areas such as architecture design, long-context and efficient attention, optimization and the science of scaling. Ideal candidates have a deep understanding of LLM architectures, a sophisticated understanding of model inference, and a hands-on empirical approach. A good fit for this role will be equally happy coming up with a creative breakthrough, investing in strengthening a baseline, designing an eval, debugging a thorny regression, or tracking down a bottleneck. This role is based in London. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, prototype and scale up new architectures to improve model intelligence Execute and analyze experiments autonomously and collaboratively Study, debug, and optimize both model performance and computational performance Contribute to training and inference infrastructure You might thrive in this role if you: Have experience landing contributions to major LLM training runs Can thoroughly evaluate and improve deep learning architectures in a self-directed fashion Are motivated by safely deploying LLMs in the real world Are well-versed in the state of the a

AWSRestAIGo
W
📍 London, England, United Kingdom· Full-time· Remote
✓ High-confidence listingCompany trend +37.5%
Quick readStrong listing-quality and freshness signals

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Join WRITER's security team as a staff detection and response engineer and help protect the AI infrastructure that's transforming how the world works. You'll build sophisticated detection systems that identify attacks targeting our AI platform, training data, and model deployments while creating automated response capabilities that scale with our explosive growth. This isn't just traditional security work – you're defending cutting-edge AI/AGI systems against adversaries who are evolving their tactics as fast as AI itself advances. This role combines hands-on security engineering with strategic thinking to stay ahead of novel threats that don't exist in textbooks yet. You'll be the operational arm of our security function, translating threat intelligence into real-time detections, coordinating incident response across multiple teams, and hunting for sophisticated attacks across GPU clusters and distributed training environments. If you're excited by the challen

W
📍 London, England, United Kingdom· Full-time· Remote
✓ High-confidence listingCompany trend +37.5%
Quick readStrong listing-quality and freshness signals

🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role We are looking for a visionary and execution-oriented AI transformation lead to serve as a high-impact partner for our most strategic Fortune 100 accounts. In this role, you’ll be the architect of change, moving beyond simple software implementation to help the world's largest organizations fundamentally rethink how they work. You’ll sit at the intersection of business strategy and cutting-edge technology, translating the power of the WRITER platform into measurable P&L impact and helping executives navigate the shift to an AI-first operating model. This is a rare opportunity to build the playbook for enterprise AI transformation. You won’t just be managing accounts; you’ll be driving the next industrial revolution by helping C-suite leaders move from AI experimentation to full-scale value realization. Your work will directly influence WRITER's product roadmap and help define how the world’s biggest brands use superintelligence to expand human capacity. Thi

Other cities to consider

More places hiring for this role

🔔

Get new model behavior engineer jobs in London, United Kingdom by email

Daily job updates · Unsubscribe anytime