Jobiba hiring network

Model Behavior Engineer Jobs

4,916 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current model behavior engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
1mo ago

About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in Sydney. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role you will Own technical delivery across multiple deployments from first prototype to stable production. Build full-stack systems that deliver customer value and sharpen how we learn. Embed closely with customer teams, understand their needs, and guide adoption of what you build. Scope work, sequence delivery, and remove blockers early. Make trade-offs between scope, speed, and quality; adjust plans to protect delivery. Contribute directly in the code when progress or clarity depends on it. Codify working patterns into tools, playbooks, or building blocks that others can use. Share field feedback that helps Research and Product understand where the models succeed and where they can improve. Keep teams moving through clarity and follow-through. You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work. Have scoped and delivered complex systems in fast-moving or ambiguous environments. Write and review production-grade code across frontend and backend using Python, JavaScr

javascriptpythonjava
View job →
O
1mo ago

About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in Singapore. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. 50% travel is expected. In this role you will Own technical delivery across multiple deployments from first prototype to stable production. Build full-stack systems that deliver customer value and sharpen how we learn. Embed closely with customer teams, understand their needs, and guide adoption of what you build. Scope work, sequence delivery, and remove blockers early. Make trade-offs between scope, speed, and quality; adjust plans to protect delivery. Contribute directly in the code when progress or clarity depends on it. Codify working patterns into tools, playbooks, or building blocks that others can use. Share field feedback that helps Research and Product understand where the models succeed and where they can improve. Keep teams moving through clarity and follow-through. You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work. Have scoped and delivered complex systems in fast-moving or ambiguous environments. Write and review production-grade code across frontend and backend using P

javascriptpythonjava
View job →
O
OpenAI
📍 Singapore• Full-time
1mo ago

About the Team The Consumer Products team at OpenAI brings groundbreaking AI technology to life through world-class hardware. Our Operations group ensures that every product moves seamlessly from concept to customer — integrating supply chain strategy, manufacturing operations, and quality systems to enable rapid, reliable, and scalable production. We partner closely with engineering, design, and manufacturing teams around the world to deliver exceptional products at launch and beyond. About the Role As an Operations Manager, you will own the execution and strategy that transforms early product concepts into high-quality, manufacturable consumer devices. You’ll lead cross-functional planning, manage vendor relationships, and ensure supply, cost, and schedule alignment across fast-moving hardware programs. This role is based in Singapore. We follow a hybrid model (four days per week in the office) and offer relocation support for new employees. Occasional travel to supplier and manufacturing sites is expected. In this role, you will: Develop and execute operations strategies that bridge product design and large-scale manufacturing. Drive build readiness and production planning from prototype through mass production. Partner with global suppliers to ensure materials, tooling, and capacity meet program goals. Collaborate with hardware, quality, and supply chain teams to establish and maintain robust operational systems. Track cost, yield, and schedule performance to ensure efficient and predictable product delivery. Build scalable processes that support rapid iteration and future product launches. You might thrive in this role if you: Have 6+ years of experience in hardware operations, manufacturing, or supply chain for consumer electronics or similar industries. Have successfully taken at least one product from concept through mass production. Are fluent in NPI processes, contract manufacturing dynamics, and global operations management. Bring strong analytical and co

awsrestai
View job →

About the Team The Statsig team within OpenAI owns the experimentation, rollout, dynamic configuration, and analytics infrastructure that sits on the launch path for OpenAI products. Our systems help teams ship safely, evaluate product and model changes in production, and make high-confidence decisions from real-world usage. Statsig began as an independent company built around experimentation, feature management, and product analytics at scale. After Statsig joined OpenAI, the team began the next chapter: bringing that platform expertise and infrastructure into OpenAI as the experimentation and rollout foundation for every product we ship. This is infrastructure with a very direct product consequence. Teams working on ChatGPT, Codex, model measurement, consumer experiences including ads, business subscriptions, developer products, and shared platform systems depend on Statsig to evaluate configurations, move traffic safely, ingest experiment data, serve analytics, and roll changes forward or back when production reality demands it. We are at a critical point in the platform journey. Adoption is accelerating quickly across OpenAI, and the systems that were already important are becoming load-bearing for how the company launches. The infrastructure needs to stay fast under sharply increasing evaluation volume, reliable when more services depend on it, observable enough to debug quickly, and efficient enough to support OpenAI-wide scale. Recent SDK and server-side infrastructure work has already produced measurable wins in latency, reliability, memory usage, and compute efficiency across important services. The next phase is to make those gains systematic: a platform that can absorb rapidly growing product velocity while preserving low latency, data quality, operational safety, and developer trust. Based out of OpenAI's Bellevue office, we are a close-knit team that values in-person collaboration, technical depth, operational ownership, and building infrastructure that

vueawsrest
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Artifacts team is building the AI-native creation layer for documents, spreadsheets, slide decks, dashboards, reports, analyses, and new forms of interactive work products. We are rethinking what creation looks like when models can move from an ambiguous user goal to a polished, editable artifact with strong structure, taste, correctness, and speed. This is a high-agency team working across product, infrastructure, and research. We partner closely with model training teams to shape how frontier models create artifacts, and with ChatGPT product teams to turn those capabilities into experiences that millions of people can use. The work spans full-stack product engineering, model integration, rendering and editing systems, collaboration, storage, evaluation loops, and production reliability. Our ambition is to build the premier product experience for AI-generated artifacts: starting with familiar work products like slides, sheets, and docs, then expanding into new artifact types that are only possible in an AI-native world. About the Role As Engineering Manager, Artifacts, you will lead and grow the engineering team responsible for building this product and technical foundation. You will manage a team of full-stack and infrastructure-oriented engineers, set technical direction, and stay hands-on enough to shape architecture and debug hard problems. This role sits at the intersection of product engineering, research, and infrastructure. You will partner with researchers on how models are trained and evaluated for artifact creation, with product and design on the user experience. This is a strong fit for a technical manager who wants to build and ship, not only coordinate. The team has a fast trajectory, so you will help define both the product surface and the team that builds it. In this role, you will: Lead, manage, and grow a team building AI-native artifact creation experiences across documents, spreadsheets, slide decks, and emerging artifact form

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI’s GTM Partnerships team builds a strategic, global partner ecosystem designed to accelerate customer success, enable enterprise AI adoption, and drive durable growth in support of OpenAI’s mission. We work closely across Product, Research, Sales, Legal, Communications, and regional leadership to ensure a cohesive strategy and disciplined execution with our most strategic partners. About the Role We are hiring a Partner Director, Global McKinsey Alliance to lead and scale OpenAI’s relationship with McKinsey & Company. This role will serve as the single accountable executive owner of the alliance globally. You will define the partnership strategy and operating model, translating executive alignment into scaled commercial and transformation outcomes across regions and industries. You will work closely with McKinsey’s global leadership, industry and functional practices, and technology organizations—including QuantumBlack, AI by McKinsey—to develop differentiated enterprise AI offerings, pursue complex transformation opportunities, and help customers move from strategy to production deployment and measurable value. The role carries responsibility for the health, performance, and long-term expansion of the partnership. Success requires senior executive presence, sound judgment, commercial discipline, technical fluency, and the ability to operate effectively across a complex, partner-led organization. In this role, you will: Own the global McKinsey relationship, including alliance strategy, pipeline growth, enterprise impact, deployment scale, and long-term expansion across priority industries. Define and run the partnership’s operating model across executive governance, joint planning, regional execution, opportunity management, resource allocation, and escalation. Serve as OpenAI’s senior executive counterpart to McKinsey’s global leadership and relevant industry, functional, technology, and regional leaders. Develop a differentiated alliance s

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team OpenAI's Industrial Compute organization is responsible for planning, delivering, operating, and optimizing the compute infrastructure that powers frontier AI. As OpenAI scales toward becoming an intelligence utility, Industrial Compute coordinates a complex lifecycle spanning infrastructure strategy, capacity planning, provider partnerships, fleet operations, product demand, and financial planning. The organization manages one of the largest and fastest-growing compute footprints in the world, where decisions around capacity allocation, deployment readiness, utilization, reliability, and product demand directly impact product availability, customer experience, and business performance. The Capacity Systems team builds the software platforms, data systems, and automation frameworks that connect these functions into a shared operating model. We transform fragmented planning workflows into scalable systems that enable teams to understand what compute was contracted, delivered, healthy, allocated, and ultimately converted into business and research outcomes. About the Role We are seeking a Capacity Systems Software Engineer to build the platforms and services that power Industrial Compute planning, forecasting, optimization, and operational decision-making. In this role, you will design and develop software systems that connect infrastructure delivery, fleet health, capacity allocation, demand forecasting, deployment readiness, financial planning, and product consumption into a unified system of record. Your work will help OpenAI make better decisions about where compute should be deployed, how capacity should be allocated, and how infrastructure investments translate into business value. You will partner closely with Capacity Planning, Fleet Operations, Infrastructure Engineering, Product, Finance, Supply Chain, and Strategic Sourcing teams to replace spreadsheet-driven workflows with scalable software systems that enable visibility, automation, and dec

typescriptpythonjava
View job →
O
1mo ago

About the Team OpenAI’s Stargate and 3P Engineering teams are responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems. We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute. Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy. As OpenAI’s infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads. About the Role We are seeking a Technical Program Manager, Token-as-a-Service (TaaS) to lead delivery of external compute capacity that directly serves OpenAI model workloads. In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale. You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput. This role sits at the intersection of infrastructure execution, systems readiness, and business impact. Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners. This is a high-visibility role with direct impact on OpenAI’s ability to scale model training and inference globally. This role is based in San Francisco, CA, with a hybrid work model of 3 days in office per week. Relocation assistance is available. Key Responsibilities Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply. Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments. Build

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Foundations Research team works on high-risk, high-reward ideas that could shape the next decade of AI. Our goal is to advance the science and data that enable our training and scaling efforts, with a particular focus on future frontier models. Pushing the boundaries of data, scaling laws, optimization techniques, model architectures, and efficiency improvements to propel our science. The Search team sits within Foundations, building agentic search by co-designing model–system interfaces with the core search stack (serving, indexing, retrieval) to translate model intent into reliable, real-world actions. Operating at the frontier of AI and information retrieval, the team develops large-scale systems that transform and index vast corpora, enabling models to reason over global knowledge and act dependably. In close partnership with researchers, we rapidly bring modeling breakthroughs into production and redefine how intelligent systems discover, retrieve, and synthesize information at planetary scale. About the Role We’re looking for a Software Engineer focused on building and scaling retrieval systems. You’ll work with a team of researchers and engineers to develop infrastructure that enables models to retrieve and act on the right information at the right time. This includes designing and operating indexing systems, retrieval pipelines, and serving layers. This work supports retrieval across OpenAI products and research, with direct impact on system performance, reliability, and scale. Responsibilities Build and scale retrieval infrastructure across indexing, serving, and query execution. Develop low-latency, high-throughput systems for real-time model interaction. Partner with research to productionize embedding and retrieval techniques. Support dense, sparse, and hybrid retrieval pipelines. Own system performance, reliability, and observability at scale. Collaborate across Pretraining, Inference, and Product teams to integrate retrieval end-to-e

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the Team The Integrity team builds the systems OpenAI uses to understand, prevent, and respond to misuse across our products. We partner with Product, Policy, Safety Systems, User Operations, Security, Legal, Privacy, OpenAI for Government, and research teams to turn policy and threat models into product controls, review workflows, measurement systems, and enforcement paths. About the Role We are hiring a Product Manager to own Integrity's product strategy for sensitive deployments: contexts where model capability, customer or deployment setting, privacy constraints, and misuse potential make the operating bar unusually high. This includes government and other high stakes deployments, regulated or high-trust enterprise contexts, zero data retention and privacy-constrained environments, and agentic workflows where harm can unfold across many steps rather than a single prompt. The Sensitive Deployments PM will focus on high-consequence use cases relevant to government deployments and broader deployment-readiness questions for high-risk domains (e.g., healthcare), while building reusable Integrity capabilities for agentic detection and enforcements in these environments. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Own the roadmap for sensitive deployment Integrity controls and readiness criteria. Define how we measure residual risk, decision quality, review quality, and mitigation effectiveness. Build agentic investigation and context-assembly workflows for complex, multi-step misuse patterns. Partner with Policy, Legal, Privacy, Safety Systems, User Ops, Government, and product teams to build shared capabilities. Create launch and deployment playbooks for sensitive customer, capability, and product contexts. You might thrive in this role if you: Have shipped complex product systems in AI, integrity, trust and safety, security, privacy, risk, government, regulated enterprise, or AI safety. Can turn am

awsrestai
View job →
N
Notion
📍 San Francisco• Full-time• From $180K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Millions of people use Notion — and this number is increasing every day. That means millions of people trust us to deliver a fast, reliable, and secure experience, and we value this more than anything. We want to keep earning trust, while also continuing to amaze our users with the tools they can build in Notion. The AI Platform team is responsible for building the shared foundations that let Notion ship AI products quickly and operate them safely at scale. You’ll join a team of talented engineers focused on making speed and quality compatible: reliability and availability through provider changes, quality and correctness systems like evals and release gates, observability that makes failures explainable, and shared primitives for model integrations, context management, long-running actions, and cost/performance tradeoffs. Notion’s AI platform is vital to helping product teams move faster with production-grade guardrails as models, providers, and AI capabilities rapidly evolve. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days)

node.jsrestai
View job →

About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. Role Overview & Key Responsibilities This is a high-leverage leadership role that spans architecture, execution, and org-building, and will shape the direction of our AI / ML initiatives at Ema. We are seeking an AI / ML technical leader who can take a vision and build it. As a Principal ML Engineer at Ema, you will be a senior technical leader responsible for shaping the machine learning roadmap, architecting large-scale ML systems, driving innovation, and ensuring our mixture of expert models (LLM + SLM + Custom Model) is accurate and performant at scale. You will collaborate across teams (research, product, infra, data, etc.), mentor senior engineers, and influence strategy and execution at company-wide levels. Responsibilities Lead the technical direction of GenAI and agentic ML systems that power enterprise-grade AI agents — spanning reasoning, retrieval, tool use, and integrations across various SaaS products. Architect, design, and implement scalable production pipelines for model training, fine-tuning, retrieval (RAG), agent orchestration, and evaluation — ensuring robustness, latency efficiency, and continuous learning. Define and own the multi-year ML roadmap for GenA

pythonjavamachine learning
View job →

About ElevenLabs ElevenLabs is a research and product company defining the frontier of Audio AI. Millions of individuals use ElevenLabs to read articles, voice over their videos, and reclaim voices lost from disability. And the leading developers and enterprises use ElevenLabs to create AI agents for support, sales, and education. ElevenLabs launched in January 2023 with the first AI model to cross the threshold of human-like speech. In January 2025, we raised a $180 million Series C round, valuing ElevenLabs at $3.3 billion. The round was co-led by Andreessen Horowitz and ICONIQ Growth, with continued support from the leading names in tech, including Nat Friedman, Daniel Gross, Instagram co-founder Mike Krieger, Oculus VR co-founder Brendan Iribe, DeepMind and Inflection co-founder Mustafa Suleyman, and many others. ElevenLabs is only 2 years old and scaling rapidly. We are just getting started. If you want to work hard and have an incredible impact, we would love to hear from you. How we work High-velocity: Rapid experimentation, lean autonomous teams, and minimal bureaucracy. Impact not job titles: We don’t have job titles. Instead, it’s about the impact you have. No task is above or beneath you. AI first: We use AI to move faster with higher-quality results. We do this across the whole company—from engineering to growth to operations. Excellence everywhere: Everything we do should match the quality of our AI models. Global team: We prioritize your talent, not your location. We are remote first with optional in-person offices in London, New York, San Francisco, Tokyo, and Warsaw. What we offer Learning & development : Annual discretionary stipend towards professional development. Social travel : Annual discretionary stipend to meet up with colleagues each year, however you choose. Annual company offsite: We bring the entire company together at a new location every year. Coworking : If you’re not located near one of our main hubs, we offer a monthly coworking

pythonrestai
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don't just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. The Role Snowflake's Frontier Engineering (FE) Operations organization is seeking a Director/Sr. Manager, Business Planning & GTM to design the operating and financial architecture of the business. Rather than operating in team-aligned silos, this function aligns to the entire business, specializing in a unique swimlane with a deep cross-functional mindset and flow. In short: this team designs the business — think architecture. Reporting to the Sr. Director, Global FE Operations, you will own the planning frameworks, operating model, and analytics foundation that the rest of the operations organization and the FE business run on. You will translate revenue and delivery ambitions into capa

sqlrestai
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are hiring an Enterprise Account Executive in Perth to partner with our Global Account Manager on one of Australia's most strategically important mining and resources accounts. This is a team-selling model: you will own the Perth stakeholder relationships, run your own pipeline, and close deals end to end. The GAM leads the executive strategy layer; you own the ground game. This is a quota-carrying role, not a support or overlay function. If you are a commercially hungry, technically curious salesperson ready to make the step into enterprise, with the mentorship of a senior account team behind you, this is that role. AS AN ENTERPRISE ACCOUNT EXECUTIVE AT SNOWFLAKE YOU WILL: Own the Perth stakeholder relationships at a major mining and resources account, building trust and pipeline across technology and data teams from Director level down Execute full-cycle deals, from prospecting and discovery through business case, procurement, and close Partner with the Global Account Manager on account strategy, contributing field intelligence, customer feedback, and deal insight to inform the broader account plan Build and manage pipeline with rigour: prospecting, partner engagement, accurate forecasting, and tight CRM hygiene Run structured deal cycles using MEDDPICC, including disc

aigorust
View job →
🔔

Get new model behavior engineer jobs by email

Daily job updates · Unsubscribe anytime