Jobiba hiring network

Human Evaluator Jobs

3,920 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current human evaluator jobs. Use filters to narrow by work mode, employment type, experience and date posted.

About the Team The Personalization-Memory team, within OpenAI's broader Personal AGI organization, is focused on developing agents that can learn from prior interactions in order to become more helpful and efficient over time. We build general-purpose memory and personalization capabilities that transfer across ChatGPT and other agentic products, and we collaborate with applied engineering on the product surfaces that allow users to interact with memory. About the Role As a Research Engineer / Research Scientist on the Personalization-Memory team, you will research and develop improvements to memory usage and personalization in OpenAI's frontier models. Our team works on reinforcement learning, dataset creation, evaluations, and other post-training methods. We partner closely with research and product teams across the company to realize the vision of a truly personalized ChatGPT. We're looking for individuals who have a background in frontier model post-training, are able to iterate quickly, and who are passionate about product-driven research. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own and pursue a research agenda for improving memory use and personalization in frontier models. Build robust evaluations for tracking modeling improvements. Design, implement, test, and debug code across our research stack. Collaborate closely with the research and product teams to influence the shape of technical solutions in the product. You might thrive in this role if you: Are passionate about personalization and building personalized assistants. Have experience working with user signals and human data to turn feedback into reliable signals for training and evaluation. Have a deep understanding of frontier model post-training and machine learning applications. Value principled approaches and research craftsmanship. Are comfortable diving into a lar

awsrestmachine learning
View job →
N
Notion
📍 New York• Full-time• $196K – $230K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: We’re seeking an experienced User Researcher to shape growth-related product decisions by delivering insights that fuel AI experiences. This role blends traditional growth research, such as adoption, monetization, and pricing and packaging, with fast-evolving AI features like Notion Agent and Chat that permeate experiences like onboarding and Workspace creation. It requires a unique mix of user research expertise, strategic foresight, and technical fluency in AI and AI product development. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: Design and execute studies that evaluate how people experience Notion with AI in the loop, including how users discover, trust, and get value from AI, and identify barriers to adoption and willingness to pay. Run mixed-methods research (qual + quant), including interviews, concept testing, prototype/usability testing, diary

restaigo
View job →
N
Notion
📍 San Francisco• Full-time• From $135K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Notion is looking for an Early Career Recruiter who wants to reimagine what early career hiring can look like. This is a role for a builder: someone energized by the question of how to find and cultivate exceptional emerging talent, who gets excited about crafting programs and experiences that don't yet exist, and who sees employer brand, candidate experience, and assessment design in an AI era as creative problems worth obsessing over. You'll own Notion's early career strategy end-to-end alongside a team of talented teammates — from how we reach communities to how we evaluate potential, develop talent, and convert interns and new grads into the next generation of Notinos. You'll work closely with our Engineering, Product, Design, Sales, and People teams. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: A differentiated early career engine — design and run

restaigo
View job →
N
Notion
📍 New York• Full-time• $98K – $140K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role You'll own the quality bar for Notion AI products. You’ll work with product and engineering teams to build systems to define what “good” looks like, measure our progress, and drive changes to deliver reliable and high-quality AI experiences. Your work directly shapes how Notion's AI products behave for millions of users. This isn't a traditional software engineering role. It’s an art & science role . You won't spend your days writing code. Instead, you'll focus on understanding and shaping how our AI products behave through context engineering, designing evaluation systems, and analyzing data. This team sits in our AI engineering team, working directly with engineering, product, design, and data. This role is a unique blend of ops, strategy, and product thinking. Day to day, you'll live in production data, ship prompt fixes, run evals and, in effect, shape our quality strategy. As part of that you'll shape Notion's model strategy and work directly with frontier AI labs (OpenAI, Anthropic, Google) to evaluate and launch new models. We're looking for problem-seeking generalists interested in 0 → 1 : curious people wi

sqlrestai
View job →
N
Notion
📍 Hyderabad• Full-time
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role The Hyderabad Infra team builds and maintains Notion's internal async task runner and configuration management platform. The async task runner plays a critical role in ensuring our millions of users have a fast, reliable, and secure experience. The configuration management platform enables safe, explicit configuration management for Notion's product and infrastructure engineers. As part of the Hyderabad Infra Team, you’ll have a unique opportunity to shape how Notion manages and scales its async task runner and configuration management platform, enabling innovation across the company. What You'll Achieve You will contribute to the evolution and maintenance of our async task runner to meet the needs of over 100 million global users and support the rapid growth of our product and business. With guidance from senior team members, you'll help ensure our systems remain reliable, efficient, and scalable. You'll evaluate and integrate new technologies to keep us ahead of emerging challenges. Your work will empower our engineering team to build features confidently while you grow your skills in distributed systems and infrastr

restaigo
View job →
N
Notion
📍 San Francisco• Full-time• From $152K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About The Role We’re looking for an AI Applications Engineer to help drive Notion’s business transformation efforts. In this role, you’ll be a strategic partner to our internal stakeholders (primarily GTM, Finance and People teams) and deliver and scale creative AI-driven solutions to multi-faceted problems with measurable business impact. You’ll also build reusable components, evaluation patterns, and operational guardrails that make AI delivery repeatable across teams. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You’ll Achieve Work with stakeholders to discover opportunities from ambiguous problem statements, translate them into scoped solutions, and drive iterative releases from idea to adoption Build and ship end‑to‑end AI solutions—from problem framing through data readiness, modeling, evaluation, and production rollout Establish evaluation and production-readiness patterns (met

restaigo
View job →
N
Notion
📍 London• Full-time
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: The Notion EMEA Recruiting team acts as trusted business partners who craft a thorough, inclusive, and curated interviewing experience. We believe hiring exceptional talent, at the right time, is the best path forward to achieve our company’s ambitious goals. We’re looking for a Go-to-Market Recruiter to support our growth across key Go-to-Market (GTM) teams in EMEA, with an emphasis on Sales. In this role, you'll help shape our recruiting strategy, manage full-cycle hiring efforts, and ensure a delightful yet rigorous hiring process. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: Serve as a trusted recruiting partner to leaders across the GTM org, owning the full recruiting lifecycle from sourcing through successful close Design and implement creative sourcing strategies to attract diverse top talent across EMEA markets Ensure scalable and structured hiring plans with consistent evaluation processes while creating tho

restaigo
View job →
N
Notion
📍 San Francisco• Full-time• $140K – $195K/yr
1mo ago

Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Notion is looking for a Recruiting Program Manager to help design, build, and scale the systems and programs that power how we hire. This is a role for someone who moves fast, takes ownership, and doesn't wait to be told what to do — someone who operates with both strategic altitude and operational precision, and knows how to bring clarity and momentum to complex cross-functional work. We're at an inflection point in how we think about recruiting at Notion, and this role sits at the center of it. You'll partner closely with Recruiters, Recruiting Operations, People Operations, and senior leadership to drive some of our highest-priority initiatives — building the programs and processes that define how Notion finds, evaluates, and brings in exceptional talent. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: The scope of this role will evolve with the team's

restaigo
View job →
J
11 days ago

Who We Are At Justworks, you’ll enjoy a welcoming and casual environment, great benefits, wellness program offerings, company retreats, and the ability to interact with and learn from leaders in the startup community. We work hard and care about our most prized asset - our people. We’re helping businesses get off the ground by enabling them to focus on running their business. We solve HR issues. We’re data-driven and never stop iterating. If you’d like to work in a supportive, entrepreneurial environment, are interested in building something meaningful and having fun while doing it, we’d love to hear from you. We're united by shared goals and shared motivations at Justworks. These are best summed up in our company values, which are reflected in our product and in our team. Our Values If this sounds like you, you’ll fit right in. We are looking for a Manager, Digital Sales to build and lead this new commercial motion from our Phoenix office. This is a builder role. You will lead a team of Digital Sales Advisors responsible for converting high-intent small business customers while also helping design and continuously improve the system around them—including customer signals, routing, sales plays, workflows, automation, experimentation, and feedback into our Self-Service experience. You will be accountable for revenue and conversion, but this role is broader than managing a sales number. You will constantly evaluate why customers need human help, which interventions improve outcomes, what our Advisors should do differently, and which interactions we can eventually eliminate through better product experiences, signals, or automation. The right leader combines strong commercial judgment with systems thinking. You are equally comfortable coaching a seller through a customer conversation, analyzing funnel performance, designing an experiment, and partnering with Product or GTM Engineering to improve how customers buy Justworks. Your Success Profile What You Will Work

aiHRpayroll
View job →
N
12 days ago

We're building the platform that lets long-running autonomous agents operate safely inside NVIDIA's enterprise. These are not assistants on a developer's laptop. They are fleets of agents deployed in the cloud, running continuously at scale on shared accelerated compute. They take on real work across enterprise systems, so people get far more done than they could before. This role defines the constructs that agents are built from: the blueprints they start from, the tools, skills, and plugins that power them against enterprise data, the runtime safety harness that keeps them in bounds, and the connections into credential management, sandbox, memory, and observability. The team designs and ships these building blocks so that agent developers across the company can stand up a new agent, wire it in, and run it for days or weeks. Security and safe execution come out of the box, not something each team has to get right on its own. Today an agent runs inside a single harness. Claude, Codex, and open-source agent harnesses each work differently underneath, with their own execution model, tool interface, and telemetry shape. The platform smooths over those differences, so a single skill, safety policy, or trace works the same no matter which harness is running. We want to enable agents that act on a person's behalf, governed and secure, continuously evaluated and self-improving. These agents coordinate and hand work off to each other, with identity and policy following every hop. They route and tune themselves across harnesses from live eval signals, and get better from their own production telemetry instead of waiting on a human to retrain them. Have you run agents on a harness like Claude or Codex and hit the walls that show up when they run for real, for days, against live systems — and wanted them to learn from it on their own? We're building the platform that solves those problems once, for every team. What you'll b

pythonaifinance
View job →
A
13 days ago

This is Adyen Adyen provides payments, data, and financial products in a single solution for customers like Meta, Uber, H&M, and Microsoft - making us the financial technology platform of choice. At Adyen, everything we do is engineered for ambition. For our teams, we create an environment with opportunities for our people to succeed, backed by the culture and support to ensure they are enabled to truly own their careers. We are motivated individuals who tackle unique technical challenges at scale and solve them as a team. Together, we deliver innovative and ethical solutions that help businesses achieve their ambitions faster. The Opportunity Adyen is building a top-tier AI engineering organization in Amsterdam, San Francisco and Madrid to drive our next chapter of innovation using AI globally across the entire company. This is a highly technical, hands-on role focused on the exploration and application of cutting-edge AI research within the financial technology sector. As a Senior AI Research Engineer, you will operate with a high degree of autonomy and responsibility , delivering strategic, high-impact outcomes that bridge the gap between advanced AI research and production-grade applications at a global scale, potentially impacting trillions of dollars in transactions annually. What You'll Do: Innovate and Deploy: Drive the execution of Adyen's AI strategy , focusing on the practical application of Generative AI (GenAI) and other AI methodologies in finance . This includes contributing to Adyen's efforts in key research areas such as AI agents for data analysis and operational workflows , human-in-the-loop for integrity risk , and development of foundation models . For instance, you might contribute to initiatives like the Data Agent Benchmark for Multi-step Reasoning (DABStep) , which evaluates AI agents on real-world data analysis tasks, including those from the financial sector. Build Production-grade Applications: Bridge the gap between cutti

machine learningaiExcel
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”

pythonsqlai
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”

pythonsqlai
View job →
I
14 days ago

Job Details: Job Description: Intel's Design Quality and Reliability organization is seeking an AI Platform Engineer to architect and build an enterprise-grade AI platform for mission-critical engineering work. This platform will enable Intel engineers to analyze complex design, qualification, and reliability data; automate engineering workflows; access organizational knowledge; and make faster, evidence-based decisions throughout the product lifecycle. The successful candidate will combine strong software engineering fundamentals with expertise in AI-native and agentic development. They will be highly proficient with Agentic AI coding assistants and able to use these tools responsibly to accelerate architecture, implementation, testing, debugging, and documentation. This role requires close collaboration with Design, Quality and Reliability, Product Engineering, Manufacturing, IT, Information Security, and other Intel stakeholders. Responsibilities 1. Architect and develop Intel's reusable AI platform for Design Quality and Reliability. 2. Build AI agents and workflows for engineering data analysis, qualification planning, risk assessment, knowledge retrieval, reporting, and process automation. 3. Apply Agentic AI coding assistants to accelerate software development while maintaining rigorous engineering review and validation. 4. Integrate AI capabilities with Intel engineering databases, quality-management systems, internal APIs, spreadsheets, documentation repositories, and workflow tools. 5. Develop production-grade backend services, APIs, data pipelines, model gateways, and agent-orchestration components. 6. Establish shared platform capabilities for identity, access control, tool authorization, memory, observability, evaluation, and auditability. 7. Implement human approval, deterministic validation, and rollback controls for consequential engineering actions. 8.

pythonairecruitment
View job →

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”

pythonsqlai
View job →
🔔

Get new human evaluator jobs by email

Daily job updates · Unsubscribe anytime