Jobs in United States

Human Evaluator in United States

1,781 active opportunities · Updated October 2026

Explore current human evaluator jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 Austin, TX, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

POSITION SUMMARY: The Manager, Clinical Lab is a key leader in Natera’s lab operations who oversees and directs the development, planning, implementation and maintenance of manufacturing methods, processes and operations for new or existing products and technologies. The Manager works closely with groups to ensure strict compliance with good manufacturing practices and CLIA/CAP guidelines at all times. Ensures the effective use of materials, equipment and employees in producing quality products. Monitor and control staffing / labor, capital expenditures, and manufacturing performance. Develops budget and monitors expenditures. Formulates and recommends manufacturing policies, procedures and programs. Selects, develops and evaluates line-management to ensure the efficient operation of the function. PRIMARY RESPONSIBILITIES: Under the direction of Senior lab leadership, and in close collaboration with other Laboratory leadership, this individual provides operational and strategic leadership to laboratory products assigned to the individual. Generates and routinely provides executive-level production/departmental summaries. Monitors workflow and ensures production goals and deadlines are met (e.g. TAT, SLA) Engages in future planning and scales employees and equipment appropriately Gathers and monitors metrics on production deadlines, efficiency, scrap waste, and error rates Monitors the quality control/assurance programs, test results, and equipment (if applicable) Uses metrics in decision-making and to drive changes across the team (e.g. cost reduction, quality improvement) Manages sufficient headcount levels to meet production goals by proactively driving the interview process and managing employee PTO and overtime when necessary. Refines current lab process through continuous improvement projects. Oversees team-specific lab operations. Go-to person when Senior lab leadership is not present; go-to person for knowledge of pr

ExcelHuman Resources
T
Golf. A golf role or an employer dedicated to golf.
📍 Tg, Dallas Office, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

The Accounting/Finance IT Project Manager leads delivery of financial systems initiatives across Accounting, Finance, Treasury, AP, Tax, Internal Controls, and FP&A. This role converts finance and accounting needs into requirements, project plans, and measurable outcomes, while managing scope, timeline, budget, and cross-functional coordination from initiation through go-live. Prior Accounting/Finance experience is required to evaluate process impacts, manage risk, and drive adoption across enterprise financial platforms. This role is based at Topgolf HQ in Dallas, TX, with an expectation of four days per week in the office and one day remote. Responsibilities Governance: Establish operating cadence, roadmap visibility, prioritization, decision logs, risks, and executive-ready updates for Finance/Accounting technology work. Requirements & Business Cases: Partner with functional leaders to document pain points, future-state needs, savings opportunities, and success measures before work begins. Delivery: Coordinate delivery across SAP, OneStream, Anaplan, Coupa, Concur, Pathlock, Workday, and Crunchtime, tax, treasury, billing, and reporting platforms, owning scope, dependencies, testing, cutover, and release readiness end-to-end. Stakeholder Facilitation: Lead working sessions across IT, Accounting, Finance, Operations, vendors, and business owners to resolve roadblocks and align on scope and timelines. Controls & Adoption: Ensure projects account for accounting controls, SOX/access impacts, data integrity, training, UAT, and post-go-live adoption. Reporting: Track capacity, vendor involvement, dependencies, benefits, and risk, reprioritizing as business needs shift. Artifacts: Create and maintain project charters, plans, schedules, and RAID logs across 2-3 concurrent projects. PMO Alignment:</b

SapAccountingFinanceHuman Resources
T
Golf. A golf role or an employer dedicated to golf.
📍 Tg, Dallas Office, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

We're looking for a Manager of Restaurant Back Office Systems to own and lead our CrunchTime platform strategy, administration, and integration ecosystem. This role is responsible for ensuring CrunchTime operates reliably as the backbone of restaurant-level inventory, labor, and operations management, while managing a small team and partnering closely with IT, Finance, and Operations stakeholders. This is a hands-on leadership role: you'll set direction for the platform, develop and cross-train two direct reports so the team can cover for one another, and remain deeply technical enough to troubleshoot integration issues, guide configuration decisions, and evaluate system changes yourself. ***Musts be in the office 4 days a week*** What You'll Do Own the CrunchTime platform end-to-end, including configuration, maintenance, upgrades, and long-term roadmap Lead, coach, and develop a team of 2 direct reports, setting clear goals, providing regular feedback, and supporting their career growth Cross-train direct reports across CrunchTime modules and integration points so the team has full coverage and no single point of failure during absences, escalations, or turnover Build individual development plans for direct reports, identifying growth opportunities and stretch assignments across platform, integration, and vendor management work Manage and troubleshoot integrations between CrunchTime and: Point of Sale (POS) systems SAP (financial/ERP data flows) Distribution partners such as Sysco (EDI/ordering, invoicing, and inventory feeds) Serve as the primary escalation point for data discrepancies, integration failures, and system outages affecting inventory, ordering, or labor data Partner with Finance, Supply Chain, and Operations teams to ensure back office data supports accurate COGS, inventory valuation, and labor cost reporting

SQLSapProject ManagementFinance
T
Golf. A golf role or an employer dedicated to golf.
📍 Tg, Dallas Office, United States
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

WHAT YOU WILL DO Key Responsibilities Lead cash and liquidity forecasting by integrating operating cash flow, capital expenditures, financing activity, seasonality, and site-level dynamics Improve finance and treasury processes, controls, and governance appropriate for a large, distributed operating footprint Support management of the company's revolving credit facility and debt portfolio, including availability, covenant compliance, and lender reporting Lead execution of the end-to-end internal capital expenditure governance process, including business case development, approval workflows, and investment standards Partner with functional and operating leaders to ensure capital requests are clearly scoped, return-based, and aligned to strategic priorities Track capital deployment and performance versus approved plans and lead post-investment reviews to assess realized returns Provide decision-oriented analysis and presentation materials with clear, cash-focused recommendations for executive leadership, board, lenders, and private-equity sponsors Lead execution and project management of ad-hoc finance and treasury transformation initiatives to improve EBITDA, working capital, liquidity, operational efficiency, and forecast accuracy Partner with the Treasurer on capital structure and balance-sheet strategy, including leverage, liquidity buffers, and financing alternatives Evaluate capital tradeoffs and support refinancing, recapitalization, and transaction-related initiatives as needed Help shape and scale treasury, capex governance, and strategic finance capabilities as the company grows Perform other duties as assigned CORE COMPETENCIES FOR SUCCESS Strategic & Analytical Thinking Thinking strategically to identify trends, evaluate risks and opportunities, and

ExcelProject ManagementFinanceHuman Resources
C
📍 Hartford, United States
✓ Quality checkedCompany trend +340.2%

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary The Senior Manager, Corporate Development — M&A Integration plays a critical role in advancing the organization’s mergers, acquisitions, divestitures, and strategic growth initiatives. This individual will support multiple transactions simultaneously and partner closely with business leaders and cross-functional teams to plan and execute due diligence, integrations, and separations. As a key member of the M&A Integration team, the Senior Manager will develop and manage value chain analyses, entanglement assessments, stranded-cost analyses, one-time cost evaluations, service grids, and transition services agreements for acquisitions and divestitures. The ideal candidate brings strong M&A execution experience, strategic thinking, financial acumen, stakeholder-management expertise, and the ability to deliver results in a complex, matrixed environment. Key Responsibilities Transaction Execution and Integration Support multiple acquisitions, divestitures, and strategic transactions concurrently. Partner with Corporate Development leadership to evaluate, plan, and execute complex transactions. Lead critical integration and separation planning activities to support smooth transitions and achieve transaction objectives. Develop and maintain transaction workplans, timelines, milestones, risk logs,

Project ManagementPmpFinanceProcurement
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83.9%

About the Team The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust. About the Role As a Research Program Manager, you will lead initiatives that test the safety and robustness of OpenAI’s models through creative experimentation and structured evaluation. You’ll coordinate efforts across research and engineering teams to transform ambiguous risks into concrete research programs and influence future model development and deployment. We’re looking for people who are technically savvy, comfortable with ambiguity, and excited about shaping the future of safe AI. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Lead programs that explore unexpected model behaviors and identify failure modes. Translate vague or emergent risk signals into clear priorities and actionable research plans. Design and run creative evaluations, experiments, and red-teaming campaigns. Collaborate with research, product, and deployment teams to integrate findings into model training and deployment cycles. Develop repeatable systems for tracking model performance and understanding emerging behavior patterns. You might thrive in this role if you: Have strong experience in technical program management, with excellent organizational and communication skills. Are familiar with large language models, prompt engineering, or model evaluation techniques. Are comfortable managing fast-paced, high-uncertainty projects and shaping them from the ground up. Are creative and resourceful in devising new methods for testing model behavior and performance. Can effectively coordinate across technical and non-technical stakeholders to drive alignment and execution. About OpenAI OpenAI is an AI resear

AWSRestAIRust
R
📍 New York, NY, United States· Full-time
✓ High-confidence listingCompany trend -99.2%

From $10K/yr

Quick readStrong listing-quality and freshness signals

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role As the Product Manager for Tax, you will own the vision, strategy, and roadmap for Ramp’s next generation of AI-powered tax products. You will build solutions to automate tax calculation, validation, and filing for our existing suite of products (card, bill pay, banking, etc), for both domestic and international use cases. You’ll also be accountable for **expanding** Stack - our agentic accounting product - into the tax automation domain, to automate the work of CPAs. You’ll partner closely with engineering, design, data, and cross-functional stakeholders (Finance, Legal, Security, CX, Sales) to build intelligent experiences that help customers save money, reduce risk, and eliminate manual work. This role blends AI-native product development (LLM-powered workflows, evaluation systems, human-in-the-loop design) and deep domain expertise in domestic/international tax and CPA’s workflows. You will define and execute a strategy that will unlock new markets for Ramp and help us save our customers even more time and money. Please note that

GitRestAIGo
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $243.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Software Engineer on the Foundation AI organization, you will sit at the epicenter of our foundation model efforts. While the research world is focused on architecture, you will be the architect of the data flywheel that makes VideoGen and 3DGen possible. You aren't just building pipelines; you are building the infrastructure that defines how our models perceive and generate virtual worlds in three dimensions and across time. In this role, you will partner directly with our AI researchers to advance beyond experimental datasets and into the realm of dynamic, high-fidelity data synthesis and evaluation. You will bridge the gap between research prototypes working locally to scaling for millions of users. You will design, implement, and scale robust, high-performance infrastructure to crawl, create, curate, store, and serve the massive datasets required for these models. We are seeking accomplished software engineers with a passion for data, experience building large distributed systems, and a commitment to writing high-quality, well-tested code to solve complex data challenges at scale. Your contributions will ensure that our foundation models receive the highest quality dat

PythonAWSGitAI
H
📍 Kentucky, United States· Remote
✓ High-confidence listingCompany trend +310%
Quick readStrong listing-quality and freshness signals

Become a part of our caring community The Senior Process Improvement Professional analyzes and measures the effectiveness of existing business processes and develops sustainable, repeatable and quantifiable business process improvements. The Senior Process Improvement Professional work assignments involve moderately complex to complex issues where the analysis of situations or data requires an in-depth evaluation of variable factors. The Senior Process Improvement Professional researches best business practices within and outside the organization to establish benchmark data. Collects and analyzes process data to initiate, develop and recommend business practices and procedures that focus on enhanced safety, increased productivity and reduced cost. Determines how new information technologies can support re-engineering business May specialize in one or more of the following areas: benchmarking, business process analysis and re-engineering, change management and measurement, and/or process-driven systems requirements. Begins to influence department’s strategy. Makes decisions on moderately complex to complex issues regarding technical approach for project components, and work is performed without direction. Exercises considerable latitude in determining objectives and approaches to assignments. Use your skills to make an impact Required Qualifications The Sr. Process Improvement Professional must meet one, (1) of the following requirements: Licensed Professional (LSW) in the state of Kentucky OR ability to obtain, Master’s degree from an accredited university or college in social work, psychology or related health and human services area with a minimum of one, (1) year professional experience in case management.

ExcelPower BiRecruitment
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse, misalignment, and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. Within Safety Systems, the Model Policy team works to ensure that frontier models behave safely and reliably in real-world environments by designing policies that define safe model behavior. Our relevant publications include: Safety at every step OpenAI GPT6 System Card OpenAI Model Spec GPT-Live ChatGPT Images 2.5 About the Role We are hiring a Model Policy Manager to focus on the safety of multimodal models. In this role, you will shape how OpenAI identifies, evaluates, and addresses risks in multimodal AI models - such as GPT-Live and ChatGPT Images - as well as multimodal capabilities in frontier AI models. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design and maintain model policies for audio, image, video, and omni-modal behavior. Translate theories of harm and threat models into behavioral safety policies, evaluation criteria, grading guidance, and safeguards. Identify and analyze safety regressions and failure patterns to identify gaps in existing policies and inform policy iteration. Develop policy artifacts that support model training, evaluation, and deployment, including behavior i

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse, misalignment, and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. Within Safety Systems, the Model Policy team works to ensure that frontier models behave safely and reliably in real-world environments by designing policies that define safe model behavior. Some of our publications include: Safety at every step OpenAI GPT6 System Card OpenAI Model Spec About the Role We’re hiring a Model Policy Manager to shape model behavior for U.S. government use, with a focus on national security applications. You’ll define nuanced policies and translate them into training and evaluation criteria, helping models navigate high-stakes scenarios while preserving their usefulness and capabilities. In this role, you will: Develop model policies that guide safe and useful behavior. Build evaluations, identify policy gaps and model failures, and use findings to improve policies and training. Work with research, engineering, and domain experts to support safe, reliable deployment. You might thrive in this role if you: Bring relevant experience in AI safety, policy, or risk assessment. Have strong judgment and can turn complex safety questions into clear, practical policies. Have the technical fluency to work hands-on with model data and evaluations. Are motivated by OpenAI’s mission and the responsible use of

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Team Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model training to evaluation and red-teaming—all aimed at mitigating misuse, misalignment, and maintaining our high bar for safety. We lead OpenAI's commitment to developing and deploying safe Artificial General Intelligence (AGI), fostering a culture of trust, responsibility, and transparency. Our goal is to continuously learn from deployments, distribute AI’s benefits widely, and ensure that powerful tools remain aligned with human values and safety considerations. About the Role We are hiring a Product Manager to focus on risk related to multimodal models. In this role, you will drive initiatives which ensure that OpenAI’s audio, image, and video deployments are safe, impactful, and aligned with user needs and technical innovation. You will clarify strategic priorities, develop safety-focused product roadmaps, and collaborate closely with AI researchers, software engineers, policy experts, and cross-functional partners. This role suits a proactive, technically skilled product manager adept at adversarial thinking and excited to tackle challenging, ambiguous problems through structured analysis and collaborative decision-making. This position is based in San Francisco, CA, with relocation assistance available. In this role, you will: Partner closely with AI research, engineering, data science, policy teams, and other stakeholders to embed safety throughout the development and deployment of multimodal AI models - such as GPT-Live and ChatGPT Images - as well as multimodal capabilities in frontier AI models. Develop comprehensive frameworks for understanding and mitigating deployment safety risks, drawing on data analysis, expert consultation, and adversarial assessments. Define strate

AWSRestAIGo
H
📍 New York, NY, United States
✓ Quality checkedCompany trend +310%

Become a part of our caring community You have shipped AI products before. You understand the difference between a demo and a production system. You have strong opinions about evaluation frameworks because you have experienced the consequences of operating without them. You are at your best when you own architecture decisions while continuing to build and deliver critical code yourself. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to source documents, and route complex cases to human experts. The output of these systems supports healthcare decisions that impact real members. As a Lead AI Applied Engineer, you will provide technical leadership for AI-enabled products and platforms, define architectural direction, establish engineering standards, and personally design and build the most critical components of our systems. You will lead through both technical expertise and execution, helping the team deliver reliable, scalable, and auditable AI solutions in a highly regulated healthcare environment. Why Join Us Lead the architecture of production AI systems where LLMs are foundational to the product experience. Make key technical decisions regarding model selection, system boundaries, platform architecture, and build-versus-buy strategies. Own the highest-risk and highest-impact technical challenges involving reliability, explainability, and correctness. Influence engineering culture and establish standards that shape how the team builds and ships AI products. Work on systems operating at meaningful scale, processing millions of documents and supporting healthcare decisions across a large member population. Partner

JavaScriptTypeScriptPythonReact
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83.9%

About the Team The Health team, within OpenAI’s broader Personal AGI organization, has a mission to ensure AGI improves health for all humanity. Improving human health will be one of the defining impacts of AGI. Hundreds of millions of people already turn to ChatGPT for questions about their health and millions of clinicians use it weekly to support care delivery. Increasingly capable models create an opportunity to make high-quality medical intelligence more accessible across patients and clinicians—raising the floor of human health—and accelerate the new capabilities and scientific advances that raise the ceiling of human health. Our job is to make those benefits real. We work across the full model stack—pretraining, midtraining, reinforcement learning, post-training, evaluations, harnessing, and deployment—and connect that research to the patients, clinicians, and real-world outcomes we aim to improve. About the Role We’re looking for an exceptional, hands-on researcher who wants to build frontier health capabilities and turn them into impact at scale. This is a role for someone who can take an important, underdefined problem from 0→1: identify the right bet, build what’s needed to test it, and drive it all the way to a measurable improvement in the models and products we actually ship. We’re especially excited about two kinds of people: researchers with the technical depth to move the frontier in pretraining, reinforcement learning (RL) / post-training, or evals; and researchers with real depth in developing frontier biomedical AI capabilities. Prior experience in healthcare is helpful but not required. Research excellence, velocity, ownership, and alignment with the mission are most important to us. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own a high-leverage research direction end to end—from deciding which problem matters and h

Machine LearningArtificial IntelligenceAI
S
📍 Bellevue, WA, United States· Full-time· Remote
✓ High-confidence listingCompany trend -91.7%
Quick readStrong listing-quality and freshness signals

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. We are looking for a VP of Product Management to lead AI product strategy at Smartsheet. You will lead a product team responsible for defining our comprehensive AI vision and roadmap, ensuring we build a cohesive, coordinated approach to how AI transforms our product. Your team will establish the strategic direction for specialized AI agents, evaluation infrastructure, and AI-powered capabilities, from platform foundations through to customer-facing features. You will report to our Chief Product and Technology Officer located in our Bellevue, WA office, or you may work remotely from anywhere in the US where Smartsheet is a registered employer. You Will: Lead a high-performing product team focused on defining and executing Smartsheet's comprehensive Applied AI product strategy. Define the strategic roadmap for AI product initiatives — from conversational experiences, MCP, automations, specialized agents, and evaluation infrastructure through to AI-powered features in customer-facing products — ensuring a cohesive, coordinated vision. Own the business outcomes of the AI portfolio — define the success metrics (adoption, retention, expansion, and revenue impact), set the targets, and hold the team accountable to them. Instrument, measure, and iterate: build the data and evaluation foundation that tells us whether AI features are actually working, and use it to make hard prioritization calls — including what to sunset. Move fast in a fast-moving space — ship, learn, and adjust in short cycles rather than waiting for per

VueAWSAIGo
🔔

Get new human evaluator jobs in United States by email

Daily job updates · Unsubscribe anytime