Applications close Friday, October 9, 2026 at 12:00 p.m. Pacific. We are selecting the first three applicants who qualify for the initial pilot.
$2,200per task
Evaluation is the core of AI training work: reading what a model wrote, scoring it against a rubric and saying exactly why. These roles are named for that job, as evaluators, raters, graders or reviewers, and most ask for a field or a language you already know well.
Counts are recalculated from the roles open right now, whenever roles open or close. Newest role added .
Page 1 of 2
Applications close Friday, October 9, 2026 at 12:00 p.m. Pacific. We are selecting the first three applicants who qualify for the initial pilot.
$2,200per task
Help a leading AI lab ensure its models respond to people with care, balance and sound judgment.
$45–$70/hr
The hiring partner is engaging Hospitalist Reviewers to contribute to a high-impact project with a customer focused on inpatient quality and documentation.
$100–$150/hr
The hiring partner is engaging Generalist — U.S. Tax Workflow Evaluations to participate in a short-term customer project focused on evaluating U.S. tax and financial workflows. In this…
$30–$110/hr
We are looking for experienced U.S.-based Pharmacists and Pharmacy Technicians to support an AI-focused healthcare project.
$70/hr
The hiring partner is seeking Clinical Mental Health Experts to support an AI research initiative focused on evaluating the realism and quality of simulated clinical conversations.
$100/hr
The hiring partner is engaging Data Analysts to contribute their knowledge and expertise to a customer’s data quality project.
$30–$60/hr
The hiring partner is engaging Gmail & Google Calendar AI Assistant Evaluators to collaborate on a customer-driven project enhancing AI assistant quality through real-world task evaluation.
$15–$30/hr
The hiring partner is engaging Research Engineers to participate in a project focused on code generation and model evaluation for a customer's initiative.
$50–$100/hr
The hiring partner is engaging Cantonese Language Evaluators to contribute to a language-focused project supporting a valued customer.
$30–$40/hr
The hiring partner is engaging Senior Hospitalist Clinical Reviewers to contribute to a high-impact project with a customer focused on inpatient quality and documentation.
$70–$100/hr
The hiring partner is one of the world’s leading AGI infrastructure companies, working with frontier AI labs to accelerate model development through high-quality training data, evaluations, and engineering talent.
$100–$150/hr
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
Based in San Francisco, California, the hiring partner is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems.
Rate on request
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems.
Rate on request
Of the 35 roles on this page open now, 20 publish an hourly rate. The median is $73 an hour, half of them pay between $54 and $103, and the highest pays $180. Every role shows its terms before you apply, and applying is free.
3 of 35 are open worldwide. 11 name the countries they accept, most often United States. 21 do not say, so check the role's own terms before applying. With a profile, Tier1 checks each role's countries and languages against yours.
No. These roles hire AI evaluators and raters for what they already know. Where a role needs a particular degree, licence, tool or language, its page says so.