AI Evaluator & Trainer jobs
Evaluator and trainer roles score model outputs against rubrics, compare responses and write the feedback that shapes how AI behaves.
52 open roles · Micro1, Mercor. Listed pay runs from $10 to $350 per hour. Current openings include Cantonese Language Evaluator, Subject Matter Expert – Financial & Operations Chart Analysis, Research Engineer - Code Generation & Model Evaluation, AI Domain Expert.
Each listing below links to a full description with requirements, pay and how to apply. New AI Evaluator & Trainer roles are added as they open, so check back regularly.
Open roles
Cantonese Language Evaluator
Cantonese Language Evaluator — remote, paid AI training/evaluation work (language-audio) ($30-$40/hr).
Subject Matter Expert – Financial & Operations Chart Analysis
Subject Matter Expert – Financial & Operations Chart Analysis — remote, paid AI training/evaluation work (business-operations) ($25-$50/hr).
Research Engineer - Code Generation & Model Evaluation
Research Engineer - Code Generation & Model Evaluation — remote, paid AI training/evaluation work (software-engineering) ($50-$100/hr).
AI Domain Expert
AI Domain Expert — remote, paid AI training/evaluation work (software-engineering) ($140-$200/hr).
Senior AI Trainer
Senior AI Trainer — remote, paid AI training/evaluation work (ai-machine-learning) ($50-$90/hr).
Subject Matter Expert – Medicine & Biostatistics Chart Analysis
Subject Matter Expert – Medicine & Biostatistics Chart Analysis — remote, paid AI training/evaluation work (sciences-research) ($25-$50/hr).
Subject Matter Expert – Chart & Data Visualization Analysis
Subject Matter Expert – Chart & Data Visualization Analysis — remote, paid AI training/evaluation work (data-analysis) ($25-$50/hr).
Evaluation Specialist/Recent Grad
Evaluation Specialist/Recent Grad — remote, paid AI training/evaluation work (generalist) ($20-$60/hr).
AI Consulting Domain Expert
AI Consulting Domain Expert — remote, paid AI training/evaluation work (business-operations) ($100-$200/hr).
Naval Subject Matter Expert (Cleared)
Naval Subject Matter Expert (Cleared) — remote, paid AI training/evaluation work (other) ($50-$100/hr).
AI Software Engineering Domain Expert
AI Software Engineering Domain Expert — remote, paid AI training/evaluation work (business-operations) ($100-$200/hr).
Medical Evaluation Specialist (students, residents, physicians)
Medical Evaluation Specialist (students, residents, physicians) — remote, paid AI training/evaluation work (medicine) ($40-$90/hr).
Subject Matter Expert – Data & Statistical Chart Analysis
Subject Matter Expert – Data & Statistical Chart Analysis — remote, paid AI training/evaluation work (sciences-research) ($25-$50/hr).
AI Data Science Domain Expert
AI Data Science Domain Expert — remote, paid AI training/evaluation work (business-operations) ($100-$200/hr).
AI Finance Domain Expert
AI Finance Domain Expert — remote, paid AI training/evaluation work (business-operations) ($100-$200/hr).
AI Training Data Contributor (College Students)
AI Training Data Contributor (College Students) — remote, paid AI training/evaluation work (robotics) ($10-$15/hr).
Research Evaluation Specialist (PhD / Researcher / Professor)
Research Evaluation Specialist (PhD / Researcher / Professor) — remote, paid AI training/evaluation work (sciences-research) ($40-$90/hr).
Senior AI Trainer
Senior AI Trainer — remote, paid AI training/evaluation work in the ai machine learning domain ($14-$36/hr).
Military Operations & Reporting SME
Military Operations & Reporting SME — remote, paid AI training/evaluation work in the other domain ($40-$80/hr).
Data Analyst (AI Evaluation)
Data Analyst (AI Evaluation) — remote, paid AI training/evaluation work in the data analysis domain ($30-$60/hr).
Senior DoD Staff Writing SME
Senior DoD Staff Writing SME — remote, paid AI training/evaluation work in the other domain ($40-$80/hr).
Gmail & Google Calendar AI Assistant Evaluator
Gmail & Google Calendar AI Assistant Evaluator — remote, paid AI training/evaluation work in the generalist domain ($15-$30/hr).
Website Designer (UK-Based) — AI Training
Contribute world-class website design expertise to a frontier AI training project — a remote, own-schedule independent contractor role for UK-based designers, up to 40 hours/week.
User/Customer Research and Feedback Synthesis Evaluator
Put your user and customer research expertise to work evaluating AI-generated work products for accuracy, rigor, and quality — a flexible, remote, hourly contract with Mercor.
UX/UI Product Designer (UK-Based) — AI Training
Bring your end-to-end digital product design experience to a frontier AI research project — a remote, flexible independent contractor role for UK-based UX/UI designers.
Spreadsheet QA / Workbook Maintenance Evaluator
Put your spreadsheet QA and workbook maintenance expertise to work evaluating AI-generated work products for accuracy, rigor, and quality — a flexible, remote, hourly contract with Mercor.
Real Estate Appraisal Expert — Visual Document Understanding (AI Training)
Shape how frontier AI systems interpret appraisal reports and property documents — a $59-60/hour, part-time contract role for licensed real estate appraisers in the US or Canada.
Public Health Communications Evaluator
Put your public health communications expertise to work evaluating AI-generated work products for accuracy, rigor, and quality — a flexible, remote, hourly contract with Mercor.
Psychiatry Expert - AI Training
Shape how the next generation of AI reasons about mental health care — design clinical scenarios and grade model outputs in this remote, hourly role with Mercor, paying $150-$350/hour.
Program Management - Implementation Planning Evaluator
Leverage your program management and implementation planning expertise to evaluate AI-generated content — a remote, hourly contractor role with Mercor, paid weekly.
Product Management - Roadmap - PRD Evaluator
Apply your product management, roadmap, and PRD expertise to evaluate AI-generated content — a remote, hourly contractor role with Mercor, paid weekly.
Product Launch - Experiment Readiness Evaluator
Use your product launch and experiment readiness expertise to evaluate AI-generated content — a remote, hourly contractor role with Mercor, paid weekly.
Process Improvement - SOPs Evaluator
Bring your process improvement and SOP expertise to evaluating AI-generated content — a remote, hourly contractor role with Mercor, paid weekly.
Personal Finance - Consumer Planning Evaluator
Apply your personal finance and consumer planning expertise to evaluate AI-generated content — a remote, hourly contractor role with Mercor, paid weekly.
People Ops - Recruiting Evaluator
Put your people ops and recruiting expertise to work evaluating AI-generated content for accuracy and quality — a remote, hourly contractor role with Mercor, paid weekly.
Operations, Inventory & Capacity Planning Evaluator
Apply your operations, inventory, and capacity planning expertise to grade AI-generated work products. Remote, hourly, paid weekly.
Marketing Expert — AI Grading & Evaluation
Use your marketing and advertising expertise to define grading criteria and score AI-generated creative work. Remote, independent contractor, paid weekly.
Legal Expert (Employment Law) — AI Training
Help train frontier AI models on real-world employment law reasoning. Fully remote, $100-$150/hour, flexible 6-15 hours a week.
Investment Analysis, Valuation & Credit Evaluator
Apply your investment analysis, valuation, and credit expertise to grading AI-generated work products in this flexible, remote, hourly contractor role.
Healthcare Operations Evaluator
Bring your healthcare operations expertise to grading AI-generated work products in this flexible, remote, hourly contractor role.
General Finance & Accounting Evaluator
Use your finance and accounting expertise to grade AI-generated work products in this flexible, remote, hourly contractor role.
General Business Strategy & Management Evaluator
Bring your business strategy and management expertise to grading AI-generated work products in this flexible, remote, hourly contractor role.
Finance Operations & Audit Support Evaluator
Put your finance operations and audit expertise to work grading AI-generated documents, spreadsheets, and decks in this flexible, remote, hourly contractor role.
Design Expert — AI Grading & Evaluation
Evaluate how well AI systems perform real-world design work: define grading criteria and score work samples for $80-150/hr, fully remote.
FP&A - Corporate Finance Evaluator
Apply your FP&A and corporate finance expertise to evaluate AI-generated work products for accuracy and rigor, on a flexible, remote, hourly contract with Mercor.
Education - School Evaluator
Apply your education and school administration expertise to evaluate AI-generated work products for accuracy and rigor, on a flexible, remote, hourly contract with Mercor.
Document-Deck Production QA Evaluator
Apply your document and deck production QA expertise to evaluate AI-generated work products for accuracy and rigor, on a flexible, remote, hourly contract with Mercor.
Data Quality - CRM Operations Evaluator
Apply your data quality and CRM operations expertise to evaluate AI-generated work products for accuracy and rigor, on a flexible, remote, hourly contract with Mercor.
Data Analysis - Quantitative Readouts Evaluator
Apply your data analysis and quantitative readouts expertise to evaluate AI-generated work products for accuracy and rigor, on a flexible, remote, hourly contract with Mercor.
Cybersecurity - IT GRC Evaluator
Apply your Cybersecurity / IT GRC expertise to evaluate AI-generated work products for accuracy and rigor, on a flexible, remote, hourly contract with Mercor.
BI Dashboards & Performance Reporting Evaluator
Apply your BI dashboards and performance reporting expertise to evaluate AI-generated work products, in a remote hourly engagement.
Accounting Expert — AI Grading & Evaluation
Define what excellent accounting work looks like by designing grading criteria and scoring AI-generated work samples, as a remote independent contractor.
