Short answer
What is model evaluation?
Model evaluation is the process of checking how well an AI system follows instructions, reasons, avoids unsafe claims, and handles domain-specific tasks. Human reviewers often judge examples against rubrics.
Context
Evaluation can be general or specialist. A clinician, lawyer, engineer, finance analyst, researcher, editor, or language specialist may review different kinds of model behavior.
Current research-evaluation examples
These current research roles may include model-evaluation work. Read each listing to confirm the tasks, required credentials, and screening process.
Mercor · research
English Language and Literature Expert
$50 per hour
Micro1 · research
Physics Expert (Professor / Principal Investigator)
$80 - $160 / hour
Mercor · research
Applied Physics Benchmark Specialist
$61-$77 per hour