Short answer
What is model evaluation?
Model evaluation is the process of checking how well an AI system follows instructions, reasons, avoids unsafe claims, and handles domain-specific tasks. Human reviewers often judge examples against rubrics.
Context
Evaluation can be general or specialist. A clinician, lawyer, engineer, finance analyst, researcher, editor, or language specialist may review different kinds of model behavior.
Current research-evaluation examples
These current research roles may include model-evaluation work. Read each listing to confirm the tasks, required credentials, and screening process.
Mercor · research
Research Physics Expert
$80 - $135 / hour
Micro1 · research
Physics Expert (Professor / Principal Investigator)
$80 - $160 / hour
Mercor · research
Physical Scientist Talent Network
$60 - $80 / hour