Materials Science Domain Expert
You will review the quality of materials knowledge work tasks, write the instruction specs and golden solutions that define what "correct" looks like, and build the benchmarks…
Pay
$70–$110/hr
Location
Bay Area, CA
Browse roles/Mercor
Browse roles
Mercor can offer professional AI evaluation work across software, legal, medical, finance, research, language, writing, and other fields. Compare each current role by requirements and pay before applying.
This page shows only Mercor roles; use the main opportunities page to compare platforms.
Browse all 910 current opportunitiesCurrent listings tracked
Evidence for comparison
Sorted by recently reviewed
This platform currently includes 35 language-specific roles. Search by language or use the language guide before applying.
You will review the quality of materials knowledge work tasks, write the instruction specs and golden solutions that define what "correct" looks like, and build the benchmarks…
Pay
$70–$110/hr
Location
Bay Area, CA
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
About the work CritPt is a public benchmark of research-level physics challenges, built to test whether frontier AI models can carry out genuine physics research reasoning rather…
Pay
$80–$110/hr
Location
Remote
Contribute high-quality voice recordings for training and evaluating cutting-edge speech models.
Pay
$50–$100/hr
Location
Remote
This is an evaluation framework for frontier AI models in drug research and development.
Pay
$60–$90/hr
Location
Remote
You will write and verify rigorous multiple-choice questions across core biology domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance…
Pay
$60–$75/hr
Location
Remote
You will write and verify rigorous multiple-choice questions across core physics domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance…
Pay
$61–$77/hr
Location
Remote
You will write and verify rigorous multiple-choice questions across core mathematics domains, evaluate solution quality, and help establish gold-standard benchmarks used to…
Pay
$61–$77/hr
Location
Remote
You'll apply your scientific expertise to evaluate and strengthen how these models handle specialized technical topics in Portuguese.
Pay
$50–$54/hr
Location
Remote — Western Europe preferred
you will do Forecast the timing and content of a named analyst's post-catalyst research note Reason from filings, earnings, transcripts, estimates, valuation and ratings to an…
Pay
$150–$250/hr
Location
Remote
You will be writing prompts that sit exactly on that line, then judging whether the model held it.
Pay
$65-$75 per task
Needs pay review
Location
Remote
You will be writing prompts that sit exactly on that line, then judging whether the model held it.
Pay
$65-$75 per task
Needs pay review
Location
Remote
No outside research or real work product is required.
Pay
$100/hr
Location
Remote
This role centers on evaluating how effectively a large language model (LLM) performs in English-language tasks and on curating high-quality English linguistic data to train and…
Pay
$50/hr
Location
Remote
We are looking for individuals who have professionally designed, fielded, managed, or evaluated surveys and understand what separates a rigorous, decision-useful survey from one…
Pay
$120/hr
Location
Remote
You'll apply your scientific expertise to evaluate and strengthen how these models handle specialized technical topics in Ukrainian.
Pay
$48–$52/hr
Location
Remote — Eastern Europe preferred
You'll apply your scientific expertise to evaluate and strengthen how these models handle specialized technical topics in Thai.
Pay
$24–$28/hr
Location
Remote — Southeast Asia preferred
Specialist AI Work is independent from Mercor. Some Apply links may be referral links, and we may be paid if the platform credits the referral. This does not change the pay shown in a role or how roles are ordered here.
Applicant workflow
Live roles remain the purpose of this page. This sourced summary helps you understand what can happen before and after a Mercor application.
Create or update a Mercor profile, open a current role, and submit the role-specific application requested on the official site.
Mercor documents role applications, profile information, and project-specific screening. Requirements can differ by opportunity.
Some applications use an AI interview or a role assessment. Follow the current instructions and use only assistance the platform explicitly permits.
An offer, platform acceptance, or completed interview does not by itself mean a project has started. Project onboarding can include additional documents or checks.
Mercor documents project onboarding, agreement signing, background checks for some work, and payment setup as separate steps.
A suitable project may not be available after profile or interview completion. Public documentation does not promise continuous matching, hours, or project duration.
Payment details are a separate intent from advertised listing rates. Compare documented providers, approval rules, and unknowns in the contractor payment guide.
Compare the current tracker snapshot, role mix, application steps, and platform cautions.
Compare current role coverage, pay evidence, fields, and screening context.
Compare current reviewed opportunities across the supported platforms.
Prepare for resumes, assessments, onboarding, identity checks, and unstable project access.
See current task-family counts, common requirements, and representative roles.
See which platforms have current jobs here and which have review information only.