STEM PhD evaluators
Mathematics, physics, chemistry, biology and microbiology. PhD holders who write and grade graduate-level reasoning tasks.
Talent bench
For AI labs and data vendors. We recruit, interview and verify domain experts and native-speaker linguists, then send you a shortlist with the evidence behind each person.
Weekly-paid contractors · Proctored interviews · Verified credentials · 16 roles hiring now
Role families
The role families AI teams ask us for most. Each is vetted against the credentials that matter for that field.
Mathematics, physics, chemistry, biology and microbiology. PhD holders who write and grade graduate-level reasoning tasks.
CFA and CPA holders, investment banking, private equity, credit, FP&A, accounting and insurance practitioners.
Bar-admitted attorneys, including US-licensed, for legal reasoning, drafting and rubric design.
Licensed physicians and board-certified specialists, including radiology and dermatology, for clinical reasoning evaluation.
Senior engineers in Python and C++ for agentic, long-horizon coding tasks and RL environment authoring.
CBRN background, cyber threat intelligence, OSINT, jailbreak and prompt-injection testing, child-safety research.
Gulf and Najdi, Levantine, Egyptian, Maghrebi and Iraqi. Native speakers who write their dialect accurately.
QA, transcription and text validation in South and Southeast Asian, African and European languages, by native speakers.
12 roles hiring now →Energy and grid operations, mechanical engineering and data engineering practitioners.
Leads for agentic and RAG evaluation: rubric design, reviewer calibration and inter-annotator agreement.
How we vet
A structured, role-specific interview in the language set for the role, scored against a rubric.
Camera, screen and integrity checks throughout the interview, with a proctoring score on record.
Government ID and liveness check through Didit, with duplicate-account detection.
Licences, bar admission, board certification and degrees checked with the issuing body where a public register exists.
Interview, domain test and trial results combine into an Expert Score and a tier: Vetted, Expert or Outlier.
Borderline candidates get a peer interview with a senior practitioner in the same field before any decision.
Coverage
Remote and global by default. Region-locked roles, such as US bar admission or North American generalists, are sourced only in that region.
Engagement
Experts are contractors paid in USD every week. You get one invoice.
Tell us the role, domain and languages. We send a shortlist with an evidence pack for each expert.
Replacement if an expert is not a fit, and trial work measured against your quality bar.
Talent bench
Share the field, languages, volume and timeline. We post and vet the role at no cost to you.
Working on it…
This usually takes a few seconds.
Popular: Email code · Application statuses · Payments