← all roles

Mercornew this week

ML Challenge Task Auditor

Pay
$70–$90/hr
Type
Hourly
Spots left
3
Location
Remote
Posted
August 28, 2026

Good fit for: data annotators, labelers, and detail oriented generalists.

Apply via referral →

Listing verified on the platform’s official board. Last checked September 25, 2026.

Evaluate the quality, correctness, and methodological rigor of applied machine-learning tasks used to train and evaluate a frontier AI lab's models. You'll assess experiment design, model-selection reasoning, and evaluation methodology — and provide clear, rubric-based written feedback.

Basic Qualifications

• 3+ years hands-on applied/experimental ML (experiment design, model selection, hyperparameter tuning, evaluation methodology)

• Strong grasp of data-quality rigor: leakage detection, metric gaming, and train/test/CV hygiene

• Proficiency with standard ML frameworks (PyTorch, TensorFlow, scikit-learn, XGBoost)

• Ability to critique ML claims against evidence and reproduce results

Preferred Qualifications

• Competition / benchmark experience (e.g., Kaggle)

• Graduate research or publication record in applied ML

• Prior task-grading or peer-review experience

Note: this role evaluates applied/experimental ML rigor — it is not an LLM-application-building or MLOps role.

Applying via Mercor: Mercor pays twice a week once you are hired and billing work. Sign-up is free; expect a resume upload and an AI interview before matching.

Apply via referral →

Similar live roles

Mercor
Voice Actor: CX Agent Voice Cloning (South African English)
$50–$100/hr
Mercor
African & Afrikaans-Accented English Experts
$15–$20/hr
Mercor
Cybersecurity Practitioner: Paid Expert Interviews (SOC, Incident Response, Detection, AppSec)
$125–$175/hr
Mercor
Cybersecurity Practitioner: Paid Expert Interviews (SOC, Incident Response, Detection, AppSec)
$125–$175/hr