← All jobs
Mercor Verified
remote · hourly

ML Challenge Task Auditor

$70–$90/hr
Share
Data Analysis hourly remote United States
Posted Aug 28, 2026

Evaluate the quality, correctness, and methodological rigor of applied machine-learning tasks used to train and evaluate a frontier AI lab's models. You'll assess experiment design, model-selection reasoning, and evaluation methodology — and provide clear, rubric-based written feedback.

Basic Qualifications • 3+ years hands-on applied/experimental ML (experiment design, model selection, hyperparameter tuning, evaluation methodology) • Strong grasp of data-quality rigor: leakage detection, metric gaming, and train/test/CV hygiene • Proficiency with standard ML frameworks (PyTorch, TensorFlow, scikit-learn, XGBoost) • Ability to critique ML claims against evidence and reproduce results

Preferred Qualifications • Competition / benchmark experience (e.g., Kaggle) • Graduate research or publication record in applied ML • Prior task-grading or peer-review experience

Note: this role evaluates applied/experimental ML rigor — it is not an LLM-application-building or MLOps role.

Pay range
$70–$90/hr
40h/wk
Share
Earn $1,440 for each successful referral on this role.
Similar roles

You might also like

micro1 Verified New
remote · hourly
Big Data Engineer
Data Analysis
Posted Sep 15, 2026
$30–$80/hr
Mercor Verified New
remote · part-time
Data Analysis Expert
Data Analysis
Posted Sep 15, 2026
$70–$120/hr 40h/wk
Mercor Verified New
remote · part-time
Machine Learning Expert
Data Analysis
Posted Sep 15, 2026
$70–$120/hr 40h/wk
Turing Verified New
remote
LLM Annotator - Master's Degree
Data Analysis
Posted Sep 14, 2026
Pay on listing