← All jobs

ML Challenge Task Auditor (Train AI Models Part Time!)

up to $400k/year Remote Full-time
Research ScientistArtificial Intelligence EngineerMLOps EngineerPrompt EngineerMachine Learning EngineerData ScientistAI Researcher

Archer is matching candidates to this role at Mercor. Create a free profile and Archer will check you against this role and every other live role, showing you exactly where you match.

Evaluate the quality, correctness, and methodological rigor of applied machine-learning tasks used to train and evaluate a frontier AI lab's models. You'll assess experiment design, model-selection reasoning, and evaluation methodology — and provide clear, rubric-based written feedback. Basic Qualifications • 3+ years hands-on applied/experimental ML (experiment design, model selection, hyperparameter tuning, evaluation methodology) • Strong grasp of data-quality rigor: leakage detection, metric gaming, and train/test/CV hygiene • Proficiency with standard ML frameworks (PyTorch, TensorFlow, scikit-learn, XGBoost) • Ability to critique ML claims against evidence and reproduce results Preferred Qualifications • Competition / benchmark experience (e.g., Kaggle) • Graduate research or publication record in applied ML • Prior task-grading or peer-review experience Note: this role evaluates applied/experimental ML rigor — it is not an LLM-application-building or MLOps role.

Apply knowing you're qualified

One free profile is all it takes. Archer checks you against this role and every other live role we list, and shows you exactly which requirements you meet before you apply.

More roles like this

See all matching roles

Not quite the right role?

Archer scans thousands of live roles and surfaces the ones you genuinely match, each with a clear explanation of why. It keeps working after you apply, so you hear about roles you would never have found by searching.

Create your free profile