
ML Challenge Task Auditor | $70–90/hr | Remote (US)
Join a frontier AI lab's quality assurance effort by evaluating the rigor, correctness, and methodology of applied machine learning tasks used to train and benchmark cutting-edge AI models. This is a high-impact contractor role for experienced ML practitioners who can think critically about experiment design and evaluation standards.
Please note: This role focuses exclusively on applied and experimental ML rigor — it is not an LLM application development or MLOps position.