Mercor
$60–90/hr
139 people hired
You will serve as a ground-truth expert for advanced AI evaluation benchmarks involving real-world data science and quantitative analysis. | The role involves designing complex analytical tasks covering data cleaning, statistical analysis, method comparison, interpretation, and reporting. | You will create reproducible reference analyses and evaluate where frontier AI models succeed or fail. | This is a full-time W-2 role, fully remote within the United States.
Data Science Benchmark Expert