Open roleExternal
AI Evaluation Scientist, Math PhD — Frontier Benchmark
- london, england, United Kingdom
- Permanent·On-site
- Part time
- £69 - £124 Per Hour
Job Description
Mercor is hiring PhD and Master’s scientists to author AI evaluation tasks (Sci Code). You will author original, executable research problems that today’s frontier models cannot solve.
Engage with leading AI labs to build benchmarks for scientific computing and craft robust grading criteria. You will work on multiple subdomains with a coding focus, using Python or R, and Docker-based workflows. Start date is immediate for a 6-week, part-time engagement.
#J-18808-Ljbffr

