AI Evaluation Scientist, Math PhD โ€” Frontier Benchmark

AI Evaluation Scientist, Math PhD โ€” Frontier Benchmark

Full-Time No working from home possible
M

Mercor is hiring PhD and Masterโ€™s scientists to author AI evaluation tasks (Sci Code). You will author original, executable research problems that todayโ€™s frontier models cannot solve.

Engage with leading AI labs to build benchmarks for scientific computing and craft robust grading criteria. You will work on multiple subdomains with a coding focus, using Python or R, and Docker-based workflows. Start date is immediate for a 6-week, part-time engagement.

#J-18808-Ljbffr

AI Evaluation Scientist, Math PhD โ€” Frontier Benchmark employer: Mercor

Mercor is an exceptional employer that champions innovation and creativity in the AI sector, offering a fully remote work environment that promotes flexibility and independence. With a strong focus on employee growth, team members are encouraged to take ownership of their projects while benefiting from a supportive culture that values collaboration and continuous learning. Joining Mercor means being part of a forward-thinking company backed by industry leaders, where your contributions directly impact cutting-edge search technologies.

M

Contact Details:

Mercor Recruitment Team

ยฉ StudySmarter