Mercor is hiring PhD and Masterโs scientists to author AI evaluation tasks (Sci Code). You will author original, executable research problems that todayโs frontier models cannot solve.
Engage with leading AI labs to build benchmarks for scientific computing and craft robust grading criteria. You will work on multiple subdomains with a coding focus, using Python or R, and Docker-based workflows. Start date is immediate for a 6-week, part-time engagement.
#J-18808-Ljbffr
AI Evaluation Scientist, Math PhD โ Frontier Benchmark employer: Mercor
Mercor is an exceptional employer that champions innovation and creativity in the AI sector, offering a fully remote work environment that promotes flexibility and independence. With a strong focus on employee growth, team members are encouraged to take ownership of their projects while benefiting from a supportive culture that values collaboration and continuous learning. Joining Mercor means being part of a forward-thinking company backed by industry leaders, where your contributions directly impact cutting-edge search technologies.