Mercor is seeking PhD and Master's scientists to author AI evaluation tasks for Sci Code. You will create original, executable research problems that frontier models cannot solve, contributing to a new benchmark in scientific computing.
Responsibilities include sourcing material, writing prompts, and building grading criteria, with calibration against frontier models to ensure challenge and rigor.
#J-18808-Ljbffr
Computational Mathematician: AI Benchmark Problem Designer in London employer: Mercor
Mercor is an exceptional employer that champions innovation and creativity in the AI sector, offering a fully remote work environment that promotes flexibility and independence. With a strong focus on employee growth, team members are encouraged to take ownership of their projects while benefiting from a supportive culture that values collaboration and continuous learning. Joining Mercor means being part of a forward-thinking company backed by industry leaders, where your contributions directly impact cutting-edge search technologies.