Physics AI Benchmark Scientist: Prompt Design & Evaluation in London

Physics AI Benchmark Scientist: Prompt Design & Evaluation in London

London Full-Time No working from home possible
M

Mercor is seeking PhD or Master’s scientists to author AI evaluation tasks for Sci Code in collaboration with leading AI labs on a new benchmark for scientific computing. You will source material, design executable research problems, and craft scoring criteria to challenge frontier models.

The role requires deep knowledge in at least two physics subdomains, strong Python or R skills, and experience with Git/GitHub and Docker workflows. This six-week, part-time engagement starts immediately.

#J-18808-Ljbffr

Physics AI Benchmark Scientist: Prompt Design & Evaluation in London employer: Mercor

Mercor is an exceptional employer that champions innovation and creativity in the AI sector, offering a fully remote work environment that promotes flexibility and independence. With a strong focus on employee growth, team members are encouraged to take ownership of their projects while benefiting from a supportive culture that values collaboration and continuous learning. Joining Mercor means being part of a forward-thinking company backed by industry leaders, where your contributions directly impact cutting-edge search technologies.

M

Contact Details:

Mercor Recruitment Team