Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) as part of a collaboration with leading AI labs to establish a new benchmark for scientific computing. You will author original, executable research problems that frontier models cannot solve.
Based in Greater London, you will source material, write prompts, and build grading criteria, calibrating against frontier models so that tasks ship only when models struggle.
#J-18808-Ljbffr
AI Benchmark Scientist: Biochemistry & Genetics employer: Obsidian
Obsidian is an exceptional employer located in the vibrant Greater London area, offering a dynamic work culture that fosters innovation and collaboration among experts in the field. Employees benefit from a fast-start program with opportunities for growth and extension, alongside a commitment to quality in AI model training that makes a meaningful impact in genomics. With a focus on professional development and a supportive environment, Obsidian is dedicated to empowering its team members to excel in their careers.