Mercor is seeking PhD- and Master’s-level scientists to author AI evaluation tasks as part of a new Sci Code benchmark. You will craft original, executable research problems that current frontier models cannot solve and contribute to a benchmark for scientific computing with a focus on chemistry subdomains.
You will source material, write prompts, define grading criteria, and calibrate tasks against models that fail more often than succeed, in a collaborative experiment with leading AI labs.
#J-18808-Ljbffr
AI Benchmark Scientist - Quantum & Computational Chemistry employer: Obsidian
Obsidian is an exceptional employer located in the vibrant Greater London area, offering a dynamic work culture that fosters innovation and collaboration among experts in the field. Employees benefit from a fast-start program with opportunities for growth and extension, alongside a commitment to quality in AI model training that makes a meaningful impact in genomics. With a focus on professional development and a supportive environment, Obsidian is dedicated to empowering its team members to excel in their careers.