At a Glance
- Tasks: Define grading criteria and evaluate AI performance in real-world data science tasks.
- Company: Join a leading AI research organisation with a focus on innovation.
- Benefits: Competitive salary, flexible work arrangements, and opportunities for professional growth.
- Other info: Collaborative environment with a chance to influence cutting-edge AI developments.
- Why this job: Shape the future of AI by ensuring high standards in data science evaluation.
- Qualifications: 5+ years in data science, strong communication skills, and experience with AI projects.
The predicted salary is between 59400 - 72600 Β£ per year.
Mercor is partnering with a leading AI research organization to engage experienced data scientists for a project focused on evaluating how well AI systems perform real-world data science work. Rather than producing deliverables yourself, you will define what excellent work looks like: designing task-specific grading criteria and scoring completed work samples with rigorous, well-reasoned written justifications.
Key Responsibilities
- Design precise, task-specific grading criteria for real-world data science deliverables (analyses, models, dashboards, experiment readouts, and written recommendations)
- Score AI-generated and human work samples against those criteria, with detailed written justifications for every score
- Apply consistent, evidence-based judgment so that scores are reproducible and defensible
- Incorporate structured feedback from senior reviewers and iterate quickly on your work
Ideal Qualifications
- 5+ years of professional data science experience in industry
- Background in business operations, product, or growth data science at top-tier technology companies
- Deep fluency in experiment design and A/B testing, metric definition, SQL/Python analysis, and communicating findings to executive stakeholders
- Exceptionally strong written communication
- Detail-oriented, consistent, and comfortable having your judgment reviewed and calibrated against peers
- Prior experience with AI training, evaluation, or human-data projects is a strong plus
Data Science Expert - AI Evaluation employer: Obsidian
Obsidian is an exceptional employer located in the vibrant Greater London area, offering a dynamic work culture that fosters innovation and collaboration among experts in the field. Employees benefit from a fast-start program with opportunities for growth and extension, alongside a commitment to quality in AI model training that makes a meaningful impact in genomics. With a focus on professional development and a supportive environment, Obsidian is dedicated to empowering its team members to excel in their careers.