AI Safety Practitioner - Expert Evaluator

AI Safety Practitioner - Expert Evaluator

Full-Time 59400 - 72600 Β£ / year (est.) No working from home possible
O

At a Glance

  • Tasks: Evaluate AI models for safety, quality, and alignment on complex topics.
  • Company: Join a leading AI organisation focused on safety and innovation.
  • Benefits: Competitive salary, flexible work options, and opportunities for professional growth.
  • Other info: Collaborate with top researchers in a dynamic and impactful environment.
  • Why this job: Make a real impact on AI safety and help shape technology for millions.
  • Qualifications: Bachelor's degree and 5+ years in AI Safety or related fields required.

The predicted salary is between 59400 - 72600 Β£ per year.

We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behaviour through structured evaluations and feedback.

Responsibilities

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
  • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
  • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
  • Provide structured feedback to improve model alignment and safety performance.
  • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Required Qualifications

  • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.
  • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field.
  • Excellent written English, critical thinking, and analytical reasoning skills.
  • Ability to consistently evaluate nuanced and policy-sensitive scenarios.

Preferred Qualifications

  • Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
  • Familiarity with safety policies, content moderation, or evaluation rubric development.
  • Experience reviewing complex, high-risk, or ambiguous content.

Why Join?

  • Shape the safety and behaviour of frontier AI models used by millions worldwide.
  • Work on challenging, real-world safety evaluations across nuanced and high-impact domains.
  • Collaborate with leading AI researchers, engineers, and safety teams.

AI Safety Practitioner - Expert Evaluator employer: Obsidian

At Obsidian, we pride ourselves on fostering a collaborative and innovative work culture that empowers our ML Scientists to push the boundaries of machine learning research. Located in a vibrant tech hub, we offer competitive compensation, flexible project-based work, and ample opportunities for professional growth, making us an excellent employer for those looking to make a meaningful impact in the field of AI.

O

Contact Details:

Obsidian Recruitment Team

We think you need these skills to ace AI Safety Practitioner - Expert Evaluator

Communication Skills
Problem-Solving Skills
Compassion
Flexibility
Teamwork
Organizational Skills
Adaptability