Frontier AI Safety Evaluator

Frontier AI Safety Evaluator

Full-Time 49500 - 60500 Β£ / year (est.) No working from home possible
O

At a Glance

  • Tasks: Evaluate frontier AI models for safety and quality while providing structured feedback.
  • Company: Obsidian, a leader in AI safety with a focus on innovation.
  • Benefits: Competitive salary, flexible work options, and opportunities for professional growth.
  • Other info: Collaborative environment with a focus on advancing AI safety standards.
  • Why this job: Join a mission-driven team to shape the future of AI safety and ethics.
  • Qualifications: Experience in journalism, public policy, or related fields is preferred.

The predicted salary is between 49500 - 60500 Β£ per year.

Obsidian seeks experienced AI Safety Practitioners to evaluate frontier models on safety, quality, and alignment across grey-area topics.

You will assess AI responses, apply safety policies, and contribute to improvements through structured evaluations and feedback.

Responsibilities include reviewing high-risk content, refining RLHF/SFT rubrics, and collaborating with researchers to advance safety benchmarking.

A background in journalism, public policy, or related fields is ideal.

#J-18808-Ljbffr

Frontier AI Safety Evaluator employer: Obsidian

Obsidian is an exceptional employer located in the vibrant Greater London area, offering a dynamic work culture that fosters innovation and collaboration among experts in the field. Employees benefit from a fast-start program with opportunities for growth and extension, alongside a commitment to quality in AI model training that makes a meaningful impact in genomics. With a focus on professional development and a supportive environment, Obsidian is dedicated to empowering its team members to excel in their careers.

O

Contact Details:

Obsidian Recruitment Team

We think you need these skills to ace Frontier AI Safety Evaluator

AI Safety Evaluation
Quality Assessment
Alignment Analysis
Safety Policy Application
Structured Evaluations
Feedback Contribution
High-Risk Content Review