Frontier AI Safety Evaluator & Model Alignment Expert

Frontier AI Safety Evaluator & Model Alignment Expert

Full-Time 63000 - 77000 Β£ / year (est.) No working from home possible
O

At a Glance

  • Tasks: Evaluate frontier AI models and ensure they meet safety and policy standards.
  • Company: Obsidian, a leader in AI safety and innovation.
  • Benefits: Competitive salary, flexible work options, and opportunities for professional growth.
  • Other info: Dynamic role with a focus on impactful, real-world applications.
  • Why this job: Join a mission-driven team to shape the future of AI safety and ethics.
  • Qualifications: Experience in AI safety evaluation and strong analytical skills.

The predicted salary is between 63000 - 77000 Β£ per year.

Obsidian is seeking experienced AI Safety Practitioners to evaluate frontier AI models across complex, policy-sensitive, and ambiguous thoughts. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.

Responsibilities include:

  • Evaluating safety, factual accuracy, and policy compliance
  • Reviewing content on misinformation, political persuasion, self-harm, violence, cyber, and biosecurity

Frontier AI Safety Evaluator & Model Alignment Expert employer: Obsidian

Obsidian is an exceptional employer located in the vibrant Greater London area, offering a dynamic work culture that fosters innovation and collaboration among experts in the field. Employees benefit from a fast-start program with opportunities for growth and extension, alongside a commitment to quality in AI model training that makes a meaningful impact in genomics. With a focus on professional development and a supportive environment, Obsidian is dedicated to empowering its team members to excel in their careers.

O

Contact Details:

Obsidian Recruitment Team

We think you need these skills to ace Frontier AI Safety Evaluator & Model Alignment Expert

AI Safety Evaluation
Model Alignment
Policy Compliance Assessment
Factual Accuracy Evaluation
Content Review
Misinformation Analysis
Political Persuasion Assessment