At a Glance
- Tasks: Evaluate AI models for safety and quality, providing feedback to improve performance.
- Company: Join a leading tech firm focused on AI safety and innovation.
- Benefits: Earn up to $70/hr, enjoy remote work, and develop your skills.
- Other info: Collaborate with top researchers in a dynamic, impactful environment.
- Why this job: Make a real impact on AI safety for millions of users worldwide.
- Qualifications: Bachelor's degree and 5+ years in AI Safety or related fields.
The predicted salary is between 60 - 70 Β£ per hour.
We are seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive, and ambiguous ("grey-area") topics. You will assess AI-generated responses, apply safety policies, and help improve model behaviour through structured evaluations and feedback.
Responsibilities
- Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
- Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains.
- Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
- Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
- Provide structured feedback to improve model alignment and safety performance.
- Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.
Required Qualifications
- Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.
- 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field.
- Excellent written English, critical thinking, and analytical reasoning skills.
- Ability to consistently evaluate nuanced and policy-sensitive scenarios.
Preferred Qualifications
- Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation.
- Familiarity with safety policies, content moderation, or evaluation rubric development.
- Experience reviewing complex, high-risk, or ambiguous content.
Why Join?
- Shape the safety and behaviour of frontier AI models used by millions worldwide.
- Work on challenging, real-world safety evaluations across nuanced and high-impact domains.
- Collaborate with leading AI researchers, engineers, and safety teams.
AI Safety Specialist - Fully Remote | Upto $70/hr employer: Obsidian
At Obsidian, we pride ourselves on fostering a collaborative and innovative work culture that empowers our ML Scientists to push the boundaries of machine learning research. Located in a vibrant tech hub, we offer competitive compensation, flexible project-based work, and ample opportunities for professional growth, making us an excellent employer for those looking to make a meaningful impact in the field of AI.