AI Safety Red Teamer Expert

AI Safety Red Teamer Expert

Full-Time 85500 - 104500 Β£ / year (est.) No working from home possible
O

At a Glance

  • Tasks: Identify vulnerabilities in AI systems through adversarial testing and design challenging prompts.
  • Company: Join a leading team focused on AI safety and innovation.
  • Benefits: Competitive salary, flexible work options, and opportunities for professional growth.
  • Other info: Collaborate with top researchers in a dynamic and impactful environment.
  • Why this job: Make a real impact on the future of AI safety and security.
  • Qualifications: 5+ years in AI Safety or related fields with strong analytical skills.

The predicted salary is between 85500 - 104500 Β£ per year.

We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behaviour across complex, high-risk, and ambiguous ("grey-area") topics.

Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Required Qualifications

  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Preferred Qualifications

  • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.

Why Join?

  • Help secure and strengthen the next generation of frontier AI models.
  • Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams.
  • Influence how AI systems respond to complex, real-world safety challenges.

AI Safety Red Teamer Expert employer: Obsidian

At Obsidian, we pride ourselves on fostering a collaborative and innovative work culture that empowers our ML Scientists to push the boundaries of machine learning research. Located in a vibrant tech hub, we offer competitive compensation, flexible project-based work, and ample opportunities for professional growth, making us an excellent employer for those looking to make a meaningful impact in the field of AI.

O

Contact Details:

Obsidian Recruitment Team

We think you need these skills to ace AI Safety Red Teamer Expert

Adversarial Testing
Prompt Design
Model Evaluation
Analytical Reasoning
Written Communication Skills
AI Safety
Cybersecurity