AI Safety Research Engineer β€” ML, RL & LLMs (London)

AI Safety Research Engineer β€” ML, RL & LLMs (London)

Full-Time No working from home possible
Anthropic

Anthropic in London, UK, seeks a full-time Alignment Scientist to develop and execute ML experiments, build evaluation tooling for LLM safety, and explore multi-agent RL like AI Debate. Occasional travel to San Francisco is expected; base location London.

You will contribute to safety research papers, blogs, and internal docs, working with a collaborative team. A strong background in software engineering or ML, plus proficiency in Python, is required.

#J-18808-Ljbffr

AI Safety Research Engineer β€” ML, RL & LLMs (London) employer: Anthropic

Anthropic is an exceptional employer for those passionate about advancing reinforcement learning in a collaborative and innovative environment. With competitive compensation, generous vacation and parental leave, and flexible working hours, employees enjoy a supportive work culture that prioritises both personal and professional growth. Located in a vibrant office space, team members have the unique opportunity to engage directly with cutting-edge research while making meaningful contributions to the responsible scaling of AI.

Anthropic

Contact Details:

Anthropic Recruitment Team