Anthropic in London, UK, seeks a full-time Alignment Scientist to develop and execute ML experiments, build evaluation tooling for LLM safety, and explore multi-agent RL like AI Debate. Occasional travel to San Francisco is expected; base location London.
You will contribute to safety research papers, blogs, and internal docs, working with a collaborative team. A strong background in software engineering or ML, plus proficiency in Python, is required.
#J-18808-LjbffrAI Safety Research Engineer β ML, RL & LLMs (London) employer: Anthropic
Anthropic is an exceptional employer for those passionate about advancing reinforcement learning in a collaborative and innovative environment. With competitive compensation, generous vacation and parental leave, and flexible working hours, employees enjoy a supportive work culture that prioritises both personal and professional growth. Located in a vibrant office space, team members have the unique opportunity to engage directly with cutting-edge research while making meaningful contributions to the responsible scaling of AI.