Reinforcement Learning Research Engineer – Safe AI in London

Reinforcement Learning Research Engineer – Safe AI in London

London Full-Time No working from home possible
A

Anthropic is seeking a Research Engineer within Reinforcement Learning to advance capabilities and safety of large language models. You will implement novel approaches and contribute to research direction while collaborating with researchers and engineers.

You will design and optimize RL infrastructure, create training environments, and push state-of-the-art methods in agentic model development, tool use, and reasoning.

#J-18808-Ljbffr

Reinforcement Learning Research Engineer – Safe AI in London employer: Anthropic

At Anthropic, we pride ourselves on being an exceptional employer that fosters a culture of innovation and collaboration. Our team-oriented environment encourages personal growth and empowers employees to take ownership of their projects, making a meaningful impact in the tech landscape. Located in a vibrant area, we offer competitive benefits and unique opportunities for professional development, ensuring that our engineers thrive both personally and professionally.

A

Contact Details:

Anthropic Recruitment Team