Anthropic is seeking a Research Engineer within Reinforcement Learning to advance capabilities and safety of large language models. You will implement novel approaches and contribute to research direction while collaborating with researchers and engineers.
You will design and optimize RL infrastructure, create training environments, and push state-of-the-art methods in agentic model development, tool use, and reasoning.
#J-18808-Ljbffr
Reinforcement Learning Research Engineer β Safe AI in London employer: Anthropic
At Anthropic, we pride ourselves on being an exceptional employer that fosters a culture of innovation and collaboration. Our team-oriented environment encourages personal growth and empowers employees to take ownership of their projects, making a meaningful impact in the tech landscape. Located in a vibrant area, we offer competitive benefits and unique opportunities for professional development, ensuring that our engineers thrive both personally and professionally.