Research Engineer: Large-Scale RL & Production Recipes

Research Engineer: Large-Scale RL & Production Recipes

Full-Time No working from home possible
M

Anthropic is seeking a Research Engineer for the RL Scaling Science team to design and run large-scale reinforcement learning experiments and ship validated findings into production training. This role sits at the research/engineering boundary, requiring rigorous experimentation, benchmarks, and collaboration with adjacent RL teams to advance our scalable AI systems.

You will work with Python on frontier-scale training, translating insights into practical training recipes while considering

#J-18808-Ljbffr

M

Contact Details:

Mat Vin Recruitment Team