Anthropic is seeking a Research Engineer for the RL Scaling Science team to design and run large-scale reinforcement learning experiments and ship validated findings into production training. This role sits at the research/engineering boundary, requiring rigorous experimentation, benchmarks, and collaboration with adjacent RL teams to advance our scalable AI systems.
You will work with Python on frontier-scale training, translating insights into practical training recipes while considering
#J-18808-Ljbffr