Speedrun Talent Network seeks an expert to design, build, and scale distributed RL post-training systems for large-scale models. You will orchestrate trainer, rollout, environment, and reward workloads across thousands of GPUs and integrate inference engines such as vLLM and SGLang.
You will develop RL environments and reward infrastructures, verifiers, and evaluation harnesses, ensuring stable, high-throughput training and rollout across complex, asynchronous pipelines.
#J-18808-Ljbffr
Lead RL Infrastructure Engineer - Scale Training Systems employer: Speedrun Talent Network
At Harvey, we pride ourselves on being an exceptional employer that fosters a dynamic and innovative work culture. Our commitment to employee growth is evident through our focus on leadership development and the unique opportunity to shape the future of AI in legal and professional services. Located in the vibrant EMEA region, we offer a collaborative environment where your contributions directly impact our rapid scaling and success.