Axiōma Search is seeking a driven RL research scientist to advance post-training methods for robotics through scalable, multimodal models. The role focuses on designing and running RL post-training pipelines, leveraging GPU-native simulation environments and cutting-edge diffusion and autoregressive approaches.
You will collaborate with training infra and simulation teams, craft robust reward signals, and push results from simulation to real-world robotic behaviour.
#J-18808-Ljbffr
Senior RL Research Scientist — Robotics, GPU Simulation employer: Axiōma Search
Axiōma Search is an exceptional employer for those passionate about machine learning and time-series forecasting. With a collaborative work culture that prioritises innovation and employee growth, team members are encouraged to explore new ideas and technologies while benefiting from a global network of experts. Located in Europe, the company offers unique opportunities for professional development and the chance to make a significant impact in the field of AI.