Research Scientist: Post-Training & RL for Agents (London)

Research Scientist: Post-Training & RL for Agents (London)

Full-Time 60000 - 80000 Β£ / year (est.) No working from home possible
Talanto

At a Glance

  • Tasks: Conduct cutting-edge research in AI/ML, running experiments and analysing complex systems.
  • Company: Join Google DeepMind, a leader in AI innovation and research.
  • Benefits: Competitive salary, collaborative environment, and opportunities for professional growth.
  • Other info: Fast-paced, dynamic team environment located in London.
  • Why this job: Make a real impact in AI by advancing post-training and reinforcement learning.
  • Qualifications: PhD in ML or equivalent experience, with strong communication skills.

The predicted salary is between 60000 - 80000 Β£ per year.

Google Deep Mind is seeking a Research Scientist (AI/ML) to advance post-training and RL for agents and LLM-based systems.

You will run end-to-end experiments, design evaluations, and analyze complex failure modes, collaborating with engineering and research teammates in a fast-paced environment.

The role emphasizes rigorous experimentation, scaling, and evaluation, with onsite work in London.

A Ph D in ML or equivalent experience is preferred and strong communication is essential.

#J-18808-Ljbffr

Research Scientist: Post-Training & RL for Agents (London) employer: Talanto

Almedia is an exceptional employer that fosters a dynamic and innovative work culture, perfect for those passionate about machine learning and real-time personalization. With a strong emphasis on employee growth, you will have the opportunity to mentor fellow engineers while collaborating with cross-functional teams in the vibrant city of London. The hybrid working arrangements and commitment to impactful projects make Almedia a rewarding place to advance your career in AdTech.

Talanto

Contact Details:

Talanto Recruitment Team

We think you need these skills to ace Research Scientist: Post-Training & RL for Agents (London)

Machine Learning (ML)
Reinforcement Learning (RL)
Experimental Design
Data Analysis
Failure Mode Analysis
Collaboration Skills
Communication Skills