Senior Research Scientist: Steerable Translation Models

Senior Research Scientist: Steerable Translation Models

Full-Time 60000 - 80000 Β£ / year (est.) Home office (partial)
D

At a Glance

  • Tasks: Lead the development of next-gen translation models and mentor a dynamic team.
  • Company: Join DeepL, a leader in AI-driven translation technology.
  • Benefits: Competitive salary, flexible working arrangements, and opportunities for professional growth.
  • Other info: Work in a hybrid setup with a passionate team focused on innovation.
  • Why this job: Make a significant impact on language technology with cutting-edge research.
  • Qualifications: Expertise in machine learning and experience in leading research projects.

The predicted salary is between 60000 - 80000 Β£ per year.

Deep L is hiring a Senior Research Scientist to lead fine-tuning, post-training, and RL for its next-generation translation models.

You will blend human expert data with synthetic data to steer models according to user instructions and context, working on models with hundreds of billions of parameters.

You will own the full lifecycle from prototyping to real-time production, establish robust evaluation, and mentor a fast-moving team of researchers and engineers in a hybrid UK-based setup.

#J-18808-Ljbffr

Senior Research Scientist: Steerable Translation Models employer: DeepL

DeepL is an exceptional employer that fosters a collaborative and innovative work culture, where engineers are empowered to take ownership of their projects and contribute to meaningful identity solutions. Located in a vibrant tech hub, employees benefit from continuous growth opportunities, competitive compensation, and a supportive environment that values creativity and teamwork. Join us to be part of a forward-thinking company that prioritises both personal and professional development.

D

Contact Details:

DeepL Recruitment Team

We think you need these skills to ace Senior Research Scientist: Steerable Translation Models

Fine-Tuning
Post-Training
Reinforcement Learning (RL)
Data Blending
Model Evaluation
Prototyping
Real-Time Production