At a Glance
- Tasks: Lead cutting-edge research in AI model steering and fine-tuning for translation models.
- Company: Join DeepL, a global leader in AI technology with a mission to enhance communication.
- Benefits: Enjoy flexible hours, hybrid work, competitive salary, and 30 days of annual leave.
- Other info: Be part of a diverse team with regular team events and opportunities for personal growth.
- Why this job: Shape the future of AI while working on impactful projects in a collaborative environment.
- Qualifications: Proven experience in LLM post-training and strong coding skills in Python.
The predicted salary is between 70000 - 90000 £ per year.
Meet DeepL. DeepL is a global AI product and research company focused on building secure, intelligent solutions to complex business problems. Over 200,000 business customers and millions of individuals across 228 global markets today trust DeepL's Language AI platform for human-like translation, improved writing and real-time voice translation. Founded in 2017 by CEO Jaroslaw “Jarek” Kutylowski, DeepL now has around 1,000 passionate employees and is supported by world-renowned investors including Benchmark, IVP, and Index Ventures. Our goal is to become the global leader in trusted, intelligent AI technology, building products that drive better communication, foster connections, and create a meaningful impact.
What sets us apart is our blend of cutting-edge AI technology, meaningful work, and a culture where people truly thrive. We’re a team of innovators, researchers, and creators driven by a shared purpose to unlock human potential by making work simpler, smarter, and more connected. When we share what it’s like to work at DeepL, the reactions are overwhelmingly positive. This might be because of our technology that helps millions of people and businesses communicate and work better every day, or because of the trust, curiosity, and care that shape our culture. What we know for sure is this: being part of DeepL means joining a team dedicated to innovation, growth, and well-being.
Our Language AI teams form the foundation of DeepL's success. We are a dedicated group of researchers who collaborate closely with engineers, product managers, and designers. Our mission is to build the world's leading language AI system to deliver perfect translations for the most demanding use cases. To that end, we take responsibility for the entire life cycle of the machine learning models that power our language AI products. This includes data, training, quality assurance, and operational aspects. In our highly collaborative teams, each person has the scope to drive impact across the company.
Your responsibilities include:
- Driving the development of translation models that are steerable conditioned on user preferences, rules and context.
- Conducting hands-on research and development on post-training for our core translation models: supervised fine-tuning, knowledge distillation, preference optimisation, and reinforcement learning tuned to translation quality.
- Building reward models and evaluator models for translation, including rubric- and reference-based grading, and investigating and mitigating reward hacking and quality-estimation failure modes.
- Driving an agenda toward models that ingest multimodal content and context to increase translation quality.
- Owning the full lifecycle of model delivery: prototyping, ablations, training, evaluation, optimisation, and production deployment, working closely with engineering to ship into real-time systems at scale.
- Establishing strong practices for evaluation, reproducibility, monitoring, and continuous model improvement in production.
- Mentoring researchers and engineers, promoting hands-on collaboration, and raising the bar for model quality.
Qualities we look for:
- Proven experience making large models steerable and instruction-following by identifying the most effective method to instill a given behaviour, drawing from instruction tuning, latent space methods, steering vectors, and/or constrained encoding and decoding methods.
- Deep, hands-on expertise in LLM post-training (SFT, DPO), knowledge distillation (teacher-student training), and/or reinforcement learning (RLHF/RLAIF, PPO/GSPO, and reward modelling).
- Strong data-centric instincts for building synthetic-data and preference-data pipelines, LLM-as-judge generation, data curation and filtering, and reasoning about data mixtures and ablations.
- Experience designing evaluation and reward signals using automatic metrics, LLM-as-judge evaluation, non-verifiable rewards, and human-in-the-loop evaluation.
- A hands-on builder who enjoys training models, running experiments, debugging pipelines, and integrating ML systems into production while staying grounded in product impact and real-world quality.
- Ownership of a substantial research direction with strong execution, and experience mentoring others on a fast-moving, applied research team.
- Strong coding and experimentation skills (Python, PyTorch/JAX/Tensorflow), and the ability to communicate clearly and align research with product and engineering priorities.
Nice to have:
- Demonstrated experience fine-tuning and training large models at scale, including distributed/multi-node training (e.g. FSDP, DeepSpeed, or Megatron-style frameworks) and efficient training techniques.
- Experience fine-tuning existing reasoning models for specific tasks and behaviours without degrading their reasoning capabilities.
- Experience with machine translation, multilingual NLP, or language quality estimation.
- Familiarity with inference and serving at scale (e.g. via vLLM, SGLang, TensorRT-LLM, etc) and long-context modelling.
- Publications at top-tier venues.
What we offer:
- Diverse and internationally distributed team: joining our team means becoming part of a large, global community with people of more than 90 nationalities.
- Open communication, regular feedback: as a language-focused company, we value the importance of clear, honest communication.
- Hybrid work, flexible hours: we offer a hybrid work schedule, with team members coming into the office twice a week.
- Virtual Shares: an ownership mindset in every role.
- Regular in-person team events: we bond over vibrant events that are as unique as our team.
- Monthly full-day hacking sessions: every month, we have Hack Fridays.
- 30 days of annual leave: we value your peace of mind.
- Competitive benefits: we’ve crafted it to reflect the diversity of our team.
If this role and our mission resonate with you, but you’re hesitant because you don’t check all the boxes, don’t let that hold you back. At DeepL, it’s all about the value you bring and the growth we can foster together. We can’t wait to meet you!
We are an equal opportunity employer. You are welcome at DeepL for who you are - we appreciate authenticity here. Our product is for everyone, and so is our workplace. The more voices we have represented and amplified in our business, the more we will all succeed, contribute, and think forward!
Senior Research Scientist | Model Steering in London employer: DeepL SE
DeepL is an exceptional employer that fosters a dynamic and inclusive work culture, where innovation thrives and employees are empowered to lead impactful projects. With a strong focus on professional development, the company offers ample growth opportunities and a hybrid work model that promotes work-life balance, making it an ideal place for those looking to make a meaningful contribution in the fast-evolving field of AI and payments.
StudySmarter Expert Advice🤫
We think this is how you could land Senior Research Scientist | Model Steering in London
✨Get Involved in Research Communities
Dive headfirst into the scientific research world by joining relevant communities and forums. Engage in discussions, share your insights, and even attend conferences or seminars in your field. This not only boosts your visibility but can also lead to potential job opportunities—don't forget to connect with like-minded folks!
✨Show Off Your Research Projects
Have you worked on any cool research projects? Make it easy for potential employers to see your work by creating a portfolio or a personal website. This way, when you apply for roles like the one at DeepL SE, you can point them to your projects and publications, showcasing your expertise directly.
✨Utilise Professional Networks
Networking is key in scientific research. Join professional bodies or organisations related to your field. They often have job boards and resources tailored for job seekers. Make connections with professionals who may know about openings or can give you tips on landing a full-time position.
✨Keep Your Eyes on Openings & Apply Directly
Don’t just rely on job boards! Keep an eye on the careers section of the websites of companies like DeepL SE. Apply directly through their website because sometimes they post jobs there before anywhere else. Plus, it shows your proactive approach!
We think you need these skills to ace Senior Research Scientist | Model Steering in London
Some tips for your application 🫡
Highlight Your Research Experience:When applying for a full-time role in scientific research, make sure to emphasise your research experience prominently in your CV. Share specific projects you’ve worked on, the methodologies you used, and any significant findings. If you’ve published papers or presented at conferences, definitely include that too – it shows you’re on it in the academic world!
Tailor Your Cover Letter to the Research Area:Your cover letter should reflect your passion for the specific area of research at DeepL SE. Mention relevant experiences that align with the organisation’s goals or projects. This shows that you’ve done your homework and are genuinely interested in the position – plus, it helps us see how you’d fit into the team dynamics.
Showcase Your Data Analysis Skills:In scientific research, data analysis skills are a big deal! Make sure to detail any relevant analytical tools or software you’re familiar with, like R, Python, or statistical packages. Employers are keen to know you can handle the data-heavy elements of the role, so add specific examples where you’ve used these skills effectively.
Discuss Your Future Research Goals:In your motivation section, it’s a great idea to talk about your future research goals and how they align with the work being done at DeepL SE. This shows that you’re not just looking for any job, but rather a chance to contribute meaningfully to the field. We love to see applicants who are forward-thinking and enthusiastic about their research journey!
How to prepare for a job interview at DeepL SE
✨Showcase Your Research Skills
In scientific research, it’s crucial to demonstrate your ability to design and conduct experiments. Come armed with examples of past projects where you've developed hypotheses, collected data, and analysed results. Be ready to discuss any specific methodologies or tools you’ve used, like PCR techniques or statistical software.
✨Prepare for Technical Questions
Expect some technical questions specific to your field. Make sure you're up to speed with recent advancements in scientific research related to the role at DeepL SE. Brush up on concepts relevant to their projects and be prepared to discuss how you would approach a specific research problem or challenge they might face.
✨Know Your Publications
If you've authored or co-authored any papers, be prepared to discuss them! Highlighting your contributions to published research can really set you apart. It shows not only your expertise but also your ability to communicate complex ideas clearly, which is key in scientific research roles.
✨Exhibit Your Team Spirit
In full-time roles, collaboration is often at the heart of scientific research. Prepare examples that show how you've successfully worked in teams, dealt with conflicts, or contributed to group projects. We want to know how you can work effectively with the team at DeepL SE to drive research projects forward.