At a Glance
- Tasks: Design and build AI solutions for complex legal document understanding.
- Company: Join a leading tech firm transforming the legal industry.
- Benefits: Enjoy flexible work, competitive pay, and comprehensive wellness programs.
- Other info: Collaborative culture with opportunities for career growth and social impact.
- Why this job: Make a real impact on how legal professionals research and analyse documents.
- Qualifications: PhD or Master's in relevant fields with hands-on AI experience.
The predicted salary is between 63000 - 77000 £ per year.
This position is based in either Zug, Switzerland or London, UK. Want to use your experience of building search-led AI solutions to enhance our leading products in the tax, legal and professional services industries? Document understanding is a foundational intelligence layer that powers every major capability across our legal AI platform, from search and information extraction to agentic reasoning in products like Westlaw, PracticalLaw, and CoCounsel.
In this new role, you'll build state-of-the-art semantic chunking, document enrichment, and knowledge graph construction systems that serve as the cognitive foundation multiple product teams depend on, working across authoritative legal, tax and accounting content and extraordinarily diverse customer data. This is a rare opportunity to solve publishing-quality research problems with immediate production impact—your innovations will directly shape how millions of legal professionals research, analyze, and reason over complex legal documents while advancing the capabilities that enable the next generation of intelligent legal AI agents.
About The Role
- Innovate & Deliver: Design, build, test, and deploy end-to-end AI solutions for complex document understanding tasks in the legal domain. Develop advanced models for semantic chunking of lengthy, non-uniformly structured legal documents with adjustable granularity levels for different use cases. Build document enrichment systems that classify documents according to legal and customer-defined taxonomies and extract rich metadata. Create LLM-based knowledge graph construction pipelines that extract and link heterogeneous legal knowledge including citations, entities, and legal concepts across diverse legal content. Develop scalable synthetic data generation systems to support model training, simulate complex legal research queries and generate hallucination-free answers. Work in collaboration with engineering to ensure well-managed software delivery and reliability at scale.
- Evaluate & Optimize: Develop comprehensive data and evaluation strategies for both component-level and end-to-end quality, leveraging expert human annotation and synthetic data generation. Apply robust training and evaluation methodologies that balance model performance with latency requirements, particularly for SLM-based solutions. Apply knowledge distillation techniques to compress large models into efficient SLMs suitable for production deployment.
- Drive Technical Decisions: Independently determine appropriate architectures for challenging document understanding problems including: semantic chunking strategies that handle diverse document formats, preserve legal document structure, and adapt to different granularity needs; document classification approaches that work across varying legal taxonomies and generalize to customer-defined schemas; LLM-based knowledge extraction methods that handle challenges like citation recognition errors and contextual references; multi-document reasoning architectures for generating synthetic multi-hop queries that reflect complex legal research patterns. Balance accuracy, efficiency, and scalability while solving real-world challenges like handling diverse document formats and content types.
- Align & Communicate: Partner closely with Engineering and Product teams to translate complex legal document understanding challenges into scalable, production-ready solutions. Engage stakeholders across multiple product lines to deeply understand use case requirements, shaping objectives that align document understanding capabilities with diverse business needs including next-generation search and deep legal research.
- Advance the Field: Maintain scientific and technical expertise in one or more relevant areas as demonstrated through product deliverables, published research at top venues (e.g., ACL, EMNLP, ICLR, NeurIPS, SIGIR, KDD), and intellectual property.
About You
- PhD in Computer Science, AI, NLP, or a related field, or a Master’s with equivalent research/industry experience.
- Demonstrable hands-on experience building and deploying document understanding systems, information extraction pipelines, or knowledge graph construction using deep learning, LLMs and NLP methods.
- Proven ability to translate complex document understanding problems into innovative AI applications that balance accuracy and efficiency.
- Professional experience scaling yourself and leading through others, in an applied research setting.
- Strong programming skills (e.g., Python) and experience with modern deep learning frameworks (e.g., PyTorch, Hugging Face Transformers, DeepSpeed).
- Publications at relevant venues such as ACL, EMNLP, ICLR, NeurIPS, SIGIR, KDD.
Technical Qualifications
- Deep understanding of document understanding fundamentals: document layout analysis, semantic chunking approaches beyond fixed-size or paragraph-based methods, document classification handling hierarchical taxonomies, imbalanced multi-label classification, and adapting to domain-specific schemas.
- Expertise in knowledge extraction and knowledge graph construction: entity recognition and linking, relation extraction, citation parsing, and building graph representations from unstructured text.
- Expertise in LLM-based information extraction, few-shot and multi-task learning, post-training and knowledge distillation.
- Solid understanding of synthetic data generation techniques for NLP, including query-answer generation with verification and scalable data augmentation for training specialized models.
- Solid understanding of efficiency optimization including knowledge distillation, model compression, and designing SLM-based solutions that balance performance with computational constraints.
- Solid understanding of DL/ML approaches used for NLP tasks.
- Experience designing annotation workflows, creating high-quality labeled datasets with clear guidelines, and developing evaluation frameworks for document understanding tasks.
What's in it For You?
- Hybrid Work Model: We've adopted a flexible hybrid working environment (2-3 days a week in the office depending on the role) for our office-based roles while delivering a seamless experience that is digitally and physically connected.
- Flexibility & Work-Life Balance: Flex My Way is a set of supportive workplace policies designed to help manage personal and professional responsibilities, whether caring for family, giving back to the community, or finding time to refresh and reset. This builds upon our flexible work arrangements, including work from anywhere for up to 8 weeks per year, empowering employees to achieve a better work-life balance.
- Career Development and Growth: By fostering a culture of continuous learning and skill development, we prepare our talent to tackle tomorrow’s challenges and deliver real-world solutions. Our Grow My Way programming and skills-first approach ensures you have the tools and knowledge to grow, lead, and thrive in an AI-enabled future.
- Industry Competitive Benefits: We offer comprehensive benefit plans to include flexible vacation, two company-wide Mental Health Days off, access to the Headspace app, retirement savings, tuition reimbursement, employee incentive programs, and resources for mental, physical, and financial wellbeing.
- Culture: Globally recognized, award-winning reputation for inclusion and belonging, flexibility, work-life balance, and more. We live by our values: Obsess over our Customers, Compete to Win, Challenge (Y)our Thinking, Act Fast / Learn Fast, and Stronger Together.
- Social Impact: Make an impact in your community with our Social Impact Institute. We offer employees two paid volunteer days off annually and opportunities to get involved with pro-bono consulting projects and Environmental, Social, and Governance (ESG) initiatives.
- Making a Real-World Impact: We are one of the few companies globally that helps its customers pursue justice, truth, and transparency. Together, with the professionals and institutions we serve, we help uphold the rule of law, turn the wheels of commerce, catch bad actors, report the facts, and provide trusted, unbiased information to people all over the world.
Senior Applied Scientist, Search - NLP/GenAI employer: Thomson Reuters
As an Editor at The Insurer, you will thrive in a dynamic and supportive work environment located in the heart of London, where innovation meets collaboration. Our hybrid work model promotes flexibility, allowing you to balance personal and professional commitments while benefiting from comprehensive career development programmes tailored to help you excel in your role. With a strong emphasis on social impact and employee wellbeing, we offer competitive benefits that include mental health days, volunteer opportunities, and resources for your overall wellness, making us an exceptional employer for those seeking meaningful and rewarding employment.
StudySmarter Expert Advice🤫
We think this is how you could land Senior Applied Scientist, Search - NLP/GenAI
✨Get Involved in Data Science Meetups
Tap into local data science meetups or workshops to connect with fellow enthusiasts and professionals. These events are goldmines for networking, and sometimes even lead directly to job openings at companies like Thomson Reuters!
✨Show Off Your Projects
Start building a public portfolio showcasing your data science projects on platforms like GitHub or personal websites. Highlight unique analyses or models you've developed. This not only demonstrates your skills but also gets your name out there for roles like Senior Applied Scientist, Search - NLP/GenAI at Thomson Reuters.
✨Leverage Professional Networks
Join professional bodies related to data science, like the Data Science Society or similar organisations. Getting involved can lead to mentorship opportunities and insider knowledge about full-time positions at companies like Thomson Reuters.
✨Apply Directly through Our Website
When you find a suitable opening like Senior Applied Scientist, Search - NLP/GenAI at Thomson Reuters, make sure to apply directly through our website. It gives you an edge and shows you're keen to join our team. Plus, who doesn’t love a direct application? It’s easier than navigating through job boards!
We think you need these skills to ace Senior Applied Scientist, Search - NLP/GenAI
Some tips for your application 🫡
Show Off Your Projects:In the world of data science, your projects can speak volumes about your skills. Make sure to showcase a few key projects in your CV or portfolio, especially those that highlight your ability to work with data sets, build models, or use relevant tools like Python, R, or SQL. Don’t forget to include links to any GitHub repositories if applicable!
Quantify Your Achievements:Employers love numbers! When drafting your CV, highlight your achievements with quantifiable results. For instance, mention how your data analysis led to a certain percentage increase in efficiency or revenue at a previous job or project. These details can really make your application pop!
Craft a Tailored Cover Letter:For a full-time role at Thomson Reuters, your cover letter should reflect your passion for data science and your excitement about the specific projects or values of the company. Dive into why you’re a good fit, how your skills align with their needs, and any unique perspectives you can bring to the team.
Stand Out with Relevant Courses and Certifications:Although experience talks, relevant courses or certifications can be your ticket to impressing hiring managers at Thomson Reuters. Mention any standout courses you've completed that equipped you with essential skills, such as machine learning certifications or data visualisation courses. This shows your commitment to continuously developing your skills in the field!
How to prepare for a job interview at Thomson Reuters
✨Brush Up on Your Statistics
For a data science role, we need to seriously sharpen our statistics skills. Get ready to tackle technical questions on probability distributions, hypothesis testing, and regression analysis. These are often the bread and butter of data science interviews, so don't just skim over them!
✨Showcase Your Projects
Prepare a killer portfolio showcasing your data science projects. We should include details about the datasets used, the tools and techniques applied, and the impact of your findings. If we can walk them through a particularly challenging project or a cool visualisation that had real-world implications, it’ll really make us stand out!
✨Get Comfortable with Python and R
Most data science positions require us to be proficient in programming languages like Python and R. We should practice common libraries like pandas, NumPy, and scikit-learn, and be ready for live coding exercises or algorithm questions. Showing off our coding chops can really impress the interviewers at Thomson Reuters!
✨Prepare for Case Studies
Expect to encounter real-world case studies during the interview. We might be asked how we’d approach a data problem or analyse a dataset to extract insights. It's essential to think out loud and demonstrate our problem-solving process so that the interviewer can see our logical thinking in action.