At a Glance
- Tasks: Design and build NLP pipelines to extract insights from scientific texts at scale.
- Company: Join a leading publisher transforming knowledge into impactful discoveries.
- Benefits: Competitive salary, flexible work hours, and opportunities for professional growth.
- Other info: Collaborative culture with a focus on innovation and continuous learning.
- Why this job: Make a real impact in science and learning with cutting-edge AI technologies.
- Qualifications: Deep Python skills and strong background in NLP techniques required.
The predicted salary is between 59100 - 84633 £ per year.
We believe in bold ideas, diverse perspectives, and the drive to transform knowledge into impact. Here, your curiosity fuels progress, your voice shapes innovation, and your ambition helps redefine what’s possible within science and learning. We are a culture that obsesses over impact, challenges, and drives what’s next to power infinite possibilities for our customers, colleagues and society at large.
About the Role:
We're building the systems that turn one of the world's largest scientific corpora into research intelligence. That means production NLP pipelines running over millions of journal articles, extracting entities, classifications, claim tuples, and summaries optimized for use by downstream agentic applications. We're looking for a senior data scientist to own domain-specific content modeling work end to end, from the eval set through the pipeline stage that ships it. You'll join a small, senior team where data scientists own their models in production. You'll write the code, own the evaluations, ship the changes, and stay accountable for the outcomes. This is a hands-on role for someone who wants to see their models through to real users in a rapidly evolving market.
What you'll do:
- Design and build NLP enrichment pipelines that extract entities, classifications, claims, and summaries from scientific full-text at scale.
- Compare NLP approaches to extraction and enrichment against LLM-based approaches, and pick the right tool for each task.
- Own evaluation. Build the golden sets in consultation with SMEs and vendors, choose the metrics, and make productive tradeoffs between speed, quality, and cost.
- Write production-quality Python.
- Manage concurrency and cost for high-volume LLM workloads.
- Collaborate with a team of data engineers to orchestrate work in data pipeline and data build tools like Airflow and Dagster.
- Contribute to agentic AI application work: tool-using systems that reason over the enriched corpus.
- Work directly with editors, product managers, and engineers.
What you'll bring:
- Deep Python experience in production, at scale.
- Strong NLP background across modern and classical approaches.
- A track record of shipping systems that deliver value to real users.
Nice to have:
- Experience working with scientific or scholarly text.
- Familiarity with AWS and data lake patterns.
- Experience running LLMs under real cost and latency budgets in production.
Why us:
We publish some of the world's most-read research, applying modern AI to a corpus of trusted scientific knowledge. Researchers will use the systems you build here to move faster and get closer to the answers they came for. That's the work: from knowledge to impact. We power infinite possibilities.
Wiley is an equal opportunity/affirmative action employer. We evaluate all qualified applicants without regard to race, colour, religion, sex, sexual orientation, gender identity or expression, national origin, disability, protected veteran status, genetic information, or based on any individual's status in any group or class protected by applicable laws.
We are proud that our workplace promotes continual learning and internal mobility. We offer meeting-free Friday afternoons allowing more time for heads down work and professional development.
We are committed to fair, transparent pay, and we strive to provide competitive compensation in addition to a comprehensive benefits package.
Principal Data Scientist (NLP + Applied AI) employer: Wiley Global Technology
Wiley is an exceptional employer that champions bold ideas and diverse perspectives, fostering a culture where your curiosity and ambition can drive meaningful impact in science and learning. With a commitment to continual learning and internal mobility, employees enjoy unique benefits such as meeting-free Friday afternoons for focused work and professional development, alongside a comprehensive benefits package and competitive compensation. Here, your contributions are valued, and you have the opportunity to grow within a supportive environment that prioritises health and well-being.
StudySmarter Expert Advice🤫
We think this is how you could land Principal Data Scientist (NLP + Applied AI)
✨Get Involved in Data Science Meetups
Tap into local data science meetups or workshops to connect with fellow enthusiasts and professionals. These events are goldmines for networking, and sometimes even lead directly to job openings at companies like Wiley Global Technology!
✨Show Off Your Projects
Start building a public portfolio showcasing your data science projects on platforms like GitHub or personal websites. Highlight unique analyses or models you've developed. This not only demonstrates your skills but also gets your name out there for roles like Principal Data Scientist (NLP + Applied AI) at Wiley Global Technology.
✨Leverage Professional Networks
Join professional bodies related to data science, like the Data Science Society or similar organisations. Getting involved can lead to mentorship opportunities and insider knowledge about full-time positions at companies like Wiley Global Technology.
✨Apply Directly through Our Website
When you find a suitable opening like Principal Data Scientist (NLP + Applied AI) at Wiley Global Technology, make sure to apply directly through our website. It gives you an edge and shows you're keen to join our team. Plus, who doesn’t love a direct application? It’s easier than navigating through job boards!
We think you need these skills to ace Principal Data Scientist (NLP + Applied AI)
Some tips for your application 🫡
Show Off Your Projects:In the world of data science, your projects can speak volumes about your skills. Make sure to showcase a few key projects in your CV or portfolio, especially those that highlight your ability to work with data sets, build models, or use relevant tools like Python, R, or SQL. Don’t forget to include links to any GitHub repositories if applicable!
Quantify Your Achievements:Employers love numbers! When drafting your CV, highlight your achievements with quantifiable results. For instance, mention how your data analysis led to a certain percentage increase in efficiency or revenue at a previous job or project. These details can really make your application pop!
Craft a Tailored Cover Letter:For a full-time role at Wiley Global Technology, your cover letter should reflect your passion for data science and your excitement about the specific projects or values of the company. Dive into why you’re a good fit, how your skills align with their needs, and any unique perspectives you can bring to the team.
Stand Out with Relevant Courses and Certifications:Although experience talks, relevant courses or certifications can be your ticket to impressing hiring managers at Wiley Global Technology. Mention any standout courses you've completed that equipped you with essential skills, such as machine learning certifications or data visualisation courses. This shows your commitment to continuously developing your skills in the field!
How to prepare for a job interview at Wiley Global Technology
✨Brush Up on Your Statistics
For a data science role, we need to seriously sharpen our statistics skills. Get ready to tackle technical questions on probability distributions, hypothesis testing, and regression analysis. These are often the bread and butter of data science interviews, so don't just skim over them!
✨Showcase Your Projects
Prepare a killer portfolio showcasing your data science projects. We should include details about the datasets used, the tools and techniques applied, and the impact of your findings. If we can walk them through a particularly challenging project or a cool visualisation that had real-world implications, it’ll really make us stand out!
✨Get Comfortable with Python and R
Most data science positions require us to be proficient in programming languages like Python and R. We should practice common libraries like pandas, NumPy, and scikit-learn, and be ready for live coding exercises or algorithm questions. Showing off our coding chops can really impress the interviewers at Wiley Global Technology!
✨Prepare for Case Studies
Expect to encounter real-world case studies during the interview. We might be asked how we’d approach a data problem or analyse a dataset to extract insights. It's essential to think out loud and demonstrate our problem-solving process so that the interviewer can see our logical thinking in action.