NLP Research Engineer

NLP Research Engineer

Full-Time 59400 - 72600 Β£ / year (est.) Home office (partial)
F

At a Glance

  • Tasks: Design and train NLP models while collaborating with a dynamic team.
  • Company: Join a startup-like tech company focused on innovative research solutions.
  • Benefits: Enjoy a 4-day work week, hybrid work, and flexible benefits.
  • Other info: Collaborative environment with opportunities to mentor and learn.
  • Why this job: Make a real impact in the NLP field and grow your skills.
  • Qualifications: Experience in NLP and a passion for clean, reproducible research.

The predicted salary is between 59400 - 72600 Β£ per year.

Our Mission is to democratise the global research industry by providing unbiased business intelligence through analysis of mass public data, using Machine Learning and NLP techniques. We have recently been acquired by Beroe, but have kept the 'startup feeling' within our team, which is now looking to grow.

About the role

The role sits at the front of our data pipeline, designing and training the NLP components that our ETL team productionizes downstream. While we use agentic LLMs throughout our stack, this role is rooted in traditional NLP craft: you should be as comfortable reasoning about vector comparison, clustering behaviour, and tokenisation edge cases as you are prompting a model. You'll work primarily in vector space β€” building, tuning, and evaluating embedding-based systems that transform raw text into structured, meaningful signal. Output from this role β€” models, embedding pipelines, evaluation harnesses, labelled datasets and classifiers etc β€” becomes the spec that our ETL Engineers build into production services, so clear documentation and handoff discipline matter as much as the research itself.

We recognise the importance of proprietary data, and are in the process of creating an extensive network graph capturing decades of analyst insights as well as public information, allowing Graph RAG into the brains of an all-knowing procurement expert. We work a 4-day week, with on-call every other Friday (business hours only).

Responsibilities

  • Design, train, and evaluate NLP models and embedding pipelines, using both traditional techniques (TF-IDF, topic modelling, rule-based/statistical NLP) and Sentence Transformers where they earn their complexity.
  • Own the vector space end-to-end: procuring training data, model/architecture choice, dimensionality trade-offs, indexing strategy in FAISS, drift monitoring, and rigorous evaluation using precision/recall, clustering quality, human-in-the-loop review rather than relying on vibes or leaderboard scores.
  • Right-size every model decision: default to CPU-friendly and lightweight approaches, and build the evidence-based case required before reaching for GPU infrastructure.
  • Build and maintain labelled datasets β€” including guidelines, annotation QA, and versioning β€” that are clean and well-documented enough for ETL Engineers to productionize without guesswork.
  • Prototype using agentic LLMs, while maintaining independent and conscious judgment on when an LLM-based approach is the right/wrong tool for the job.
  • Partner with ETL/DevOps Engineers to define handoff contracts: model artifacts, inference expectations, latency/cost budgets, and schema for outputs (via Pydantic) that plug cleanly into the existing pipeline.
  • Mentor junior team members on NLP fundamentals and evaluation rigor, and contribute to code reviews on the research side of the pipeline.

Qualities/skills we are looking for

  • Vector-Native: You have a deep, intuitive understanding of embedding spaces β€” how to construct them, index and query them efficiently (e.g. FAISS), and critically evaluate whether they're actually capturing what you think they are.
  • Lean by Instinct: You find elegance in solving a problem with the smallest model that reliably works, and you treat GPU spend and spin-up time as a cost to justify, not a resource to reach for by default.
  • ML-Skeptical, Not ML-Averse: You reach for machine learning when it's the right tool, and you're equally willing to argue against it β€” or against a heavier model β€” when a simpler statistical or rule-based method will do the job more reliably or cheaply.
  • Rigorous Evaluator: You don't trust a model until you've stress-tested it. You design evaluation frameworks before you design the model, and you're honest about failure modes.
  • AI Capable, but not Reliant: You use agentic LLMs and coding assistants to move faster, but your judgment β€” not the model's confidence β€” is the final word on quality.
  • Discipline: You value clean, reproducible research code, appreciate versioned datasets, and document your methodology so others can build on it without reverse-engineering your notebook.
  • Self Driven: You take ownership of open-ended research questions, and you're comfortable defining the problem, but stay clear-headed to back out of a rabbit hole to find a second opinion.
  • Collaborative Handoff Mindset: You understand that your output is someone else's input β€” you write for the ETL Engineer who has to productionize your work, not just for yourself.
  • Curious and Rigorous: You stay current with NLP research, but you evaluate new techniques on evidence, not hype, as well as considering maintenance before recommending they enter the pipeline.

Advice for applicants

If you believe you are up for the challenge of being a Forestreeter, we have the following advice before you apply:

  • We do not believe in numeric years of experience - we care about how you think and build.
  • Github portfolio > CV – we love to see your Github. It does not need to be perfect – progression is what we want to see.
  • We value expertise outside of what we require above – whether it’s a different industry you came from, or a second programming language – they all contain value you can bring to the table.

Company Benefits

  • 4 day work-week (on-call every other Friday during business hours).
  • Hybrid working environment with two days per week in a well-provisioned company office.
  • Pension plan and flexible benefits.
  • 20 days holiday per year.

Interview Process

  • Discovery call
  • First technical interview
  • Take-home task preceding second technical interview
  • Meet the team and leadership at our office in Farringdon / Chancery Lane
  • Values interview with Co-Founders
  • Offer

We are a small team, and we insist on Engineers hiring Engineers: the interviewers are the same people you will be working alongside. However, this may mean that we may not have lightning response time in the middle of a deployment!

NLP Research Engineer employer: Forestreet

At our company, we pride ourselves on fostering a collaborative and innovative work culture that empowers our employees to take ownership of their projects. With a unique 4-day work week and a hybrid working environment in the vibrant Farringdon area, we offer a perfect balance of professional growth and personal well-being. Our commitment to mentorship and continuous learning ensures that every team member has the opportunity to thrive and contribute meaningfully to our mission of democratizing research through cutting-edge NLP techniques.

F

Contact Details:

Forestreet Recruitment Team

StudySmarter Expert Advice🀫

We think this is how you could land NLP Research Engineer

✨Get Involved in Data Science Meetups

Tap into local data science meetups or workshops to connect with fellow enthusiasts and professionals. These events are goldmines for networking, and sometimes even lead directly to job openings at companies like Forestreet!

✨Show Off Your Projects

Start building a public portfolio showcasing your data science projects on platforms like GitHub or personal websites. Highlight unique analyses or models you've developed. This not only demonstrates your skills but also gets your name out there for roles like NLP Research Engineer at Forestreet.

✨Leverage Professional Networks

Join professional bodies related to data science, like the Data Science Society or similar organisations. Getting involved can lead to mentorship opportunities and insider knowledge about full-time positions at companies like Forestreet.

✨Apply Directly through Our Website

When you find a suitable opening like NLP Research Engineer at Forestreet, make sure to apply directly through our website. It gives you an edge and shows you're keen to join our team. Plus, who doesn’t love a direct application? It’s easier than navigating through job boards!

We think you need these skills to ace NLP Research Engineer

NLP Techniques
Machine Learning
Vector Comparison
Clustering Behaviour
Tokenization
Embedding Pipelines
TF-IDF

Some tips for your application 🫑

Show Off Your Projects:In the world of data science, your projects can speak volumes about your skills. Make sure to showcase a few key projects in your CV or portfolio, especially those that highlight your ability to work with data sets, build models, or use relevant tools like Python, R, or SQL. Don’t forget to include links to any GitHub repositories if applicable!

Quantify Your Achievements:Employers love numbers! When drafting your CV, highlight your achievements with quantifiable results. For instance, mention how your data analysis led to a certain percentage increase in efficiency or revenue at a previous job or project. These details can really make your application pop!

Craft a Tailored Cover Letter:For a full-time role at Forestreet, your cover letter should reflect your passion for data science and your excitement about the specific projects or values of the company. Dive into why you’re a good fit, how your skills align with their needs, and any unique perspectives you can bring to the team.

Stand Out with Relevant Courses and Certifications:Although experience talks, relevant courses or certifications can be your ticket to impressing hiring managers at Forestreet. Mention any standout courses you've completed that equipped you with essential skills, such as machine learning certifications or data visualisation courses. This shows your commitment to continuously developing your skills in the field!

How to prepare for a job interview at Forestreet

✨Brush Up on Your Statistics

For a data science role, we need to seriously sharpen our statistics skills. Get ready to tackle technical questions on probability distributions, hypothesis testing, and regression analysis. These are often the bread and butter of data science interviews, so don't just skim over them!

✨Showcase Your Projects

Prepare a killer portfolio showcasing your data science projects. We should include details about the datasets used, the tools and techniques applied, and the impact of your findings. If we can walk them through a particularly challenging project or a cool visualisation that had real-world implications, it’ll really make us stand out!

✨Get Comfortable with Python and R

Most data science positions require us to be proficient in programming languages like Python and R. We should practice common libraries like pandas, NumPy, and scikit-learn, and be ready for live coding exercises or algorithm questions. Showing off our coding chops can really impress the interviewers at Forestreet!

✨Prepare for Case Studies

Expect to encounter real-world case studies during the interview. We might be asked how we’d approach a data problem or analyse a dataset to extract insights. It's essential to think out loud and demonstrate our problem-solving process so that the interviewer can see our logical thinking in action.