At a Glance
- Tasks: Lead groundbreaking research on AI peer review evaluation and design innovative methodologies.
- Company: Join OxSci, a pioneering credit rating agency for science, blending AI with expert insights.
- Benefits: Competitive pay, equity options, flexible hours, and a generous LLM token budget.
- Other info: Founding seat with equity, access to expert networks, and opportunities for rapid personal growth.
- Why this job: Shape the future of AI in peer review and make a real impact in scientific evaluation.
- Qualifications: PhD or near completion in CS, ML, NLP, or related fields; strong Python skills required.
The predicted salary is between 60000 - 80000 £ per year.
OxSci is building a credit rating agency for science: a certification layer that combines AI with expert peer review, so researchers, institutions, and AI developers can assess research quality quickly and at scale. The open scientific question at the heart of this company is yours to own: where does the frontier of AI peer review actually stand? What can AI reviewers catch that human experts miss, what do they still get wrong, and how do you prove it rigorously? The standards we set now, including what "quality" even means and how we demonstrate our AI reviewers are actually good, will define both the company and, we believe, the field.
The tech side is led by OxSci's cofounder, a senior tech lead from one of the world's largest technology companies, with deep experience building and operating systems at global scale. You'd work with both founders day to day, shaping the evaluation methodology and research culture from the ground up.
What you’ll do:
- Own the research agenda on AI-reviewer evaluation.
- Track the frontier (AI-scientist, automated-review, LLM-as-a-judge, and scholarly-NLP literature), position our system against it, and decide what we measure next and why.
- Design meta-evaluations that expose weaknesses, not just measure agreement.
- Build fine-grained, criticism-level evaluations of AI review agents (correctness, factual grounding, significance, sufficiency of evidence, hallucination rate, and venue/journal matching) that reveal where and why they fail, going beyond verdict-matching.
- Run expert-annotation studies at scale.
- Design the protocols, rubrics, inter-annotator agreement, and statistics needed to compare AI and human reviewers credibly, including head-to-head evaluations against other AI review systems, and defend the numbers to a skeptical scientific audience.
- Build a living taxonomy of AI-reviewer failure modes such as subfield blind spots, long-context degradation, over-anchoring, and spurious criticism, and turn each into a regression benchmark that guards against backsliding as models and prompts change.
- Calibrate the combined rating.
- Define quality-scoring rubrics for human review reports and calibrate how expert and AI judgment fuse into a single, defensible rating: the core of what universities and publishers buy from us.
- Translate benchmark findings into concrete improvements to our review agents (retrieval, context engineering, orchestration, model choice) and prove the gains with the same rigor you used to find the gaps.
What we’re looking for:
- A PhD (or near completion) in CS, ML, NLP, or a related field, or an equivalent research track record.
- You’re likely already working on LLM evaluation, LLM-as-a-judge, AI for science, automated peer review, or a nearby frontier.
- A fascination with the boundary between AI and human reviewers.
- A track record of rigorous evaluation of ML/LLM systems: benchmark or eval-framework design, evaluator/judge models, expert-annotation study design, hallucination and factuality measurement, uncertainty quantification, or RAG evaluation.
- Fluency in evaluation methodology and statistics: sampling, inter-annotator agreement, significance testing, and the discipline to distinguish a real effect from a lucky prompt.
- Strong Python and hands-on habits.
- Bonus: publications in NLP/ML evaluation or automated peer review; open-source benchmarks or evaluator models the community actually uses; experience with scholarly content at scale.
What we offer:
- Founding seat with meaningful equity and a direct line to the founders.
- Ownership of a genuinely open research question, with encouragement to publish and present the work.
- A standing expert-reviewer network as your annotation infrastructure.
- A proprietary, growing dataset of paired human and AI reviews of real submissions.
- The rare chance to define the standard by which AI reviewers themselves are judged.
- Genuinely competitive pay; equity discussed openly.
- Generous LLM token budget for your daily work.
- Flexible working hours; fast personal growth with broad ownership from day one.
AI Research Scientist in London employer: OxSci
OxSci is an exceptional employer that fosters a dynamic and innovative work culture in the heart of London. With a focus on cutting-edge research and development, employees enjoy flexible hours, equity opportunities, and the chance to contribute to groundbreaking advancements in AI peer review. The company prioritises employee growth through collaborative projects and encourages exploration of open research questions, making it an ideal place for those seeking meaningful and rewarding careers.
StudySmarter Expert Advice🤫
We think this is how you could land AI Research Scientist in London
✨Get Involved in Data Science Meetups
Tap into local data science meetups or workshops to connect with fellow enthusiasts and professionals. These events are goldmines for networking, and sometimes even lead directly to job openings at companies like OxSci!
✨Show Off Your Projects
Start building a public portfolio showcasing your data science projects on platforms like GitHub or personal websites. Highlight unique analyses or models you've developed. This not only demonstrates your skills but also gets your name out there for roles like AI Research Scientist at OxSci.
✨Leverage Professional Networks
Join professional bodies related to data science, like the Data Science Society or similar organisations. Getting involved can lead to mentorship opportunities and insider knowledge about full-time positions at companies like OxSci.
✨Apply Directly through Our Website
When you find a suitable opening like AI Research Scientist at OxSci, make sure to apply directly through our website. It gives you an edge and shows you're keen to join our team. Plus, who doesn’t love a direct application? It’s easier than navigating through job boards!
We think you need these skills to ace AI Research Scientist in London
Some tips for your application 🫡
Show Off Your Projects:In the world of data science, your projects can speak volumes about your skills. Make sure to showcase a few key projects in your CV or portfolio, especially those that highlight your ability to work with data sets, build models, or use relevant tools like Python, R, or SQL. Don’t forget to include links to any GitHub repositories if applicable!
Quantify Your Achievements:Employers love numbers! When drafting your CV, highlight your achievements with quantifiable results. For instance, mention how your data analysis led to a certain percentage increase in efficiency or revenue at a previous job or project. These details can really make your application pop!
Craft a Tailored Cover Letter:For a full-time role at OxSci, your cover letter should reflect your passion for data science and your excitement about the specific projects or values of the company. Dive into why you’re a good fit, how your skills align with their needs, and any unique perspectives you can bring to the team.
Stand Out with Relevant Courses and Certifications:Although experience talks, relevant courses or certifications can be your ticket to impressing hiring managers at OxSci. Mention any standout courses you've completed that equipped you with essential skills, such as machine learning certifications or data visualisation courses. This shows your commitment to continuously developing your skills in the field!
How to prepare for a job interview at OxSci
✨Brush Up on Your Statistics
For a data science role, we need to seriously sharpen our statistics skills. Get ready to tackle technical questions on probability distributions, hypothesis testing, and regression analysis. These are often the bread and butter of data science interviews, so don't just skim over them!
✨Showcase Your Projects
Prepare a killer portfolio showcasing your data science projects. We should include details about the datasets used, the tools and techniques applied, and the impact of your findings. If we can walk them through a particularly challenging project or a cool visualisation that had real-world implications, it’ll really make us stand out!
✨Get Comfortable with Python and R
Most data science positions require us to be proficient in programming languages like Python and R. We should practice common libraries like pandas, NumPy, and scikit-learn, and be ready for live coding exercises or algorithm questions. Showing off our coding chops can really impress the interviewers at OxSci!
✨Prepare for Case Studies
Expect to encounter real-world case studies during the interview. We might be asked how we’d approach a data problem or analyse a dataset to extract insights. It's essential to think out loud and demonstrate our problem-solving process so that the interviewer can see our logical thinking in action.