AI Security Institute in London is seeking a Research Engineer to join the Human Influence team. The role blends research and engineering, focusing on scalable post-training and fine-tuning of LLMs, with reinforcement learning and validation on large compute clusters.
You will design RL environments to curb deceptive interactions, apply interpretability methods to reveal model risks, and build the systems that deliver repeatable evaluations and benchmarks.
#J-18808-Ljbffr
Research Engineer, AI Human Influence & Safety employer: AISI
The AI Security Institute in London is an exceptional employer, offering a dynamic work environment where cutting-edge research meets real-world impact. With a strong focus on employee growth and collaboration, you will have the opportunity to work alongside leading experts in virology and AI, contributing to vital biosecurity initiatives. Our inclusive culture fosters innovation and encourages professional development, making it a rewarding place for those passionate about advancing science and policy.