AI Safety Benchmark Engineer

AI Safety Benchmark Engineer

Full-Time On-site
M

White Circle in London is seeking a research engineer to build and maintain an internal benchmark suite spanning single/multi-turn content and agentic guardrails. You will collaborate with the core safety team to study agent behaviours and push evaluation tooling into production.

Applicants should have hands-on Python production experience, strong benchmark design skills, and a track record of shipping reliable code for ML systems. Hybrid work with offices in London and Paris is offered.

#J-18808-Ljbffr

AI Safety Benchmark Engineer employer: Moonfire

Mindstone is an exceptional employer for those looking to thrive in the AI transformation space. With a strong focus on employee growth, competitive salaries, and meaningful equity opportunities, we foster a collaborative and innovative work culture in our London office. Our commitment to diversity and inclusion ensures that every team member's voice is valued, making it a rewarding environment for all.

M

Contact Details:

Moonfire Recruitment Team