AI Harness Engineer: LLM Judge & Benchmarking in London

AI Harness Engineer: LLM Judge & Benchmarking in London

London Full-Time No working from home possible
V

Valarian Technologies Limited is seeking an experienced AI Harness Engineer to own the experimental design, evaluation methodologies, and benchmark infrastructure for AI models and autonomous workloads. You will bridge data science, statistical validation, and production agent scaffolding, collaborating across teams in a hybrid London-based setup.

The role emphasizes rigorous hypothesis testing, robust metric design, and hands-on Python tooling to advance evaluation pipelines and guardrails for

#J-18808-Ljbffr

AI Harness Engineer: LLM Judge & Benchmarking in London employer: Valarian

Valarian Technologies is an exceptional employer, offering a dynamic work culture that prioritises innovation and security in the fast-paced tech landscape of London. Employees benefit from comprehensive growth opportunities, including hands-on experience with cutting-edge technologies and a commitment to professional development, all while contributing to meaningful projects that enhance secure software delivery in regulated environments.

V

Contact Details:

Valarian Recruitment Team