AI Harness Engineer β€” LLM Evaluation & Benchmark Architect in London

AI Harness Engineer β€” LLM Evaluation & Benchmark Architect in London

London Full-Time No working from home possible
Valarian Technologies

Valarian Technologies is seeking an AI Harness Engineer to own experimental design and evaluation infrastructure for AI models and autonomous workloads. You will bridge data science, statistics, and production scaffolds, designing robust datasets, calibrated LLM-judge pipelines, and scalable Python harnesses.

You will drive metric integrity, error analysis, and governance of tool orchestration, ensuring safe, reproducible evaluations while collaborating with London-based teams in a hybrid setup.

#J-18808-Ljbffr

AI Harness Engineer β€” LLM Evaluation & Benchmark Architect in London employer: Valarian Technologies

Valarian Technologies is an exceptional employer that fosters a dynamic work culture, encouraging innovation and collaboration among its team members. With a strong focus on employee growth, the company offers numerous opportunities for professional development in the rapidly evolving field of information security. Located in London, the hybrid work model provides flexibility while ensuring that employees are part of a global network dedicated to maintaining the highest standards of security and compliance.

Valarian Technologies

Contact Details:

Valarian Technologies Recruitment Team