Appnovation in London is seeking a QA / AI Evaluation Engineer to join a forward-leaning team that validates platform improvements through large-scale evaluations. You will measure factual grounding and accuracy lift while building reusable metrics to show progress over time.
The role requires 4+ years in QA for data/ML systems, strong Python skills, and experience with LLM evaluation frameworks. Collaboration with engineers to reproduce fixes is essential.
#J-18808-Ljbffr