Cohere is seeking an experienced role focused on creating next-generation evaluation methods and infrastructure to measure LLM progress. You will work across teams to translate model feedback into trustworthy evaluations and build scalable tools to analyze performance.
The role emphasizes rigorous measurement of AI capabilities, prototype-driven work, and strong software engineering to support scalable evaluation pipelines. Remote-friendly with global offices.
#J-18808-Ljbffr
Senior LLM Evaluation Scientist employer: Cohere
Cohere is an exceptional employer that fosters a dynamic and inclusive work culture, where innovation thrives and employees are empowered to make a real impact in the AI landscape. With generous benefits such as a weekly lunch stipend, comprehensive health coverage, and a robust education stipend, we prioritise employee well-being and growth. Our London office offers a collaborative environment with opportunities for meaningful engagement with enterprise clients, making it an ideal place for those looking to advance their careers in cutting-edge technology.