Machine Learning Evaluation Specialist

Machine Learning Evaluation Specialist

Full-Time On-site
A

About The Role

The quality of AI depends entirely on the quality of the problems used to test it. We're looking for researchers and domain experts with deep machine learning knowledge to design the evaluation challenges that define β€” and push β€” the limits of today's most capable AI systems.

About The Role

The quality of AI depends entirely on the quality of the problems used to test it. We're looking for researchers and domain experts with deep machine learning knowledge to design the evaluation challenges that define β€” and push β€” the limits of today's most capable AI systems.

  • Organization: Alignerr
  • Type: Hourly Contract
  • Location: Fully Remote
  • Commitment: 10-40 hours/week

What You'll Do

  • Design complex, original machine learning problems rooted in your specific domain of expertise
  • Craft evaluation tasks that require advanced domain knowledge well beyond standard ML pipelines
  • Draw from your own research experience to create problems that genuinely challenge state-of-the-art AI
  • Define rigorous problem statements, evaluation criteria, and gold-standard solutions
  • Assess AI-generated ML solutions for correctness, creativity, and methodological soundness
  • Document problem difficulty levels, required domain knowledge, and expected AI failure modes
  • Collaborate asynchronously with a global team of researchers and engineers

Who You Are

  • Graduate-level expertise (MS or PhD preferred) in a scientific or technical domain that intersects with machine learning
  • Strong working knowledge of ML methods - model selection, feature engineering, evaluation metrics, and pipeline design
  • Deep familiarity with active research problems in your field
  • Able to identify precisely where general ML knowledge falls short and specialized domain insight becomes critical
  • Experience publishing or conducting original research is highly valued
  • Excellent written communication - you can articulate complex problems clearly and precisely
  • Self-motivated and comfortable working independently on intellectually demanding tasks

Example Domains (Not Exhaustive)

  • Computational biology, genomics, or bioinformatics
  • Climate science and environmental modeling
  • Medical imaging and healthcare ML
  • Materials science and computational chemistry
  • Astrophysics and signal processing
  • Natural language processing for low-resource or specialized corpora
  • Robotics, control theory, or reinforcement learning in complex environments
  • Financial modeling and quantitative analysis

Why Join Us

  • Work at the true frontier of AI evaluation and safety research
  • Collaborate with top research labs pushing the boundaries of what AI can do
  • Finally put your specialized domain expertise to use in a high-impact, meaningful way
  • Full autonomy over your schedule - work when and how you do your best thinking
  • Flexible, fully remote contract with potential for ongoing work and deeper research involvement
  • Build your profile as a recognized contributor to cutting-edge AI development
  • Join a global community of researchers and engineers who take this work seriously

#J-18808-Ljbffr

Machine Learning Evaluation Specialist employer: Alignerr Corp.

Join a leading global AI research firm that values innovation and collaboration, offering a dynamic remote work environment for an Applied Physicist. With competitive pay, flexible scheduling, and a strong emphasis on employee growth, this role provides the opportunity to engage in cutting-edge projects while contributing to the advancement of AI technology. Experience a supportive culture that encourages creativity and professional development, making it an ideal place for those passionate about physics and AI.

A

Contact Details:

Alignerr Corp. Recruitment Team