We are looking for Evaluation Specialist/Recent Grad candidates for a project delivered through micro1.
What you'll do
- Develop original, challenging question-and-answer pairs that test the limits of advanced AI models across diverse subjects.
- Conduct in-depth research, applying rigorous source triangulation to ensure each answer is accurate, comprehensive, and well-documented.
- Create multi-step questions that require synthesis and analytical reasoning — not just single-source lookups.
- Test questions against AI models, analyze outcomes, and iterate to increase or adjust difficulty as needed.
- Document research process, providing clear citations and logical explanations for each answer.
- Refine content based on reviewer input and maintain alignment with project guidelines and quality standards.
What you need
- Proven ability to conduct thorough, independent research and critically evaluate information from multiple sources.
- Exceptional attention to detail and written precision in English (fluency required, but it does not need to be your first language).
- Talent for crafting nuanced and well-structured questions that probe deep understanding.
- Strong self-direction, reliability, and accountability in remote, independent project work.
- Interest in cutting-edge AI and a passion for testing the boundaries of current model capabilities.
Nice to have
- Prior experience in AI training, evaluation, or content creation is a plus, but not required.
Expertise
Also listed: attention to detail, analytical thinking.
Who you work with
Project and contracting process: micro1. Applications continue on the provider's website.

