Other open roles
AI Evaluation Analyst
Work Expert (WE) is an independent publishing and referral website. We are not a recruiter, hiring manager, agent or employer, and we are not affiliated with or endorsed by micro1. Applying takes you to the platform's own website, where we may be recorded as the referring source. We may receive a referral fee at no additional cost to you. Read our full affiliate disclosure →
About the Role
micro1 is engaging AI Evaluation Analysts to contribute to a customer’s project focused on advancing frontier language model capabilities. In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input. No prior experience in AI is required — your domain knowledge is what matters. You will play a crucial role in producing evaluation and training data that directly influence how powerful AI models understand and interact. This i
What You'll Do
- Author detailed, task-based multi-turn conversations and rubrics aligned with project specifications.
- Test and refine conversation drafts against frontier large language models, iterating to meet quality and difficulty requirements.
- Deliver comprehensive evaluation assets including transcripts, target behaviors, binary rubrics, and supporting evidence.
- Ensure strict fidelity to evolving project specs while maintaining high throughput and attention to detail.
- Validate and calibrate outputs with team leads and quality control as guidelines change.
- Work independently and consistently, meeting expected output rates for deliverable completion.
You're a Good Fit If You
- Required skills: Working knowledge of frontier LLM behavior, Data Annotation, Written English clarity and structure, Spec fidelity at volume, Self-direction to a detailed spec
- Native-level written English with exceptional clarity, structure, and attention to detail.
- Prior experience in data annotation, RLHF, SFT, evaluation, or prompt engineering for AI systems.
- Working knowledge of frontier LLM behaviors and common model failure patterns.
- Demonstrated ability to interpret and apply highly detailed specifications without supervision.
- Strong critical thinking and analytical skills in writing-heavy or analysis-heavy domains.
- Experience authoring evaluation items, rubrics, or conducting deep analysis of technology outputs.
- Background in research, editorial, technical writing, or quality assurance is a plus.
Role Highlights
Pay & Payout
Engaged directly through micro1. Apply via the link below - micro1 handles onboarding and payment.
Apply on micro1Opens micro1 in a new tab. You pay nothing; we may earn a referral fee.
Similar roles
More research work you may qualify for.
Enterprise Client Partner, Frontier AI
The Role We work with leading AI labs and advanced enterprise teams building frontier models. Our focus is on evaluation systems, reinforcement learning (RL) environments, and high
Human Data Manager
In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world inpu
Community Manager
Join our team as a dynamic Community Manager and play a pivotal role in creating a positive, collaborative, and engaged online environment. You will be at the forefront of customer