Other open roles

View all opportunities

AI Evaluation Analyst

$30
per hour
Contractor Remote Last verified 21 Jul 2026
Affiliate disclosure

Work Expert (WE) is an independent publishing and referral website. We are not a recruiter, hiring manager, agent or employer, and we are not affiliated with or endorsed by micro1. Applying takes you to the platform's own website, where we may be recorded as the referring source. We may receive a referral fee at no additional cost to you. Read our full affiliate disclosure →

About the Role

micro1 is engaging AI Evaluation Analysts to contribute to a customer’s project focused on advancing frontier language model capabilities. In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input. No prior experience in AI is required — your domain knowledge is what matters. You will play a crucial role in producing evaluation and training data that directly influence how powerful AI models understand and interact. This i

What You'll Do

  • Author detailed, task-based multi-turn conversations and rubrics aligned with project specifications.
  • Test and refine conversation drafts against frontier large language models, iterating to meet quality and difficulty requirements.
  • Deliver comprehensive evaluation assets including transcripts, target behaviors, binary rubrics, and supporting evidence.
  • Ensure strict fidelity to evolving project specs while maintaining high throughput and attention to detail.
  • Validate and calibrate outputs with team leads and quality control as guidelines change.
  • Work independently and consistently, meeting expected output rates for deliverable completion.

You're a Good Fit If You

  • Required skills: Working knowledge of frontier LLM behavior, Data Annotation, Written English clarity and structure, Spec fidelity at volume, Self-direction to a detailed spec
  • Native-level written English with exceptional clarity, structure, and attention to detail.
  • Prior experience in data annotation, RLHF, SFT, evaluation, or prompt engineering for AI systems.
  • Working knowledge of frontier LLM behaviors and common model failure patterns.
  • Demonstrated ability to interpret and apply highly detailed specifications without supervision.
  • Strong critical thinking and analytical skills in writing-heavy or analysis-heavy domains.
  • Experience authoring evaluation items, rubrics, or conducting deep analysis of technology outputs.
  • Background in research, editorial, technical writing, or quality assurance is a plus.

Role Highlights

Pay
$20-$30/h
Contract
Contractor
Location
Remote
Category
Research
Posted
12 days ago
Job ID
WE-7146

Pay & Payout

Engaged directly through micro1. Apply via the link below - micro1 handles onboarding and payment.

Apply on micro1

Opens micro1 in a new tab. You pay nothing; we may earn a referral fee.