What the work involves

You receive AI-generated work products of the kind a brand or creative lead would normally produce or approve: pitch decks, brand guidelines, messaging frameworks, campaign briefs, launch one-pagers, and occasionally spreadsheets like media plans or naming matrices. Your job is to grade them against a rubric and explain your scores in writing. A typical task means reading the prompt and its constraints, assessing whether the output actually answers the brief, then flagging what a client would reject — invented statistics, a positioning statement that contradicts the stated audience, tone that drifts from the brand voice, hierarchy that collapses on slide four, filler copy dressed up as strategy.

Strong evaluators here do two things at once. They catch craft-level problems (typographic inconsistency, misused grid, captions that repeat the headline) and they catch judgment-level problems (a differentiation claim no competitor scan would support, a CTA that doesn't match the funnel stage). Written feedback needs to be specific and actionable: which element, why it fails, what the correct standard is. Vague notes like "feels off-brand" don't clear review.

What the platform screens for

  • Verifiable depth. Expect follow-ups on specific brands, campaigns, or collateral you've owned, and on how decisions were made — not just what shipped.
  • Rubric discipline. Whether you can separate personal aesthetic preference from defensible quality criteria, and score consistently across similar artifacts.
  • Slide fluency. Real working proficiency in Google Slides and PowerPoint, since much of the reviewed output arrives in those formats.
  • Writing. Feedback is the deliverable; clarity and structure matter as much as the assessment itself.

Logistics

Fully remote and asynchronous. Work is drawn from a queue with no fixed shifts, so you set your own hours, though sustained availability of roughly 10–20 hours per week tends to keep task flow steady. Engagements are hourly and project-scoped; volume fluctuates with the client's pipeline. Pay in the $80–120/hr band reflects rates observed on similar Mercor evaluation projects and is not guaranteed for any individual assignment.