What the work involves

You receive AI-produced artifacts — a curriculum unit plan, a parent-communication deck, an assessment blueprint, a school budget or attendance spreadsheet, a professional-development slide deck — and evaluate them against a rubric supplied by the platform. Most tasks ask for a score per dimension plus written justification: what is factually wrong, what is pedagogically unsound, what would not survive review by a principal, department head, or district office. Feedback needs to be specific enough that an engineer who has never taught can act on it, which usually means citing the exact slide, cell, or paragraph and stating the correct version.

Expect a mix of formats. Slides carry disproportionate weight in this queue, so comfort with Google Slides and PowerPoint conventions — layout, hierarchy, alt text, chart choice, accessibility of text density — is not optional. Spreadsheet tasks may involve checking formula logic, grading scales, gradebook weighting, or enrollment projections. Document tasks lean toward standards alignment, differentiation, IEP/504 language, and policy accuracy.

What the platform screens for

  • Verifiable experience. Five or more years in K–12 or higher-ed teaching, instructional coaching, curriculum design, school administration, or ed policy. Screens probe for institution names, grade bands, subject areas, and what you actually owned.
  • Depth under follow-up. A general answer about "differentiation" invites a harder second question. Be ready to name standards frameworks, assessment types, and specific decisions you made.
  • Evaluation judgment. Can you separate a stylistic preference from a defensible error, apply a rubric you disagree with, and stay calibrated across dozens of similar samples?
  • Writing. Feedback is the deliverable. Clarity and structure matter more than volume.

Logistics

Fully remote, asynchronous, hourly, contractor status. There is no fixed schedule; work appears in batches and you claim what you can complete. Most contributors report 5–20 hours per week, with volume that fluctuates by project. Rates in the $80–120/hr band are what contributors have observed on similar Mercor evaluation projects — the actual offer depends on credentials, assessed rubric accuracy, and project budget, and is not guaranteed. An advanced degree helps but does not substitute for practitioner experience.