What the work involves
You receive AI-generated business artifacts and grade them the way a partner or VP would review a junior's draft. In practice that means market-entry memos, competitive landscapes, org-design recommendations, pricing analyses, board decks, and the spreadsheets behind them. You check whether the logic actually holds: does the market sizing reconcile with its own assumptions, is the framework applied or just name-dropped, does the recommendation follow from the evidence presented, are the slide headlines assertions or empty labels. You then write feedback that a model trainer can act on — specific, cited to the artifact, and ranked by severity rather than a general impression.
A meaningful share of tasks are presentation-heavy, which is why Slides and PowerPoint fluency is a hard requirement rather than a nice-to-have. You will be asked to judge layout, chart selection, data-label accuracy, message-title quality, and internal consistency across a 15-slide deck — and to distinguish cosmetic issues from substantive ones.
What the platform screens for
- Verifiable depth. Mercor's AI-led interview probes specific engagements, decisions, and numbers. Vague seniority claims tend to fail on follow-up.
- Evaluation judgment. Can you separate a confidently written but analytically hollow output from one that is rough but directionally sound? Can you apply someone else's rubric instead of your own taste?
- Written clarity. Feedback quality matters as much as grading accuracy; terse, structured, and unambiguous beats long.
- Tooling reality. Expect questions about how you actually build and critique decks and models.
Logistics
Fully remote and asynchronous, with work claimed from a task queue rather than assigned on a schedule. Most contributors treat it as part-time — 10 to 20 hours a week is typical, and volume fluctuates with project cycles, so treat it as supplemental income rather than a stable book of hours. Pay is hourly and the $80–120/hr band reflects observed rates for this category; actual offers depend on screening outcome, project, and seniority, and are not guaranteed.