What the work involves

The customer is training a model to create, read, and edit Office Open XML files — not just describe spreadsheets, but produce workbooks with working formulas, coherent structure, and correct formatting. Your job is to supply the hard cases. You design scenarios drawn from your own professional history (a three-statement model refresh, a headcount plan with scenario toggles, a client-ready deck built from raw exports), run them as multi-turn conversations with the model, then open the resulting files and grade them.

Most tasks are pairwise: two AI-generated solutions, one decision, and a written rationale. The rationale is the deliverable that matters. "Option B is better" is worthless; "Option B hardcodes the growth rate in row 14, breaking the sensitivity table, while Option A wires it to the assumptions tab" is the kind of note that changes model behavior. Expect to check formula logic, circular references, number formatting, named ranges, chart source data, pivot structure, and whether the output would survive a review from a VP.

What the platform screens for

  • Verifiable depth. Where you worked, what you built, and what broke. Micro1's screen is AI-led and follows up on specifics — naming INDEX/MATCH is not the same as explaining when you'd choose it over XLOOKUP in a shared model.
  • Scenario design ability. Can you invent a task that is genuinely hard for a competent model, rather than a toy exercise?
  • Written precision. Feedback is read by people who don't know your domain. Ambiguity is a disqualifier.
  • Calibrated judgment. Distinguishing cosmetic differences from substantive ones, and saying so when two outputs are equally acceptable.

Logistics

Fully remote, contractor engagement, asynchronous task queues with occasional live calls or verbal walkthroughs. Observed pay for this listing runs $20–60/hr, set by demonstrated depth and the specialization of the queue — not guaranteed, and rates on document-generation projects vary with task complexity. Volume is project-driven and can pause between phases; treat it as supplemental rather than a full schedule. No AI background required, though prior prompt-writing experience shortens the ramp.