What the work actually is
This is a talent network application, not a single assignment. Once verified, you become eligible for contracts that labs open with Mercor — most of them some mix of authoring hard mathematical problems with verifiable answers, writing reference solutions, grading model-generated proofs line by line, and annotating exactly where a chain of reasoning goes wrong. Some projects skew toward adversarial prompt writing (find problems the model confidently botches); others are pure evaluation, where you rank two model outputs and justify the ranking in prose a researcher can act on. Statistics-heavy briefs show up regularly: checking whether a model's inference, estimator choice, or simulation setup is actually defensible rather than merely plausible.
What the screen looks for
The AI interview probes whether your mathematical background is load-bearing. Expect follow-ups that push past your first answer — naming a subfield is not enough; you'll be asked what you proved, what technique you used, and why. Graders care about your ability to distinguish a proof that is wrong from one that is merely unfamiliar or terse, since false-positive error flags are expensive to labs. Clear written English matters as much as the mathematics: your rationale is the deliverable, and it has to be legible to someone who is not a specialist in your area.
Logistics
- Fully remote and largely asynchronous; you work against deadlines, not shift blocks.
- Typical commitments run 15–30 hours per week, though scope and duration vary by project.
- Pay in the $60–80/hr range has been observed for mathematics work on this platform; actual rates are set per project and are not guaranteed.
- Matching is rolling — approval makes you eligible, and an invitation may follow weeks later. Expect a second, project-specific interview before you start.