What the work involves
You record voice samples against supplied scripts — conversational turns, narration, instructional copy, and customer-service style exchanges — following a detailed spec covering microphone distance, room treatment, sample rate, file naming, and take structure. Sessions typically ask for multiple takes per line with controlled variation in emotion, emphasis, and pacing, so the model learns range rather than a single flat read. Consistency matters more than performance flair here: your timbre, accent placement, and energy level need to match across sessions that may be weeks apart, since inconsistent source audio degrades the cloned voice.
The central condition is stated plainly in the listing: your voice may be cloned for the client's internal CX AI agent. The platform states recordings will not be sold, licensed, or reused beyond that use case. Read the consent and usage terms carefully before applying — this is not a role to take with reservations about synthetic voice reproduction.
What the platform screens for
Mercor's screen is an AI-led interview plus a submitted audio sample. Expect it to test:
- Accent neutrality — internationally intelligible Latin American Spanish without strong regional markers (rioplatense, caribeño, chilango, etc. tend to be flagged)
- Verifiable credits — dubbing, narration, IVR, audiobook, broadcast, or prior AI voice dataset work you can name and point to
- Technical setup — your specific microphone, interface, treatment, and noise floor, described concretely rather than as "professional studio"
- Direction-following — whether you can hit a described read (calm corporate, warm conversational, energetic) on request and reproduce it later
- Informed consent — clear, unhesitant understanding of what voice cloning means for you
Logistics and pay
Remote and largely asynchronous, with recording blocks you schedule yourself against submission deadlines. The listing asks for reliable 5–10 hours per week across the project duration; you must be currently based in Latin America. Observed pay for this listing is $50/hr — a rate, not a guarantee, and subject to the platform's confirmation of hours and deliverable acceptance. Familiarity with Audacity, Audition, or Reaper is preferred, mainly for basic cleanup and format compliance rather than heavy post-production.