speech model

The most expressive synthetic voice available, and the best voice cloning.
ElevenLabs remains the quality benchmark for synthetic speech, particularly for emotional range and for cloning a specific person’s voice from a short sample. In hiring, the cloning capability is the interesting and dangerous part: a recruiter’s real voice can front thousands of outbound calls, which is powerful and requires explicit consent and disclosure.
Variants
Same family, different quality/cost rung. Picking the wrong one is the most common way teams overpay.
eleven-v3Most expressive. Best for produced audio.
eleven-flash-v2-5~75ms model latency for live conversation.
eleven-turbo-v2-5Balance of expressiveness and speed.
Fit
The same model is a different proposition depending on what you point it at.
Recruitment fit
The best-sounding option for candidate-facing voice — with the heaviest disclosure obligations attached.
Best for
Strengths
Watch out for
The numbers
HumanLike evals
Run on our own harness.