Side by side
Up to 4 models, normalised onto the same cost basis and scored through the lens you pick.
| Cost of talking | ||
| Per minute | $1.00Estimate | $0.500Estimate |
| Per hour | $60.00 | $30.00 |
| 30-min interview | $30.00 | $15.00 |
| List price | $1/minEstimate | $0.5/minEstimate |
| Recruitment fit | ||
| Overall | 90/100 | 72/100 |
| Verdict | The most complete option for conversational screening — the perception layer is worth the premium when a real person is on the other end. | Good for employer-brand content, weaker as a live interviewer than Tavus. |
| Best for | First-round conversational screeningEmployer-brand videos personalised per candidateProctored assessments where attention signals matter | Employer-brand videoMultilingual job adsOnboarding content |
| Watch out for |
|
|
| What it animates | ||
| Coverage | ||
| Face | Full | Full |
| Lip-sync | Full | Full |
| Upper torso | Partial | Full |
| Hands | None | Partial |
| Full body | None | Partial |
| Capabilities | ||
| Modalities | avatar, video, speech | avatar, video |
| Context window | — | — |
| Variants | Phoenix-3 (render), Raven-0 (perception), Sparrow-0 (turn-taking), Hummingbird-0 (lip-sync) | Interactive, Studio (rendered) |
| On HumanLike | AI Personas, Interview | Not yet |
| HumanLike evals | ||
| Wired into HumanLike avatars | Yes — face_id providerOur eval | — |
| Published benchmarks | ||
| Lip-sync accuracy | 90%Estimate | 92%Estimate |
| Utterance-to-utterance latency | 1000 msEstimate | 1200 msEstimate |
Cost per hour assumes 150 wpm and a 40% AI speaking share. Hover a figure for its full derivation.