Side by side

Synthesia vs Tavus CVI

Up to 4 models, normalised onto the same cost basis and scored through the lens you pick.

Cost of talking
Per minuteNot conversational$1.00Estimate
Per hour$60.00
30-min interview$30.00
List pricePlan-basedEstimate$1/minEstimate
Recruitment fit
Overall55/10090/100
VerdictExcellent for onboarding and policy video; structurally unable to do live screening.The most complete option for conversational screening — the perception layer is worth the premium when a real person is on the other end.
Best for
Onboarding modulesCompliance trainingCareers-page brand film
First-round conversational screeningEmployer-brand videos personalised per candidateProctored assessments where attention signals matter
Watch out for
  • No real-time path at all
  • Credit pricing makes per-candidate cost awkward to model
  • At roughly $1/min it is 200× the cost of Homo 1 — model your per-interview cost before committing
  • Recording a candidate’s webcam raises consent and data-retention obligations; several jurisdictions require explicit notice for automated video assessment
  • A trained replica of a real employee needs that employee’s written consent
What it animates
Coverage
FaceFullFull
Lip-syncFullFull
Upper torsoFullPartial
HandsPartialNone
Full bodyFullNone
Capabilities
Modalitiesavatar, videoavatar, video, speech
Context window
VariantsStudioPhoenix-3 (render), Raven-0 (perception), Sparrow-0 (turn-taking), Hummingbird-0 (lip-sync)
On HumanLikeNot yetAI Personas, Interview
HumanLike evals
Wired into HumanLike avatarsYes — face_id providerOur eval
Published benchmarks
Lip-sync accuracy94%Estimate90%Estimate
Time to outputminutesEstimate1000 msEstimate

Cost per hour assumes 150 wpm and a 40% AI speaking share. Hover a figure for its full derivation.