Side by side
Up to 4 models, normalised onto the same cost basis and scored through the lens you pick.
| Cost of talking | ||
| Per minute | $0.012Estimate | $0.096Estimate |
| Per hour | $0.698 | $5.76 |
| 30-min interview | $0.349 | $2.88 |
| List price | $1.75 / $14 per 1MEstimate | $15 / $75 per 1MEstimate |
| Recruitment fit | ||
| Overall | 93/100 | 91/100 |
| Verdict | The strongest general model for hiring work — but only worth its price on the judgement-heavy steps, not on parsing. | The best choice for rubric-based scoring, where consistency across a batch matters more than peak score. |
| Best for | Structured interview scoring against a rubricComparing a shortlist against a job specDrafting candidate feedback that a human will edit | Structured scoringBias-audit passes over model outputCandidate feedback drafting |
| Watch out for |
|
|
| Capabilities | ||
| Modalities | text, audio | text |
| Context window | 400K | 200K |
| Variants | Max effort, High, Medium, Low, Mini | Opus 5, Sonnet 5, Haiku 4.5 |
| On HumanLike | Not yet | Not yet |
| HumanLike evals | ||
| CV field extraction (F1) | 0.94Estimate | — |
| Time to first token (medium effort) | 640 msEstimate | — |
| Published benchmarks | ||
| MMLU-Pro | TBDEstimate | — |
| SWE-bench Verified | TBDEstimate | — |
Cost per hour assumes 150 wpm and a 40% AI speaking share. Hover a figure for its full derivation.