Side by side
Up to 4 models, normalised onto the same cost basis and scored through the lens you pick.
| Cost of talking | ||
| Per minute | $0.0031 | $0.012Estimate |
| Per hour | $0.187 | $0.698 |
| 30-min interview | $0.094 | $0.349 |
| List price | $0.5 / $1.5 per 1MPublished | $1.75 / $14 per 1MEstimate |
| Recruitment fit | ||
| Overall | 22/100 | 93/100 |
| Verdict | Do not screen candidates with this. Its judgement is not good enough to put in front of a hiring decision. | The strongest general model for hiring work — but only worth its price on the judgement-heavy steps, not on parsing. |
| Best for | Nothing candidate-facing | Structured interview scoring against a rubricComparing a shortlist against a job specDrafting candidate feedback that a human will edit |
| Watch out for |
|
|
| Capabilities | ||
| Modalities | text | text, audio |
| Context window | 16.385K | 400K |
| Variants | Turbo (16K), Instruct | Max effort, High, Medium, Low, Mini |
| On HumanLike | Not yet | Not yet |
| HumanLike evals | ||
| CV field extraction (F1) | — | 0.94Estimate |
| Time to first token (medium effort) | — | 640 msEstimate |
| Published benchmarks | ||
| MMLU | 70%Published | — |
| MMLU-Pro | — | TBDEstimate |
| SWE-bench Verified | — | TBDEstimate |
Cost per hour assumes 150 wpm and a 40% AI speaking share. Hover a figure for its full derivation.