text model

Huge context and native audio/video understanding at aggressive prices.
Gemini 3.5 is the value pick when the input is large or non-textual. Its native audio and video understanding is why the same family powers our transcription tier — you can hand it a raw interview recording rather than a transcript.
Variants
Same family, different quality/cost rung. Picking the wrong one is the most common way teams overpay.
gemini-3.5-proFrontier tier with the full 1M context.
gemini-3.5-flashThe price/performance sweet spot.
gemini-3.5-flash-liteCheapest tier for mechanical work.
Fit
The same model is a different proposition depending on what you point it at.
Recruitment fit
The cheapest credible way to process raw interview recordings without a separate transcription step.
Best for
Strengths
Watch out for
The numbers
HumanLike evals
Run on our own harness.
Measured on gemini-3.5-transcribe-live, the audio sibling of this family.