
OpenAI TTS
speech model · Available
OpenAISteerable voice, you describe the delivery in a prompt rather than tuning sliders.
Fit and price
Recruitment fit
76
Per hour
$0.308
Per minute
$0.0051
30-min interview
$0.154
About
OpenAI’s TTS models are less expressive than ElevenLabs at the top end but introduce a genuinely useful idea: you instruct the delivery in natural language ("warm, unhurried, like you are explaining to a nervous candidate"). For screening flows where tone should shift with context, that is easier to control than a fixed voice preset.
- Prompted delivery
- Streaming
- Low cost
- Preset voices
Strengths and watch-outs
Strengths
- Prompted tone per stage
- Cheaper than ElevenLabs at volume
- No cloning, no consent issue
Watch out for
- Less expressive at the top
- No custom voice cloning
Best for Automated candidate updates · Interview instructions · High-volume notifications
2 variants

Mini TTSBalanced
Steerable delivery at low cost.
400ms
latency
1×
rel. cost
82
quality

TTS-1 HDHigh
Higher-fidelity preset voices, no steering.
700ms
latency
2×
rel. cost
78
quality