Flux TTS
Maker: Deepgram · speech synthesis
Text-to-speech built for voice agents: it keeps conversational state across turns, starts speaking in around 80 ms and reports the exact text it was saying when interrupted. Model strings look like flux-{voice}-{language}, English only for now.
Understands
text
This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.
Produces
speech
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.