Mist v3
Maker: Rime · speech synthesis
Low-latency text-to-speech: about 37 ms time to first audio, 78 voices, English, French, German and Spanish. No inline pronunciation control.
Understands
text
This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.
Produces
speech
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.