Home / Models / Mist v3

Mist v3

Maker: Rime · speech synthesis

Low-latency text-to-speech: about 37 ms time to first audio, 78 voices, English, French, German and Spanish. No inline pronunciation control.

Understands

text

This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.

Produces

speech

Contents verified with the vendor: 2026-08-18

Included in subscriptions

No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.