Coda
Maker: Rime · speech synthesis
Rime's flagship text-to-speech model: 253 voices across nine languages, sub-100 ms latency and word-level timestamps.
Understands
text
This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.
Produces
speech
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.