Home / Models / Coda

Coda

Maker: Rime · speech synthesis

Rime's flagship text-to-speech model: 253 voices across nine languages, sub-100 ms latency and word-level timestamps.

Understands

text

This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.

Produces

speech

Contents verified with the vendor: 2026-08-18

Included in subscriptions

No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.