Home / Models / Seed Audio 1.0
Seed Audio 1.0
Maker: ByteDance · speech synthesis
Audio creation model that treats sound as a unified scene: speech with emotional range and voice consistency, sound effects and ambience in one pass, 100 ms dialogue timing control, 20+ languages, up to 2 minutes per generation with continuation, voice creation from a text description or a reference sample. Vendor announcement 2026-07-20, available via BytePlus.
Understands
text, audio
Produces
speech, sound effects
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.