Home / Models / Qwen3.5-Omni-Plus
Qwen3.5-Omni-Plus
Maker: Alibaba · language
Qwen's omni-modal model: it understands text, images, up to 3 hours of audio and 1 hour of video, and replies in text and speech. 113 voices, output in 74 languages and 39 dialects, with volume, rate and emotion set by instruction.
Understands
text, images, audio, video
Produces
text, speech
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.