Home / Models / Qwen3.5-Omni-Plus

Qwen3.5-Omni-Plus

Maker: Alibaba · language

Qwen's omni-modal model: it understands text, images, up to 3 hours of audio and 1 hour of video, and replies in text and speech. 113 voices, output in 74 languages and 39 dialects, with volume, rate and emotion set by instruction.

Understands

text, images, audio, video

Produces

text, speech

Contents verified with the vendor: 2026-08-18

Included in subscriptions

No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.