Youtu-VITA
Maker: Tencent · language
A Tencent multimodal model accepting images and video with a 128K-token context.
предлагается на платформе Tencent Cloud TokenHub
Understands
text, images, video
Produces
text
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.