GLM-4.6V
Maker: Z.ai · language
A Z.ai multimodal model with a 128K-token context accepting text, images, video and files. It handles roughly 150 document pages, 200 slides or an hour of video, with native function calling.
Understands
text, images, video, documents
Produces
text, code
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.