Home / Models / Llama 3.2 11B Vision
Llama 3.2 11B Vision
Maker: Meta · language
A multimodal Meta model with 11B parameters: accepts text and images, answers with text. Recognizes and reasons over pictures, and can work as a plain text model at the Llama 3.1 level. The image immediately preceding a query is the one used to answer it.
по правилам допустимого использования Meta недоступна из Европейского союза; версия предыдущего поколения
Understands
text, images
Produces
text
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.