Home / Models / Llama 3.2 11B Vision

Llama 3.2 11B Vision

Maker: Meta · language

A multimodal Meta model with 11B parameters: accepts text and images, answers with text. Recognizes and reasons over pictures, and can work as a plain text model at the Llama 3.1 level. The image immediately preceding a query is the one used to answer it.

по правилам допустимого использования Meta недоступна из Европейского союза; версия предыдущего поколения

Understands

text, images

Produces

text

Contents verified with the vendor: 2026-08-18

Included in subscriptions

No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.