-
llama3.2-vision
Llama 3.2 Vision is a collection of instruction-tuned image reasoning generative models in 11B and 90B sizes.
vision 11b 90b5.3M Pulls 9 Tags Updated 1 year ago
-
granite3.2-vision
A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.
vision tools 2b1M Pulls 5 Tags Updated 1 year ago