Llama 3.2 Vision is a collection of instruction-tuned image reasoning generative models in 11B and 90B sizes.
5.2M Pulls 9 Tags Updated 1 year ago