Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
vision · Ollama
Search for models on Ollama.
  • qwen3-vl

    The most powerful vision-language model in the Qwen model family to date.

    vision tools thinking 2b 4b 8b 30b 32b 235b

    5.9M  Pulls 57  Tags Updated  10 months ago

  • deepseek-ocr

    DeepSeek-OCR is a vision-language model that can perform token-efficient OCR.

    vision 3b

    529.2K  Pulls 3  Tags Updated  9 months ago

  • qwen2.5vl

    Flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL.

    vision 3b 7b 32b 72b

    4.8M  Pulls 17  Tags Updated  1 year ago

  • llava

    🌋 LLaVA is a novel end-to-end trained large multimodal model that combines a vision encoder and Vicuna for general-purpose visual and language understanding. Updated to version 1.6.

    vision 7b 13b 34b

    14.8M  Pulls 98  Tags Updated  2 years ago

  • llama3.2-vision

    Llama 3.2 Vision is a collection of instruction-tuned image reasoning generative models in 11B and 90B sizes.

    vision 11b 90b

    5.2M  Pulls 9  Tags Updated  1 year ago

  • minicpm-v

    A series of multimodal LLMs (MLLMs) designed for vision-language understanding.

    vision 8b

    5.5M  Pulls 17  Tags Updated  1 year ago

  • granite3.2-vision

    A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.

    vision tools 2b

    995K  Pulls 5  Tags Updated  1 year ago

  • moondream

    moondream2 is a small vision language model designed to run efficiently on edge devices.

    vision 1.8b

    1.8M  Pulls 18  Tags Updated  2 years ago

  • mistral-small3.1

    Building upon Mistral Small 3, Mistral Small 3.1 (2503) adds state-of-the-art vision understanding and enhances long context capabilities up to 128k tokens without compromising text performance.

    vision tools 24b

    786.5K  Pulls 5  Tags Updated  1 year ago

  • medgemma

    MedGemma is a collection of Gemma 3 variants that are trained for performance on medical text and image comprehension.

    vision 4b 27b

    362.5K  Pulls 9  Tags Updated  4 months ago

  • VisionVTAI/Aria-sama

    tools

    29  Pulls 1  Tag Updated  1 year ago

  • openchat

    A family of open-source models trained on a wide variety of data, surpassing ChatGPT on various benchmarks. Updated to version 3.5-0106.

    7b

    1.3M  Pulls 50  Tags Updated  2 years ago

  • aiconjured/Qwen3.5-9B-Uncensored-HauhauCS-Aggressive-MTP-GGUF-NVFP4

    Text + Vision Qwen3.5-9B-Uncensored-HauhauCS-Aggressive-MTP-GGUF-NVFP4

    vision

    690  Pulls 1  Tag Updated  4 days ago

  • mannix/omnimerge-v6

    v4's merge rebased on Qwen/Qwen3.8-27B (+ 3 Qwen3.6 fine-tunes) — +6.5pp LiveCodeBench, MTP + vision

    vision tools thinking

    461  Pulls 39  Tags Updated  9 minutes ago

  • home-rag/gemma4

    gemma4:31b-mtp-vision-bf16

    vision tools thinking

    7  Pulls 1  Tag Updated  3 days ago

  • orcarouter/Qwen3.8-27B-Uncensored

    Qwen3.8-27B tensor-level abliterated, vision tower and MTP head untouched. 0% over-refusal on XSTest, 0-6% refusal across the A/B suite, no measurable capability loss. Full mmproj vision, tool calling and thinking, 262K context.

    vision tools thinking

    145K  Pulls 20  Tags Updated  1 week ago

  • srchmnmichael/Qwen3.8-Uncensored

    Qwen3.8-27B tensor-level abliterated, vision tower and MTP head untouched. 0% over-refusal on XSTest, 0-6% refusal across the A/B suite, no measurable capability loss. Full mmproj vision, tool calling and thinking, 262K context.

    vision tools thinking

    5,838  Pulls 4  Tags Updated  1 week ago

  • jikepjikep_16HEX/qwen3.8-27b-nightshift-heretic-uncensored-q4

    😈 Uncensored Qwen3.8-27B Vision (27.3B • Q4_K_M) for Ollama. 👁️ Multimodal Vision AI with CLIP-ViT, Thinking, MTP, Tool Calling, Rust 1.98.0 Image Analysis, GGUF & local AI development. Medic🚀 #16HEX Matrix / Nightshift Heretic.

    vision

    2,460  Pulls 1  Tag Updated  2 weeks ago

  • aiconjured/Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF-Q8-NVFP4

    Text + Vision Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF-Q8-NVFP4

    vision

    830  Pulls 1  Tag Updated  2 weeks ago

  • ericli1018/Hermes3.6-35B-A3B-Uncensored-Genesis-V7

    Ollama repack of Qwen3.6-35B-A3B Genesis Hermes V7 with APEX/APEX-Compact, Vision, and 224K context for coding agents.

    vision

    859  Pulls 3  Tags Updated  3 weeks ago

© 2026 Ollama
Blog Support