Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Qwen-VL · Ollama
Search for models on Ollama.
  • qwen2.5vl

    Flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL.

    vision 3b 7b 32b 72b

    5.1M  Pulls 17  Tags Updated  1 year ago

  • qwen3-vl

    The most powerful vision-language model in the Qwen model family to date.

    vision tools thinking 2b 4b 8b 30b 32b 235b

    6.2M  Pulls 57  Tags Updated  10 months ago

  • hhao/openbmb-minicpm-llama3-v-2_5

    MiniCPM-V surpasses proprietary models such as GPT-4V, Gemini Pro, Qwen-VL and Claude 3 in overall performance, and support multimodal conversation for over 30 languages.

    vision

    66K  Pulls 8  Tags Updated  2 years ago

  • dlasher/Qwen3-VL-30B-A3B-Instruct-GGUF

    Qwen/Qwen3-VL-30B-A3B-Instruct - IQ4_NL Quant

    tools

    135  Pulls 1  Tag Updated  3 weeks ago

  • feadxus/Huihui-Qwen3-VL-4B-Instruct-abliterated

    vision

    1,233  Pulls 14  Tags Updated  1 month ago

  • junquan2k/Qwen2.5-VL-7B-local9-0414

    vision

    153  Pulls 1  Tag Updated  4 months ago

  • huihui_ai/qwen2.5-vl-abliterated

    Flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL.

    vision 3b 7b 32b

    46.3K  Pulls 16  Tags Updated  10 months ago

  • adelnazmy2002/Qwen3-VL-4B-Instruct

    vision tools

    8,001  Pulls 2  Tags Updated  9 months ago

  • MedAIBase/Qwen3-VL-Reranker

    The Qwen3-VL-Reranker model series are the latest additions to the Qwen family, built upon the recently open-sourced and powerful Qwen3-VL foundation model.

    2b

    3,701  Pulls 3  Tags Updated  8 months ago

  • MedAIBase/Qwen3-VL-Embedding

    The Qwen3-VL-Embedding model series are the latest additions to the Qwen family, built upon the recently open-sourced and powerful Qwen3-VL foundation model.

    2b

    3,007  Pulls 3  Tags Updated  8 months ago

  • adelnazmy2002/Qwen3-VL-8B-Instruct

    vision tools

    2,523  Pulls 1  Tag Updated  9 months ago

  • feadxus/Huihui-Qwen3-VL-4B

    vision

    25  Pulls 1  Tag Updated  1 month ago

  • Keyvan/german-ocr-turbo

    German-OCR-Turbo ist ein fine-tuned Vision-Language-Modell basierend auf Qwen3-VL-2B, optimiert für die präzise Texterkennung aus deutschen Rechnungen, Formularen und Geschäftsdokumenten. Das Modell extrahiert strukturierte Daten im Markdown-Format.

    vision tools

    1,873  Pulls 1  Tag Updated  9 months ago

  • Willem27/Qwen3-VL-8B-Thinking

    1,019  Pulls 1  Tag Updated  11 months ago

  • seamon67/Qwen3-VL-Heretic

    Qwen 3 VL but like Heretic.

    vision tools thinking 32b

    712  Pulls 4  Tags Updated  7 months ago

  • mikgr/doctype-classifier-vl

    A specialized document classification model based on Qwen2.5-VL-3B that automatically detects document types from PDFs and images with high accuracy and calibrated confidence scores.

    vision

    608  Pulls 2  Tags Updated  8 months ago

  • ahmadwaqar/mai-ui

    Alibaba Tongyi GUI agent on Qwen3-VL. SOTA: 73.5% ScreenSpot-Pro, 76.7% AndroidWorld. Returns bbox [x1,y1,x2,y2] for UI automation. Supports MCP tools & device-cloud collaboration. Apache 2.0. Tags: 2b (default), 8b.

    vision 2b 8b

    617  Pulls 3  Tags Updated  8 months ago

  • theoistic/Qwen-3-VL-30B-A3B-Instruct

    High Quality Vision Instruct Model

    vision

    570  Pulls 1  Tag Updated  7 months ago

  • fervent_mcclintock/Qwen3-VL-Embedding-2B

    vision

    485  Pulls 2  Tags Updated  7 months ago

  • mirage335/Qwen-3-VL-30B-A3B-Instruct-virtuoso

    Recommended for transcribing and summarizing text from screenshots.

    vision tools

    470  Pulls 1  Tag Updated  8 months ago

© 2026 Ollama
Blog Support