Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
GLM-4V · Ollama
Search for models on Ollama.
  • byczech/glm-4.7-flash-16G-UD-Q3_K_XL

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to Q3_K_XL and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 136K fitting within 16 GB in suitable configurations.

    tools thinking 30b

    218  Pulls 1  Tag Updated  1 month ago

  • byczech/glm-4.7-flash-12G-UD-IQ2_XSS

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ2_XSS and optimized for 12 GB VRAM. Supports a maximum context window of 202,752 tokens, subject to the Ollama version, GPU, backend and runtime configuration.

    tools thinking 30b

    192  Pulls 1  Tag Updated  1 month ago

  • byczech/glm-4.7-flash-16G-UD-IQ3_XSS

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_XSS and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 180K fitting within 16 GB in suitable configurations.

    tools thinking 30b

    101  Pulls 1  Tag Updated  1 month ago

  • alibilge/Huihui-GLM-4.6V-Flash-abliterated

    Abliterated (Uncensored) GLM4.6 Flash

    9,465  Pulls 11  Tags Updated  8 months ago

  • haervwe/GLM-4.6V-Flash-9B

    GLM 4.6V Flash 9B model with vision, tools, and hybrid thinking enabled. using custom template to align it to ollama and the recomended sampling settigns by default. using unsloth quants at q4K_M

    vision tools thinking

    5,477  Pulls 1  Tag Updated  8 months ago

  • gurubot/GLM-4.6V-Flash-GGUF

    tools thinking

    1,240  Pulls 1  Tag Updated  8 months ago

  • MedAIBase/GLM-4.6V-Flash

    GLM-4.6V-Flash (9B) is a lightweight model optimized for local deployment and low-latency applications. It scales its context window to 128k tokens in training and achieves SoTA performance in visual understanding among models of similar parameter scales.

    9b

    471  Pulls 3  Tags Updated  7 months ago

  • scorpion7slayer/GLM-4.6V-Flash

    model imported from hf

    264  Pulls 1  Tag Updated  8 months ago

  • rnogy/GLM-4.6V

    unsloth/GLM-4.6V 106b

    228  Pulls 1  Tag Updated  8 months ago

  • aia/GLM-4.7-Flash-REAP-23B-A3B-GGUF

    A memory-efficient compressed variant of GLM-4.7-Flash that maintains near-identical performance while being 25% lighter.

    856  Pulls 1  Tag Updated  7 months ago

  • aiasistentworld/GLM-4.6-LLM

    New version GLM-4.6

    436  Pulls 1  Tag Updated  10 months ago

  • MichelRosselli/GLM-4.6-REAP-268B-A32B

    GLM-4.6-REAP-268B-A32B (by Cerebras), a memory-efficient compressed variant of GLM-4.6 that maintains near-identical performance while being 25% lighter.

    tools thinking

    138  Pulls 9  Tags Updated  9 months ago

  • JollyLlama/GLM-4-32B-0414-Q4_K_M

    This model requires Ollama v0.6.6 or later

    6,107  Pulls 1  Tag Updated  1 year ago

  • mychen76/GLM-4-32B-cline-roocode

    Quantized version of THUDM/GLM-4-32B-0414 optimized for tool usage with Cline / Roo Code and complex problem solving.

    tools

    1,807  Pulls 3  Tags Updated  1 year ago

  • ucx0204/glm-4.6V-Flash-Q8

    vision tools thinking

    641  Pulls 1  Tag Updated  7 months ago

  • pdevine/glm-4.7-flash

    Experimental version of glm-4.7-flash

    tools thinking

    60  Pulls 2  Tags Updated  6 months ago

  • gsarti/phi3-mini-rebus-solver

    Phi-3 Mini 4K fine-tuned on verbalized rebus solving in Italian

    13  Pulls 1  Tag Updated  2 years ago

  • tinyrick/gemma-4-31B-it-uncensored-heretic-vision-llmfan46

    llmfan46/gemma-4-31B-it-uncensored-heretic-GGU with Vision

    vision tools thinking

    12K  Pulls 2  Tags Updated  1 week ago

  • aisingapore/Gemma-SEA-LION-v4-4B-VL

    Gemma-SEA-LION-v4-4B-VL is a multilingual, multimodal model which has been pretrained and instruct-tuned for the Southeast Asia region. Developed by AI Singapore and funded by National Research Foundation, Singapore.

    vision tools

    434  Pulls 5  Tags Updated  6 months ago

  • Agen/gemma-4-26B-A4B-it-uncensored-heretic

    llmfan46/gemma-4-26B-A4B-it-uncensored-heretic - quantized to q4_K_M from HF with vision capability retained

    vision tools thinking

    5,532  Pulls 1  Tag Updated  4 months ago

© 2026 Ollama
Blog Support