Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
glm-4.6 · Ollama
Search for models on Ollama.
  • zerocopia/glm-4.6

    tools thinking cloud

    50  Pulls 1  Tag Updated  4 months ago

  • sparksammy/glm-4.6v-flash-unsloth

    vision tools thinking

    805  Pulls 5  Tags Updated  7 months ago

  • ucx0204/glm-4.6V-Flash-Q8

    vision tools thinking

    657  Pulls 1  Tag Updated  8 months ago

  • doitmagic/glm-4.6v-flash

    553  Pulls 1  Tag Updated  9 months ago

  • alibilge/Huihui-GLM-4.6V-Flash-abliterated

    Abliterated (Uncensored) GLM4.6 Flash

    9,645  Pulls 11  Tags Updated  8 months ago

  • haervwe/GLM-4.6V-Flash-9B

    GLM 4.6V Flash 9B model with vision, tools, and hybrid thinking enabled. using custom template to align it to ollama and the recomended sampling settigns by default. using unsloth quants at q4K_M

    vision tools thinking

    5,643  Pulls 1  Tag Updated  8 months ago

  • MichelRosselli/GLM-4.6

    GLM-4.6 is a hybrid reasoning model that provides two modes: a thinking mode for complex reasoning and tool use, and a non-thinking mode for immediate responses.

    tools thinking

    5,883  Pulls 9  Tags Updated  9 months ago

  • gurubot/GLM-4.6V-Flash-GGUF

    tools thinking

    1,256  Pulls 1  Tag Updated  9 months ago

  • ShreyanGondaliya/s5

    A model based on the GLM-4.6v-flash:9b q5_k_m, and uncensored. For local use I recommend editing the model context in modelfile as it is set to 128k. #EDIT: New local optimised model same with context 4096 https://ollama.com/ShreyanGondaliya/s5-reduced

    490  Pulls 1  Tag Updated  7 months ago

  • MedAIBase/GLM-4.6V-Flash

    GLM-4.6V-Flash (9B) is a lightweight model optimized for local deployment and low-latency applications. It scales its context window to 128k tokens in training and achieves SoTA performance in visual understanding among models of similar parameter scales.

    9b

    482  Pulls 3  Tags Updated  7 months ago

  • aiasistentworld/GLM-4.6-LLM

    New version GLM-4.6

    436  Pulls 1  Tag Updated  11 months ago

  • scorpion7slayer/GLM-4.6V-Flash

    model imported from hf

    272  Pulls 1  Tag Updated  8 months ago

  • rnogy/GLM-4.6V

    unsloth/GLM-4.6V 106b

    234  Pulls 1  Tag Updated  8 months ago

  • MichelRosselli/GLM-4.6-REAP-218B-A32B-FP8-mixed-AutoRound

    This model is a mixed gguf q2ks format of Cerebras' GLM-4.6-REAP-218B-A32B-FP8 generated using Intel's AutoRound algorithm.

    tools thinking

    167  Pulls 1  Tag Updated  10 months ago

  • MichelRosselli/GLM-4.6-REAP-268B-A32B

    GLM-4.6-REAP-268B-A32B (by Cerebras), a memory-efficient compressed variant of GLM-4.6 that maintains near-identical performance while being 25% lighter.

    tools thinking

    138  Pulls 9  Tags Updated  9 months ago

  • byczech/glm-4.7-flash-uncensored-16G-AU-IQ3_M

    Uncensored GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_M and optimized for 16 GB VRAM. Suitable for coding, agents, automation, roleplay, analysis and unrestricted local experimentation.

    tools thinking 30b

    1,526  Pulls 1  Tag Updated  1 month ago

  • byczech/glm-4.7-flash-16G-UD-Q3_K_XL

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to Q3_K_XL and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 136K fitting within 16 GB in suitable configurations.

    tools thinking 30b

    261  Pulls 1  Tag Updated  1 month ago

  • byczech/glm-4.7-flash-16G-UD-IQ3_XSS

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_XSS and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 180K fitting within 16 GB in suitable configurations.

    tools thinking 30b

    117  Pulls 1  Tag Updated  1 month ago

  • JollyLlama/GLM-4-32B-0414-Q4_K_M

    This model requires Ollama v0.6.6 or later

    6,126  Pulls 1  Tag Updated  1 year ago

  • rhundt/GLM-4-0414-32b-128k-Q4_K_M

    GLM-4-0414 32B with 128k context (YaRN RoPE scaling). Needs ollama 0.6.6

    tools

    1,091  Pulls 1  Tag Updated  1 year ago

© 2026 Ollama
Blog Support