Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Glm 9B · Ollama
Search for models on Ollama.
  • gemma2

    Google Gemma 2 is a high-performing and efficient model available in three sizes: 2B, 9B, and 27B.

    2b 9b 27b

    32M  Pulls 94  Tags Updated  2 years ago

  • sllm/glm-z1-9b

    https://huggingface.co/THUDM/GLM-Z1-9B-0414

    1,092  Pulls 1  Tag Updated  1 year ago

  • haervwe/GLM-4.6V-Flash-9B

    GLM 4.6V Flash 9B model with vision, tools, and hybrid thinking enabled. using custom template to align it to ollama and the recomended sampling settigns by default. using unsloth quants at q4K_M

    vision tools thinking

    5,589  Pulls 1  Tag Updated  8 months ago

  • yeahdongcn/AutoGLM-Phone-9B

    https://huggingface.co/zai-org/AutoGLM-Phone-9B

    1,851  Pulls 1  Tag Updated  9 months ago

  • MedAIBase/GLM-4.6V-Flash

    GLM-4.6V-Flash (9B) is a lightweight model optimized for local deployment and low-latency applications. It scales its context window to 128k tokens in training and achieves SoTA performance in visual understanding among models of similar parameter scales.

    9b

    482  Pulls 3  Tags Updated  7 months ago

  • ac11274/GLM-Z1-9B

    38  Pulls 1  Tag Updated  7 months ago

  • anthony-maio/slipstream

    A finetuned version of GLM-Z1-9B-0414 trained on the Slipstream protocol - a semantic quantization system that achieves 82% token reduction in multi-agent AI communication. More info at: https://huggingface.co/collections/anthonym21/streamlined-inter-age

    13  Pulls 1  Tag Updated  8 months ago

  • milkey/GLM-4-9B-0414

    GLM-4-0414 series models.

    964  Pulls 1  Tag Updated  1 year ago

  • lsm03624/GLM-Z1-9B-0414-Q8_0

    GLM-4-Z1-9B-0414 9B小参数模型,其整体表现已处于同尺寸开源模型中的领先水平,此模型使用Ollama v0.6.6版本制作生成,需要Ollama v0.6.6及以上版本才能运行推理。

    550  Pulls 1  Tag Updated  1 year ago

  • transkatgirl/GLM-4-9B

    Base model

    133  Pulls 3  Tags Updated  1 year ago

  • EntropyYue/longwriter-glm4

    LongWriter-glm4-9b is trained based on glm-4-9b, and is capable of generating 10,000+ words at once.

    9b

    3,271  Pulls 1  Tag Updated  2 years ago

  • NitrAI/OpenGCM-v2

    Open-source distilled model based on Qwen3.5-9B.

    27  Pulls 1  Tag Updated  1 month ago

  • wangshenzhi/gemma2-9b-chinese-chat

    The official ollama model for Gemma-2-9B-Chinese-Chat (https://huggingface.co/shenzhi-wang/Gemma-2-9B-Chinese-Chat).

    4,960  Pulls 1  Tag Updated  2 years ago

  • aisingapore/Gemma-SEA-LION-v3-9B-IT

    Gemma-SEA-LION-v3-9B-IT is a multilingual model which has been pretrained and instruct-tuned for the Southeast Asia region. Developed by AI Singapore and funded by National Research Foundation, Singapore.

    709  Pulls 10  Tags Updated  1 year ago

  • lucasalmeida/gemma-2-9b-it-sppo-iter3

    Original model: bartowski/Gemma-2-9B-It-SPPO-Iter3-GGUF

    79  Pulls 1  Tag Updated  2 years ago

  • mdq100/qwen3.5

    Custom Qwen3.5 variants optimized for 128GB unified memory systems, such as AMD Ryzen AI Max+ 395. On Windows 11, GPU is limited to 96GB (32GB reserved for OS/CPU), requiring context window capped at 131072 tokens (128K) to fit within GPU memory limits.

    vision tools thinking

    447  Pulls 2  Tags Updated  5 months ago

© 2026 Ollama
Blog Support