Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
glm 5.3 · Ollama
Search for models on Ollama.
  • glm-5.3-flash

    Z.ai's first natively multimodal model, approaching Claude Opus 4.8 on coding and agentic benchmarks with just 18B active parameters.

    vision tools thinking cloud

    77.5K  Pulls 1  Tag Updated  1 week ago

  • glm-5.3

    Z.ai's flagship model and the most capable open-weights model for coding, with major gains on long-horizon agentic tasks.

    tools thinking cloud

    32.8K  Pulls 1  Tag Updated  1 week ago

  • frob/glm-5.3

    753b

    67  Pulls 9  Tags Updated  5 days ago

  • kck4156/glm-5.3-opencode

    cloud

    14  Pulls 1  Tag Updated  6 days ago

  • kck4156/glm-5.3-flash-opencode

    cloud

    34  Pulls 1  Tag Updated  1 week ago

  • n0404n0404/qwen3.6-finetune-qwen3.8-max-glm5.2-kimi-k3-distillation-a56-1168cb-heretic

    602  Pulls 4  Tags Updated  3 weeks ago

  • yanjia/Qwen3.5-27B-GLM5.1-Distill-v1

    Quantization based on Jackrong / Qwen3.5-27B-GLM5.1-Distill-v1

    219  Pulls 1  Tag Updated  4 months ago

  • second_constantine/yandex-gpt-5-lite

    Instruct version of the large language model YandexGPT 5 Lite with 8B parameters with a context length of 32k tokens. (quantised version of Q5_K_M)

    8b

    614  Pulls 3  Tags Updated  12 months ago

  • 4skl/gemma4-e4b-mtp

    Multimodal 4.5B Gemma 4 with Unsloth Dynamic QAT & MTP (5.3GB). Features near-lossless vision, audio, text reasoning, an active 4-token speculative window, and an expanded 64K context size. Best for edge execution on systems with 8GB to 16GB RAM.

    vision tools thinking

    805  Pulls 1  Tag Updated  1 month ago

  • jewelzufo/LFM2.5-350M-GGUF

    Source: https://huggingface.co/LiquidAI/LFM2.5-350M-GGUF

    327  Pulls 1  Tag Updated  5 months ago

© 2026 Ollama
Blog Support