Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
glm 4.7 · Ollama
Search for models on Ollama.
  • glm-4.7-flash

    As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.

    tools thinking

    1.6M  Pulls 4  Tags Updated  3 months ago

  • jewelzufo/GLM4.7-Distill-LFM2.5-1.2B

    71  Pulls 1  Tag Updated  5 months ago

  • yasserrmd/GLM4.7-Distill-LFM2.5-1.2B

    281  Pulls 1  Tag Updated  7 months ago

  • byczech/glm-4.7-flash-uncensored-16G-AU-IQ3_M

    Uncensored GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_M and optimized for 16 GB VRAM. Suitable for coding, agents, automation, roleplay, analysis and unrestricted local experimentation.

    tools thinking 30b

    1,414  Pulls 1  Tag Updated  1 month ago

  • rafw007/glm-4.7-flash-opencode

    A family of custom models built on **GLM-4.7-Flash** (MoE, 30B total / 3B active), tuned to act as autonomous coding agents — **each variant targeting a specific harness**

    tools thinking

    633  Pulls 1  Tag Updated  3 months ago

  • byczech/glm-4.7-flash-16G-UD-Q3_K_XL

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to Q3_K_XL and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 136K fitting within 16 GB in suitable configurations.

    tools thinking 30b

    232  Pulls 1  Tag Updated  1 month ago

  • huihui_ai/glm-4.7-flash-abliterated

    As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.

    tools thinking

    299.5K  Pulls 5  Tags Updated  7 months ago

  • byczech/glm-4.7-flash-12G-UD-IQ2_XSS

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ2_XSS and optimized for 12 GB VRAM. Supports a maximum context window of 202,752 tokens, subject to the Ollama version, GPU, backend and runtime configuration.

    tools thinking 30b

    197  Pulls 1  Tag Updated  1 month ago

  • dhiltgen/glm-4.7-flash

    tools thinking

    203  Pulls 10  Tags Updated  1 month ago

  • byczech/glm-4.7-flash-16G-UD-IQ3_XSS

    GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_XSS and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 180K fitting within 16 GB in suitable configurations.

    tools thinking 30b

    102  Pulls 1  Tag Updated  1 month ago

  • kalikorpz/glm-4.7-flash

    tools

    182  Pulls 1  Tag Updated  4 months ago

  • zerocopia/glm-4.7

    tools thinking cloud

    60  Pulls 1  Tag Updated  4 months ago

  • gag0/glm-4.7-flash

    tools thinking

    33  Pulls 1  Tag Updated  5 months ago

  • sparksammy/glm-4.7-flash-unsloth

    tools thinking

    1,230  Pulls 5  Tags Updated  6 months ago

  • ogecromgames/glm-4.7-flash

    tools thinking

    28  Pulls 1  Tag Updated  4 months ago

  • huihui_ai/glm-4.7-abliterated

    Advancing the Coding Capability

    tools thinking

    882  Pulls 1  Tag Updated  7 months ago

  • frob/glm-4.7

    tools thinking 358b

    768  Pulls 3  Tags Updated  2 months ago

  • bergencvv/cerebras-glm-4.7-reap-218b-a32b

    13  Pulls 1  Tag Updated  1 month ago

  • abmateen/glm-4.7cloud

    glm-4.7cloud model

    522  Pulls 1  Tag Updated  8 months ago

  • tender_mayer_860/glm-4.7-flash-abliterated

    tools thinking

    276  Pulls 1  Tag Updated  7 months ago

© 2026 Ollama
Blog Support