Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
GLM-4 Air · Ollama
Search for models on Ollama.
  • HammerAI/GLM-4.5-Air-Derestricted

    ArliAI/GLM-4.5-Air-Derestricted

    tools thinking

    332  Pulls 2  Tags Updated  2 months ago

  • gurubot/GLM-4.5-Air-Derestricted

    uncensored GLM Air

    tools thinking

    2,092  Pulls 1  Tag Updated  8 months ago

  • aiasistentworld/GLM-4.5-Air-LLM

    The Conscious Partner Your Dynamic AI Collaborator

    385  Pulls 1  Tag Updated  11 months ago

  • MichelRosselli/GLM-4.5-Air

    GLM-4.5-Air is a hybrid reasoning model that provides two modes: a thinking mode for complex reasoning and tool use, and a non-thinking mode for immediate responses.

    tools thinking

    114.9K  Pulls 9  Tags Updated  1 year ago

  • polklori0/glm-4.5-air-derestricted

    17  Pulls 1  Tag Updated  2 months ago

  • aia/GLM-4.7-Flash-REAP-23B-A3B-GGUF

    A memory-efficient compressed variant of GLM-4.7-Flash that maintains near-identical performance while being 25% lighter.

    856  Pulls 1  Tag Updated  7 months ago

  • haervwe/GLM-4.6V-Flash-9B

    GLM 4.6V Flash 9B model with vision, tools, and hybrid thinking enabled. using custom template to align it to ollama and the recomended sampling settigns by default. using unsloth quants at q4K_M

    vision tools thinking

    5,470  Pulls 1  Tag Updated  8 months ago

  • MichelRosselli/GLM-4.6-REAP-218B-A32B-FP8-mixed-AutoRound

    This model is a mixed gguf q2ks format of Cerebras' GLM-4.6-REAP-218B-A32B-FP8 generated using Intel's AutoRound algorithm.

    tools thinking

    167  Pulls 1  Tag Updated  10 months ago

  • byczech/glm-4.7-flash-uncensored-16G-AU-IQ3_M

    Uncensored GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_M and optimized for 16 GB VRAM. Suitable for coding, agents, automation, roleplay, analysis and unrestricted local experimentation.

    tools thinking 30b

    1,340  Pulls 1  Tag Updated  1 month ago

© 2026 Ollama
Blog Support