Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
glm · Ollama
Search for models on Ollama.
  • glm-ocr

    GLM-OCR is a multimodal OCR model for complex document understanding, built on the GLM-V encoder–decoder architecture.

    vision tools

    6.2M  Pulls 3  Tags Updated  5 months ago

  • glm-5.1

    GLM-5.1 is our next-generation flagship model for agentic engineering, with significantly stronger coding capabilities than its predecessor. It achieves state-of-the-art performance on SWE-Bench Pro and leads GLM-5 by a wide margin.

    tools thinking cloud

    2.3M  Pulls 1  Tag Updated  3 months ago

  • glm-5.2

    GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks.

    tools thinking cloud

    265.5K  Pulls 1  Tag Updated  1 month ago

  • glm-4.7-flash

    As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.

    tools thinking

    1.4M  Pulls 4  Tags Updated  1 month ago

  • glm4

    A strong multi-lingual general language model with competitive performance to Llama 3.

    9b

    1.2M  Pulls 32  Tags Updated  2 years ago

  • gemma4

    Gemma 4 models are designed to deliver frontier-level performance at each size. They are well-suited for reasoning, agentic workflows, coding, and multimodal understanding.

    vision tools thinking audio cloud e2b e4b 12b 26b 31b

    19.5M  Pulls 49  Tags Updated  3 weeks ago

  • granite4.1-guardian

    Granite Guardian 4.1 is a specialized safety and judging model from IBM Research that evaluates whether LLM prompts and responses meet specified harm criteria.

    tools thinking 8b

    5,760  Pulls 16  Tags Updated  1 month ago

  • gemma

    Gemma is a family of lightweight, state-of-the-art open models built by Google DeepMind. Updated to version 1.1

    2b 7b

    7.3M  Pulls 102  Tags Updated  2 years ago

  • granite3.1-moe

    The IBM Granite 1B and 3B models are long-context mixture of experts (MoE) Granite models from IBM designed for low latency usage.

    tools 1b 3b

    3M  Pulls 33  Tags Updated  1 year ago

  • granite3.3

    IBM Granite 2B and 8B models are 128K context length language models that have been fine-tuned for improved reasoning and instruction-following capabilities.

    tools 2b 8b

    1.1M  Pulls 3  Tags Updated  1 year ago

  • granite3.2-vision

    A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.

    vision tools 2b

    951K  Pulls 5  Tags Updated  1 year ago

  • granite3.1-dense

    The IBM Granite 2B and 8B models are text-only dense LLMs trained on over 12 trillion tokens of data, demonstrated significant improvements over their predecessors in performance and speed in IBM’s initial testing.

    tools 2b 8b

    994.3K  Pulls 33  Tags Updated  1 year ago

  • granite3.2

    Granite-3.2 is a family of long-context AI models from IBM Granite fine-tuned for thinking capabilities.

    tools 2b 8b

    445.7K  Pulls 9  Tags Updated  1 year ago

  • goliath

    A language model created by combining two fine-tuned Llama 2 70B models into one.

    467K  Pulls 16  Tags Updated  2 years ago

  • granite4.1

    IBM Granite Models are a family of enterprise-ready, open foundation models that support multilingual capabilities, coding, retrieval-augmented generation (RAG), tool use, and structured JSON output. Released under Apache 2.0 license.

    tools 3b 8b 30b

    279.5K  Pulls 48  Tags Updated  2 months ago

  • minicpm-v4.5

    A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone

    vision 8b

    17.2K  Pulls 13  Tags Updated  1 month ago

  • granite4

    Granite 4 features improved instruction following (IF) and tool-calling capabilities, making them more effective in enterprise applications.

    tools 350m 1b 3b

    1.4M  Pulls 17  Tags Updated  8 months ago

  • glennjammin/log-doctor

    Model to analyze log files and help troubleshoot errors based on ministral-3:3b model

    tools

    69  Pulls 1  Tag Updated  5 months ago

  • qwen3

    Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models.

    tools thinking 0.6b 1.7b 4b 8b 14b 30b 32b 235b

    32.8M  Pulls 58  Tags Updated  9 months ago

  • gemma3

    The current, most capable model that runs on a single GPU.

    vision 270m 1b 4b 12b 27b

    38.9M  Pulls 26  Tags Updated  11 months ago

© 2026 Ollama
Blog Contact