Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Gemma 4 2B · Ollama
Search for models on Ollama.
  • cajina/gemma4_e2b-q4_k_s

    Gemma4 model 2B parameters with Q4_K_S quatization. Thinking removed. Context set to 4096

    tools thinking

    1,229  Pulls 1  Tag Updated  4 months ago

  • ttempvnn/HauhauCS-Gemma4-26B-A4B-Uncensored-HauhauCS-Balanced-Q4-K-M

    vision

    202  Pulls 1  Tag Updated  3 days ago

  • vidraft/pocket-26b

    # POCKET-26B **On-device Korean AI, based on Google Gemma4-26B-A4B.** Runs on your PC or phone with **no GPU** — in Ollama, LM Studio, PocketPal, or any llama.cpp app. ## Run

    tools thinking

    33  Pulls 1  Tag Updated  1 week ago

  • julienp79/occitan-gemma-4-e2b-it-lora-sfttrainer

    Gemma 4 2B fine-tuned on Occitan via standard LoRA (r=16) with SFTTrainer. Q2_K, Q4_K_M, Q5_K_M, Q8_0, f16 quants available.

    38  Pulls 6  Tags Updated  2 months ago

  • julienp79/occitan-gemma-4-e2b-it-rslora-sfttrainer

    Gemma 4 2B fine-tuned on Occitan (Lengadocian) via RS-LoRA (r=32) with SFTTrainer. Q2_K, Q4_K_M, Q5_K_M, Q8_0, f16 quants available.

    6  Pulls 6  Tags Updated  2 months ago

  • 4skl/gemma4-e2b-mtp

    Ultra-fast multimodal 2.3B Gemma 4 for on-device edge AI (3.7GB). Adds native vision/audio parsing to Unsloth Dynamic QAT with a 2-token MTP pipeline and a mobile-safe 32K context window. Perfect for smartphones and laptops sharing 8GB of total RAM.

    vision tools thinking

    444  Pulls 1  Tag Updated  3 weeks ago

  • sonct988/gemma4-26b-a4b-it-q4km-256k

    Gemma 4 26B A4B Instruct GGUF Q4_K_M build for Ollama, configured with a 256K context window. This is a text-focused local model built from google/gemma-4-26B-A4B-it and intended for chat, summarization, tool-style workflows, and long-context testing.

    tools thinking

    116.9K  Pulls 1  Tag Updated  1 month ago

  • aratan/gemma4-26B-Q4-update

    vision tools thinking

    100  Pulls 1  Tag Updated  2 weeks ago

  • VladimirGav/gemma4-26b-16GB-VRAM

    Gemma 4 26B (IQ4_XS) - Optimized for 16GB VRAM

    tools thinking

    21.4K  Pulls 1  Tag Updated  4 months ago

  • jahangircse/Gemma4_12B

    tools thinking

    44  Pulls 1  Tag Updated  2 weeks ago

  • VladimirGav/gemma4-26b-16GB-VRAM-Uncensored

    Gemma 4 Uncensored 26B (IQ4_XS) - Optimized for 16GB VRAM

    tools thinking

    8,511  Pulls 1  Tag Updated  3 months ago

  • batiai/gemma4-26b

    Gemma 4 26B MoE quantized by BatiAI. 77 t/s on M4 Max. Requires 24GB+ Mac.

    tools thinking

    6,678  Pulls 6  Tags Updated  4 months ago

  • satgeze/gemma4-26b-uncensored-1m

    Gemma 4 26B-A4B uncensored MoE: 1M context, vision, fast. ~91% recall, honestly documented.

    vision tools thinking

    1,873  Pulls 1  Tag Updated  1 month ago

  • bjoernb/gemma4-26b-think

    Gemma 4 26B MoE (Google DeepMind) with thinking mode enabled. Mixture-of-Experts — 25.2B total / 3.8B active parameters, 256K context. Supports text and image input. Knowledge cutoff: January 2025.

    vision tools thinking

    2,756  Pulls 1  Tag Updated  4 months ago

  • bjoernb/gemma4-e2b-fast

    Gemma 4 E2B (Google DeepMind) with thinking mode disabled. Compact multimodal model — 2.3B effective / 5.1B total parameters. Supports text, image and audio input. Designed for edge devices and local deployment. Knowledge cutoff: January 2025.

    vision tools thinking

    1,799  Pulls 1  Tag Updated  4 months ago

  • SetneufPT/Gemma4-12B-IT-QAT_Q4_64K_16GB-GPU

    Custom model for coding with agents to use locally with 16gb or 2x8gb GPUs (working fine...)

    vision tools thinking

    988  Pulls 1  Tag Updated  2 months ago

  • mannix/gemma4-98e-v6-coder

    The best 20b coding model just got better! Beats the bigger 26b brother in Python and code reasoning

    vision tools thinking

    1,115  Pulls 62  Tags Updated  1 week ago

  • aravhawk/gemma4

    Gemma 4 26B Optimized for 16GB VRAM via Q3 Quantization

    tools thinking 26b

    1,075  Pulls 2  Tags Updated  3 months ago

  • rafw007/gemma4-26b-claude-coder

    Niestandardowy model Gemma 4 26B (~25,8B parametrów), dostrojony do działania jako niezależny agent kodowania i administracji . Obsługuje API zgodne z Anthropic, dzięki czemu obsługuje Claude Code, Codex i Opencode

    tools thinking

    740  Pulls 1  Tag Updated  2 months ago

  • bjoernb/gemma4-26b-fast

    Gemma 4 26B MoE (Google DeepMind) with thinking mode disabled. Mixture-of-Experts — 25.2B total / 3.8B active parameters, 256K context. Supports text and image input. Knowledge cutoff: January 2025.

    vision tools thinking

    1,035  Pulls 1  Tag Updated  4 months ago

© 2026 Ollama
Blog Contact