Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Gemma 4 4B · Ollama
Search for models on Ollama.
  • medgemma1.5

    MedGemma 1.5 4B is an updated version of the MedGemma 4B model.

    vision 4b

    121.5K  Pulls 5  Tags Updated  3 months ago

  • luisppb16/gemma4-e4b-SecOps

    Fine-tune of Gemma 4 E4B (4B parameters) specialized in offensive and defensive cybersecurity. Trained with LoRA (r=16) on a dataset covering vulnerability analysis, secure code review, and Java insecure code remediation.

    tools thinking

    594  Pulls 1  Tag Updated  3 months ago

  • rubinmaximilian/Monk-Router-Gemma4e2b

    A lightweight, hardware-aware router built on Gemma 4 E2B. It acts as a dispatcher for local AI setups, automatically deciding whether a prompt should run on edge hardware (like a Jetson Nano), a local GPU, or the cloud based on task complexity.

    vision tools thinking

    201  Pulls 1  Tag Updated  4 months ago

  • MedAIBase/MedGemma1.5

    MedGemma 1.5 4B is an updated version of the MedGemma 1 4B model, delivers improved accuracy on medical text reasoning and modest improvement on standard 2D image interpretation compared to MedGemma 1 4B. The 4b-it-q4_0 has overfitting! Avoid it.

    4b

    6,059  Pulls 5  Tags Updated  6 months ago

  • aisingapore/Gemma-SEA-LION-v4-4B-VL

    Gemma-SEA-LION-v4-4B-VL is a multilingual, multimodal model which has been pretrained and instruct-tuned for the Southeast Asia region. Developed by AI Singapore and funded by National Research Foundation, Singapore.

    vision tools

    430  Pulls 5  Tags Updated  6 months ago

  • AntAngelMed/MedGemma1.5

    MedGemma 1.5 4B is an updated version of the MedGemma 1 4B model, delivers improved accuracy on medical text reasoning and modest improvement on standard 2D image interpretation compared to MedGemma 1 4B. The 4b-it-q4_0 has overfitting! Avoid it.

    4b

    228  Pulls 5  Tags Updated  6 months ago

  • boboba/Nidum-Gemma-3-4B-it-Uncensored-GGUF

    https://huggingface.co/nidum/Nidum-Gemma-3-4B-it-Uncensored-GGUF

    1,335  Pulls 1  Tag Updated  1 year ago

  • ttempvnn/HauhauCS-Gemma4-26B-A4B-Uncensored-HauhauCS-Balanced-Q4-K-M

    vision

    202  Pulls 1  Tag Updated  3 days ago

  • mo-shakib/gemma4-e4b-uncensored

    An uncensored Gemma 4 8B model for local inference with zero built-in refusal behavior

    vision tools thinking

    1,801  Pulls 1  Tag Updated  1 week ago

  • 4skl/gemma4-e4b-mtp

    Multimodal 4.5B Gemma 4 with Unsloth Dynamic QAT & MTP (5.3GB). Features near-lossless vision, audio, text reasoning, an active 4-token speculative window, and an expanded 64K context size. Best for edge execution on systems with 8GB to 16GB RAM.

    vision tools thinking

    567  Pulls 1  Tag Updated  3 weeks ago

  • R4C3R/gemma-3-4b-it-heretic

    Gemma-3-4B-IT fine-tuned for uncensored creative writing and roleplay. Made by RACER IS OP.

    475  Pulls 2  Tags Updated  3 weeks ago

  • sonct988/gemma4-26b-a4b-it-q4km-256k

    Gemma 4 26B A4B Instruct GGUF Q4_K_M build for Ollama, configured with a 256K context window. This is a text-focused local model built from google/gemma-4-26B-A4B-it and intended for chat, summarization, tool-style workflows, and long-context testing.

    tools thinking

    116.9K  Pulls 1  Tag Updated  1 month ago

  • byczech/gemma4-it-mmproj-16G-Q3_K_S

    Gemma 4 31B multimodal instruct model with vision support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports a practical context window of approximately 36K–40K tokens, depending on the Ollama version, GPU, backend and runtime configuration.

    vision tools thinking 31b

    193  Pulls 1  Tag Updated  3 weeks ago

  • dna5rm/gemma4

    Context-optimized variant of Google's Gemma 4 12B model, tuned for GPU-constrained inference.

    vision tools thinking

    61  Pulls 1  Tag Updated  1 week ago

  • edtorre/gemma4-31b-hermes

    Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.

    vision tools thinking

    103  Pulls 1  Tag Updated  4 weeks ago

  • maxwellb/gemma4-12b-it-oym

    Quantized GGUF of hf.co/OpenYourMind/gemma-4-12B-it-abliterated-uncensored

    vision tools thinking

    13.3K  Pulls 3  Tags Updated  2 months ago

  • VladimirGav/gemma4-26b-16GB-VRAM

    Gemma 4 26B (IQ4_XS) - Optimized for 16GB VRAM

    tools thinking

    21.4K  Pulls 1  Tag Updated  4 months ago

  • VladimirGav/gemma4-26b-16GB-VRAM-Uncensored

    Gemma 4 Uncensored 26B (IQ4_XS) - Optimized for 16GB VRAM

    tools thinking

    8,511  Pulls 1  Tag Updated  3 months ago

  • satgeze/gemma4-12b-uncensored-1.5m

    Gemma 4 12B uncensored, 1.5M max context: certified 10/10 to 1.31M on a 32GB card. Vision included.

    vision tools thinking

    3,266  Pulls 1  Tag Updated  1 month ago

  • vidraft/pocket-26b

    # POCKET-26B **On-device Korean AI, based on Google Gemma4-26B-A4B.** Runs on your PC or phone with **no GPU** — in Ollama, LM Studio, PocketPal, or any llama.cpp app. ## Run

    tools thinking

    33  Pulls 1  Tag Updated  1 week ago

© 2026 Ollama
Blog Contact