Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
gemma-4-31b · Ollama
Search for models on Ollama.
  • tinyrick/gemma-4-31B-it-uncensored-heretic-vision-llmfan46

    llmfan46/gemma-4-31B-it-uncensored-heretic-GGU with Vision

    vision tools thinking

    12.3K  Pulls 2  Tags Updated  2 weeks ago

  • juilpark/gemma-4-31B-it-uncensored-heretic

    llmfan46/gemma-4-31B-it-uncensored-heretic-GGUF

    tools thinking

    6,692  Pulls 1  Tag Updated  5 months ago

  • SiliconBasedWorld/Gemma-4-31B-JANG_4M-CRACK

    tools thinking

    3,008  Pulls 5  Tags Updated  5 months ago

  • HammerAI/gemma-4-31b-heretic

    coder3101/gemma-4-31B-it-heretic

    tools thinking

    878  Pulls 2  Tags Updated  2 months ago

  • austinlaw076/gemma-4-31B-it-Mystery-Fine-Tune-HERETIC-UNCENSORED-Thinking-Instruct-GGUF-Q6_K

    From DavidAU/gemma-4-31B-it-Mystery-Fine-Tune-HERETIC-UNCENSORED-Thinking-Instruct-GGUF:Q6_K

    vision tools thinking

    1,106  Pulls 1  Tag Updated  4 months ago

  • its_the_jisoo/gemma-4-31b-it-uncensored

    llmfan46/gemma-4-31B-it-uncensored-heretic-GGUF

    tools thinking

    860  Pulls 1  Tag Updated  4 months ago

  • cleex/gemma-4-31B-it-Claude-Opus-Distill-GGUF

    vision tools thinking

    797  Pulls 1  Tag Updated  5 months ago

  • ttempvnn/TrevorJS-gemma-4-31B-it-uncensored-GGUF-Q4-K-M

    📝 A copy by author TrevorJS. I tested it, the quality is excellent! 👍👍👍

    373  Pulls 1  Tag Updated  1 month ago

  • bjoernb/gemma4-31b-think

    Gemma 4 31B (Google DeepMind) with thinking mode enabled. Best for complex reasoning, math, coding, and multi-step analysis. Knowledge cutoff: January 2025. Sampling: temperature 1.0 / top_p 0.95 / top_k 64.

    vision tools thinking cloud

    653  Pulls 1  Tag Updated  5 months ago

  • bjoernb/gemma4-31b-fast

    Gemma 4 31B (Google DeepMind) with thinking mode disabled. Best for quick questions, chat, and straightforward tasks. Knowledge cutoff: January 2025. Sampling: temperature 1.0 / top_p 0.95 / top_k 64.

    vision tools thinking cloud

    583  Pulls 1  Tag Updated  5 months ago

  • byczech/gemma4-it-mmproj-16G-Q3_K_S

    Gemma 4 31B multimodal instruct model with vision support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports a practical context window of approximately 36K–40K tokens, depending on the Ollama version, GPU, backend and runtime configuration.

    vision tools thinking 31b

    296  Pulls 1  Tag Updated  3 weeks ago

  • Jarcgon/Gemma-4-31B-it-abliterated-Q4_K_M

    391  Pulls 1  Tag Updated  3 months ago

  • justingtzk/gemma-4-31B-it-qat-GGUF

    Unsloth's Gemma4 QAT models with enforced context windows.

    vision tools thinking

    294  Pulls 2  Tags Updated  3 months ago

  • edtorre/gemma4-31b-hermes

    Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.

    vision tools thinking

    197  Pulls 1  Tag Updated  2 months ago

  • Barbatos/gemfable-agent

    Agentic coding model for 24 GB GPUs: TeichAI's Gemma-4-31B Fable-5 agent distill (vision + thinking + tools) with a disciplined coding-agent system prompt baked in. Inspect → reproduce → smallest fix → re-verify.

    vision tools thinking 31b

    161  Pulls 1  Tag Updated  1 month ago

  • ttempvnn/DavidAU-gemma-4-31B-it-The-DECKARD-HERETIC-UNCENSORED-Thinking-Q4-K-M

    📝 A copy by author DavidAU──── ୨୧ ────🔄 Convert to GGUF by Mradermacher

    vision

    147  Pulls 1  Tag Updated  1 month ago

  • jetelain/Gemma-4-31B

    Gemma-4-31B UD-Q4_K_XL MTP from Unsloth

    vision tools thinking

    160  Pulls 1  Tag Updated  2 months ago

  • rlaglovesme/luminar

    Luminar is an advanced conversational engine built on Gemma 4 31B, optimized for dynamic persona roleplay, authentic instant messaging, with 262k context length.

    cloud

    93  Pulls 1  Tag Updated  1 month ago

  • byczech/gemma4-it-16G-Q3_K_S

    Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.

    tools thinking 31b

    85  Pulls 1  Tag Updated  1 month ago

  • moophlo/gemma-4-31B-it-GGUF

    vision

    138  Pulls 1  Tag Updated  5 months ago

© 2026 Ollama
Blog Support