2,708 4 months ago

4-bit K-quntized version of google/gemma-4-E4B-it. Generated using llama.cpp

tools thinking
ollama run su_robin/gemma-4-E4B-it-Q4_K_M

Details

4 months ago

75c5c8502945 · 5.3GB ·

gemma4
·
7.52B
·
Q4_K_M
{ "stop": [ "<turn|>" ] }

Readme

Gemma-4-E4B-it-Q4_K_M

Lightweight redistribution of the Gemma-4-E4B-it instruction-tuned model. Conveted to gguf format with 4-bit K-Quantization (Q4_K_M) using llama.cpp.

Attribution

All credit belongs to the original authors and contributors.

Lisence

Model weights from Google Deepmind under the Apache 2.0 Lisence