585 1 month ago

Ultra-fast multimodal 2.3B Gemma 4 for on-device edge AI (3.7GB). Adds native vision/audio parsing to Unsloth Dynamic QAT with a 2-token MTP pipeline and a mobile-safe 32K context window. Perfect for smartphones and laptops sharing 8GB of total RAM.

vision tools thinking
ollama run 4skl/gemma4-e2b-mtp

Details

1 month ago

1221fe8109e5 · 3.7GB ·

gemma4
·
4.63B
·
Q4_0
clip
·
476M
·
F16
GGUF���1�������+��������������general.architecture����������gemma4-assistant �������general.type
{ "draft_num_predict": 2, "num_ctx": 32768, "stop": [ "<turn|>" ], "temp

Readme

No readme