-
gemma4-e4b-mtp
Multimodal 4.5B Gemma 4 with Unsloth Dynamic QAT & MTP (5.3GB). Features near-lossless vision, audio, text reasoning, an active 4-token speculative window, and an expanded 64K context size. Best for edge execution on systems with 8GB to 16GB RAM.
vision tools thinking762 Pulls 1 Tag Updated 1 month ago
-
gemma4-12b-mtp
High-tier 12B Gemma 4 model with Unsloth Dynamic QAT & MTP. Packs deep multimodal reasoning, an aggressive 4-token speculative draft window, and a massive 128K context capacity. Tailored for heavy codebase ingestion and reasoning tasks.
vision tools thinking741 Pulls 1 Tag Updated 1 month ago
-
gemma4-e2b-mtp
Ultra-fast multimodal 2.3B Gemma 4 for on-device edge AI (3.7GB). Adds native vision/audio parsing to Unsloth Dynamic QAT with a 2-token MTP pipeline and a mobile-safe 32K context window. Perfect for smartphones and laptops sharing 8GB of total RAM.
vision tools thinking585 Pulls 1 Tag Updated 1 month ago