39 2 weeks ago

Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.

vision tools thinking
b507b9c2f6ca · 13B
{{ .Prompt }}