Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
edtorre
/
gemma4-31b-hermes
:latest
39
Downloads
Updated
2 weeks ago
Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.
Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.
Cancel
vision
tools
thinking
gemma4-31b-hermes:latest
...
/
template
b507b9c2f6ca · 13B
{{ .Prompt }}