Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
edtorre
/
gemma4-31b-hermes
39
Downloads
Updated
2 weeks ago
Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.
Gemma 4 31B (Q4_0) optimized for Hermes Agent — 64K context, 8192 max tokens, flash attention + Q8 KV cache. Fits 24 GB VRAM with optimizations.
Cancel
vision
tools
thinking
Name
1 model
Size / Usage
Context
Input
gemma4-31b-hermes:latest
615f1a4c56f6
• 19GB • 256K context window •
Text, Image input • 2 weeks ago
Text, Image input • 2 weeks ago
gemma4-31b-hermes:latest
19GB
256K
Text, Image
615f1a4c56f6
· 2 weeks ago