Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
byczech
/
gemma4-it-16G-Q3_K_S
:31b
55
Downloads
Updated
3 weeks ago
Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.
Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.
Cancel
tools
thinking
31b
gemma4-it-16G-Q3_K_S:31b
...
/
template
b507b9c2f6ca · 13B
{{ .Prompt }}