Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
byczech
/
gemma4-it-16G-Q3_K_S
:31b
57
Downloads
Updated
3 weeks ago
Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.
Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.
Cancel
tools
thinking
31b
gemma4-it-16G-Q3_K_S:31b
...
/
params
56380ca2ab89 · 42B
{
"temperature": 1,
"top_k": 64,
"top_p": 0.95
}