Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
byczech
/
gemma4-it-16G-Q3_K_S
57
Downloads
Updated
3 weeks ago
Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.
Gemma 4 31B text-only instruct model with tool-calling and thinking support, quantized to Q3_K_S and optimized for 16 GB VRAM. Supports up to 256K context, with approximately 64K tokens fitting fully within 16 GB VRAM in tested configurations.
Cancel
tools
thinking
31b
Name
1 model
Size / Usage
Context
Input
gemma4-it-16G-Q3_K_S:31b
a3085fcdc8ec
• 13GB • 256K context window •
Text input • 3 weeks ago
Text input • 3 weeks ago
gemma4-it-16G-Q3_K_S:31b
13GB
256K
Text
a3085fcdc8ec
· 3 weeks ago