Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
byczech
/
glm-4.7-flash-12G-UD-IQ2_XSS
30
Downloads
Updated
4 days ago
GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ2_XSS and optimized for 12 GB VRAM. Supports a maximum context window of 202,752 tokens, subject to the Ollama version, GPU, backend and runtime configuration.
GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ2_XSS and optimized for 12 GB VRAM. Supports a maximum context window of 202,752 tokens, subject to the Ollama version, GPU, backend and runtime configuration.
Cancel
tools
thinking
30b
Name
1 model
Size / Usage
Context
Input
glm-4.7-flash-12G-UD-IQ2_XSS:30b
6ed1c9fa5068
• 11GB • 198K context window •
Text input • 4 days ago
Text input • 4 days ago
glm-4.7-flash-12G-UD-IQ2_XSS:30b
11GB
198K
Text
6ed1c9fa5068
· 4 days ago