5 3 days ago

GLM-4.7 Flash 30B text model with tool-calling support, quantized to IQ3_XSS and optimized for 16 GB VRAM. The model supports up to 202,752 tokens context window, with approximately 180K fitting within 16 GB in suitable configurations.

tools thinking 30b
b507b9c2f6ca · 13B
{{ .Prompt }}