GLM 4.6V Flash 9B model with vision, tools, and hybrid thinking enabled. using custom template to align it to ollama and the recomended sampling settigns by default. using unsloth quants at q4K_M
5,432 Pulls 1 Tag Updated 7 months ago
GLM-4.6V-Flash (9B) is a lightweight model optimized for local deployment and low-latency applications. It scales its context window to 128k tokens in training and achieves SoTA performance in visual understanding among models of similar parameter scales.
469 Pulls 3 Tags Updated 7 months ago
LongWriter-glm4-9b is trained based on glm-4-9b, and is capable of generating 10,000+ words at once.
3,247 Pulls 1 Tag Updated 1 year ago
GLM-4-0414 series models.
957 Pulls 1 Tag Updated 1 year ago
GLM-4-Z1-9B-0414 9B小参数模型,其整体表现已处于同尺寸开源模型中的领先水平,此模型使用Ollama v0.6.6版本制作生成,需要Ollama v0.6.6及以上版本才能运行推理。
549 Pulls 1 Tag Updated 1 year ago
Base model
132 Pulls 3 Tags Updated 1 year ago
Generic all purpose model. Occasionally may have notable logic, usually Llama-3_3-Nemotron-Super-49B-v1_5 is preferred.
126 Pulls 1 Tag Updated 9 months ago
125 Pulls 1 Tag Updated 9 months ago