Abliterated (Uncensored) GLM4.6 Flash
9,459 Pulls 11 Tags Updated 8 months ago
GLM 4.6V Flash 9B model with vision, tools, and hybrid thinking enabled. using custom template to align it to ollama and the recomended sampling settigns by default. using unsloth quants at q4K_M
5,469 Pulls 1 Tag Updated 8 months ago
1,240 Pulls 1 Tag Updated 8 months ago
GLM-4.6V-Flash (9B) is a lightweight model optimized for local deployment and low-latency applications. It scales its context window to 128k tokens in training and achieves SoTA performance in visual understanding among models of similar parameter scales.
471 Pulls 3 Tags Updated 7 months ago
model imported from hf
264 Pulls 1 Tag Updated 8 months ago
640 Pulls 1 Tag Updated 7 months ago