-
glm-4.6v-flash
578 Pulls 1 Tag Updated 10 months ago
-
wedlm-7b-base
Tencent WeDLM-7B-Base converted to GGUF (Q4_K_M). A text-diffusion model based on Qwen2.5 architecture, optimized for efficient parallel decoding.
219 Pulls 1 Tag Updated 9 months ago
-
qwen3-ro-rel-extract
Specialized Romanian Relation Extraction (Qwen3 4B). Structured JSON output. Tags: f16 (high-precision) & q4_k_m (fast).
87 Pulls 3 Tags Updated 9 months ago