Ollama model based on Unsloth's UD-Q2_K_XL quantization of Qwen3-235B-A22B-Instruct-2507.
334 Pulls 1 Tag Updated 8 months ago
2-bit Q2_K_XL quantized GGUF version of Qwen3-235B-A22B-Thinking-2507 (MoE, 22B active), optimized for deep reasoning with a 262K context window. Runs on Ollama with ~86.5 GiB RAM.
1,423 Pulls 1 Tag Updated 1 year ago
566 Pulls 1 Tag Updated 1 year ago
(c) https://huggingface.co/unsloth/Qwen3-235B-A22B-Instruct-2507-GGUF, non-thinking model
486 Pulls 1 Tag Updated 1 year ago
286 Pulls 1 Tag Updated 1 year ago
Qwen3-235B-A22B-Instruct-2507 nothink Q8
154 Pulls 1 Tag Updated 1 year ago
dynamic quants 2.0 from unsloth, merged
133 Pulls 1 Tag Updated 1 year ago