To run DeepSeek-V4-Flash in full precision lossless, run IQ1 (UD-IQ1_S), It is 82GB. vision tools thinking
292 Pulls 1 Tag Updated 1 month ago
DeepSeek-V4-Flash 284B MoE (13B active), single-file IQ2_M merged from AtomicChat's GGUF shards so Ollama can pull it. ~33 tok/s on a 128 GB Apple Silicon Mac. Model by DeepSeek (MIT), quantization by AtomicChat.
585 Pulls 1 Tag Updated 3 weeks ago