DeepSeek-V4.1-Flash is an advanced tool designed to enhance search capabilities, providing users with faster and more accurate results.
61.1K Pulls 1 Tag Updated 3 weeks ago
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash - is an efficient reasoning model distilled using high-quality data from DeepSeek-V4. + vision. ollama v.0.30.0-rc20 +
6,843 Pulls 1 Tag Updated 4 months ago
pull from hf.co/Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
3,866 Pulls 1 Tag Updated 4 months ago
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash - is an efficient reasoning model distilled using high-quality data from DeepSeek-V4.
2,680 Pulls 1 Tag Updated 4 months ago
DeepSeek-V4-Flash 284B MoE (13B active), single-file IQ2_M merged from AtomicChat's GGUF shards so Ollama can pull it. ~33 tok/s on a 128 GB Apple Silicon Mac. Model by DeepSeek (MIT), quantization by AtomicChat.
947 Pulls 1 Tag Updated 1 month ago
To run DeepSeek-V4-Flash in full precision lossless, run Q4 (UD-Q4_K_XL), It is 155GB. vision tools thinking
740 Pulls 1 Tag Updated 2 months ago
DeepSeek-V4-Flash-Fast is a custom agent build based on DeepSeek-V4-Flash (Open Weights, Apache-2.0, MoE, ~285B total / ~20B active). Weights in low-bit quantization for fully CPU-only deployment:
432 Pulls 1 Tag Updated 1 month ago
To run DeepSeek-V4-Flash in full precision lossless, run IQ1 (UD-IQ1_S), It is 82GB. vision tools thinking
378 Pulls 1 Tag Updated 2 months ago
ollama run hf.co/unsloth/DeepSeek-V4-Flash-GGUF:UD-Q4_K_XL
365 Pulls 1 Tag Updated 2 months ago
To run DeepSeek-V4-Flash in full precision lossless, run Q3 (UD-Q3_K_M), It is 129 GB. vision tools thinking
154 Pulls 1 Tag Updated 2 months ago
479 Pulls 1 Tag Updated 1 month ago