DeepSeek-V4-Flash is a preview of the DeepSeek-V4 series, a Mixture-of-Experts model with 284B total parameters and 13B activated, built for efficient reasoning across a 1M-token context window.
368.4K Pulls 3 Tags Updated 2 weeks ago
DeepSeek-V4-Pro is a frontier Mixture-of-Experts model with a large context window and three reasoning modes.
332K Pulls 3 Tags Updated 2 days ago
An open-source Mixture-of-Experts code language model that achieves performance comparable to GPT4-Turbo in code-specific tasks.
3M Pulls 64 Tags Updated 1 year ago
ollama run hf.co/unsloth/DeepSeek-V4-Flash-GGUF:UD-Q4_K_XL
245 Pulls 1 Tag Updated 1 month ago
DeepSeek-V4-Flash 284B MoE (13B active), single-file IQ2_M merged from AtomicChat's GGUF shards so Ollama can pull it. ~33 tok/s on a 128 GB Apple Silicon Mac. Model by DeepSeek (MIT), quantization by AtomicChat.
24 Pulls 1 Tag Updated yesterday
DeepSeek-V4-Flash-Fast is a custom agent build based on DeepSeek-V4-Flash (Open Weights, Apache-2.0, MoE, ~285B total / ~20B active). Weights in low-bit quantization for fully CPU-only deployment:
188 Pulls 1 Tag Updated 1 week ago
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash - is an efficient reasoning model distilled using high-quality data from DeepSeek-V4. + vision. ollama v.0.30.0-rc20 +
5,213 Pulls 1 Tag Updated 2 months ago
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash - is an efficient reasoning model distilled using high-quality data from DeepSeek-V4.
2,206 Pulls 1 Tag Updated 3 months ago
22 Pulls 1 Tag Updated yesterday
5 Pulls 1 Tag Updated yesterday
2,592 Pulls 15 Tags Updated 1 week ago
90 Pulls 1 Tag Updated 1 week ago
53 Pulls 1 Tag Updated 3 weeks ago
deepseek v4 pro on max reasoning using the official system prompt
48 Pulls 1 Tag Updated 2 weeks ago
22 Pulls 1 Tag Updated 2 weeks ago
783 Pulls 15 Tags Updated 1 week ago
22 Pulls 1 Tag Updated 3 weeks ago
664 Pulls 1 Tag Updated 2 months ago
238 Pulls 1 Tag Updated 1 month ago
343 Pulls 1 Tag Updated 2 months ago