DeepSeek-V4-Flash is a preview of the DeepSeek-V4 series, a Mixture-of-Experts model with 284B total parameters and 13B activated, built for efficient reasoning across a 1M-token context window.
346.8K Pulls 3 Tags Updated 1 week ago
DeepSeek-V4-Pro is a frontier Mixture-of-Experts model with a large context window and three reasoning modes.
318K Pulls 1 Tag Updated 3 months ago
DeepSeek-OCR is a vision-language model that can perform token-efficient OCR.
509.2K Pulls 3 Tags Updated 8 months ago
DeepSeek-V3.1-Terminus is a hybrid model that supports both thinking mode and non-thinking mode.
719.8K Pulls 7 Tags Updated 10 months ago
DeepSeek-R1 is a family of open reasoning models with performance approaching that of leading models, such as O3 and Gemini 2.5 Pro.
91.2M Pulls 35 Tags Updated 1 year ago
DeepSeek Coder is a capable coding model trained on two trillion code and natural language tokens.
4.5M Pulls 102 Tags Updated 2 years ago
A fully open-source family of reasoning models built using a dataset derived by distilling DeepSeek-R1.
1.2M Pulls 15 Tags Updated 1 year ago
A version of the DeepSeek-R1 model that has been post trained to provide unbiased, accurate, and factual information by Perplexity.
415.1K Pulls 9 Tags Updated 1 year ago
An upgraded version of DeekSeek-V2 that integrates the general and coding abilities of both DeepSeek-V2-Chat and DeepSeek-Coder-V2-Instruct.
284.2K Pulls 7 Tags Updated 1 year ago
A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.
3.8M Pulls 5 Tags Updated 1 year ago
DeepSeek-V4-Flash-Fast is a custom agent build based on DeepSeek-V4-Flash (Open Weights, Apache-2.0, MoE, ~285B total / ~20B active). Weights in low-bit quantization for fully CPU-only deployment:
73 Pulls 1 Tag Updated 4 days ago
To run DeepSeek-V4-Flash in full precision lossless, run Q4 (UD-Q4_K_XL), It is 155GB. vision tools thinking
519 Pulls 1 Tag Updated 1 week ago
No Oficial 29 t/s en 8 VRAM - Calidad a escala fronteriza como GPT-5.5, DeepSeek-V4-pro y Kimi-K2.6.
784 Pulls 1 Tag Updated 3 weeks ago
To run DeepSeek-V4-Flash in full precision lossless, run IQ1 (UD-IQ1_S), It is 82GB. vision tools thinking
145 Pulls 1 Tag Updated 1 week ago
ollama run hf.co/unsloth/DeepSeek-V4-Flash-GGUF:UD-Q4_K_XL
224 Pulls 1 Tag Updated 4 weeks ago
To run DeepSeek-V4-Flash in full precision lossless, run Q3 (UD-Q3_K_M), It is 129 GB. vision tools thinking
55 Pulls 1 Tag Updated 1 week ago
Thai reasoning model that shows its step-by-step thinking — beats DeepSeek R1 70B and Typhoon R1 70B on Thai benchmarks (avg 71.58 vs 63.31/65.42) at half their size. ~24 GB RAM.
41 Pulls 2 Tags Updated 2 weeks ago
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash - is an efficient reasoning model distilled using high-quality data from DeepSeek-V4. + vision. ollama v.0.30.0-rc20 +
5,010 Pulls 1 Tag Updated 2 months ago
pull from hf.co/Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
3,287 Pulls 1 Tag Updated 2 months ago
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash - is an efficient reasoning model distilled using high-quality data from DeepSeek-V4.
2,135 Pulls 1 Tag Updated 2 months ago