DeepSeek-V4-Pro is a frontier Mixture-of-Experts model with a large context window and three reasoning modes.
281.6K Pulls 1 Tag Updated 3 months ago
DeepSeek-V4-Flash is a preview of the DeepSeek-V4 series, a Mixture-of-Experts model with 284B total parameters and 13B activated, built for efficient reasoning across a 1M-token context window.
277.7K Pulls 1 Tag Updated 3 months ago
No Oficial 29 t/s en 8 VRAM - Calidad a escala fronteriza como GPT-5.5, DeepSeek-V4-pro y Kimi-K2.6.
373 Pulls 1 Tag Updated 1 week ago
pull from hf.co/Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
3,116 Pulls 1 Tag Updated 2 months ago
DeepSeek-V3-Pruned-Coder-411B is a pruned version of the DeepSeek-V3 reduced from 256 experts to 160 experts, The pruned model is mainly used for code generation.
1,395 Pulls 5 Tags Updated 1 year ago
77 Pulls 1 Tag Updated 1 month ago