DeepSeek-R1 is a family of open reasoning models with performance approaching that of leading models, such as O3 and Gemini 2.5 Pro.
92.2M Pulls 35 Tags Updated 1 year ago
A fully open-source family of reasoning models built using a dataset derived by distilling DeepSeek-R1.
1.2M Pulls 15 Tags Updated 1 year ago
A version of the DeepSeek-R1 model that has been post trained to provide unbiased, accurate, and factual information by Perplexity.
419.5K Pulls 9 Tags Updated 1 year ago
Adapted for Cline tool / Roo Code use in VS Code fused model , hybrid of DeepSeekR1 and Qwen2.5 coder, from FuseAI/FuseO1-DeepSeekR1-Qwen2.5-Coder-32B-Preview.
4,937 Pulls 2 Tags Updated 1 year ago
`DeepSeekR1-QwQ-SkyT1-32B-Fusion` is a mixed model that combines the strengths of three powerful Qwen-based models: huihui-ai/DeepSeek-R1-Distill-Qwen-32B-abliterated, huihui-ai/QwQ-32B-Preview-abliterated and huihui-ai/Sky-T1-32B-Preview-abliterated
2,519 Pulls 14 Tags Updated 1 year ago
The quantized versions of the FuseO1 + DeepSeek R1 + QwQ + Sky T1 Fusion Model
520 Pulls 8 Tags Updated 1 year ago
434 Pulls 1 Tag Updated 1 year ago
32 Pulls 1 Tag Updated 4 days ago
A fine-tuned DeepSeek-R1-70B model for CVE-worthiness review, exploit-path analysis, and JSON-only vulnerability triage.
1,905 Pulls 1 Tag Updated 2 weeks ago
DeepSeek-R1-0528-Qwen3-8B
1,155 Pulls 1 Tag Updated 4 months ago
Based on DeepSeek R1 because OpenCode tries to verify on the registry for tool compatibility
236 Pulls 1 Tag Updated 5 months ago
Thai reasoning model that shows its step-by-step thinking — beats DeepSeek R1 70B and Typhoon R1 70B on Thai benchmarks (avg 71.58 vs 63.31/65.42) at half their size. ~24 GB RAM.
88 Pulls 2 Tags Updated 1 month ago
Huggingface link - https://huggingface.co/iradukunda-dev/law-finetuned-DeepSeek-R1-Distill-Qwen-7B
522 Pulls 1 Tag Updated 7 months ago
基于 DeepSeek-R1-Distill-Qwen-1.5B 微调的中文轻量对话模型,自带猫娘口癖与亲昵风格。
531 Pulls 1 Tag Updated 10 months ago
SmallCoder is a compact reasoning-focused coding model, fine-tuned from DeepSeek-R1 1.5B using a code dataset that includes step-by-step reasoning.
381 Pulls 1 Tag Updated 6 months ago
Unsloth's DeepSeek-R1 , I just merged the thing and uploaded it here. This is the full 671b model. MoE Bits:1.58bit Type:UD-IQ1_S Disk Size:131GB Accuracy:Fair Details:MoE all 1.56bit. down_proj in MoE mixture of 2.06/1.56bit
171K Pulls 2 Tags Updated 1 year ago
DeepSeek-R1-Distill models are fine-tuned based on open-source models, using samples generated by DeepSeek-R1. We slightly change their configs and tokenizers. Please use our setting to run these models.
149.1K Pulls 2 Tags Updated 1 year ago
65 Pulls 2 Tags Updated 6 months ago
SmallCoder is a compact reasoning-focused math model, fine-tuned from DeepSeek-R1 1.5B using a math dataset that includes step-by-step reasoning.
63 Pulls 1 Tag Updated 6 months ago
Unsloth's DeepSeek-R1 1.58-bit, I just merged the thing and uploaded it here. This is the full 671b model, albeit dynamically quantized to 1.58bits.
101.6K Pulls 1 Tag Updated 1 year ago