A fine-tuned version of Deepseek-R1-Distilled-Qwen-1.5B that surpasses the performance of OpenAI’s o1-preview with just 1.5B parameters on popular math evaluations.
1.3M Pulls 5 Tags Updated 1 year ago
1,967 Pulls 3 Tags Updated 1 year ago
基于 DeepSeek-R1-Distill-Qwen-1.5B 微调的中文轻量对话模型,自带猫娘口癖与亲昵风格。
531 Pulls 1 Tag Updated 10 months ago
SmallCoder is a compact reasoning-focused coding model, fine-tuned from DeepSeek-R1 1.5B using a code dataset that includes step-by-step reasoning.
379 Pulls 1 Tag Updated 6 months ago
SmallCoder is a compact reasoning-focused math model, fine-tuned from DeepSeek-R1 1.5B using a math dataset that includes step-by-step reasoning.
62 Pulls 1 Tag Updated 6 months ago
DeepSeek-R1-Distill-Qwen-1.5B
4,988 Pulls 1 Tag Updated 1 year ago
1,552 Pulls 5 Tags Updated 1 year ago
382 Pulls 1 Tag Updated 1 year ago
265 Pulls 1 Tag Updated 1 year ago
DeepScaleR-1.5B-Preview is a language model fine-tuned from DeepSeek-R1-Distilled-Qwen-1.5B using distributed reinforcement learning (RL)
110 Pulls 1 Tag Updated 1 year ago
86 Pulls 1 Tag Updated 1 year ago
1.5b model
77 Pulls 1 Tag Updated 1 year ago
这是一个测试模型,模型不会算数了
17 Pulls 1 Tag Updated 1 year ago
15 Pulls 1 Tag Updated 1 year ago
Unsloth's DeepSeek-R1 1.58-bit, I just merged the thing and uploaded it here. This is the full 671b model, albeit dynamically quantized to 1.58bits.
101.6K Pulls 1 Tag Updated 1 year ago
deepseek r1 1.5b dengan bahasa indonesia
318 Pulls 1 Tag Updated 1 year ago
Brainstorm 40x by DavidAU available in [F16, q8_0, q6_K, q4_K_S]
293 Pulls 4 Tags Updated 1 year ago
deepseek-r-11.5b-nolimits - is ai model that have no limit, no restrictions and fully easy to use and deploy on CPU 8GB RAM.
312 Pulls 2 Tags Updated 1 month ago