Llama 2 based model fine tuned on an Orca-style dataset. Originally called Free Willy.
946.4K Pulls 49 Tags Updated 2 years ago
DeepSeek-V4-Flash is the official release of DeepSeek-V4-Flash, built for efficient reasoning across a 1M-token context window, outperforming DeepSeek-V4-Pro (Preview).
449.4K Pulls 2 Tags Updated 1 month ago
An update to Mistral Small that improves on function calling, instruction following, and less repetition errors.
2.5M Pulls 5 Tags Updated 1 year ago
1,736 Pulls 4 Tags Updated 5 months ago
490 Pulls 3 Tags Updated 4 months ago
329 Pulls 3 Tags Updated 4 months ago
296 Pulls 6 Tags Updated 4 months ago
155 Pulls 2 Tags Updated 4 months ago
55 Pulls 4 Tags Updated 4 months ago
54 Pulls 1 Tag Updated 5 months ago
762 Pulls 1 Tag Updated 6 months ago
314 Pulls 1 Tag Updated 6 months ago
93.3K Pulls 3 Tags Updated 12 months ago
63 Pulls 1 Tag Updated 10 months ago
Alternative quantization levels, no fine-tuning
909 Pulls 11 Tags Updated 1 year ago
277 Pulls 11 Tags Updated 1 year ago
5 Pulls 1 Tag Updated 1 year ago
DeepSeek-V4-Pro is a frontier Mixture-of-Experts model with a large context window and three reasoning modes.
399.2K Pulls 2 Tags Updated 4 weeks ago
SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
4M Pulls 49 Tags Updated 1 year ago
The powerful family of models by Nous Research that excels at scientific discussion and coding tasks.
1.1M Pulls 33 Tags Updated 2 years ago