Our most capable model to date, designed for long-horizon work. 70.2% on Terminal-Bench 2.1 at 118B-A8B.
129.5K Pulls 8 Tags Updated 1 week ago
Laguna XS 2.1 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine.
106.2K Pulls 7 Tags Updated 6 days ago
Laguna XS.2 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine.
28.6K Pulls 7 Tags Updated 1 month ago
This model extends LLama-3 8B's context length from 8k to over 1m tokens.
1M Pulls 35 Tags Updated 2 years ago
LFM2.5-8B-A1B, an edge model built for fast, reliable tool calling on consumer hardware.
140.3K Pulls 5 Tags Updated 3 months ago
LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.
1.1M Pulls 6 Tags Updated 6 months ago
Llama 3.1 is a new state-of-the-art model from Meta available in 8B, 70B and 405B parameter sizes.
119.2M Pulls 93 Tags Updated 1 year ago
New state of the art 70B model. Llama 3.3 70B offers similar performance compared to the Llama 3.1 405B model.
4.1M Pulls 14 Tags Updated 1 year ago
Meta's latest open model built for always-on local agents. 30B parameters, licensed under Apache 2.0 and runs on a single GPU — tuned for tool use, long tasks, and failure recovery.
191.9K Pulls 15 Tags Updated 1 week ago
NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
656.9K Pulls 4 Tags Updated 4 months ago
DeepSeek-V4-Pro is a frontier Mixture-of-Experts model with a large context window and three reasoning modes.
385.7K Pulls 2 Tags Updated 3 weeks ago
The most powerful vision-language model in the Qwen model family to date.
5.8M Pulls 57 Tags Updated 10 months ago
LFM2.5 is a new family of hybrid models designed for on-device deployment.
1.3M Pulls 5 Tags Updated 7 months ago
DeepSeek-OCR is a vision-language model that can perform token-efficient OCR.
527.7K Pulls 3 Tags Updated 9 months ago
Olmo is a series of Open language models designed to enable the science of language models. These models are pre-trained on the Dolma 3 dataset and post-trained on the Dolci datasets.
455.5K Pulls 15 Tags Updated 8 months ago
288.8K Pulls 10 Tags Updated 8 months ago
A general-purpose multimodal mixture-of-experts model for production-grade tasks and enterprise workloads.
102.6K Pulls 1 Tag Updated 9 months ago
Meta's Llama 3.2 goes small with 1B and 3B models.
82.6M Pulls 63 Tags Updated 1 year ago
Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models.
36.4M Pulls 58 Tags Updated 11 months ago
Qwen2.5 models are pretrained on Alibaba's latest large-scale dataset, encompassing up to 18 trillion tokens. The model supports up to 128K tokens and has multilingual support.
39.5M Pulls 133 Tags Updated 1 year ago