OLMo 2 is a new family of 7B and 13B models trained on up to 5T tokens. These models are on par with or better than equivalently sized fully open models, and competitive with open-weight models such as Llama 3.1 on English academic benchmarks.
3.8M Pulls 9 Tags Updated 1 year ago
Olmo is a series of Open language models designed to enable the science of language models. These models are pre-trained on the Dolma 3 dataset and post-trained on the Dolci datasets.
460.5K Pulls 15 Tags Updated 9 months ago
291.9K Pulls 10 Tags Updated 9 months ago
OpenCoder is an open and reproducible code LLM family which includes 1.5B and 8B models, supporting chat in English and Chinese languages.
651.3K Pulls 9 Tags Updated 1 year ago
Orca 2 is built by Microsoft research, and are a fine-tuned version of Meta's Llama 2 models. The model is designed to excel particularly in reasoning.
946.3K Pulls 33 Tags Updated 2 years ago
NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
665.1K Pulls 4 Tags Updated 4 months ago
LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.
1.1M Pulls 6 Tags Updated 7 months ago
DeepSeek-OCR is a vision-language model that can perform token-efficient OCR.
537.5K Pulls 3 Tags Updated 10 months ago
New state of the art 70B model. Llama 3.3 70B offers similar performance compared to the Llama 3.1 405B model.
4.1M Pulls 14 Tags Updated 1 year ago
The TinyLlama project is an open endeavor to train a compact 1.1B Llama model on 3 trillion tokens.
5.8M Pulls 36 Tags Updated 2 years ago
Open-source medical large language model adapted from Llama 2 to the medical domain.
891.8K Pulls 22 Tags Updated 2 years ago
OLMo model
578 Pulls 2 Tags Updated 2 years ago
OLMo 2 32B Instruct March 2025
222 Pulls 2 Tags Updated 1 year ago
A Mixture-of-Experts LLM with 1B active and 7B total parameters.
217 Pulls 6 Tags Updated 1 year ago
Copy of https://huggingface.co/allenai/OLMoE-1B-7B-0125-Instruct-GGUF
201 Pulls 1 Tag Updated 1 year ago
This model is a fine-tuned version of allenai/OLMo-1B-hf on the HuggingFaceH4/ultrachat_200k dataset
54 Pulls 1 Tag Updated 2 years ago
Base model
26 Pulls 3 Tags Updated 1 year ago
20 Pulls 1 Tag Updated 1 year ago
8 Pulls 9 Tags Updated 1 year ago
7 Pulls 3 Tags Updated 1 year ago