A family of efficient AI models under 10B parameters performant in science, math, and coding through innovative training techniques.
2.6M Pulls 17 Tags Updated 1 year ago
A large language model built by the Technology Innovation Institute (TII) for use in summarization, text generation, and chat bots.
1.2M Pulls 38 Tags Updated 2 years ago
Falcon2 is an 11B parameters causal decoder-only model built by TII and trained over 5T tokens.
545.2K Pulls 17 Tags Updated 2 years ago
Granite-3.2 is a family of long-context AI models from IBM Granite fine-tuned for thinking capabilities.
459.2K Pulls 9 Tags Updated 1 year ago
DeepSeek-V4-Flash is the official release of DeepSeek-V4-Flash, built for efficient reasoning across a 1M-token context window, outperforming DeepSeek-V4-Pro (Preview).
469.6K Pulls 2 Tags Updated 1 month ago
LFM2.5-8B-A1B, an edge model built for fast, reliable tool calling on consumer hardware.
157.2K Pulls 5 Tags Updated 3 months ago
MistralLite is a fine-tuned model based on Mistral with enhanced capabilities of processing long contexts.
528.3K Pulls 17 Tags Updated 2 years ago
Qwen3-4B Q5_K_XL Unsloth UD 2.0 adaptively quantized model, much better for coding than vanilla Q4_K_M quants without taking up the VWAM of an 8-bit Q8_0 model. From https://huggingface.co/unsloth/Qwen3-4B-GGUF/tree/main
447.2K Pulls 1 Tag Updated 1 year ago
1,704 Pulls 5 Tags Updated 1 year ago
https://huggingface.co/cognitivecomputations/Dolphin3.0-R1-Mistral-24B
1,554 Pulls 1 Tag Updated 1 year ago
llava-NousResearch_Nous-Hermes-2-Vision-GGUF_Q4_0 with function calling
1,770 Pulls 1 Tag Updated 2 years ago
https://huggingface.co/cognitivecomputations/Dolphin3.0-Mistral-24B
693 Pulls 1 Tag Updated 1 year ago
https://huggingface.co/cognitivecomputations/Dolphin3.0-Llama3.1-8B-GGUF
588 Pulls 1 Tag Updated 1 year ago
The latest and greatest model of the Falcon LLM series.
606 Pulls 1 Tag Updated 1 year ago
319 Pulls 1 Tag Updated 1 year ago
This repo contains the instruction-tuned 3B Qwen2.5-Coder model in the GGUF Format: https://huggingface.co/Qwen/Qwen2.5-Coder-3B-Instruct-GGUF/tree/main
295 Pulls 1 Tag Updated 1 year ago
Jina AI ReaderLM-v2
315 Pulls 1 Tag Updated 1 year ago
239 Pulls 4 Tags Updated 1 year ago
hugging face model: https://huggingface.co/tiiuae/Falcon3-7B-Instruct-1.58bit-GGUF#usage
292 Pulls 1 Tag Updated 1 year ago
Falcon 3 10b for tool usage and function call
145 Pulls 2 Tags Updated 1 year ago