A family of efficient AI models under 10B parameters performant in science, math, and coding through innovative training techniques.
2.6M Pulls 17 Tags Updated 1 year ago
A large language model built by the Technology Innovation Institute (TII) for use in summarization, text generation, and chat bots.
1.1M Pulls 38 Tags Updated 2 years ago
Falcon2 is an 11B parameters causal decoder-only model built by TII and trained over 5T tokens.
528K Pulls 17 Tags Updated 2 years ago
Granite-3.2 is a family of long-context AI models from IBM Granite fine-tuned for thinking capabilities.
447.2K Pulls 9 Tags Updated 1 year ago
LFM2.5-8B-A1B, an edge model built for fast, reliable tool calling on consumer hardware.
96.7K Pulls 5 Tags Updated 2 months ago
MistralLite is a fine-tuned model based on Mistral with enhanced capabilities of processing long contexts.
512K Pulls 17 Tags Updated 2 years ago
Qwen3-Coder-Next is a coding-focused language model from Alibaba's Qwen team, optimized for agentic coding workflows and local development.
1.9M Pulls 3 Tags Updated 5 months ago
18 Pulls 1 Tag Updated 5 months ago
417 Pulls 5 Tags Updated 9 months ago
Qwen3-4B Q5_K_XL Unsloth UD 2.0 adaptively quantized model, much better for coding than vanilla Q4_K_M quants without taking up the VWAM of an 8-bit Q8_0 model. From https://huggingface.co/unsloth/Qwen3-4B-GGUF/tree/main
447.1K Pulls 1 Tag Updated 1 year ago
5 Pulls 1 Tag Updated 4 months ago
1,627 Pulls 5 Tags Updated 12 months ago
1,902 Pulls 21 Tags Updated 1 year ago
https://huggingface.co/cognitivecomputations/Dolphin3.0-R1-Mistral-24B
1,374 Pulls 1 Tag Updated 1 year ago
llava-NousResearch_Nous-Hermes-2-Vision-GGUF_Q4_0 with function calling
1,759 Pulls 1 Tag Updated 2 years ago
https://huggingface.co/cognitivecomputations/Dolphin3.0-Llama3.1-8B-GGUF
555 Pulls 1 Tag Updated 1 year ago
https://huggingface.co/cognitivecomputations/Dolphin3.0-Mistral-24B
540 Pulls 1 Tag Updated 1 year ago
The latest and greatest model of the Falcon LLM series.
583 Pulls 1 Tag Updated 1 year ago
306 Pulls 1 Tag Updated 1 year ago
This repo contains the instruction-tuned 3B Qwen2.5-Coder model in the GGUF Format: https://huggingface.co/Qwen/Qwen2.5-Coder-3B-Instruct-GGUF/tree/main
271 Pulls 1 Tag Updated 1 year ago