A new small reasoning model fine-tuned from the Qwen 2.5 3B Instruct model.
262.9K Pulls 5 Tags Updated 1 year ago
SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.
4.1M Pulls 49 Tags Updated 1 year ago
Nex-N2.5-mini (nex-agi) : MoE 35B / ~3B actifs post-trained on Qwen3.5-35B-A3B. Vision, tool-calling, reasoning, 256k context. Quants Q5_K_M, Q4_K_M (latest), Q3_K_M, Q2_K.
276 Pulls 4 Tags Updated 2 weeks ago
TNEA-Advisor is a hyper-localized, data-driven admissions strategist mathematically fused into Qwen 2.5 3B and quantized to 4-bit GGUF for consumer hardware. It bypasses cloud dependencies to deliver hyper-accurate, offline engineering cutoff predictions.
4 Pulls 1 Tag Updated 2 weeks ago
Base model
59 Pulls 3 Tags Updated 1 year ago
1,614 Pulls 5 Tags Updated 1 year ago
A new small reasoning model fine-tuned from the Qwen 2.5 3B Instruct model. I-Quants models, abliterated with uncensored prompt.
993 Pulls 15 Tags Updated 1 year ago
A new small reasoning model fine-tuned from the Qwen 2.5 3B Instruct model. I-Quants models.
133 Pulls 14 Tags Updated 1 year ago
llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-GGUF with Vision
3,165 Pulls 2 Tags Updated 1 month ago
Typhoon2.5 30B A3B - 30B parameters with 3B active Thai / English bilingual LLM build based on Qwen3.
1,756 Pulls 1 Tag Updated 1 year ago
32B reasoning model trained from Qwen2.5-32B-Instruct with 17K data with performance on par with o1 preview.
1,645 Pulls 2 Tags Updated 1 year ago
Llama 3.1 8B Instruct trained on 9,000,000 Claude Opus/Sonnet tokens
1,776 Pulls 1 Tag Updated 2 years ago
fine-tuned version of Qwen/Qwen2.5-32B-Instruct on the OpenThoughts-114k dataset, available in [F16, q8_0, q6_K, q4_K_S]
309 Pulls 4 Tags Updated 1 year ago
This is a 32B reasoning model trained from Qwen2.5-32B-Instruct with 17K data. The performance is on par with o1-preview model on both math and coding.
267 Pulls 2 Tags Updated 1 year ago
219 Pulls 1 Tag Updated 1 year ago
Code-targeted 184-expert prune of Ornith-1.5-35B-A3B (256e→184e, ~27B, A3B active, top-8). Beats the 256-expert teacher on LiveCodeBench v6 (+6.5pp) and HumanEval+ (+2.4pp) at ~28% fewer experts. Native MTP head + vision tower
178 Pulls 39 Tags Updated 3 weeks ago
A fine-tuned Qwen 3.5 35B model for vulnerability finding review, false-positive reduction, and JSON-only security triage.
7,029 Pulls 1 Tag Updated 1 month ago
A text-only, thinking-capable variant of Qwen3.5-35B-A3B — leaner and faster by removing the CLIP vision projector. Based on Unsloth's Q4_K_M quantization of Alibaba's Qwen3.5-35B-A3B.
2,373 Pulls 2 Tags Updated 6 months ago
Deductive Reasoning Qwen 32B is a reinforcement fine-tune of Qwen 2.5 32B Instruct to solve challenging deduction problems
248 Pulls 7 Tags Updated 1 year ago
from gmonsoon/MiniCPM-3B-Turangga-v3-ep50-GGUF
109 Pulls 1 Tag Updated 2 years ago