22 Pulls 2 Tags Updated 2 weeks ago
81 Pulls 1 Tag Updated 3 months ago
14 Pulls 1 Tag Updated 2 months ago
This model is Qwen3-32b-Instruct fine-tuned on ~13k Climsight data which are pairs of prompts/completions extracted from climsight pipeline. Below are all the technical specs of the fine-tuning process along with the evaluation used.
15 Pulls 1 Tag Updated 5 months ago
Revolutionary model with unique thinking/non-thinking modes, delivering superior reasoning performance with seamless mode switching for any task.
181 Pulls 1 Tag Updated 9 months ago
137 Pulls 1 Tag Updated 7 months ago
Zhi-Create-Qwen3-32B is a fine-tuned model derived from Qwen/Qwen3-32B, with a focus on enhancing creative writing capabilities. Through careful optimization, the model shows promising improvements in creative writing performance, as evaluated using the W
805 Pulls 2 Tags Updated 1 year ago
包含2个量化版本GGUF:Qwen3-32B-Q8_0,Qwen3-32B-Q5_K_M
804 Pulls 2 Tags Updated 1 year ago
这是unsloth的Q4动态量化版本,精度第一的量化版本!Unsloth Dynamic 2.0 实现了卓越的准确性,并超越了其他领先的量化模型。
824 Pulls 1 Tag Updated 1 year ago
这是unsloth的Q8动态量化版本,精度第一的量化版本!Unsloth Dynamic 2.0 实现了卓越的准确性,并超越了其他领先的量化模型。
225 Pulls 1 Tag Updated 1 year ago
GGUF Ruadapt версии модели Qwen/Qwen3-32B (квантизованная версия Q4_K_M)
142 Pulls 1 Tag Updated 1 year ago
Model with tweaked params optimized for agent use
60 Pulls 1 Tag Updated 1 year ago
24 Pulls 1 Tag Updated 1 year ago
9 Pulls 1 Tag Updated 1 year ago
6 Pulls 1 Tag Updated 1 year ago
Qwen3.8-27B in Q4_K_M quantization (32GB+ VRAM required). Dense 27.8B parameters with hybrid attention for long context (256K tokens). Apache 2.0 license. Ideal for high-VRAM setups (RTX 5090/4090 dual, M4 Ultra, etc.).
855 Pulls 4 Tags Updated 2 days ago
32k token limit, for running on mac book pro with 16GB (~ 10 GB available for AI without anything else running)
12 Pulls 1 Tag Updated 1 week ago
Qwen-SEA-LION-v4-32B-IT is a multilingual model which has been pretrained and instruct-tuned for the Southeast Asia region. Developed by AI Singapore and funded by National Research Foundation, Singapore.
344 Pulls 3 Tags Updated 10 months ago
Recommended for transcribing and summarizing text from screenshots.
147 Pulls 1 Tag Updated 8 months ago
Qwen2.5 coder tools model can work with Cline (prev. Claude Dev). Update 0.5b, 1.5b, 3b, 7b, 14b, 32b coder models.
60K Pulls 15 Tags Updated 1 year ago