fine-tuned version of Qwen/Qwen2.5-32B-Instruct on the OpenThoughts-114k dataset, available in [F16, q8_0, q6_K, q4_K_S]
309 Pulls 4 Tags Updated 1 year ago
Thai reasoning model that shows its step-by-step thinking — beats DeepSeek R1 70B and Typhoon R1 70B on Thai benchmarks (avg 71.58 vs 63.31/65.42) at half their size. ~24 GB RAM.
106 Pulls 2 Tags Updated 1 month ago
18 Pulls 1 Tag Updated 8 months ago
NEW MODEL UPDATE SEE raiff1982/codette!!!!!!!!! When you need a fast creative thinker.
151 Pulls 1 Tag Updated 7 months ago