1,668 yesterday

Mirror of Prism ML's Bonsai 2 27B ternary GGUF packs (PQ2_0 and PTQ1_0): a 27B hybrid-attention reasoning model in 6.70 GiB or 5.53 GiB, 262K context, vision and tool calling, Apache-2.0. Requires Prism ML's llama.cpp fork — stock Ollama cannot load

vision tools thinking 27b
c8835af3a6b5 · 212B
Apache-2.0. Copyright Prism ML. Base model: Qwen/Qwen3.8-27B (Apache-2.0). Original weights and GGUF pack: https://huggingface.co/prism-ml/Ternary-Bonsai-2-27B-gguf -- Reproduced here for local research use only.