1,668
Downloads
Updated
yesterday
Mirror of Prism ML's Bonsai 2 27B ternary GGUF packs (PQ2_0 and PTQ1_0): a 27B hybrid-attention reasoning model in 6.70 GiB or 5.53 GiB, 262K context, vision and tool calling, Apache-2.0. Requires Prism ML's llama.cpp fork — stock Ollama cannot load
vision
tools
thinking
27b