1 1 month ago

ollama run cmsmanhattan/jirack-ultra-32b-q3

Details

1 month ago

abd05f7c2277 · 16GB

qwen2
·
32.8B
·
Q3_K_M
<|User|>{{ .Prompt }}<|Assistant|>
You are JiRack, a helpful and honest AI assistant.
{ "num_ctx": 8192, "repeat_penalty": 1.1, "stop": [ "<|end▁of▁sentence|>

Readme

JiRack Ultra 32B Q3_K_M

Large ~32B model from CMS Manhattan. BitNet-style ternary path, extended tokenizer with Routing, Tool-call and Robotics tags. Production Q3_K_M build for balanced size and quality.

HF: https://huggingface.co/CMSManhattan/JiRackUltra_32b

Run

ollama run cmsmanhattan/jirack-ultra-32b-q3

What it is

  • Flagship mid/large size in the Ultra line — stronger reasoning, coding, and tool use than 14B
  • Same JiRack tokenizer tags: routing, tools, robotics
  • Q3_K_M quant — tighter than Q4, better quality retention than Q2
  • Fits serious RAG, multi-agent, and Spring / Java tool-call stacks
  • American ternary stack from CMS Manhattan (New York)

Hardware

Q3_K_M: roughly 14–16 GB on disk.
Recommend 20–24 GB VRAM (GPU) or 32–40 GB system RAM (CPU).

Contact

grabko@cmsmanhattan.com
+1 (516) 777-0945
New York, USA

License

MIT on weights unless stated otherwise on the HF card. Commercial terms for UI / enterprise on request.