7 1 month ago

ollama run cmsmanhattan/jirack-ultra-14b-q4

Details

1 month ago

030f9bccecad · 9.0GB

qwen2
·
14.8B
·
Q4_K_M
<|User|>{{ .Prompt }}<|Assistant|>
You are JiRack, a helpful and honest AI assistant.
{ "num_ctx": 8192, "repeat_penalty": 1.1, "stop": [ "<|end▁of▁sentence|>

Readme

JiRack Ultra 14B Q4_K_M

Mid-size ~14B model from CMS Manhattan. BitNet-style ternary path, extended tokenizer with Routing, Tool-call and Robotics tags. Production Q4_K_M build for local and cloud inference.

HF: https://huggingface.co/CMSManhattan/JiRackUltra_14b

Run

ollama run cmsmanhattan/jirack-ultra-14b-q4

What it is

  • Stronger general + coding / tool-call capacity than the 1B line
  • Same JiRack tokenizer tags: routing, tools, robotics
  • Q4_K_M quant for practical GPU or high-RAM CPU boxes
  • Fits RAG, agents, and Spring / Java tool-call stacks
  • American ternary stack from CMS Manhattan (New York)

Hardware

Q4_K_M: roughly 8–9 GB on disk.
Recommend 12–16 GB VRAM (GPU) or 24+ GB system RAM (CPU).

Contact

grabko@cmsmanhattan.com
+1 (516) 777-0945
New York, USA

License

MIT on weights unless stated otherwise on the HF card. Commercial terms for UI / enterprise on request.