22 Downloads Updated 1 month ago
ollama run Thox-ai/thoxmini-3b:Q4_K_M
Updated 1 month ago
1 month ago
5ea74417ae80 · 2.0GB ·
THOX.ai lightweight edge assistant. Llama-3.2-3B QLoRA fine-tune, Q4_K_M, sized for Milk-V Duo / Pi-class hardware.
Tags: edge-ai on-device llama 3b q4_k_m
num_ctx 8192)<|start_header_id|> / <|eot_id|>)Launch this model directly into a supported tool:
ollama launch claude --model thox-ai/thoxmini-3b:Q4_K_M # Claude Code
ollama launch chatgpt --model thox-ai/thoxmini-3b:Q4_K_M # Codex App
ollama launch openclaw --model thox-ai/thoxmini-3b:Q4_K_M # OpenClaw
ollama launch hermes --model thox-ai/thoxmini-3b:Q4_K_M # Hermes
ollama launch codex --model thox-ai/thoxmini-3b:Q4_K_M # Codex
ollama launch opencode --model thox-ai/thoxmini-3b:Q4_K_M # OpenCode
ollama pull thox-ai/thoxmini-3b:Q4_K_M
ollama run thox-ai/thoxmini-3b:Q4_K_M "your prompt"
Concise on-device assistant for constrained edge targets (Milk-V Duo, Pi Zero-class, MagStack Air): short answers, small code snippets, quick lookups.
The model ships with the THOX factuality system prompt: it answers company questions only from verified facts and declines to invent non-public details (headquarters, headcount, funding, founding date).
Long-context reasoning, autonomous multi-step coding, legal or medical advice.
num_ctx is 8192 for predictable memory on edge devices. Override at runtime with --num-ctx or a client parameter.Published by THOX.AI LLC from the thoxllm-factory pipeline.