136 Downloads Updated 2 months ago
ollama run R4C3R/qwen2.5-coder-0.5b-heretic
ollama launch claude --model R4C3R/qwen2.5-coder-0.5b-heretic
ollama launch opencode --model R4C3R/qwen2.5-coder-0.5b-heretic
ollama launch hermes --model R4C3R/qwen2.5-coder-0.5b-heretic
ollama launch openclaw --model R4C3R/qwen2.5-coder-0.5b-heretic
Name
4 models
qwen2.5-coder-0.5b-heretic:latest
398MB · 32K context window · Text · 2 months ago
qwen2.5-coder-0.5b-heretic:q4_k_m
398MB · 32K context window · Text · 2 months ago
qwen2.5-coder-0.5b-heretic:q5_k_m
420MB · 32K context window · Text · 2 months ago
qwen2.5-coder-0.5b-heretic:q8_0
531MB · 32K context window · Text · 2 months ago
Heretic Series by RACER IS OP
Qwen2.5-Coder-0.5B-Heretic is a decensored variant of Qwen’s tiny 0.5B coding model. Refusals dropped from 52⁄100 to 8⁄100 (KL 0.12) — the world’s smallest uncensored code model.
Who this is for — developers who want a coding model that runs anywhere. Perfect for CPU-powered coding agents, on-device IDE assistants, and edge AI code generation on a Raspberry Pi or phone.
ollama run R4C3R/qwen2.5-coder-0.5b-heretic
| Quant | VRAM | When to use |
|---|---|---|
q4_k_m |
~0.4 GB | Best balance — runs on anything |
q5_k_m |
~0.5 GB | Higher quality, still tiny |
q8_0 |
~0.6 GB | Maximum quality, CPU-friendly |
Every quant is tiny — pick whichever you like.
q8_0costs almost nothing in VRAM.
Hugging Face · sdad.pro · Heretic
License: qwen-research
Made with love by RACER IS OP — follow for more uncensored models in the Heretic Series.