310 Downloads Updated 2 months ago
ollama run R4C3R/qwen2.5-0.5b-heretic
ollama launch claude --model R4C3R/qwen2.5-0.5b-heretic
ollama launch opencode --model R4C3R/qwen2.5-0.5b-heretic
ollama launch hermes --model R4C3R/qwen2.5-0.5b-heretic
ollama launch openclaw --model R4C3R/qwen2.5-0.5b-heretic
Name
4 models
qwen2.5-0.5b-heretic:latest
398MB · 32K context window · Text · 2 months ago
qwen2.5-0.5b-heretic:q4_k_m
398MB · 32K context window · Text · 2 months ago
qwen2.5-0.5b-heretic:q5_k_m
420MB · 32K context window · Text · 2 months ago
qwen2.5-0.5b-heretic:q8_0
531MB · 32K context window · Text · 2 months ago
Heretic Series by RACER IS OP
Qwen2.5-0.5B-Heretic is a decensored variant of Qwen’s ultra-compact 0.5B instruct model. The smallest general-purpose uncensored model in the Heretic Series — runs on virtually anything.
Who this is for — developers who need the absolute smallest uncensored model possible. Ideal for IoT devices, Raspberry Pi, CPU inference, and memory-constrained edge deployment.
ollama run R4C3R/qwen2.5-0.5b-heretic
| Quant | VRAM | When to use |
|---|---|---|
q4_k_m |
~0.4 GB | Best balance — runs on anything |
q5_k_m |
~0.5 GB | Higher quality, still tiny |
q8_0 |
~0.6 GB | Maximum quality, CPU-friendly |
Every quant is tiny — pick whichever you like.
q8_0costs almost nothing in VRAM.
License: qwen-research
Made with love by RACER IS OP — follow for more uncensored models in the Heretic Series.