1,060 Downloads Updated 7 months ago
ollama run richardyoung/qwen3-8b-abliterated:Q4_K_M
Updated 7 months ago
7 months ago
6d7230a92342 ยท 5.0GB ยท
Abliterated (uncensored) version of Qwen3-8B for unrestricted conversations, reasoning, and creative writing.
This model is an abliterated version of Qwen/Qwen3-8B, with refusal behavior reduced through targeted weight modification. The abliteration process uses the Heretic library with conservative parameters to preserve model coherence while removing restrictions. It retains the full Qwen3 feature set, including seamless switching between thinking mode (for complex reasoning, math, and code) and non-thinking mode (for efficient, general-purpose dialogue).
| Metric | Before | After |
|---|---|---|
| Refusals | TBD | TBD |
| Reduction | , | TBD |
| KL Divergence | , | TBD |
Refusal metrics pending re-measurement.
A low KL divergence (< 1.0) indicates the model maintains its original capabilities and coherence.
/think and /no_think| Tag | Size | BPW | Notes |
|---|---|---|---|
| latest / Q4_K_M | 5.0 GB | 4.85 | Recommended, balanced quality/size |
Only the
Q4_K_Mbuild is currently published. The BPW guide below shows where additional quants sit if released:
Quant BPW Profile IQ3_M 3.66 Smallest, for low VRAM IQ4_XS 4.25 Great quality/size balance Q4_K_M 4.85 Recommended Q5_K_M 5.68 Higher quality Q6_K 6.56 Very high quality Q8_0 8.5 Near-lossless
ollama run richardyoung/qwen3-8b-abliterated
ollama run richardyoung/qwen3-8b-abliterated:latest
| VRAM | Performance |
|---|---|
| 6 GB | Slow, may swap to CPU |
| 8 GB | Good performance |
| 12 GB+ | Excellent performance |
This model has reduced safety guardrails. The reduction of refusal behavior means the model will engage with a wider range of prompts. Use responsibly and in accordance with applicable laws and regulations.
Built & maintained by Richard Young ยท DeepNeuro