191 Downloads Updated 2 days ago
ollama run richardyoung/deepseek-r1-0528-qwen3-8b-heretic:Q5_K_M
Updated 2 days ago
2 days ago
1c1d59bff4ae · 5.9GB
Uncensored DeepSeek-R1-0528-Qwen3-8B via Heretic: 4⁄100 refusals (from 89), KL 0.041. Reasoning model, MIT.
This is an abliterated build of deepseek-ai/DeepSeek-R1-0528-Qwen3-8B, DeepSeek’s R1-0528 reasoning distilled into Qwen3-8B: a reasoning model that emits chains before answering. Refusal behavior was reduced using the Heretic library with KL-targeted parameters that preserve the model’s coherence.
| Metric | Before | After |
|---|---|---|
| Refusals | 89⁄100 | 4⁄100 |
| Reduction | – | 96% |
| KL Divergence | – | 0.041 |
The very low KL divergence (0.041, far below the 0.5 “damage” threshold) means the model retains essentially all of its original capabilities and coherence.
| Tag | Size | BPW | Notes |
|---|---|---|---|
| latest / Q4_K_M | 5.0 GB | 4.85 | Recommended |
| Q5_K_M | 5.9 GB | 5.68 | Higher quality |
| Q6_K | 6.7 GB | 6.56 | Very high quality |
| Q8_0 | 8.7 GB | 8.5 | Near-lossless |
ollama run richardyoung/deepseek-r1-0528-qwen3-8b-heretic # recommended (Q4_K_M)
ollama run richardyoung/deepseek-r1-0528-qwen3-8b-heretic:Q8_0 # near-lossless
| VRAM | Recommended tier |
|---|---|
| 6 GB+ | Q4_K_M (5.0 GB file) |
| 10 GB+ | Q8_0 (8.7 GB file) |
This model has reduced safety guardrails. The removal of refusal behavior means it will engage with a wider range of prompts. Use responsibly and in accordance with applicable laws and regulations.
Built & maintained by Richard Young · DeepNeuro