147 Downloads Updated 1 week ago
ollama run richardyoung/phi-4-mini-instruct-heretic:Q4_K_M
Updated 1 week ago
1 week ago
778c7720d65d · 2.5GB
Uncensored Phi-4-mini-instruct via Heretic: 5⁄100 refusals (from 99), KL 0.038. 3.8B with 128K context, MIT.
This is an abliterated build of microsoft/Phi-4-mini-instruct, Microsoft’s 3.8B Phi-4-mini, a compact model strong at reasoning and math for its size. Refusal behavior was reduced using the Heretic library with KL-targeted parameters that preserve the model’s coherence.
| Metric | Before | After |
|---|---|---|
| Refusals | 99⁄100 | 5⁄100 |
| Reduction | – | 95% |
| KL Divergence | – | 0.038 |
The very low KL divergence (0.038, far below the 0.5 “damage” threshold) means the model retains essentially all of its original capabilities and coherence.
| Tag | Size | BPW | Notes |
|---|---|---|---|
| latest / Q4_K_M | 2.5 GB | 4.85 | Recommended |
| Q5_K_M | 2.8 GB | 5.68 | Higher quality |
| Q6_K | 3.2 GB | 6.56 | Very high quality |
| Q8_0 | 4.1 GB | 8.5 | Near-lossless |
ollama run richardyoung/phi-4-mini-instruct-heretic # recommended (Q4_K_M)
ollama run richardyoung/phi-4-mini-instruct-heretic:Q8_0 # near-lossless
| VRAM | Recommended tier |
|---|---|
| 4 GB+ | Q4_K_M (2.5 GB file) |
| 6 GB+ | Q8_0 (4.1 GB file) |
This model has reduced safety guardrails. The removal of refusal behavior means it will engage with a wider range of prompts. Use responsibly and in accordance with applicable laws and regulations.
Built & maintained by Richard Young · DeepNeuro