7 Downloads Updated yesterday
ollama run richardyoung/lfm2.5-1.2b-instruct-heretic:Q4_K_M
Updated yesterday
yesterday
9ed3ca55c32b ยท 731MB
Uncensored LFM2.5-1.2B-Instruct via Heretic: 3โ100 refusals (from 98), KL 0.059. Tiny and fast; runs on almost anything.
This is an abliterated build of LiquidAI/LFM2.5-1.2B-Instruct, Liquid AIโs 1.2B edge model with a hybrid convolution + attention architecture, built for fast on-device inference. Refusal behavior was reduced using the Heretic library with KL-targeted parameters that preserve the modelโs coherence.
| Metric | Before | After |
|---|---|---|
| Refusals | 98โ100 | 3โ100 |
| Reduction | โ | 97% |
| KL Divergence | โ | 0.059 |
The very low KL divergence (0.059, far below the 0.5 โdamageโ threshold) means the model retains essentially all of its original capabilities and coherence.
| Tag | Size | BPW | Notes |
|---|---|---|---|
| latest / Q4_K_M | 0.7 GB | 4.85 | Recommended |
| Q5_K_M | 0.8 GB | 5.68 | Higher quality |
| Q6_K | 1.0 GB | 6.56 | Very high quality |
| Q8_0 | 1.2 GB | 8.5 | Near-lossless |
ollama run richardyoung/lfm2.5-1.2b-instruct-heretic # recommended (Q4_K_M)
ollama run richardyoung/lfm2.5-1.2b-instruct-heretic:Q8_0 # near-lossless
| VRAM | Recommended tier |
|---|---|
| 4 GB+ | Q4_K_M (0.7 GB file) |
| 4 GB+ | Q8_0 (1.2 GB file) |
This model has reduced safety guardrails. The removal of refusal behavior means it will engage with a wider range of prompts. Use responsibly and in accordance with applicable laws and regulations.
Built & maintained by Richard Young ยท DeepNeuro