2,087 2 months ago

Decensored Qwen2.5-3B-Instruct with 2/100 refusals via Heretic abliteration. General-purpose 3B model for local use.

tools
ollama run R4C3R/qwen2.5-3b-heretic

Details

2 months ago

084a47c8753d Β· 6.2GB

qwen2
Β·
3.09B
Β·
BF16
{{- if .Tools }}<|im_start|>system {{ .System }} # Tools You may call one or more functions to assis
{ "repeat_penalty": 1.15, "stop": [ "<|im_start|>", "<|im_end|>", "<

Readme

🏴 Qwen2.5-3B-Heretic

Heretic Series by RACER IS OP


Qwen2.5-3B-Heretic is a decensored variant of Qwen’s popular 3B instruct model. Refusals dropped from 96⁄100 to 2⁄100 with minimal capability loss (KL 0.13) β€” one of the most effective ablations in the series. Runs comfortably on CPU with Q4 quantization.

Who this is for β€” developers who want a compact 3B general-purpose model that answers directly instead of refusing. Great for local agents, roleplay, edge deployment, or anything blocked by RLHF over-refusal.


Quickstart

ollama run R4C3R/qwen2.5-3b-heretic

VRAM Guide (choose your quant)

Quant VRAM When to use
q4_k_m ~1.8 GB Best balance β€” runs on any system
q5_k_m ~2.1 GB Higher quality, still lightweight
q8_0 ~3.2 GB Maximum quality, CPU-friendly

Have a 4GB GPU? Start with q4_k_m. On CPU only? q8_0 gives you full quality.

Links

Hugging Face Β· sdad.pro Β· Heretic

License: qwen-research


Made with love by RACER IS OP β€” follow for more uncensored models in the Heretic Series.