1,244 1 month ago

Decensored Qwen2.5-Coder-3B-Instruct with 3/100 refusals and near-zero KL divergence via Heretic abliteration.

tools
ollama run R4C3R/qwen2.5-coder-3b-heretic

Details

1 month ago

09173b435ee9 · 6.2GB ·

qwen2
·
3.09B
·
BF16
{{- if .Tools }}<|im_start|>system {{ .System }} # Tools You may call one or more functions to assis
{ "repeat_penalty": 1.15, "stop": [ "<|im_start|>", "<|im_end|>", "<

Readme

🏴 Qwen2.5-Coder-3B-Heretic

Heretic Series by RACER IS OP


Qwen2.5-Coder-3B-Heretic is a decensored variant of Qwen’s coding-specialized 3B instruct model. Refusals dropped from 100100 to 3100 with near-zero capability loss (KL 0.016) — the base model’s coding abilities are virtually untouched.

Who this is for — developers who want a compact 3B code model that writes what you ask without refusals. Perfect for local coding agents, code generation, and edge-deployed dev tools.


Quickstart

ollama run R4C3R/qwen2.5-coder-3b-heretic

VRAM Guide (choose your quant)

Quant VRAM When to use
q4_k_m ~1.8 GB Best balance — runs on any system
q5_k_m ~2.1 GB Higher quality, still lightweight
q8_0 ~3.2 GB Maximum quality, CPU-friendly

Have a 4GB GPU? Start with q4_k_m. On CPU only? q8_0 gives you full quality.

Links

Hugging Face · sdad.pro · Heretic

License: qwen-research


Made with love by RACER IS OP — follow for more uncensored models in the Heretic Series.