64 1 week ago

A decensored multilingual 1B Llama 3.2 model with suppressed refusal behavior. Made by RACER IS OP.

tools
ollama run R4C3R/llama3.2-1b-heretic:q8_0

Details

1 week ago

a0d378bbcd80 · 1.3GB ·

llama
·
1.24B
·
Q8_0

Readme

🏴 Llama-3.2-1B-Heretic

Heretic Series by RACER IS OP


Llama-3.2-1B-Heretic is a decensored variant of Meta’s latest 1B instruct model. Retains multilingual support (English, German, French, Italian, Portuguese, Hindi, Spanish, Thai) while removing refusal guardrails via targeted ablation.

Who this is for — developers who want Meta’s newest small model architecture without censorship. Great for multilingual agents, edge deployment, summarization, and retrieval tasks across 8 languages.


Quickstart

ollama run R4C3R/llama3.2-1b-heretic

VRAM Guide (choose your quant)

Quant VRAM When to use
q4_k_m ~0.7 GB Best balance — runs on any system
q5_k_m ~0.8 GB Higher quality, still lightweight
q8_0 ~1.1 GB Maximum quality, CPU-friendly

On a tight system? Start with q4_k_m. Have room to spare? q8_0 gives you the full model quality.

Links

sdad.pro · Heretic

License: Llama 3.2 Community License


Made with love by RACER IS OP — follow for more uncensored models in the Heretic Series.