621 1 month ago

A decensored bilingual 1B mergeset model for roleplay and thinking tasks. Made by RACER IS OP.

tools
ollama run R4C3R/minicpm5-1b-fable5-heretic:q4_k_m

Details

1 month ago

46681bcb3603 · 688MB ·

llama
·
1.08B
·
Q4_K_M

Readme

🏴 MiniCPM5-1B-Fable5-Heretic

Heretic Series by RACER IS OP


MiniCPM5-1B-Fable5-Heretic is a decensored mergeset combining the efficiency of MiniCPM5-1B with the Fable5 thinking dataset. Refusals dropped from 93100 to 3100 with near-zero capability loss (KL 0.02), while keeping the full 128K context and bilingual EN/ZH support.

Who this is for — developers building local agents, roleplay characters, on-device chat, or anything that needs a tiny model to think and respond freely. Runs comfortably on CPU.


Quickstart

ollama run R4C3R/minicpm5-1b-fable5-heretic

VRAM Guide (choose your quant)

Quant VRAM When to use
q4_k_m ~0.7 GB Best balance — runs on any system
q5_k_m ~0.8 GB Higher quality, still lightweight
q8_0 ~1.1 GB Maximum quality, CPU-friendly

On a tight system? Start with q4_k_m. Have room to spare? q8_0 gives you the full model quality.

Links

Hugging Face · sdad.pro · Heretic

License: Apache 2.0


Made with love by RACER IS OP — follow for more uncensored models in the Heretic Series.