901 Downloads Updated 2 weeks ago
ollama run R4C3R/qwen2.5-coder-7b-instruct-heretic:q4_k_m
ollama launch claude --model R4C3R/qwen2.5-coder-7b-instruct-heretic:q4_k_m
ollama launch opencode --model R4C3R/qwen2.5-coder-7b-instruct-heretic:q4_k_m
ollama launch hermes --model R4C3R/qwen2.5-coder-7b-instruct-heretic:q4_k_m
ollama launch openclaw --model R4C3R/qwen2.5-coder-7b-instruct-heretic:q4_k_m
Name
4 models
qwen2.5-coder-7b-instruct-heretic:q4_k_m
4.7GB · 32K context window · Text · 2 weeks ago
qwen2.5-coder-7b-instruct-heretic:q5_k_m
5.4GB · 32K context window · Text · 2 weeks ago
qwen2.5-coder-7b-instruct-heretic:q6_k
6.3GB · 32K context window · Text · 2 weeks ago
qwen2.5-coder-7b-instruct-heretic:q8_0
8.1GB · 32K context window · Text · 2 weeks ago
Heretic Series by RACER IS OP
Qwen2.5-Coder-7B-Heretic is a decensored variant of Qwen’s industry-leading 7B code model. The strongest uncensored coding model in the Heretic Series — writes what you ask, no refusals.
Who this is for — developers who want a powerful 7B code model without guardrails. Perfect for local coding agents, code review, and unrestricted code generation. Runs well on 8 GB+ GPUs with Q4 quantization.
ollama run R4C3R/qwen2.5-coder-7b-instruct-heretic
| Quant | VRAM | When to use |
|---|---|---|
q4_k_m |
~4.2 GB | Best balance — fits 6-8 GB GPUs |
q5_k_m |
~4.8 GB | Higher quality, needs 8 GB GPU |
q8_0 |
~7.2 GB | Maximum quality, needs 10 GB+ |
On an 8 GB GPU?
q4_k_mis your sweet spot. Have 12 GB+? Goq5_k_morq8_0for full quality.
Hugging Face · sdad.pro · Heretic
License: qwen-research
Made with love by RACER IS OP — follow for more uncensored models in the Heretic Series.