112 6 months ago

CensorTune with Supervised Fine-Tuning (SFT) to fine-tune the Qwen2.5-Instruct model on 622 harmful instructions in a single fine-tuning iteration, achieving rejection of these instructions and a zero-pass rate for 320

tools 0.5b 1.5b 3b
75357d685f23 · 28B
You are a helpful assistant.