14 Downloads Updated 6 months ago
ollama run grenishrai/ralph1.5-think
Updated 6 months ago
6 months ago
cdd6b2b7a87b · 3.1GB ·
Ralph 1.5 Think is an experimental research model exploring whether a small language model can be fine-tuned to exhibit reasoning capabilities that outperform its non-thinking base model and compete with significantly larger models.
The model is built by fine-tuning Qwen2-1.5B from Alibaba, transforming a standard SLM into a reasoning-oriented variant through targeted supervision and training strategies. The goal is not scale, but efficiency: extracting maximal reasoning performance from minimal parameters.
This work specifically investigates:
GSM8K
Notes:
Ralph 1.5 Think demonstrates early evidence that reasoning-centric fine-tuning can substantially narrow the gap between small and large language models, reinforcing the viability of SLMs for cost-efficient reasoning tasks.