39 Downloads Updated 1 month ago
ollama run LoPld/devstral-24b-small-2505-q4_K_S
ollama launch claude --model LoPld/devstral-24b-small-2505-q4_K_S
ollama launch opencode --model LoPld/devstral-24b-small-2505-q4_K_S
ollama launch hermes --model LoPld/devstral-24b-small-2505-q4_K_S
ollama launch openclaw --model LoPld/devstral-24b-small-2505-q4_K_S
This is a quantized version (Q4_K_S) of the original:
Devstral Small – 24B (FP16)
This model was created to:
16 GB VRAM GPU:
~20 GB VRAM GPU :
Compared to FP16:
Run locally:
ollama run LoPld/devstral-24b-small-2505-q4_K_S
Created from:
FROM devstral:24b-small-2505-fp16
Quantized using:
ollama create <model> --quantize q4_K_S -f Modelfile
Original model: Apache 2.0 This quantized version inherits the same license terms.