Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
hisparshmishra1
/
fast-qwen
65
Downloads
Updated
1 week ago
Qwen2.5-1.5B-Instruct (Q4_K_M), tuned for fast local inference — 5.4x faster than the original with verified quality (perplexity +3%, task scores unchanged).
Qwen2.5-1.5B-Instruct (Q4_K_M), tuned for fast local inference — 5.4x faster than the original with verified quality (perplexity +3%, task scores unchanged).
Cancel
Name
1 model
Size / Usage
Context
Input
fast-qwen:latest
b3b78097de2b
• 1.1GB • 32K context window •
Text input • 1 week ago
Text input • 1 week ago
fast-qwen:latest
1.1GB
32K
Text
b3b78097de2b
· 1 week ago