Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
AntLing
/
Ling-3.0-flash
82
Downloads
Updated
2 weeks ago
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities.
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities.
Cancel
Name
2 models
Size / Usage
Context
Input
Ling-3.0-flash:Q4_K_M
7f4de67dae5e
• 77GB • 256K context window •
Text input • 2 weeks ago
Text input • 2 weeks ago
Ling-3.0-flash:Q4_K_M
77GB
256K
Text
7f4de67dae5e
· 2 weeks ago
Ling-3.0-flash:Q6_K
cac9f3dc6ddd
• 105GB • 256K context window •
Text input • 2 weeks ago
Text input • 2 weeks ago
Ling-3.0-flash:Q6_K
105GB
256K
Text
cac9f3dc6ddd
· 2 weeks ago