79 2 weeks ago

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities.

3e731f420f61 · 44B
{
"temperature": 0.6,
"top_k": 20,
"top_p": 0.95
}