27 3 hours ago

Qwen3.8-Flash-Next tensor-level abliterated— a latest Qwen4-architecture MoE (~177B total / ~6B active) with vision, reasoning, tool-calling, and 262K context. MLX 4/6/8-bit for Apple Silicon. Research use only.

vision tools thinking
{
"bos_token_id": 248044,
"do_sample": true,
"eos_token_id": [
248046,
248044
],
"pad_token_id": 248044,
"temperature": 1.0,
"top_k": 20,
"top_p": 0.95
}