Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Qwen: Qwen3.8 Flash · Ollama
Search for models on Ollama.
  • orcarouter/Qwen3.8-Flash-Next-Uncensored

    Qwen3.8-Flash-Next tensor-level abliterated— a latest Qwen4-architecture MoE (~177B total / ~6B active) with vision, reasoning, tool-calling, and 262K context. MLX 4/6/8-bit for Apple Silicon. Research use only.

    vision tools thinking

    6,695  Pulls 4  Tags Updated  1 week ago

  • wcamaralopes/qwen3.8-flash-iq3xs

    Base model https://huggingface.co/unsloth/Qwen3.8-Flash-Next-GGUF

    106  Pulls 1  Tag Updated  6 days ago

  • metalspork/qwen3.8-flash-next-ud

    Unsloth Dynamic Quants for qwen3.8-flash-next

    vision

    1,322  Pulls 9  Tags Updated  1 week ago

  • edtorre/qwen3.6-hermes

    Qwen 3.6 27B (Q4_K_M) optimized for Hermes Agent — 64K context, 8192 max tokens, MTP for speed, flash attention + Q8 KV cache.

    vision tools thinking

    893  Pulls 1  Tag Updated  2 months ago

© 2026 Ollama
Blog Support