Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Qwen3.5 Max · Ollama
Search for models on Ollama.
  • mdq100/qwen3.5

    Custom Qwen3.5 variants optimized for 128GB unified memory systems, such as AMD Ryzen AI Max+ 395. On Windows 11, GPU is limited to 96GB (32GB reserved for OS/CPU), requiring context window capped at 131072 tokens (128K) to fit within GPU memory limits.

    vision tools thinking

    431  Pulls 2  Tags Updated  5 months ago

  • srchmnmichael/qwen3.5-9B-uncensored

    Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF

    1,021  Pulls 4  Tags Updated  4 days ago

  • Omoeba/qwen3-2507-abliterated-max

    maximum 256k context length for coding and other long-context tasks

    tools 30b

    871  Pulls 1  Tag Updated  11 months ago

  • Havenlon/Execution-Boundary-Qwen35-27B-Q4_K_M

    A Havenlon-focused Qwen3.5 27B model for deep reasoning about execution boundaries, Adversarial Completeness, AI Agent control, evidence chains, and real-world execution.

    2  Pulls 3  Tags Updated  1 week ago

  • zdolny/qwen3-coder58k-tools

    qwen3-coder with tools calling, context 58k to match full memory 31GB on RTX5090

    tools

    586  Pulls 1  Tag Updated  1 year ago

  • fredrezones55/Qwen3.5-APEX

    Qwen3.5-35B-A3B APEX GGUF -- A Novel MoE-Aware Mixed-Precision Quantization Technique Brought to you by the LocalAI team -- the creators of LocalAI the open-source AI engine that runs any model - LLMs, vision, image - on any hardware.

    vision tools thinking

    755  Pulls 2  Tags Updated  4 months ago

  • lucloner/wen36-35b-uncensored-1m

    Qwen3.6-35B-A3B Uncensored is a Mixture-of-Experts model with 35.5B total parameters and 3B activated, extended to 1M-token context with the official MTP speculative-decoding layer, delivering 1.5x decode speedup on Ollama, with native thinking and tools.

    tools thinking

    347  Pulls 2  Tags Updated  2 weeks ago

  • rafw007/qwen36-a3b-claude-coder

    Qwen3.6-35B-A3B MoE coding agent for Claude Code / Codex / opencode, 64K context, native tool-calling, honest tool use, safety guardrails intact

    vision tools thinking

    3,559  Pulls 2  Tags Updated  2 months ago

  • oamazonasgabriel/qwen3.6-35b-a3b

    A memory-efficient model configuration of Qwen3.6-35B-A3B using an upstream imatrix-calibrated IQ4_XS quantization and q4_0 KV cache. Designed for 24 GB VRAM

    tools thinking

    2,022  Pulls 1  Tag Updated  2 months ago

© 2026 Ollama
Blog Support