Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Qwen Plus · Ollama
Search for models on Ollama.
  • qwen2.5

    Qwen2.5 models are pretrained on Alibaba's latest large-scale dataset, encompassing up to 18 trillion tokens. The model supports up to 128K tokens and has multilingual support.

    tools 0.5b 1.5b 3b 7b 14b 32b 72b

    35.6M  Pulls 133  Tags Updated  1 year ago

  • glassesglitchstudio/GulmezCetinerMax

    Monolithic AGI Special Operations Model. Powered by Qwen 3.6 Plus. Developed by Niko Software under CEO Berkay." (Veya Türkçe istersen: "Niko Software tarafından CEO Berkay liderliğinde geliştirilmiş, Qwen 3.6 Plus tabanlı monolitik AGI modeli.

    tools

    15  Pulls 1  Tag Updated  2 months ago

  • pleasecech/qwen3.6-plus

    2,578  Pulls 1  Tag Updated  3 months ago

  • r4c3r/qwen2.5-3b-heretic

    Decensored Qwen2.5-3B-Instruct with 2/100 refusals via Heretic abliteration. General-purpose 3B model for local use.

    tools

    1,322  Pulls 1  Tag Updated  1 week ago

  • sam860/dolphin3-qwen2.5

    Dolphin 3.0 Qwen 2.5 🐬 - A powerful, customizable AI model for local use.

    tools 1.5b 3b

    3,005  Pulls 9  Tags Updated  1 year ago

  • Abhisar2006/Ananya

    Ananya is a light-weight LLM based on the Qwen2 architecture with 7B Parameters. She is more AI assistant than a generic LLM. You can use her as a wrapper around complex programs.

    tools

    65  Pulls 1  Tag Updated  1 year ago

  • prakasharyan/qwen-arabic

    This project fine-tunes the Qwen2-1.5B model for Arabic language tasks using Quantized LoRA (QLoRA).

    1,117  Pulls 1  Tag Updated  1 year ago

  • treyrowell1826/qwen3-pinion

    Qwen3-1.7b Fully sft on MaggiePie 300k filtered, then lora adapter merged. Various quants available.

    10  Pulls 4  Tags Updated  4 months ago

  • nuibang/Cline_FuseO1-DeepSeekR1-Qwen2.5-Coder-32B-Preview

    Adapted for Cline tool / Roo Code use in VS Code fused model , hybrid of DeepSeekR1 and Qwen2.5 coder, from FuseAI/FuseO1-DeepSeekR1-Qwen2.5-Coder-32B-Preview.

    tools

    4,913  Pulls 2  Tags Updated  1 year ago

  • SetneufPT/Qwen3.5-9B-Coder_Q4_256k_ABL_16GB-GPU

    Custom model for coding with agents to use locally with 16gb GPUs (working very fine...)

    vision tools thinking

    527  Pulls 1  Tag Updated  2 weeks ago

  • SetneufPT/Qwen3.6-27B-CODER-MTP_Q4_105k_24GB-GPU

    Custom model for coding with agents to use locally with 24gb GPUs - BEST FOR OPENCODE!

    vision tools thinking

    629  Pulls 1  Tag Updated  2 weeks ago

  • parable/qwen3-fable

    Parable is Qwen3 fine-tuned on Claude Fable 5 and GPT-5.5 agent traces. Tool use, planning, and thinking for local agents. Sibling: Granite line at parable/granite4.1-fable.

    tools thinking 4b 8b

    355  Pulls 10  Tags Updated  6 days ago

  • ayushshah/qwen3-1.7b-chat

    Preference-aligned conversational assistant based on Qwen3-1.7B-Base. Fine-tuned using UltraChat and optimized with DPO on UltraFeedback, Intel Orca and Capybara

    43  Pulls 6  Tags Updated  1 week ago

  • SetneufPT/Qwen3.6-27B-MTP_Q3_32K_16GB-GPU

    Custom model for coding with agents to use locally with 16gb GPUs (working fine...)

    1,611  Pulls 1  Tag Updated  1 month ago

  • oamazonasgabriel/qwen3.5-9b

    A coding-optimized configuration of Qwen3.5-9B designed for 16 GB single-GPU hardware. The model uses the official Q4_K_M quantization (~6.6 GB weights), leaving ~9 GB headroom for KV cache — enabling 32K+ context windows comfortably.

    vision tools thinking

    706  Pulls 1  Tag Updated  1 month ago

  • Nehc/Qwen3-Coder

    Weights, parameters and templates are taken from unsloth. Tools and MCP servers work correctly. Tested on Continue for VS Code

    tools 30b 480b

    1,069  Pulls 4  Tags Updated  11 months ago

  • lastmass/Qwen3_Medical_GRPO

    A specialized medical model fine-tuned from Qwen3 using SFT and Group Relative Policy Optimization (GRPO) for advanced clinical case analysis.

    423  Pulls 1  Tag Updated  1 year ago

  • renchris/qwen3-coder

    Alibaba's performant long context models for agentic and coding tasks — quantized and optimized in GGUF format by Unsloth for fast local inference on consumer devices.

    tools

    2,784  Pulls 1  Tag Updated  12 months ago

© 2026 Ollama
Blog Contact