Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Router · Ollama
Search for models on Ollama.
  • phi

    Phi-2: a 2.7B language model by Microsoft Research that demonstrates outstanding reasoning and language understanding capabilities.

    2.7b

    1.5M  Pulls 18  Tags Updated  2 years ago

  • cogito

    Cogito v1 Preview is a family of hybrid reasoning models by Deep Cogito that outperform the best available open models of the same size, including counterparts from LLaMA, DeepSeek, and Qwen across most standard benchmarks.

    tools 3b 8b 14b 32b 70b

    2.1M  Pulls 20  Tags Updated  1 year ago

  • stablelm-zephyr

    A lightweight chat model allowing accurate, and responsive output without requiring high-end hardware.

    3b

    520.1K  Pulls 17  Tags Updated  2 years ago

  • rubinmaximilian/Monk-Router-Gemma4e2b

    A lightweight, hardware-aware router built on Gemma 4 E2B. It acts as a dispatcher for local AI setups, automatically deciding whether a prompt should run on edge hardware (like a Jetson Nano), a local GPU, or the cloud based on task complexity.

    vision tools thinking

    195  Pulls 1  Tag Updated  3 months ago

  • nick040791/NeuronRouter

    1. Classify the request, -- 2. Reason whether to answer directly, inspect files, run tools, or escalate. -- 3. Choose which worker gets the Task. -- 4. Expose only the minimum tool set.

    vision tools thinking

    159  Pulls 1  Tag Updated  4 months ago

  • rubinmaximilian/Monk-Router-phi4mini

    A high-precision, hardware-aware router built on Phi-4 Mini (3.8B). It acts as a dispatcher for local AI setups, automatically deciding whether a prompt should run on edge hardware (like a Jetson Nano), a local GPU, or the cloud based on task complexity.

    tools

    71  Pulls 1  Tag Updated  3 months ago

  • YuriiFominYoung/openrouter-fusion

    OpenRouter Fusion is a multi-model orchestration system

    cloud

    38  Pulls 1  Tag Updated  3 weeks ago

  • Adiptify/Router

    tools

    18  Pulls 1  Tag Updated  10 months ago

  • gleison/fastapi-router

    🚀 Setup do FastAPI Router Generator com Ollama

    thinking

    16  Pulls 1  Tag Updated  1 year ago

  • mattslarson/IngressingPromptstoRouterapril2026

    tools

    1  Tag Updated  2 months ago

  • dcostenco/prism-coder

    Local-first AI tool router. 4 sizes (2B/4B/14B/32B). 100% routing accuracy (BFCL, 115 cases x 3 seeds). 97% of traffic stays local.

    vision tools thinking 2b 4b 9b 14b 27b 32b

    231  Pulls 6  Tags Updated  1 month ago

  • fauxpaslife/arch-router

    A specialized 1.5B parameter model for intelligent routing between multiple LLMs based on domain and action preferences.

    1.5b

    316  Pulls 1  Tag Updated  5 months ago

  • peeyush_16/supra-router

    Fast 51M-parameter Llama-based router that classifies prompts and dispatches them to the right downstream model or tool.

    6  Pulls 1  Tag Updated  3 weeks ago

  • dcostenco/prism-ide

    Local-first AI tool router + coder. 4 sizes. 100% routing accuracy. 22/22 coding eval. 97% free. Beats Opus.

    14b 32b

    38  Pulls 2  Tags Updated  1 month ago

  • kengetrw/router

    tools thinking cloud

    17  Pulls 1  Tag Updated  7 months ago

  • mithunrajm06/gemma-router-3-1b-router

    8  Pulls 1  Tag Updated  7 months ago

  • mithunrajm06/qwen-router-0.6b-gguf

    4  Pulls 1  Tag Updated  7 months ago

  • jiayuan1/router_llm_v2

    9  Pulls 1  Tag Updated  1 year ago

  • jiayuan1/router_llm

    4  Pulls 1  Tag Updated  1 year ago

  • jiayuan1/router_30

    2  Pulls 1  Tag Updated  1 year ago

© 2026 Ollama
Blog Contact