Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Falcon · Ollama
Search for models on Ollama.
  • falcon3

    A family of efficient AI models under 10B parameters performant in science, math, and coding through innovative training techniques.

    1b 3b 7b 10b

    2.6M  Pulls 17  Tags Updated  1 year ago

  • falcon

    A large language model built by the Technology Innovation Institute (TII) for use in summarization, text generation, and chat bots.

    7b 40b 180b

    1.2M  Pulls 38  Tags Updated  2 years ago

  • falcon2

    Falcon2 is an 11B parameters causal decoder-only model built by TII and trained over 5T tokens.

    11b

    545.2K  Pulls 17  Tags Updated  2 years ago

  • granite3.2

    Granite-3.2 is a family of long-context AI models from IBM Granite fine-tuned for thinking capabilities.

    tools 2b 8b

    459.2K  Pulls 9  Tags Updated  1 year ago

  • deepseek-v4-flash

    DeepSeek-V4-Flash is the official release of DeepSeek-V4-Flash, built for efficient reasoning across a 1M-token context window, outperforming DeepSeek-V4-Pro (Preview).

    tools thinking cloud

    469.6K  Pulls 2  Tags Updated  1 month ago

  • lfm2.5

    LFM2.5-8B-A1B, an edge model built for fast, reliable tool calling on consumer hardware.

    tools thinking 8b

    157.2K  Pulls 5  Tags Updated  3 months ago

  • mistrallite

    MistralLite is a fine-tuned model based on Mistral with enhanced capabilities of processing long contexts.

    7b

    528.3K  Pulls 17  Tags Updated  2 years ago

  • ExpedientFalcon/Qwen3-4B-UD-Q5_K_XL

    Qwen3-4B Q5_K_XL Unsloth UD 2.0 adaptively quantized model, much better for coding than vanilla Q4_K_M quants without taking up the VWAM of an 8-bit Q8_0 model. From https://huggingface.co/unsloth/Qwen3-4B-GGUF/tree/main

    tools

    447.2K  Pulls 1  Tag Updated  1 year ago

  • ExpedientFalcon/qwen3-reranker

    1,704  Pulls 5  Tags Updated  1 year ago

  • GFalcon-UA/dolphin3-r1-mistral

    https://huggingface.co/cognitivecomputations/Dolphin3.0-R1-Mistral-24B

    1,554  Pulls 1  Tag Updated  1 year ago

  • GFalcon-UA/nous-hermes-2-vision

    llava-NousResearch_Nous-Hermes-2-Vision-GGUF_Q4_0 with function calling

    1,770  Pulls 1  Tag Updated  2 years ago

  • GFalcon-UA/dolphin3-mistral

    https://huggingface.co/cognitivecomputations/Dolphin3.0-Mistral-24B

    693  Pulls 1  Tag Updated  1 year ago

  • GFalcon-UA/dolphin3-llama3.1

    https://huggingface.co/cognitivecomputations/Dolphin3.0-Llama3.1-8B-GGUF

    588  Pulls 1  Tag Updated  1 year ago

  • Hudson/falcon-mamba-instruct

    The latest and greatest model of the Falcon LLM series.

    606  Pulls 1  Tag Updated  1 year ago

  • ExpedientFalcon/qwen3-1.7b-autocomplete

    tools thinking

    319  Pulls 1  Tag Updated  1 year ago

  • ExpedientFalcon/qwen2.5-coder-3b-instruct-q6_k

    This repo contains the instruction-tuned 3B Qwen2.5-Coder model in the GGUF Format: https://huggingface.co/Qwen/Qwen2.5-Coder-3B-Instruct-GGUF/tree/main

    tools

    295  Pulls 1  Tag Updated  1 year ago

  • GFalcon-UA/ReaderLM-v2

    Jina AI ReaderLM-v2

    315  Pulls 1  Tag Updated  1 year ago

  • ExpedientFalcon/qwen3-embedding

    embedding

    239  Pulls 4  Tags Updated  1 year ago

  • viraatdas/Falcon3-7B-Instruct-1.58bit

    hugging face model: https://huggingface.co/tiiuae/Falcon3-7B-Instruct-1.58bit-GGUF#usage

    292  Pulls 1  Tag Updated  1 year ago

  • bilel_cherif/falcon3-tools

    Falcon 3 10b for tool usage and function call

    tools

    145  Pulls 2  Tags Updated  1 year ago

© 2026 Ollama
Blog Support