Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
NVIDIA Nemotron · Ollama
Search for models on Ollama.
  • nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

    tools thinking 30b

    179.7K  Pulls 11  Tags Updated  3 weeks ago

  • nemotron3

    NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.

    vision tools thinking 33b

    665.1K  Pulls 4  Tags Updated  4 months ago

  • nemotron-3-ultra

    NVIDIA Nemotron 3 Ultra is built for high-throughput reasoning and long-running agent workflows.

    tools thinking cloud

    102.7K  Pulls 1  Tag Updated  3 months ago

  • nemotron-3-super

    NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.

    tools thinking cloud 120b

    3M  Pulls 7  Tags Updated  6 months ago

  • brnpistone/NVIDIA-Nemotron-3-Nano-30-AgentCoder-q4-k-m

    tools thinking

    15  Pulls 1  Tag Updated  1 week ago

  • samuser3/NVIDIA-Nemotron-3-Nano-4B

    samuser3/NVIDIA-Nemotron-3-Nano-4B is a quantized (Q4_K_M) small language model by NVIDIA, designed for edge deployment with strong reasoning, coding and friendly responses, suitable for local AI tasks like voice assistants and gaming NPCs.

    474  Pulls 1  Tag Updated  5 months ago

  • oamazonasgabriel/nemotron-nano-9b-v2

    Q4_K_M and BF16 quantizations of NVIDIA Nemotron Nano 9B v2, NVIDIA’s open 9B reasoning model. Q4 quantized locally from the BF16 source and tuned for a single 16 GB GPU card with maximum KV cache.

    tools thinking

    183  Pulls 3  Tags Updated  1 month ago

  • bvassie/nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

    tools

    88  Pulls 1  Tag Updated  3 weeks ago

  • kaylee12102/NVIDIA-Nemotron-3-Nano-4B-GGUF

    53  Pulls 1  Tag Updated  4 months ago

  • mirage335/NVIDIA-Nemotron-Nano-9B-v2-virtuoso

    1,381  Pulls 1  Tag Updated  9 months ago

  • FieldMouse/bartowski_NVIDIA-Nemotron-Nano

    75  Pulls 4  Tags Updated  7 months ago

  • rnogy/Nvidia_Llama-3_3-Nemotron-Super-49B-v1_5

    Llama-3.3-Nemotron-Super-49B-v1.5 is a large language model which is a derivative of Meta Llama-3.3-70B-Instruct. It is a reasoning model that is post trained for reasoning, human chat preferences, and agentic tasks, such as RAG and tool calling.

    852  Pulls 4  Tags Updated  10 months ago

  • Maoyue/AceReason-Nemotron-14B-Q4_K_M

    AceReason-Nemotron-14B by Nvidia. GGUF from MaoyueOUO/AceReason-Nemotron-14B-GGUF/

    165  Pulls 1  Tag Updated  1 year ago

  • hung_alex/Nemotron-Content-Safety-Reasoning-4B

    https://huggingface.co/nvidia/Nemotron-Content-Safety-Reasoning-4B

    29  Pulls 2  Tags Updated  5 months ago

  • huihui_ai/acereason-nemotron-abliterated

    This is an uncensored version of nvidia/AceReason-Nemotron created with abliteration Edit

    7b 14b

    2,698  Pulls 9  Tags Updated  1 year ago

  • MHKetbi/nvidia_Llama-3.3-Nemotron-Super-49B-v1

    reasoning model that is post trained for reasoning, human chat preferences, and tasks, such as RAG and tool calling.

    2,986  Pulls 6  Tags Updated  1 year ago

  • fuukeidaisuki/nvidia-nemotron-nano-9b-v2-japanese

    tools thinking

    725  Pulls 1  Tag Updated  6 months ago

  • avil/nvidia-llama-3.1-nemotron-nano-4b-v1.1

    tools

    409  Pulls 3  Tags Updated  1 year ago

  • avil/nvidia-llama-3.1-nemotron-nano-4b-v1.1-thinking

    126  Pulls 1  Tag Updated  1 year ago

  • Maoyue/mistral-nemo-minitron-8b-instruct

    Mistral-NeMo-Minitron-8B-Instruct by nvidia, GGUF file from MaoyueOUO/mistral-nemo-minitron-8b-instruct-GGUF on hf.

    tools

    263  Pulls 1  Tag Updated  1 year ago

© 2026 Ollama
Blog Support