Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
NVIDIA 3B · Ollama
Search for models on Ollama.
  • nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

    tools thinking 30b

    39.3K  Pulls 11  Tags Updated  1 week ago

  • nemotron-cascade-2

    An open 30B MoE model from NVIDIA with 3B activated parameters that delivers strong reasoning and agentic capabilities.

    tools thinking 30b

    142.5K  Pulls 3  Tags Updated  5 months ago

  • nemotron-3-super

    NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.

    tools thinking cloud 120b

    2.9M  Pulls 7  Tags Updated  5 months ago

  • nemotron-3-ultra

    NVIDIA Nemotron 3 Ultra is built for high-throughput reasoning and long-running agent workflows.

    tools thinking cloud

    52.7K  Pulls 1  Tag Updated  2 months ago

  • bvassie/nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

    30  Pulls 1  Tag Updated  5 days ago

  • samuser3/NVIDIA-Nemotron-3-Nano-4B

    samuser3/NVIDIA-Nemotron-3-Nano-4B is a quantized (Q4_K_M) small language model by NVIDIA, designed for edge deployment with strong reasoning, coding and friendly responses, suitable for local AI tasks like voice assistants and gaming NPCs.

    428  Pulls 1  Tag Updated  4 months ago

  • kaylee12102/NVIDIA-Nemotron-3-Nano-4B-GGUF

    48  Pulls 1  Tag Updated  3 months ago

© 2026 Ollama
Blog Contact