Ollama
Models Docs Pricing
Sign in Download
Models Download Docs Pricing Sign in
⇅
Nemotron 3 33B · Ollama
Search for models on Ollama.
  • nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

    tools thinking 30b

    142.2K  Pulls 11  Tags Updated  2 weeks ago

  • bvassie/nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

    58  Pulls 1  Tag Updated  2 hours ago

  • oamazonasgabriel/nemotron-3.5-lightning

    An open 30B MoE model with ~3B active parameters, packaged by impacte.tech for the execution layer of always-on agents. Uses the official Q4_K_M / IQ4_XS quantizations. Features native tool calling and thinking modes.

    tools thinking

    75  Pulls 2  Tags Updated  6 days ago

  • dlasher/Nemotron-Cascade-2-30B-A3B

    pulled from unsloth/Nemotron-Cascade-2-30B-A3B-GGUF-BF16, requant to IQ4_NL, with combined_en_huge as prompt (English + Coding focus)

    tools thinking

    12  Pulls 1  Tag Updated  2 days ago

  • mirage335/Nemotron-3-Nano-30B-A3B-virtuoso

    Possibly useful for agentic AI systems. Apparently compatible with ~16GB VRAM .

    tools

    837  Pulls 1  Tag Updated  8 months ago

  • samuser3/NVIDIA-Nemotron-3-Nano-4B

    samuser3/NVIDIA-Nemotron-3-Nano-4B is a quantized (Q4_K_M) small language model by NVIDIA, designed for edge deployment with strong reasoning, coding and friendly responses, suitable for local AI tasks like voice assistants and gaming NPCs.

    440  Pulls 1  Tag Updated  5 months ago

© 2026 Ollama
Blog Support