NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.
179.7K Pulls 11 Tags Updated 3 weeks ago
NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
665.1K Pulls 4 Tags Updated 4 months ago
NVIDIA Nemotron 3 Ultra is built for high-throughput reasoning and long-running agent workflows.
102.7K Pulls 1 Tag Updated 3 months ago
NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.
3M Pulls 7 Tags Updated 6 months ago
15 Pulls 1 Tag Updated 1 week ago
samuser3/NVIDIA-Nemotron-3-Nano-4B is a quantized (Q4_K_M) small language model by NVIDIA, designed for edge deployment with strong reasoning, coding and friendly responses, suitable for local AI tasks like voice assistants and gaming NPCs.
474 Pulls 1 Tag Updated 5 months ago
Q4_K_M and BF16 quantizations of NVIDIA Nemotron Nano 9B v2, NVIDIA’s open 9B reasoning model. Q4 quantized locally from the BF16 source and tuned for a single 16 GB GPU card with maximum KV cache.
183 Pulls 3 Tags Updated 1 month ago
88 Pulls 1 Tag Updated 3 weeks ago
53 Pulls 1 Tag Updated 4 months ago
1,381 Pulls 1 Tag Updated 9 months ago
75 Pulls 4 Tags Updated 7 months ago
Llama-3.3-Nemotron-Super-49B-v1.5 is a large language model which is a derivative of Meta Llama-3.3-70B-Instruct. It is a reasoning model that is post trained for reasoning, human chat preferences, and agentic tasks, such as RAG and tool calling.
852 Pulls 4 Tags Updated 10 months ago
AceReason-Nemotron-14B by Nvidia. GGUF from MaoyueOUO/AceReason-Nemotron-14B-GGUF/
165 Pulls 1 Tag Updated 1 year ago
https://huggingface.co/nvidia/Nemotron-Content-Safety-Reasoning-4B
29 Pulls 2 Tags Updated 5 months ago
This is an uncensored version of nvidia/AceReason-Nemotron created with abliteration Edit
2,698 Pulls 9 Tags Updated 1 year ago
reasoning model that is post trained for reasoning, human chat preferences, and tasks, such as RAG and tool calling.
2,986 Pulls 6 Tags Updated 1 year ago
725 Pulls 1 Tag Updated 6 months ago
409 Pulls 3 Tags Updated 1 year ago
126 Pulls 1 Tag Updated 1 year ago
Mistral-NeMo-Minitron-8B-Instruct by nvidia, GGUF file from MaoyueOUO/mistral-nemo-minitron-8b-instruct-GGUF on hf.
263 Pulls 1 Tag Updated 1 year ago