Models
Docs
Pricing
Sign in
Download
Models
Download
Docs
Pricing
Sign in
oamazonasgabriel
/
nemotron-nano-9b-v2
37
Downloads
Updated
6 days ago
Q4_K_M and BF16 quantizations of NVIDIA Nemotron Nano 9B v2, NVIDIA’s open 9B reasoning model. Q4 quantized locally from the BF16 source and tuned for a single 16 GB GPU card with maximum KV cache.
Q4_K_M and BF16 quantizations of NVIDIA Nemotron Nano 9B v2, NVIDIA’s open 9B reasoning model. Q4 quantized locally from the BF16 source and tuned for a single 16 GB GPU card with maximum KV cache.
Cancel
tools
thinking
Name
3 models
Size / Usage
Context
Input
nemotron-nano-9b-v2:latest
007175a6bfc3
• 18GB • 1M context window •
Text input • 1 week ago
Text input • 1 week ago
nemotron-nano-9b-v2:latest
18GB
1M
Text
007175a6bfc3
· 1 week ago
nemotron-nano-9b-v2:bf16
007175a6bfc3
• 18GB • 1M context window •
Text input • 1 week ago
Text input • 1 week ago
nemotron-nano-9b-v2:bf16
18GB
1M
Text
007175a6bfc3
· 1 week ago
nemotron-nano-9b-v2:q4-km-16gbGPU
1be50069450b
• 6.5GB • 1M context window •
Text input • 6 days ago
Text input • 6 days ago
nemotron-nano-9b-v2:q4-km-16gbGPU
6.5GB
1M
Text
1be50069450b
· 6 days ago