909 5 hours ago

NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.

tools thinking 30b
ollama run nemotron-3.5-lightning:30b-a3b

Details

5 hours ago

e7a64ff15fb1 · 25GB ·

nemotron_h_moe
·
32.9B
·
Q4_K_M
NVIDIA Open Model License Agreement Last Modified: October 24, 2025 This NVIDIA Open Model License A
{ "draft_num_predict": 2, "temperature": 1, "top_p": 0.95 }
{{ .Prompt }}

Readme

NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for the execution layer of always-on agents. It is designed for harnesses like OpenClaw and Hermes Agent—all supported by the NVIDIA NemoClaw open source security and management stack for running always-on AI agents.

Nemotron 3.5 Lightning offers 4x higher throughput and 30% lower task completion time compared to other leading open models of similar size

Use cases

  • Personal agents: Power long-running personal assistants for everyday tasks, like managing your email, calendar, projects and bookings. Run these agents locally on your PC to leverage contextualized intel.
  • Financial services workflows: Power special agents for specialized financial services tasks, such as extracting data from documents, checking policy rules, monitoring risk signals, and preparing structured summaries.
  • Cybersecurity operations: Power special agents for specialized security tasks, such as enriching alerts, classifying incidents, querying logs, validating controls, correlating indicators, and preparing structured findings for analysts.
  • Telecom: Power special agents for autonomous networks and proactive customer care, such as triaging network alarms, optimizing network configurations, and answering billing questions.
  • Retail: Power special agents for enriching product catalogs, resolving inventory and fulfillment exceptions, assisting product discovery, and answering order, return, or loyalty questions

Benchmarks

image.png