NVIDIA Nemotron 3.5 Lightning at a glance
An optimized 30B mixture-of-experts Nemotron release for accelerated inference.
- Provider: NVIDIA
- Focus: fast
- Access: Downloadable checkpoint
- Model-card license: other
Access and deployment
The checkpoint is published as nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 on Hugging Face. A downloadable checkpoint gives you a deployment option; it does not establish that your hardware has enough memory or that commercial use is allowed.
Evaluation checklist
Measure quality at your target latency and concurrency. For local deployment, check memory requirements at the context length and quantization you intend to use.
Check precision, architecture, GPU memory requirements and whether a release is a base model, reward model or optimized derivative.
Pricing and usage limits
Pricing, quotas and license conditions are not independently verified in this entry. Check the linked official source before choosing a production deployment. For a local deployment, include hardware, hosting and maintenance costs in the comparison.