Back to Model
N

NVIDIA Nemotron 3.5 Lightning

Model Sep 10, 2026

An optimized 30B mixture-of-experts Nemotron release for accelerated inference.

Visit primary link 2 sources attached

NVIDIA Nemotron 3.5 Lightning at a glance

An optimized 30B mixture-of-experts Nemotron release for accelerated inference.

  • Provider: NVIDIA
  • Focus: fast
  • Access: Downloadable checkpoint
  • Model-card license: other

Access and deployment

The checkpoint is published as nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 on Hugging Face. A downloadable checkpoint gives you a deployment option; it does not establish that your hardware has enough memory or that commercial use is allowed.

Evaluation checklist

Measure quality at your target latency and concurrency. For local deployment, check memory requirements at the context length and quantization you intend to use.

Check precision, architecture, GPU memory requirements and whether a release is a base model, reward model or optimized derivative.

Pricing and usage limits

Pricing, quotas and license conditions are not independently verified in this entry. Check the linked official source before choosing a production deployment. For a local deployment, include hardware, hosting and maintenance costs in the comparison.

Sources