Remix.run Logo
adrian_b 43 minutes ago

Yes, there are 5 Nemotron-3.5 models:

https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...

to be used for further training/fine-tuning.

https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...

main model.

https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...

quantized version of the previous.

https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...

this "DFlash" model should be used together with one of the previous two "for lower-latency speculative decoding deployments tuned for low-concurrency data center and workstation workflows".

https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...

like DFlash, the previous model above, but optimized for DGX Spark.