| ▲ | adrian_b 43 minutes ago | |
Yes, there are 5 Nemotron-3.5 models: https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-... to be used for further training/fine-tuning. https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-... main model. https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-... quantized version of the previous. https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-... this "DFlash" model should be used together with one of the previous two "for lower-latency speculative decoding deployments tuned for low-concurrency data center and workstation workflows". https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-... like DFlash, the previous model above, but optimized for DGX Spark. | ||