Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
64% Positive
Analyzed from 563 words in the discussion.
Trending Topics
#nvidia#model#nemotron#models#open#https#lightning#training#huggingface#version
Discussion Sentiment
Analyzed from 563 words in the discussion.
Trending Topics
Discussion (17 Comments)Read Original on HackerNews
While it looks "behind" the qwen equivalent model on most benchmarks, a few personal notes:
- nemotron models feel to me a bit less benchmaxxed / "stubborn". That means that they generalise a bit better, or can be tasked to solve similar but not quite identical task types to the training data (something that's hard to do w/ qwen/ds models)
- nemotron series are also open training (w/ open training recipes and some training data public)
- nvda will have an incentive to continue this kind of releases, even if other parties slowly abandon the open release of models. Whatever other incentives 3rd party labs have (i.e. meta, goog w/ gemma, the chinese labs that IPOd, etc) nvda will always want to sell hardware so their incentive to keep pushing open models is evident and will likely continue "forever".
[1] - https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...
https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...
to be used for further training/fine-tuning.
https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...
main model.
https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...
quantized version of the previous.
https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...
this "DFlash" model should be used together with one of the previous two "for lower-latency speculative decoding deployments tuned for low-concurrency data center and workstation workflows".
https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-...
like DFlash, the previous model above, but optimized for DGX Spark.
I actually copied the link from NVIDIA's Technical Blog post:
- https://developer.nvidia.com/blog/nvidia-nemotron-3-5-lightn...
You can also try the model via a free API endpoint from Openrouter, would be interesting to see if it's the BF16 or NVFP4 version:
- https://openrouter.ai/nvidia/nemotron-3.5-lightning:free
This! It's literally in their best interest for open-weights models to succeed
Hopefully Qwen follows up their 3.8 launch with a new 35b-a3b
I find these releases are bad taste.
Make 1TB DGX priced affordably, not some crap model for people to waste time on.
I suggested to my local restaurant to pay me for eating there and pointed out it was a win-win situation because it would increase the restaurant's ARR. I was kicked out.
Of course I made a blog post afterwards about the restaurant being negative luddites and backwards.