| ▲ | johnnyApplePRNG 3 hours ago |
| Their strategy is to prevent open models from proliferating, so their massive investments in these AI frauds are not completely unwound. That's my take, at least. Nemotron is TERRIBLE, and purposefully so. It must be. They cannot be THAT BAD at training AI models. I don't believe it. |
|
| ▲ | terribleperson 2 hours ago | parent | next [-] |
| I think almost the exact opposite. They want open models. Without credible open models, they only have a few customers, and those customers have leverage against nvidia. With open models, they have tons of customers, and nvidia has all the leverage. Smaller customers are also less able to develop their own hardware and threaten NVidia's business. |
| |
| ▲ | SequoiaHope 2 hours ago | parent | next [-] | | Sure but it also means these companies nvidia buys don’t work with AMD and other competitors anymore. There are myriad motivations and it’s not all one or the other but this is a classic component of Silicon Valley acquisition strategies. | |
| ▲ | SR2Z 39 minutes ago | parent | prev [-] | | They want LEVERAGE. They can't ride the hyperscaler gravy train forever; at some point between Google, AMD, and Apple NVIDIA is going to lose its monopoly on serving large customers. At that point, it would be useful if a few open models existed which were only a couple of months behind the frontier. But it's very important that the open models never be TOO good, because the AI companies are buying compute on the assumption that their software will add value. If it becomes a commodity business with frontier open models, NVIDIA won't be able to get away with such a crazy markup. | | |
| ▲ | Wowfunhappy 19 minutes ago | parent [-] | | > If it becomes a commodity business with frontier open models, NVIDIA won't be able to get away with such a crazy markup. Of course they would. You need hardware to run the model. nVidia sets the floor. |
|
|
|
| ▲ | mark_l_watson 3 hours ago | parent | prev | next [-] |
| Have you used Nemotron-3.5-lightening? I don’t use it as much as Poolside’s (excellent!!) Laguna XS 2.1 6bit, but the new Nemotron model is good. I think NVIDIA does want small open models running on-prem to explode as a market! Lots of smaller GPU installations for companies who wisely want on-prem inference. Of course NVIDIA will also keep making a ton of money selling to hyper scalers, but not forever: Chinese chips are getting better, Google, Microsoft, Amazon, etc. designing their own inference chips. NVIDIA is handling this brilliantly. |
|
| ▲ | odo1242 an hour ago | parent | prev | next [-] |
| I think this is a classic [Commoditize Your Complement](https://news.ycombinator.com/item?id=17047348) - Nvidia wants open models because their business is hardware and it's complement is AI models, so they want AI models to be commoditized so that hardware is the industry with leverage. OpenAI/Anthropic/etc want closed models so that the AI model development/data has leverage over the hardware providers. |
|
| ▲ | MangoCoffee 2 hours ago | parent | prev | next [-] |
| Nvidia is trying to increase its customer base. Look at the Mag 7, Amazon, Google, Microsoft, and even Meta are all working on their own inference chips. I don't think they can completely ditch Nvidia for LLM training, but they can make their own chips for inference. I believe that's also why OpenAI made its own. No one wants to pay the Nvidia tax |
| |
| ▲ | tayo42 24 minutes ago | parent [-] | | If you make your own chip don't you need to make your own software stack too. Which is what I thought kept everyone using Nvidia. |
|
|
| ▲ | radium3d 3 hours ago | parent | prev | next [-] |
| I believe NVIDIA's long term strategy will be to pivot from the data center to the public, and the public will utilize open models on NVIDIA hardware at home. This will come after the RAMpocalypse completes (when the new fabrication plants (China, Tesla/SpaceXAI/Intel) fully ramp up and start selling their RAM for cheap in the next few years). Data centers will be for training mostly. |
| |
| ▲ | SV_BubbleTime 2 hours ago | parent | next [-] | | There’s a lot about this that would make sense. But, not really at the current technologies. Kimi and GLM are fucking awesome, but I don’t have 3TB of VRAM to run them, and I don’t expect to even when ram prices drop. So now you’re back to the scaling issue before talking about power and compute distribution. | |
| ▲ | charcircuit 2 hours ago | parent | prev [-] | | Why do you think the public wants to self host models over using a cheaper solution hosted in the cloud? | | |
| ▲ | Geezus_42 an hour ago | parent [-] | | Subscriptions are almost never actually cheaper. | | |
| ▲ | charcircuit 14 minutes ago | parent [-] | | It is cheaper to subscribe to AI for $500 a year for the rest of your life than to buy a machine capable of running the a current frontier model with no subagents. |
|
|
|
|
| ▲ | InsideOutSanta 2 hours ago | parent | prev | next [-] |
| I doubt Nvidia wants to be fully dependent on the success of two highly unprofitable companies that could implode at any moment. It makes much more sense for them to commoditize LLMs so that their target market grows to every mid-sized or larger company. |
|
| ▲ | solarkraft 2 hours ago | parent | prev | next [-] |
| Maybe as an LLM it is, but I constantly use their streaming ASR model (called Nemotron Streaming) through Handy and it works wonderfully well. |
|
| ▲ | ReptileMan 3 hours ago | parent | prev | next [-] |
| The other way. Nvidia would love open source jevon paradoxed ai - that would run inference on their chips. |
| |
| ▲ | kimixa 3 hours ago | parent | next [-] | | But their goal would be to ensure it only runs on their chips, and not any competitors. I can't see how they could do that if the best models truely were "Open". I can see one of Nvidia's biggest fears is the inference hardware becoming commoditised. | |
| ▲ | redanddead 3 hours ago | parent | prev | next [-] | | can someone weigh in on this. what's the actual play here are they actually suppressing the western open models? china doesn't give a fuck either way imo, they see the weakness emerging at the intersection of all the labs, everybody knew there was no moat, so they're gonna control its direction and basically tell the Jev guys what they want them to work on | | |
| ▲ | mattmaroon 3 hours ago | parent [-] | | It’s just an illogical conspiracy theory. Open weights models still have to run on someone’s chips. NVIDIA is model-agnostic. They are just trying to grow the pie because they have nobody else competing for slices. My guess is they want as many frontier models using their chips as possible. The only threat to their business is companies making their own chips which Google does and the others are working toward. The last thing they want is only 3 frontier labs who are all not buying NVIDIA. | | |
| ▲ | MangoCoffee 2 hours ago | parent [-] | | yes, this is my read as well. Nvidia does not care about how many LLM provider is out there as long as they pay the toll (Nvidia tax). for now, you can't beat the Nvidia CUDA/chip for training but for inference. that's where you can gain ground. Hugging Face - distribution for model that you can run on your local Nvidia Spark Neoclouds - Nvidia setup a 500 billion investment fund with Wall Street so Nvidia can sell chip and this news about buying a LLM start up. it seem like another customer for Nvidia. correct me if i'm wrong but i remember i saw an interview with Jensen where he want more company to have their own model and country to have their own LLM model. Nvidia doesn't make money from the gold rush. Nvidia made money from selling the shovels. | | |
|
| |
| ▲ | jollyllama 3 hours ago | parent | prev [-] | | Watch what they do and not what they say. |
|
|
| ▲ | seizethecheese 3 hours ago | parent | prev | next [-] |
| I think the idea that they’d purposely spend company time and resources making a bad model is an extraordinary claim, requiring extraordinary evidence. The more likely explanation is that they aren’t willing to distill from their own customers, so they are at a disadvantage. |
|
| ▲ | frozenseven 3 hours ago | parent | prev [-] |
| Insane conspiracy theory. There's no incentive whatsoever for Nvidia to release weak models. If you bothered to pay any attention, they are aggressively trying to catch up. Whether they succeed, that's of course a separate question. |
| |
| ▲ | alexdns 3 hours ago | parent [-] | | why would Nvidia try to compete against their biggest customers - openai and anthropic ? | | |
| ▲ | cbsks an hour ago | parent | next [-] | | Dog fooding their own product. This isn’t new for Nvidia. For a long time they have been selling GPU cards, and also licensing the chips for other manufacturers to make their own cards. Similar for their automotive products, Shield, and DGX Spark. | |
| ▲ | bobthepanda 3 hours ago | parent | prev | next [-] | | the power that giveth can also taketh away. similar to how eventually AWS started making their own chips for data centers, and Apple did that for their hardware, it's not a ridiculous thing to plan for the AI companies to start making their own chips to optimize for their use cases and cut out the middleman for margins. | |
| ▲ | an hour ago | parent | prev | next [-] | | [deleted] | |
| ▲ | rapsin4 3 hours ago | parent | prev [-] | | So that nvidia gets bargaining power...? Nvidia needs to diversify its customer base |
|
|