| ▲ | somenameforme 2 hours ago | ||||||||||||||||
Another interesting potential market here will be 'LLM in a box'. All the hardware and other tooling in a prebuilt, but modular, package ready to go. Pay one up-front cost, get a system running [whatever open LLM] with a token rate of [x], optionally configured to be immediately ready for distributed usage. Basically the opposite of cloud stuff: no rent, no dependency, 100% guaranteed uptime, guaranteed security/privacy (at least subject to your own actions), and so on. | |||||||||||||||||
| ▲ | adrian_b an hour ago | parent | next [-] | ||||||||||||||||
Palantir already offers a "turnkey AI datacenter", i.e. a rack with "NVIDIA Blackwell Ultra systems with eight NVIDIA Blackwell Ultra GPUs and NVIDIA Spectrum-X™ Ethernet networking for AI training and inference". It is said that it comes with all hardware and software required to run inference or training with an open weights LLM. The existence of this product, which competes with cloud-based offerings like those of OpenAI and Anthropic, is presumably the reason why the Palantir CEO criticized very harshly some time ago the business model of OpenAI/Anthropic. While I doubt that the ethics of Palantir is any better than of OpenAI/Anthropic, in this particular case I have to agree with Alex Karp about "Sovereign AI", i.e. that only losers will make their business completely dependent on an external entity like OpenAI or Anthropic, who are certainly not trustworthy. | |||||||||||||||||
| |||||||||||||||||
| ▲ | jurgenburgen 2 hours ago | parent | prev | next [-] | ||||||||||||||||
> 100% guaranteed uptime Disagree there but I think this is an interesting idea. We would need to find some more cost-efficient hardware to run it on than Nvidia GPUs. | |||||||||||||||||
| |||||||||||||||||
| ▲ | bevekspldnw an hour ago | parent | prev | next [-] | ||||||||||||||||
“100% guaranteed downtime when you least can afford it and the support tickets are your problem.” We’ve a hybrid shop, including hosting our own ML infra, and we save a ton from cloud spend with local ML. Easily one million USD over past three years. But it’s not “free”, you are shifting a lot of labor into your plate. | |||||||||||||||||
| |||||||||||||||||
| ▲ | Godsend69 an hour ago | parent | prev [-] | ||||||||||||||||
[dead] | |||||||||||||||||