Remix.run Logo
make3 a day ago

NVidia has reasonable incentives to keep things open indefinitely, it wants people to use it's GPUs

noosphr a day ago | parent | next [-]

Issue is that llama.cpp is the best way to run models on hardware that isn't nvidias.

dannyw a day ago | parent | next [-]

A lot of llama.cpp contributions come from the community and ecosystem, like Unsloth. If something goes awry, I fully expect lots of forks.

MrDrMcCoy 10 hours ago | parent [-]

There already are a lot of forks for things they decline to implement. TurboQuant, ROCmFPX, and more. I need to set up an agent that will loop on merging them.

alightsoul a day ago | parent | prev [-]

Except when they have less than 16 gb of ram?

well_ackshually a day ago | parent | prev | next [-]

Nvidia is halving the production of its higher end GPUs, doubling prices and heavily segmenting the market. They don't give a single shit about individuals. One wafer that makes 10 5090s that might sell at 3000 each, or one wafer that it already sold 6 months ago for 200k to one of the incestuous AI companies it works with?

swozey a day ago | parent | prev [-]

It'd be really nice if I could use my egpu 4090 on my macbook pro..

richrichardsson a day ago | parent [-]

For LLM workloads this might help? https://docs.tinygrad.org/tinygpu/

swozey a day ago | parent [-]

Interesting, I've never seen this, thanks