Remix.run Logo
noosphr a day ago

Issue is that llama.cpp is the best way to run models on hardware that isn't nvidias.

dannyw a day ago | parent | next [-]

A lot of llama.cpp contributions come from the community and ecosystem, like Unsloth. If something goes awry, I fully expect lots of forks.

MrDrMcCoy 7 hours ago | parent [-]

There already are a lot of forks for things they decline to implement. TurboQuant, ROCmFPX, and more. I need to set up an agent that will loop on merging them.

alightsoul a day ago | parent | prev [-]

Except when they have less than 16 gb of ram?