Remix clone Hacker News

new | show | ask | jobs Github

	▲	EnPissant 4 days ago
		llama.cpp has support for running some of or all of the layers on the CPU. It does not swap them into the GPU as needed.