Remix.run Logo
cmrdporcupine an hour ago

A PC with a small GPU coordinating and then llama.cpp using the llama RPC stuff (over RDMA to reduce latency) talking to the two nodes.

I dunno, maybe he'll do a write-up someday.

runekaagaard 44 minutes ago | parent [-]

OK thanks! Would love to read :)