| ▲ | segmondy a day ago | |||||||
From the article, "One guide this summer was literally titled “Open Weights You Can’t Run.”" I ran Kimi a few days ago at 1/2token per second. I only get to play with it during the weekend when i have time, but I'm certain I'll be able to get it up to 5tk/sec when I'm done in a month or two. So yeah, we can run them all locally. Folks might say it's not run if it's that slow, but feh! If you can run the best AI model locally at 1tk/sec, why won't you? | ||||||||
| ▲ | konmam 19 hours ago | parent [-] | |||||||
Well considering I can just use DeepSeek for pennies, not sure what I might get out of that 1 tk/sec locally. Maybe the question is not can you, but rather should you? | ||||||||
| ||||||||