| ▲ | Axsuul 5 hours ago | |
Can anyone recommend the perfect sweet spot for someone who wants to run their own inference? | ||
| ▲ | rkangel 4 hours ago | parent | next [-] | |
Thinkstation PGX maybe? Got the recommendation from these articles: https://www.xda-developers.com/qwen-3-8-27b-reverse-engineer... https://www.xda-developers.com/lenovo-thinkstation-pgx-revie... But haven't had a chance to try it myself. | ||
| ▲ | AbsurdCensor 4 hours ago | parent | prev [-] | |
For me it's be Strix Halo, 128gb machine, especially running Qwen models. Except when I bought it, it was $1,900, now it's $4,600 for the same box. (Wow that's insane) For tinkering and learning, it's been great. Tie it into something like Hermes and you have a pretty powerful AI assistant in a box. And when you need to step up your model, you just do something like OpenRouter and it makes it pretty easy. | ||