logoalt Hacker News

Axsuultoday at 2:01 PM2 repliesview on HN

Can anyone recommend the perfect sweet spot for someone who wants to run their own inference?


Replies

rkangeltoday at 2:49 PM

Thinkstation PGX maybe?

Got the recommendation from these articles: https://www.xda-developers.com/qwen-3-8-27b-reverse-engineer... https://www.xda-developers.com/lenovo-thinkstation-pgx-revie...

But haven't had a chance to try it myself.

AbsurdCensortoday at 2:48 PM

For me it's be Strix Halo, 128gb machine, especially running Qwen models. Except when I bought it, it was $1,900, now it's $4,600 for the same box. (Wow that's insane)

For tinkering and learning, it's been great. Tie it into something like Hermes and you have a pretty powerful AI assistant in a box. And when you need to step up your model, you just do something like OpenRouter and it makes it pretty easy.