logoalt Hacker News

int_19hyesterday at 8:19 PM1 replyview on HN

That is not a counterargument to cloud inference though. You can also run open weight models in the cloud, and it's still cheaper. So privacy really is the only motivation to run on local hardware.


Replies

anon373839yesterday at 11:58 PM

Ah, no, that’s not cheaper. Renting GPUs adds up quickly and leaves you with nothing in the end.

Renting tokens from open model providers is cheaper but it incurs the same issues: unexpected changes in model quality, inconsistent speeds, service outages.