logoalt Hacker News

ttoinoutoday at 10:13 AM0 repliesview on HN

We could make LLM inference 100x cheaper to run at home efficiently, but that solution might need to be updated every 1-2 years, whereas current GPU are useful for various others tasks and last longer