logoalt Hacker News

pdyctoday at 6:01 AM1 replyview on HN

+1. i have similar workflow and i use local models on igpu so token cost is free and electricity cost is in cents.


Replies

spuztoday at 8:51 AM

I cannot imagine local models running on an igpu could get anything close to either a useful plan or execution of a plan. I've tested Qwen 3.5 27B locally and its solutions to coding problems are usually flawed and running in thinking mode is too slow and that's on a discrete GPU. How do you get anything useful done on a model that runs on an igpu?

show 1 reply