logoalt Hacker News

expedited123today at 5:49 PM1 replyview on HN

Thanks! I don't have any knowledge of running models locally.

I assume it would not be able to handle an unquantized Qwen3.6-35B or is it irrelevant as you almost always would want to run a quantized version of the model on consumer hardware?


Replies

seanmcdirmidtoday at 5:51 PM

not parent, but 4-bit quantization is generally consider a good trade off for speed/performance, so you might use it even when you aren't on consumer hardware, but definitely when you are on consumer hardware.