logoalt Hacker News

simonwtoday at 1:31 AM2 repliesview on HN

Was that with the default xhigh reasoning setting? I suggest trying again with reasoning set to low or turned off entirely.


Replies

dofmtoday at 2:43 AM

You now have me testing it with reasoning turned off, which I have never bothered much with on any other local models because it's rarely worth it.

The result appears to be almost as good as Qwen 3.6 35B A3B on medium thinking mode.

It second-guesses a little, it gives broader/more speculative answers, of course, and it missed the nuance of one of my prompts, but this gives me a lot more confidence that the Low reasoning effort is going to be as good as they say, and perhaps in some cases non-thinking looks like it would be enough.

Really useful, thanks.

SwellJoetoday at 1:48 AM

Yes, default everything, no tuning, 8_K_XL Unsloth quantization on dual Radeon V620 GPUs (which aren't blazing, but faster than the Strix Halo).

show 1 reply