logoalt Hacker News

seanmcdirmid • today at 5:59 PM • 0 replies • view on HN

I still haven't found a use case for Qwen3.8 27B that Qwen 3.6 35b A3b (MoE) is better at. I can get at most 40 tokens/second with 27B, but I get around 90 tokens/second with the MoE and it seems to be a more capable model.

I guess I should still keep experimenting though. Maybe I'm just not using a dense model correctly.