The few tests I ran were by no means comprehensive, but while kimi felt like the real deal qwen seems a bit of a benchmark princess.
Qwen3.6 is still the best agentic open weight LLM around 30b params (Gemma isn’t very good at agentic execution).
I also find the model is a lot more predictable and less “glitchy” when made to think in Chinese. You can do this in the system prompt.
[dead]
Qwen3.6 is still the best agentic open weight LLM around 30b params (Gemma isn’t very good at agentic execution).
I also find the model is a lot more predictable and less “glitchy” when made to think in Chinese. You can do this in the system prompt.