logoalt Hacker News

Measured LLM inference speeds on Apple Silicon, with raw data (CC BY 4.0)

10 pointsby bagdaerdevyesterday at 11:45 PM3 commentsview on HN

Comments

kennywinkertoday at 4:38 AM

Why do no benchmarks show qwen3.6?

As far as i can tell the state of the art small open models is qwen3.6 and gemma 4, yet they rarely appear in benchmarks - even ones made recently

show 2 replies
bagdaerdevyesterday at 11:45 PM

[flagged]