logoalt Hacker News

piyhtoday at 3:39 PM3 repliesview on HN

Qwen 3.6 is ~$2/m tok, 3.8 should be drop in replacement. Gemma 31B is $0.34/m tok. The price differential on these models is massive on openrouter.


Replies

SparkyMcUnicorntoday at 5:19 PM

Yeah, I would appreciate if someone could make sense of the pricing differences between these models. How can a provider run DSv4F at lower cost than a 27B dense or 35B A3B model?

Does it come down to utilization and/or specific model tricks and efficiencies (attention, kv cache, etc.)?

DeepInfra prices:

Qwen 3.6 27B: $0.32 in / $3.20 out

Gemma 3 27B: $0.08 in / $0.16 out

DeepSeek V4 Flash 0731: $0.08 in / $0.18 out

Qwen 3.6 35B A3B: $0.10 in / $0.95 out

https://openrouter.ai/qwen/qwen3.6-27b

https://openrouter.ai/google/gemma-3-27b-it

https://openrouter.ai/qwen/qwen3.6-35b-a3b

https://openrouter.ai/deepseek/deepseek-v4-flash-0731

show 1 reply
jjicetoday at 4:31 PM

Where do you see that? From what I can see on Open Router, Qwen 3.6 27B (the closest dense equivalent to Gemma 31) is $0.28/m. Am I missing something?

https://openrouter.ai/qwen/qwen3.6-27b

show 2 replies
satvikpendemtoday at 6:03 PM

Why are you comparing a 2.4 trillion Max model to a 31 billion model?