Since the x-axis is log-scaled, DeepSeek is much cheaper than visually implied (mousing over the raw values, it's 1/4th the cost of Luna).
Is this pricing from Deepseek with training on usage?
Is this pricing from Deepseek with training on usage?