logoalt Hacker News

DeepSeek API Pricing Update

65 pointsby mfiguieretoday at 12:49 PM25 commentsview on HN

Comments

dhxtoday at 3:27 PM

Old DeepSeek Flash 0731 prices have been independently reproduced.[1] The issue is DeepSeek being inundated and not having capacity to serve the demand, hence the price increases to significantly dampen demand. Never mind international demand either--just think about the magnitude of Chinese domestic demand. Prices for anything related to AI or computing in general (mobile phones, cloud data centre hosting, etc) will continue to climb fast as demand for computer chips _far_ exceeds supply. DeepSeek doesn't have an option other than to just work away on improving their technology in the period of time before computer chips once again become a commodity. For example, DeepSeek's cache ratio for their models apparently leads to 1/2 GPU time requirement versus the second best provider.[2]

[1] https://nitter.net/thdxr/status/2085377844515922210#m

[2] https://nitter.net/thdxr/status/2087610161636471289#m

usagisushitoday at 2:39 PM

According to their post:

  (Input / Output / Cache Read, [$/M])
  DeepSeek-V4-Flash: 
    Prev: 0.14 / 0.28 / 0.0028 
    Off-Peak: 0.22 (1.6x) / 0.66 (2.4x) / 0.007 (2.5x)
    Peak: 0.44 (3.1x) / 1.32 (4.7x) / 0.014 (5.0x)

  DeepSeek-V4-Pro: 
    Prev: 0.435 / 0.87 / 0.003625
    Off-Peak: 0.66 (1.5x) / 1.98 (2.3x) / 0.022 (6.1x) 
    Peak: 1.32 (3.0x) / 3.96 (4.6x) / 0.044 (12.1x)
gpt-5.6-luna: $0.20 / $1.20 / $0.02 / $0.25 (In / Out / Cache Read / Cache Write)

EDIT: formatting

EDIT2: giving up on the formatting :-/

show 1 reply
petercoopertoday at 2:01 PM

As well as the headline in/out changes, people heavily using agentic coding tools will want to note the 6x (off peak) and 12x (peak) increase to cache hit pricing on Pro (since cache hit can easily make up 90%+ of input on long sessions).

DeepSeek was hugely underpricing cache hit pricing before and even after this increase they're still cheaper on that metric than every other provider I'm aware of, but it will put an end to those "I used 1 billion tokens and spent $4" reports.

show 1 reply
f311atoday at 2:48 PM

Opencode said they are working on matching the old prices using their own inference.

Right now, they give 4100 credits for Luna and 63 000 for Deepseek on their prepaid plan (both are 2x)

show 2 replies
reddectoday at 3:10 PM

It doesn't make any sense unless they are going to exit from inference market. They will be literally one of the costliest option (by output, for flash) if use openrouter as source.

unified101today at 1:58 PM

About 3x increase. Luna is now a much better deal. Hope they don't increase their prices in response.

show 3 replies
kortzeustoday at 3:00 PM

Makes sense, basically increased price for peak hours when they don't have enough infra to serve everyone. Can expect the return of prices back.

timmmmmmaytoday at 2:53 PM

the neat thing about the models being open-weights is there's a dozen other providers on OpenRouter still at the old price, or lower

show 2 replies
slopinthebagtoday at 2:49 PM

Ah well. I spend about $5/month with Deepseek, so now I’ll have to find room in my budget for $15. Might have to tip my barista less or something.

pu_petoday at 2:08 PM

This now places deepseek flash v4 from DeepSeek themselves at higher prices than openrouter (depending on caching). Will be interesting to see if third party prices remain the same.

show 1 reply
schafbergtoday at 3:34 PM

[flagged]

TrustScoreAgenttoday at 2:54 PM

[flagged]