logoalt Hacker News

Gecko4072today at 4:48 PM6 repliesview on HN

Currently burning money quickly on official deepseek api. They are also increasing pricing starting today. V4 Flash 0731 still feels like the most outstanding model of the past few months and probably to come.


Replies

nolist_policytoday at 5:09 PM

DeepSeek V4 Flash is the "too cheap to meter" of AI. And you can run the full unquantized model locally for $8000 (2x DGX Spark) at full 1M context and decent speeds: https://github.com/elsung/dgx-spark-deepseek-v4-flash#-long-...

sschuellertoday at 5:48 PM

Deepseek seems to have gotten too cheap. I have been using it for a long time and it's at a point now where my credits balance barely moves even at max setting.

elitoday at 5:33 PM

The Deepseek official API is good with excellent caching.

But their privacy policy is unusually bad - they can train off your prompts and completions.

show 1 reply
Eueudhsbsj32today at 5:12 PM

What's the new pricing?

The prices on OpenRouter still look the same.

show 1 reply
igravioustoday at 5:15 PM

yup :)

i'm doing opencode <-> openrouter <-> official deepseek api (i don't get the opencode hate, i like it)

how are you doing it?

am also using Kimi K3 via kimi-code

and also GLM 5.2 via ZCode

happy with all three, they're trailing frontier but i figure if i'm running GNU/Linux then i ought to favour open weights models with my €s -- reduced my usage of claude/gpt to the ~$20 tier just to keep abreast of claude_code/codex developments

show 1 reply
Jsttantoday at 5:07 PM

What is the new price through?

show 1 reply