logoalt Hacker News

Why isn't the industry freaking out about DeepSeek 4.1 Flash?

332 points • by jonotime • today at 12:14 AM • 289 comments • view on HN

Comments

jeffrallen • today at 10:25 PM

Also, it is willing to do legitimate work I need done which other models flag as dangerous and refuse to do. (Software testing of a DHCP server to survive bad inputs.)

bitfilped • today at 10:23 PM

Because in two weeks someone will be asking why I'm not freaking out about AlphaDolphins 0.3 Zip and then in a month FrozenMonkey 2.5 Artic.

try-working • today at 10:17 PM

I have used over 40B tokens and spent over $800 on DeepSeek API over the past 30 days, mostly on V4.1 Flash.

It's good, and you can do most work with this. For complex software implementation you need to split your runs into various phases, build in verification, and use subagents so that work gets another audit and repair pass from the lead agent. You can do pretty much everything then. Frontier models can do without compelx workflows, that's the difference.

tonyhart7 • today at 10:54 PM

it literally hallucinating a lot

I dont get why people says D4.1 flash is good

robertlane0 • today at 9:27 PM

Honestly for me the intelligence gap between DS 4.1 Flash and Muse Spark 1.3 makes Muse more worth it for me, especially on a $10 OpenCode Go sub, with the caveat that everything I use it on is open source which makes the fact that I'm sharing it with Meta a little moot because it's already published permissively on GitHub anyways.

kristianp • today at 8:12 PM

> shrank the KV cache by roughly 437X

Can't you just say "shrank to 1/437th the size"? It's not that hard.

MisterMunchkin • today at 8:15 PM

I had it make 25 different things today and it cost $0.70

It’s disgustingly good value. I find it capable of doing anything I want.

Obviously can’t use it at work, but for home projects it’s awesome.

➕ show 1 reply
anguralbanish2 • today at 7:52 PM

I would love to get them more better, it's good not a bad thing.

cactusplant7374 • today at 8:54 PM

Because engineers are lusting for 1000 tokens per second. You can only achieve something like that with OpenAI.

pessimizer • today at 8:35 PM

I'm no expert, but it think that it's the pricing on GPT-6 Luna. I'm also guessing that it's been underpriced just for this reason. I also don't think it's all that great, but it's definitely very cheap.

If it's underpriced, it's a loss leader to sell the other models, so it actually can't be too good.

I really put these things through their paces because I use them to review and work with new abstract game rules and models, so they're always flying blind. Luna misses the obvious (and more importantly, the clearly explained) consistently. My second prompt is listing all of the points in its first response, and saying "No, it doesn't work like that." The third prompt is picking out the two or three suggestions it made after correcting itself on all of the original points and saying "That's how it already works." The fourth prompt is "Now that we're done going over the rules, can we start?"

I actually feel like 5.6 Luna seemed better.

sergiotapia • today at 8:25 PM

In my experience it just takes so much longer to arrive at "done" state for me. It thinks for soooooo long. I guess if you're running 12 sessions at once you don't really notice.

AIblemblio • today at 7:55 PM

No they can't.

And as long as I pay as little for claude opus 5.5 i do right now, i'm using it.

But yes i'm glad that we have alternatives.

m3kw9 • today at 8:58 PM

i thought 6.1sol copied the caching architecture so this isn't such a big deal no more

doctorpangloss • today at 8:17 PM

because it doesn't work very well?

if you have a legitimate coding application, it isn't very good. if you have some kind of inauthentic activity, which could be what it is trained for for all sorts of reasons...

➕ show 1 reply
cbeach • today at 11:05 PM

Honestly, I just don't trust the Chinese Communist Party having agentic access to my computer.

Every company in China has to abide by the 2017 National Intelligence Law: "supporting, assisting and cooperating" with state intelligence work, and keeping that cooperation secret. They have to hand prior knowledge of vulnerabilities to the state before public disclosure, in order that the state always has an exploit pipeline. No matter how ethical the company staff may be, they'll always be bound by law into being an arm of the Communist Party.

Agentic access is infinitely worse than chatbots. They can exfiltrate silently, target users, plant persistent malware, and be run by third parties through you.

You don't have to be a tin foil hat sinophobe to understand the dangers of being a Westerner granting CCP access to your files and network.

ByteDance staff accessed US journalists' TikTok data to hunt leakers (admitted in 2022). Volt Typhoon and Salt Typhoon were state operations pre-positioned in Western infrastructure and telecoms. Regulators in Italy and South Korea blocked DeepSeek's app over data handling, and analysts found its web client sending data to a China Mobile domain.

Please don't sacrifice security for cost and convenience.

verdverm • today at 2:07 AM

Why would we freak out? The systems we use have always gotten better, faster, cheaper with time

kydanet • today at 8:35 PM

[flagged]

oh_no • today at 9:58 PM

AA shows Luna at 1/4 the price, 1 point behind on intelligence matrix with a 38.

Haiku 5.5 is 23% cheaper with a 4 point intelligence lead.

I'm on subscription usage so I can't compare Flash 4.1 to them directly but the OP has his head up his ass if he thinks Opus 5.5 is the best point of comparison. Why is anyone using Opus if the new Haiku is indistinguishable /s

Just absolutely terrible post, admits to using Opus for review but claims its intelligence isn't needed, why aren't you using Haiku or Sonnet then?

CurbStomper4 • today at 9:06 PM

[dead]

distantsounds • today at 8:19 PM

because we've all figured out that AI is just a huge grift?

sroussey • today at 8:08 PM

Not comparing to gpt-6-luna which seems comparable and priced well.

wewewedxfgdf • today at 8:21 PM

You might also choose to pay money for a service that provides real value instead of actively choosing to support the Chinese deliberate effort to undermine this country.

➕ show 2 replies