logoalt Hacker News

reconnectingyesterday at 6:39 AM3 repliesview on HN

I'm sure DeepSeek isn't the point here. You can change the name to whatever you prefer and the article still holds.

Actually, I think the author put DeepSeek on purpose to avoid the obvious ChatGPT/Claude comparison — because whatever he chose, there would be a question of why model A and not B, while the point of the article isn't about models comparison at all.


Replies

terhechteyesterday at 7:07 AM

It is the point. Also open source model enthusiast tell you otherwise, there is a coding quality gap between these models. If I use DeepSeek, I do so knowing that I have to limit to simpler tasks on smaller, well specified prompts. What the author did, letting the model do the planning, is not something DeepSeek will excel at. I'm using GPT (Terra, Sol, Luna), Claude (Opus 5, Fable), Qwen 3.8 and GLM 5.3 Flash daily and have to vary which model I use where because there's a huge intelligence step function difference here. That's why this article is so useless:

Imagine someone trying to make the case that riding bicycles is a terrible experience and their whole argument is that they took a random cheapo bike with flat tires and rode it for 3min and that wasn't fun. Sure, but if you buy a 25k carbon bike you will have a different experience. I'd not trust that person. If someone told me they have 10 bikes they ride daily and can explain the differences, in detail, between their bikes, and what they excel at. I'd trust that person's opinion.

show 2 replies
paducyesterday at 6:43 AM

Maybe the failure is in trying only one model / one prompt.

show 1 reply
jonplackettyesterday at 7:05 AM

I think the point is that $10 isn’t exactly a lot of money to put where your mouth is, nor a serious effort to see if it works.