logoalt Hacker News

rjh29today at 2:26 PM6 repliesview on HN

It's crazy hearing devs on this site claim Claude is 10x better than all other AI solutions. I think it is fomo. Claude $LATEST_VERSION is perceived as the best and anything else is "missing out". New version comes out? Suddenly the old version is worthless, how on earth did anyone get work done with that?

Same reason people buy the RTX 4090 and 5090 cards - overpriced but they must have the "best". Never mind the diminishing returns trying to max out PC settings (3-4x performance hit for an almost imperceptible increase in graphics, ignoring DLSS) - it's the psychological cost of having to move a slider down a notch.

I've been using Google and now DeepSeek v4 and I am having absolutely no problems and it's a fraction of the cost. I'd love for Claude to be 10x better but it just isn't, for my use case anyway.


Replies

jnovektoday at 2:29 PM

I’ve been using DeepSeek V4 in OpenCode exclusively for about a month.

I think it’s great, but coming from Claude Code it did feel like going back in time by ~6 months in model capabilities. This isn’t a big deal to me for what I do, but the difference is definitely there.

show 1 reply
solenoid0937today at 2:58 PM

Opus 4.8 and GPT 5.5 are the best models, but people don't care about "best" anymore, until there is a big leap in capability I don't think anyone will care about point releases.

Vibes and tribalism will prevail until one of emerges as clearly and unambiguously superior to the other.

Tenemotoday at 3:44 PM

I get what you mean but the GPU comparison isn't the best here, I think. Money-is-no-object-I-want-the-best approach is questionable, definitely. But no one can argue that an old Nvidia card is objectively better for e.g. 4k gaming than a 4090 if you don't mind the wattage. You can just measure it.

With LLMs the problem is more complex, it's people getting used to how a model works and to the ecosystem. Sure, you can make all your skills harness-agnostic and deal with Anthropic's stubborn refusal to adopt the common naming/directory structure. But most people don't. So then you end up with something closer to the ancient Android vs iOS discussion. Can you prove, in isolation, that iOS is more energy efficient, the hardware is faster? Yeah. But that won't speak to someone who has been on Android for 10 years and would have to migrate and get used to iOS to experience that, first.

I've noticed myself how I get used to common failure modes of particular models in my projects. GPT5.5 tends to create some checks/booleans I don't need, it heavily overcorrects on error handling, etc. While Claude 4.7/4.8 doesn't do those as often but gets derailed on our E2E test suite, forgets to run linting despite guidance. So even assuming fully harness-agnostic working setup, a new LLM model with its own quirks can be a lot friction for heavy users who might be used to Claude specifically and all their skills/guidance pre-address common failure modes.

E.g. I might be a Prius owner, then you gift me an objectively better, more efficient, safer, newer, same-size, physical knobs car ...and I might still swear by my Prius! I'm used to how it turns, how it feels, I can repair some issues myself. Isn't that a normal reaction then?

Aurornistoday at 3:29 PM

> Same reason people buy the RTX 4090 and 5090 cards - overpriced but they must have the "best".

Or they need to run high VRAM apps like LLMs

Or they have 4K monitors and want smooth gameplay on them

Is this whole thread just dedicated to snark about other people’s personal preferences?

Hamukotoday at 2:28 PM

Hey, at least the superior performance of a 4090 or a 5090 can be objectively measured.

show 1 reply
slashdavetoday at 3:18 PM

You're projecting