logoalt Hacker News

Claude Haiku 5.5

529 points • by sfkgtbor • today at 6:01 PM • 249 comments • view on HN

Comments

simonw • today at 7:35 PM

Pelicans riding bicycles for Haiku at the different thinking levels: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

Low messes up the bicycle frame, but medium/high/xhigh/max all get the bicycle frame right.

The max one took 5 minutes 9 seconds and cost 3.3826 cents. The cheapest one (low) cost 0.0936 cents and took 7 seconds.

The most recent release of my llm-anthropic plugin queries the Anthropic model listing API directly, so I didn't have to upgrade the plugin to add support for this model:

  llm install llm-anthropic -U                                
  llm anthropic refresh
  llm -m claude-haiku-5.5 'prompt goes here'
EDIT: Here's the Haiku 4.5 pelican from a year ago for comparison, it was terrible: https://simonwillison.net/2025/Oct/15/claude-haiku-45/
➕ show 4 replies
minimaxir • today at 6:05 PM

Pricing is...a bit weird.

    Input
    $0.10 / MTok for prompts up to 100,000 tokens
    $0.50 / MTok for prompts over 100,000 tokens

    Output 
    $0.50 / MTok for prompts up to 100,000 tokens
    $2.50 / MTok for prompts over 100,000 tokens
100k tokens is an absurdly low cutoff and it is only applicable to Haiku and not Sonnet or Opus. It's a low enough cutoff that it will be quickly exceeded if you are doing anything with Agents; for typical generation or Jev-like classifiers, it's a good value and as noted in this article, that is apparently the vast majority of Haiku use.

In both cases, still much cheaper than Haiku 4.5's $1 input / $5 output and these prices better compete with GPT-6 Luna. ($0.10 input / $0.50 output, but with no token threshold [EDIT: the threshold for Luna is apparently 272k])

➕ show 17 replies
charlesabarnes • today at 6:21 PM

> Second, this week, we’ll roll out a new monthly API credit to all Max and Team subscribers for use on the Claude Platform. Max 5x users will get $100 in credits per month, Max 20x users will get $200, and Team subscribers will receive up to $500, pooled across their users

This is a very big benefit for me. I can now ship actual ai enhanced features behind my subscription without paying extra or fully relying on on-device models. I do worry that this is to soften the blow for user-unfriendly changes

➕ show 4 replies
simonw • today at 7:33 PM

My complaint about Haiku 4.5 was that it was 10x the price of GPT-6 Luna.

> Claude Haiku 5.5 is priced 90% lower than Claude Haiku 4.5 for requests up to 100,000 tokens, and 50% lower for requests over 100,000 tokens

Haiku and Luna now have the exact same price up to 100,000 tokens. Luna is now cheaper for anything after 100,000 tokens, even after Luna's own price increases at 270,000 it's still less than Haiku.

So it sounds like they've directly addressed that problem. Their self-reported benchmarks are all higher than Luna too.

➕ show 1 reply
jjcm • today at 6:30 PM

Ran image -> html tests for this. I was curious if this smaller model was good enough for complex UI. It was not.

Haiku 5.5: https://html.non.io/lcars-haiku-5.5/

Opus 5.5 for comparison: https://html.non.io/lcars-opus-5.5

Designs it was building from: https://diffui.ai/app/canvas/5093e689-1e74-4f26-b632-2a4500f...

One interesting thing is it took a look at the job at hand, and immediately delegated it to Opus 5.5. It at least knows what it isn't good at. Very fast though, and likely best used for small subagent tasks / tightly scoped work.

➕ show 4 replies
seaal • today at 6:12 PM

The monthly API credits for Max plan seems fantastic, especially considering Haiku pricing. Being able to actually use my Claude plan for other harnesses and use-cases on top of regular CC usage is everything I wanted.

Anthropic has really been doing all the right things in the past few weeks, while OpenAI continues to fumble the bag.

➕ show 3 replies
bouk • today at 6:35 PM

This is great! Been using GPT 6 Luna for decompiling my childhood favorite game (Age of Mythology) and this means I can throw Haiku into the mix as well. 17352/21965 functions matched so far...

➕ show 3 replies
wyrdcurt • today at 6:29 PM

About time Anthropic released a competitive cheap model. Haiku 4.5 has been too expensive compared to its performance for months now (in fact I don't remember being too impressed even when it was released). This one actually looks worth using in some scenarios. If it's really as much of a step up from Luna as the benchmarks they've shown indicate, it'll probably replace Luna in my workflows. 100k tokens is a pretty low threshold before the price goes up, but I tend to use these smaller models for smaller tasks anyway.

matltc • today at 8:50 PM

My weekly limit __on a Pro sub__ has not gone over 50% since before the pre-Fable promos, but usage has been pretty much the same from my point of view. Maybe I am holding it right? Anyone else getting this?

As such, I do not need to even reach for Haiku, and 4.5 was so inaccurate that it often cost more to do so in the past. Sonnet 5.5/low has been good for this kind of thing, and i didn't even touch thinking tokens or any of that. Opus 5.5 low for questions/repros, medium for implementation, basically never reaching for anything above that anymore. 5.5 has been great, so I'll try Haiku, but don't see myself going out of my way to integrate it.

➕ show 2 replies
d1l • today at 7:08 PM

At work we use haiku 4.5 for a handful of latency sensitive tasks that are fairly simple. It performs well. Just started testing 5.5 as I’ve been anticipating a nice improvement since it was teased. Results so far are trash. Prompt leakage even. And it’s slower. I guess it’s cheap but I think they got the balance wrong on this.

➕ show 1 reply
TheAmazingRace • today at 6:04 PM

I wonder if we have an AI LLM equivalent to Moore's Law. Like how often do we expect improvement in this technology and with what timing?

➕ show 6 replies
garo-pro • today at 6:23 PM

> Claude Haiku 5.5 is our fastest model to date at each model’s standard speed, although it runs less quickly than our Opus models in Fast Mode.

Opus 5.5 runs 117 tps average on Openrouter, so it must be at least 10-20 tps slower for them to mention. IDK why they mention this as it does not help for marketing though. https://openrouter.ai/anthropic/claude-opus-5.5

➕ show 2 replies
djoldman • today at 8:01 PM

https://www.anthropic.com/claude-haiku-5-5#further-updates

This section makes the reader think: why would I not pick Sonnet 5.5 instead of Haiku 5.5?

tpoacher • today at 6:13 PM

Good to see Anthropic back alternative OSes.

Topfi • today at 6:52 PM

131tok/s P50 according to OpenRouter currently, though might move up or down over the coming days. If it sticks at that speed, roughly twice the throughput of Luna and far lower latency (up to 2sec depending on provider) is impressive, though the 5x price increase beyond 100k is painful.

Was a big fan of Haiku 4.5, though understand why for most Sonnet was the far better option back then.

MisterMunchkin • today at 7:30 PM

> we’ll roll out a new monthly API credit to all Max and Team subscribers for use on the Claude Platform. Max 5x users will get $100 in credits per month, Max 20x users will get $200

They’re definitely planning to make the subscriptions API based so they can charge you full price.

waximabbax • today at 8:13 PM

Alright its still little early since there is not enough independent testing but this looks very promising and I wasn't expecting anthropic to beat GPT-6 Luna especially at the same price. Haiku 5.5 beats Luna on every shared benchmark Anthropic published, particularly computer use and agentic coding.

TomGarden • today at 6:10 PM

From these selected benchmarks, it looks like it smokes Luna capability-wise. Excited to put it through its paces

swalsh • today at 6:19 PM

Top of the page in 17 minutes? Now I know what y'all do while your agents are working.

➕ show 1 reply
dangoodmanUT • today at 7:15 PM

> Second, this week, we’ll roll out a new monthly API credit to all Max and Team subscribers for use on the Claude Platform. Max 5x users will get $100 in credits per month, Max 20x users will get $200, and Team subscribers will receive up to $500, pooled across their users. These credits are designed to allow our users to experiment with building tools, apps, and agents that call our API.

This is kind of nuts

➕ show 1 reply
skeledrew • today at 6:37 PM

The forgotten model is back on the map. I actually got OK mileage when I tried it for coding months ago. Maybe I'll try it again, with Opus guiding it, and see how it goes.

satvikpendem • today at 6:40 PM

Apparently quite a bit smarter than Luna, I wonder what use cases it can cover. I actually honestly don't need a Haiku level AI to be that smart, and looks like you pay for it in the per token cost, I need speed mainly. I might even rather have a dumber but much faster model for things like web searching and parsing to retrieve results for the app or other LLM to do things with.

patrickwdaly • today at 6:12 PM

How are y'all using Haiku though? I rarely select it.

➕ show 10 replies
hidelooktropic • today at 8:15 PM

Finally! I understand Haiku is the less intelligent model, but the gap between Sonnet and Opus has been far too wide for about a year now.

the__alchemist • today at 7:44 PM

I am perpetually confused about every name and version combination from both OpenAI and Anthropic. Especially in conjunction with the effort levels.

johnisom2001 • today at 6:52 PM

It fails the "How many r's in <word>?" test.

I ask:

> how many r's in diminished

It answers:

> Diminished has 1 r.

sroussey • today at 6:09 PM

It’s about time Haiku got an update!

declanjackson • today at 7:19 PM

According to AA benchmarks, it uses 162k output tokens per task (with max reasoning) - over double GLM-5.3 Flash for similar level of Intelligence

dhabedank • today at 9:44 PM

Really excited to use this

afrnswrth • today at 6:14 PM

The important question though...how does it do making a pelican on a bicycle?

oh_no • today at 7:27 PM

no AA benchmarks yet and the last chart in the announcement makes Haiku look useless vs new Sonnet pricing, interesting to see what 3rd party benchmarks show because i think Anthropic are costpertaskmaxxing here and it's going to look more like that bottom chart than the top ones.

3371 • today at 7:10 PM

Seeing people talking about the Agents SDK -> credits change makes me wonder does it impact Zed or likes.

sfkgtbor • today at 6:05 PM

Happy about the Sonnet cache read price cut.

➕ show 1 reply
harshitkrhere • today at 8:59 PM

request to anthropic team release haiku as os model

margorczynski • today at 6:11 PM

How does the price compare to Luna? At least looking at the numbers it is noticeably better at most tasks.

➕ show 2 replies
mattz56 • today at 7:00 PM

It's finally here ! Need to take a look at some benchmark now

simianwords • today at 6:17 PM

> Second, this week, we’ll roll out a new monthly API credit to all Max and Team subscribers for use on the Claude Platform. Max 5x users will get $100 in credits per month, Max 20x users will get $200, and Team subscribers will receive up to $500, pooled across their users. These credits are designed to allow our users to experiment with building tools, apps, and agents that call our API. They can be used on any of our models. For more information, see our Help Center article.

Did anyone read this? We get free API credits on some plans now

maz1b • today at 6:07 PM

Wow, the rate of improvements in the AI era is staggering.

GDPval-AA v2.1 as of now: 1620

GDPval-AA v2.1 for Haiku 4.5: 735

The 100k tokens pricing makes sense, looks to be a hedge against OpenAI's decisions API and Jev or its open source alternatives that are springing up.

Nice release, congrats to Anthropic.

➕ show 1 reply
justmaris • today at 7:34 PM

Finally it has arrived.

dcchambers • today at 7:14 PM

Begging Anthropic to let us use Claude subs with harnesses other than Claude Code at this point.

crooked-v • today at 6:23 PM

The important question is, does it talk in incomprehensible Claude-ese like the other Claude 5.x models?

axthauvin • today at 7:07 PM

will use it to replace luna in production !

iagocc • today at 6:09 PM

Where is Pelican? ehehhe

➕ show 1 reply
simianwords • today at 6:12 PM

I remember a friend asking me why LLMs suck so bad. She was using Haiku 4.5 and that poor model couldn't keep track of the context within 3 messages.

She said she was using Haiku 4.5 because she was advised to be careful with the spending.

I hate that model so much lol.

AtNightWeCode • today at 6:19 PM

Probably the same scam as the last Haiku update I guess. Uses more tokens to compensate for the lower price.

➕ show 1 reply
ariwilson • today at 10:04 PM

[flagged]

areoform • today at 6:48 PM

    > but they still block penetration testing and other techniques more likely to be used by attackers.
    > 
    > Haiku 5.5’s biology safeguards are the same as for Sonnet 5, Sonnet 5.5, and Opus 5. They allow research biology questions but restrict access to requests that we judge as likely to cause harm. Organizations working on wider-ranging biology and cyber activities can apply to our Life Sciences Verification Program and Cyber Verification Program.
I would like to take a moment of your time to tell you about some of the "bioweapons" Anthropic has blocked that involved Haiku!

These are the examples from "Detecting and countering misuse of AI: September 2026" - https://news.ycombinator.com/item?id=49647300

    > Importantly, because our biological safety classifiers robustly block content involving high-risk biological research (in this case, the construction of enhanced pandemic potential pathogens), all of these exchanges occurred on models in our weakest class of models (specifically, the models were Claude Sonnet 4 and Haiku 4.5, the latter of which the user began using after Sonnet 4 was deprecated). 
    >
    > Upon a detailed examination of the exchanges, we estimate that the uplift provided by Claude was primarily clerical assistance in data analysis, study ideation and design. This is consistent with our understanding of the capabilities of Sonnet 4 and Haiku 4.5, which are not able to perform expert-level biology research tasks; we estimate that the uplift provided to the researcher was limited and substantially lower than it would have been from one of our more capable models.
Anthropic then says for the above, "we estimate that the uplift provided by Claude was primarily clerical assistance in data analysis, study ideation and design"

While doing my best to avoid comment, please note, they're talking about a domain expert in a state research institution using Claude to do paperwork.

What did they save us from? What bioweapons did these filters prevent? From the front matter report,

    > The above LLM platform is not the only route via which researchers engaged in viral gain-of-function research have used our platform. In May 2026, we discovered a researcher outside the US using Claude in their research on highly-pathogenic avian influenza (“bird flu”). The research focused on viruses’ adaptation to mammals, and the mechanism by which it causes severe disease beyond the respiratory tract.
OK. Sounds serious. "Gain of function research..." but who and why?

    > The researcher pursued this work in a credible institutional context, and interacted with Claude over the course of several weeks, exchanging thousands of messages. In these exchanges, the researcher leveraged Claude’s knowledge of the scientific literature to assist the researcher in study planning and design, data analysis, and the interpretation and prioritization of experiments. The researcher also used Claude for editorial assistance in writing up the research.
So this was a researcher inside of some country's national lab ("credible institutional context") doing research on dangerous viruses using Claude for "for editorial assistance in writing up the research."

What "uplift" are you providing to scientists working at specialized global BSL-4 labs that already have – and I quote their report - "physical access to such isolates." (as in samples of viruses)? Are we uplifting their grammar?

These "safeguards" are being expanded. The scientists I know can't use Claude for grammar checks or anything serious. You can try it for yourself.

caaqil • today at 6:16 PM

> Haiku 5.5’s cybersecurity safeguards are more restrictive than Haiku 4.5’s, but somewhat less restrictive than those we’ve applied to other recent models. In cybersecurity, they permit a wider range of defensive tasks than our safeguards for Sonnet 5.5, but they still block penetration testing and other techniques more likely to be used by attackers.

If you block pentest or "other techniques more likely to be used by attackers", then what does "permit a wider range of defensive tasks" even mean?

Any defensive task that's meaningful is almost indistinguishable from legitimate red-teaming that then falls under 'likely to be used by attackers". If only they would just stop nerfing these models, that'd be great. No APT is waiting around for Anthropic's permission, so might as well let us have some cool stuff.

➕ show 1 reply
ulukaya • today at 8:14 PM

[flagged]

🔗 View 4 more comments