logoalt Hacker News

thousand_nightsyesterday at 6:43 PM4 repliesview on HN

i've never had to use control + o before but with the latest changes, i give Opus a simple task that should take a few seconds and it's like "used 15k tokens" and "thinking" for three minutes with absolutely zero indication or visibility as to what it's actually doing and i have to ESC ESC it to stop and ask what the FUCK are you actually doing claude?


Replies

misnomeyesterday at 6:58 PM

Yes, I’ve been evaluating since the start of the year and since 4.6 suddenly the most innocuous requests will sit there “thinking” for 5+ minutes and if I can get it to show me the thinking it’s just going round in circles.

Or, it decided it needs to get API documentation out and spends tens of thousands of tokens fetching every file in a repo with separate tool use instead of reading the documentation.

Profitable, if you are charging for token usage, I suspect.

But I’m reaching the point where I can’t recommend claude to people who are interesting in skeptically trying it out, because of the default model.

show 1 reply
scottyahyesterday at 7:12 PM

Yeah after my switch to Opus 4.6 I noticed a lot of this. I've been wary that eventually models are going to optimize for token usage increases, since that's how the company makes money. I told it to read the files in my directory (4 files, longest was like 380 lines) and caught it using 14 tool uses- including head -n 20 and tail -n 20 on a file. Definitely a what are you doing moment.

show 1 reply
8noteyesterday at 9:57 PM

i think yesterday it ate the whole context window in one thinking call.

i bet in a week itll eat the whole 5hour throttle in one call too:P

virtue3yesterday at 6:52 PM

I think this change is really disingenuous.

If they hide how the tool is accessing files (aka using tokens) and then charging us per token - how are we able to track loosely what our spend is?

I’m all for simplification of the UX. But when it’s helping to hide the main spend it feels shitty.