logoalt Hacker News

MisterMunchkin • today at 7:21 PM • 14 replies • view on HN

My company took away my Claude because it’s too expensive. I feel like there is a reckoning coming. The accountants are finally realising the cost of token maxing.


Replies

vablings • today at 7:27 PM

That's pretty stupid. Most people who are incurring significant costs are just tokenmaxxing rather than being efficient with usage. You can get 99% of jobs and work done with Haiku/Luna in a collaberating working enviroment.

I feel like people who are later to the AI game just like to "oneshot" and sink a bunch of usage into generating garbage

➕ show 7 replies
compiler-guy • today at 9:44 PM

That's funny. In a meeting with my manager recently, they specifically called me out for not spending enough money on tokens. It's not like I didn't use it, just apparently not enough, and apparently on too low a setting.

Since then I've had fable cranked up to 11 for even the most trivial of tasks.

233mhz • today at 10:03 PM

I know plenty of people working at very well known and very large companies who's CEOs were boasting about "not hiring anyone anymore", "all the code will be generated in 3 months", etc. they all went from "unlimited budget per dev" + public dashboard with ranking to flex how much credits everyone was burning to hard caps at $500-$1000/month/employee real fast. Some are even not allowing their devs to use the more expensive models

password54321 • today at 7:38 PM

You realise the subject here is Meta, which is all in on this stuff? Of course they are going to use Muse Spark over Claude.

>Great Depression style collapse and all the current AI companies go bankrupt.

Oh this is just a 33 day old doomer account.

➕ show 1 reply
tty456 • today at 8:59 PM

Do you know the details of the Claude Code plan you and your company are using (if not part of some enterprise deal)? Does your individual capacity out run something like Claude Max 20x ($200/mo)?

inferniac • today at 9:35 PM

taking away sounds insane, we had basically unlimited tokens (inference bought from aws) and they moved us back to the $100 sub to save money

woah • today at 8:24 PM

$200 a month is too expensive yet they employ human developers?

➕ show 3 replies
user43928 • today at 9:09 PM

Good luck to the accountant that tries to tell leadership to slash AI usage.

I'm sure investors will love it.

Now we're starting to see real impact from AI, people are learning how to use it, and OpenAI cut prices by no less than 60% like a week ago.

You think now is the time they're going to cut the spend?

yeahBoiii • today at 8:41 PM

The reckoning started years ago when we did the equivalent to token maxxing hiring coders for everything to crank LOC

Software is inherently a physics problem not all the job titles and specializations made up the last 20 years as dev job salaries kept attracting people

That was all illusory social construct to prop up jobs

Still a whole lot of that in tech but it's all at the top of the org now. Leadership sensory experience and thus innate habit to forecast future been programmed by years of yes men they refuse to accept the jig is up for them too

Sensory memory of being a useless figurehead fosters a lot of existential dread in priests, politicians, and the like. Completely aware their day to day effort is insufficient to sustain them they know how co-dependent they are. They'll dig in harder.

See Chris Matthews flame out shrieking about socialist execution squads. Dude seriously thought everyone wants to hang him from a lamp post. The reality is people just want a sense of control back and not have their perception dragged along by Chris Matthews.

AIblemblio • today at 8:43 PM

The reckoning will be throwing out all external help, then reducing team sizes.

lenerdenator • today at 8:17 PM

We're just getting put on a budget.

Our velocity is twice as high as it was before Claude, so I doubt that we'll ever go back, but I could see efficiency being a priority.

➕ show 2 replies
righthand • today at 8:15 PM

My thoughts were the reckoning would come when Infra teams started offloading AWS usage to LLMs and ended up token maxing and deploy maxing.

dyauspitr • today at 8:26 PM

I mean, we’re not far from a situation where instead of how many story points you completed per sprint the metric to optimize is going to be what was your efficiency? How many story points did you complete while minimizing your token usage. In fact, that’s a pretty good idea. I’m going try and implement it at work with some sort of complexity normalization function

sergiotapia • today at 7:26 PM

which is quite sad because opus 5.5 is really good. i say this as an anthropic hater. i wish I could move away to other models like 6.1 sol or deepseek or whatever, but they just all lack something. i _trust_ opus 5.5

i hope other labs catch up, especially chinese labs.

➕ show 1 reply