logoalt Hacker News

spyckie2today at 3:58 PM3 repliesview on HN

Google seems to have anorexia when it comes to model intelligence. They have an internal hard constraint on price per token it seems, and they are trying to squeeze out intelligence with limited compute.

I wonder if there is something with their TPU cycles that makes them want to postpone training a new model. My guess is that they have been on the same base model for 6 months and they may have waited for the next gen TPUs to train Gemini 4, which greatly limits how much intelligence they can increase and forces them to do cost efficiency increases.


Replies

JacobAsmuthtoday at 5:54 PM

Could it be that they have to serve their models to billions of users?

show 1 reply
logicchainstoday at 5:27 PM

I'd guess they did model-hardware codesign but the design ended up limiting the scaling capability of the model (i.e. they overoptimized too soon).

WarmWashtoday at 4:24 PM

Google Cloud is probably Google Deepminds biggest competitor. Big company kinda bullshit.

show 1 reply