logoalt Hacker News

deltaqueueyesterday at 7:12 PM2 repliesview on HN

That "TPU advantage" might be slowing Google down (though likely not as much as their internal bureaucracy).

Porting CUDA-based research, debugging, and overall experimentation speed is likely slower.

The GPU is still king for training.


Replies

awonghyesterday at 11:38 PM

But maybe the TPU advantage is in inference? That's what I assume because the number of compute cycles are going to be all in inference vs training. So they could train on GPUs if they want.

anthonypasqyesterday at 9:09 PM

lmao, you know all Anthropic models are trained on TPU right?

show 1 reply