logoalt Hacker News

kevincoxtoday at 11:52 AM1 replyview on HN

This isn't latency bound, it is trivially parallelize. So you want to run it on the most efficient compute you have, not the fastest.


Replies

andaitoday at 2:50 PM

OpenAI was bragging the other day about how HLE on Cerebras only took 11 hours. But you're saying if they just spun up a separate cheapo VM for each task, it would have finished faster and cheaper?