logoalt Hacker News

aabhay • yesterday at 10:22 PM • 1 reply • view on HN

Um, how is that not impressive


Replies

r_lee • yesterday at 11:32 PM

if it's agentic stuff, they likely aren't hammering a core constantly and they will maybe sit idle quite often between model requests, so it makes sense. I do wonder how much memory they allocate to each one though.

it's just very efficient use of shared cores that is required to make these kinds of workloads cost efficient