fast != practical You can get lots of tokens per second on the CPU if the entire network fits in L...

fc417fc802 • today at 4:56 AM • 0 replies • view on HN

fast != practical

You can get lots of tokens per second on the CPU if the entire network fits in L1 cache. Unfortunately the sub 64 kiB model segment isn't looking so hot.

But actually ... 3000? Did GP misplace one or two zeros there?

alt Hacker News