> "GPUs don't do deterministic matrix multiplications" is the biggest source of ra...

jstanley • today at 8:22 AM • 1 reply • view on HN

> "GPUs don't do deterministic matrix multiplications" is the biggest source of randomness in LLMs.

But this isn't a fundamental property of LLMs, it's just an implementation detail. It's pretty obvious that if you evaluate the matrix multiplications correctly and deterministically sample from the highest-probability outputs, you will have a deterministic LLM.

Replies

vbarrielle • today at 8:38 AM

It may be an implementation detail, but in practice, if the only way to get a deterministic output is to run on the CPU, then it's not going to be usable.

➕ show 2 replies

alt Hacker News

Replies