logoalt Hacker News

mh-yesterday at 6:15 PM0 repliesview on HN

To my understanding, with batched inference and other "optimizations" you wouldn't get the exact same token predictions even with temp=0.0.