logoalt Hacker News

raggi • today at 9:17 PM • 1 reply • view on HN

Some vague commentary about performance with what appears to be assumptions about GPU availability, but no clarity about which inference provider is being used. If ZAI is assumed, I believe they aren't subject to the assumptions in the post based on what they've said publicly, but if they were using some other provider, perhaps.

The second reason appeared to be simply "because we chose not to". The post seems to be pretty much content-less in any practical sense. I clicked on it because I do quite like this models average performance and I was hoping to see some kind of review content.


Replies

ThibWeb • today at 9:34 PM

Might do more of that next time! Inference was with TensorX and Neuralwatt.