logoalt Hacker News

dofmyesterday at 7:48 AM2 repliesview on HN

I really had high hopes for the larger Ternary Bonsai and it feels like there is scope to improve, but I get the sense (albeit a naïve, probably not fully informed sense) that improvement can perhaps only come by training directly into ternary.


Replies

kamranjonyesterday at 9:01 AM

I’ve actually been really impressed with the 27b model they recently released - amazing performance approaching 40 tok/s on m4 max and I didn’t run into any quality issues in the small set of tasks I tried. Haven’t gone full coding with it yet but suspect it’s better than say a 9b or 12b model.

show 2 replies
avadodinyesterday at 9:03 AM

All you need is Ternary Aware Training and for AI researchers to come up with a backronym for TIT.

show 1 reply