Remix.run Logo
dofm 2 hours ago

I really had high hopes for the larger Ternary Bonsai and it feels like there is scope to improve, but I get the sense (albeit a naïve, probably not fully informed sense) that improvement can perhaps only come by training directly into ternary.

kamranjon an hour ago | parent | next [-]

I’ve actually been really impressed with the 27b model they recently released - amazing performance approaching 40 tok/s on m4 max and I didn’t run into any quality issues in the small set of tasks I tried. Haven’t gone full coding with it yet but suspect it’s better than say a 9b or 12b model.

avadodin an hour ago | parent | prev [-]

All you need is Ternary Aware Training and for AI researchers to come up with a backronym for TIT.