Laguna S 2.1 scored really low on AlmanBench, even lower than Ternary Bonsai 27B
A 118B, 8B active parameter model, scored lower in FB8 precision, than a ternary quantized 27B model 🤨
Did the pretraining mix lack multilingual data by a lot? 🤔
https://t.co/OGBSm0LYlh
We bring you the latest updates from Onur Solmaz blog through a simple and fast subscription.
We can deliver your news in your inbox, on your phone or you can read them here on this website on your personal news page.
Unsubscribe at any time without hassle.
Onur Solmaz blog's title: Onur Solmaz blog | Explorations in software, math, languages and more.