sanitation@lemmy.today to Technology@lemmy.worldEnglish · 1 day agoIf open weight models are the future, U.S. AI companies are going to have a hard timewww.fastcompany.comexternal-linkmessage-square82linkfedilinkarrow-up1256arrow-down12
arrow-up1254arrow-down1external-linkIf open weight models are the future, U.S. AI companies are going to have a hard timewww.fastcompany.comsanitation@lemmy.today to Technology@lemmy.worldEnglish · 1 day agomessage-square82linkfedilink
minus-squareMwa@thelemmy.clublinkfedilinkEnglisharrow-up4·18 hours agoWe even got open weight models that’s 27B + 1-bit (and it still has good performance)
minus-squarebrucethemoose@lemmy.worldlinkfedilinkEnglisharrow-up3·11 hours agoBonsai? Or whatever it’s called? It’s a con, so far; it’s not better than smaller models quantized to 3-4 bits. I love, love the idea of bitnet, but it only seems to work with models trained from scratch, which no one has done at scale yet.
We even got open weight models that’s 27B + 1-bit (and it still has good performance)
Bonsai? Or whatever it’s called? It’s a con, so far; it’s not better than smaller models quantized to 3-4 bits.
I love, love the idea of bitnet, but it only seems to work with models trained from scratch, which no one has done at scale yet.
yeah bonsai