build1 publisher
Kaitchup traces Bonsai 2's 98.2% retention figure to unpacked weights on an H100
Ternary weights at 1.76 bits put a 27B model into a 5.9GB file and let a laptop decode it at 28.1 tokens a second. The retention figure comes from Prism ML's own benchmark suite, not the table on the model card.
Publishers:dev.to
Reality
- Evidence38
- Adoption42
- Hype gap+34
- Incentives74
- Confidence46