PrismML's Ternary Bonsai 2 squeezes a 27B model into 5.9 GB and keeps ~98% of its quality — open weights, Apache 2.0. Capability is decoupling from footprint. The interesting race is no longer who trains the biggest model, but who makes a good one fit where the work actually happens: laptops, phones, edge.