Caltech spinout PrismML released Bonsai 8B, a 1-bit large language model that compresses AI into just 1.15 GB of memory—14 times smaller than full-precision counterparts—while running 8 times faster and using 5 times less energy on edge hardware. Founded by Caltech professor Babak Hassibi, PrismML's architecture represents weights as simple signs {-1, +1} with shared scale factors, achieving an intelligence density score of 1.06/GB compared to 0.10/GB for competing model Qwen3 8B, though Qwen3 still leads on traditional benchmarks. The model runs natively on Apple devices and Nvidia GPUs, with PrismML positioning the technology to enable on-device AI agents, real-time robotics, and secure enterprise systems without cloud dependency.
Sep 29, 2026 · 4 sources
Sep 29, 2026 · 13 sources
Sep 28, 2026 · 6 sources
Sep 30, 2026 · 7 sources
Story comments
Loading comments…