Post #5863 263 Sep 18, 2026, 04:14 UTC 大模型一定要堆显存吗?有人把270亿参数的模型压到5.9GB:权重只留-1、0、1三种取值,跑分还剩原来的98%,一张游戏显卡就能跑,耗电比80亿参数、没压缩过的模型还低四成。想在本地用编程助手、又不想把代码传上云的,看完会有点心动。https://prismml.com/news/bonsai-2-27b Prismml PrismML — Introducing Bonsai 2 27B: Near-Lossless Compression in a 9x Smaller Footprint Ternary Bonsai 2 27B retains 98.2% of Qwen3.8 27B benchmark performance in a 5.9GB footprint, with multimodal and agentic capabilities.