←Back to NewsAI News/LlmsmodelLlmsGpusPrism ML's Ternary Bonsai 2 Squeezes a 27B Reasoning Model Into 8.6 GBPrism ML's Bonsai 2 27B compresses a 27B reasoning model to 8.6 GB using ternary weights, retaining 98.2% of FP16 benchmark quality.SourcePrismMLPublishedSep 16, 2026, 11:41 PMAuthorAlphaSignal NewsroomRead1 min readPrism ML's Bonsai 2 27B compresses a 27B reasoning model to 8.6 GB using ternary weights, retaining 98.2% of FP16 benchmark quality.Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.Read original report ↗Next readsPrismML · newsPrismML Squeezes Qwen3.8 27B Into 5.9 GB With 98% Performance RetainedAlphaSignal · modelEmpero Distills Qwen3.8 Reasoning Into a Lean 35B Open ModelGoogle Gemma · newsGoogle Ships Gemma 4 12B to Run Offline on a 16GB MacBook Air