bonsai's document reveal how much cherry picked their headlines are
Bonsai's documents reveal discrepancies in their performance claims, particularly regarding the Ternary Bonsai 2 27B model. While Bonsai initially claimed 98.2% intelligence retained, their own evaluations show Ternary Bonsai 2 27B achieving 52.8 and 60.8 on Terminal-Bench 2.1 and SWE-bench Verified, respectively. This performance is compared to Qwen3.5-27B's 69.7 and 80.6, indicating that Ternary Bonsai 2 27B retains roughly three quarters of the full-precision performance on both benchmarks, rather than the higher figure initially suggested.
This report uses Bonsai's own documents to compare its Ternary Bonsai 2 27B model's actual performance against its advertised claims, unlike other reports that only cite the company's public statements.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Sep 18, 2026, 14:00 UTC
- Ingested
- Sep 18, 2026, 14:00
- Source type
- Dev community
Discussion trend
The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.
Full text isn't available here.
Read at source →