Ternary Bonsai 2 27B Benchmarked: A 6GB Model That Actually Codes
I stress-tested Ternary Bonsai 2 27B on two Tesla T4s: 15 tokens per second on a single weak GPU, the full 262k context loading in 26GB of VRAM, and a complete todo webapp built and self-tested inside a local coding agent. The claims mostly check out, with a few catches.
