PrismML on July 14, 2026 released Bonsai 27B, a pair of extreme low-bit quantizations of Qwen3.6 27B, with a 3.9GB, 1-bit build that the company says is the first model of its size to run on a smartphone — specifically the iPhone 17 Pro at roughly 11 tokens per second.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.