AERIOXFLUX
AI Tools
open sourceNew

PrismML

Frontier-class open-weight models compressed to run on a phone

weight 0.0Open SourceLaunched 2026-08-10

💸 No earnings reported yet

What it is

PrismML builds ultra-dense quantized language models under the Bonsai line. Bonsai 27B compresses Qwen3.6-27B to 1.125 bits per weight — 3.9GB for the 1-bit build, 5.9GB for the ternary — retaining 89.5% and 94.6% of the FP16 baseline respectively, with a 262K context and a 4-bit vision tower for multimodal input. The 1-bit build runs at roughly 11 tokens/sec on an iPhone 17 Pro Max and 66 tok/s on an M5 Max. Smaller Bonsai 8B, 4B and 1.7B models plus Bonsai Image are also available. Weights ship on HuggingFace under Apache 2.0.

How AI plugs in

Alternatives & related tools

★ Reviews

No reviews yet — be the first.

Your rating

Discussion (0)

The state of AI, in flux.

The directory + magazine for AI tools and the workflows people use to make money with them.

🔥 The Sauce Drop

The week's highest-earning AI workflows, in your inbox.

Some outbound links are affiliate links — Flux may earn a commission at no cost to you; this never affects rankings. Earnings figures are self-reported and not guarantees of income; most people earn less, some earn nothing.