AI tool comparison
Bonsai-8B vs Dune
Which one should you ship with? Here is the side-by-side panel verdict, pricing read, reviewer split, and community vote comparison.
Infrastructure
Bonsai-8B
A true 1-bit 8B LLM that fits in 1.15 GB — runs on your iPhone
75%
Panel ship
—
Community
Free
Entry
Bonsai-8B is PrismML's latest model in their BitNet-inspired lineage — an 8.2B parameter language model that has been quantized end-to-end to true 1-bit precision (weights stored as -1 or +1), compressing the entire model to just 1.15 GB. That's roughly 12-14x smaller than a standard FP16 equivalent. Unlike post-training quantization hacks that lose substantial quality, PrismML trained Bonsai-8B with 1-bit arithmetic baked into the forward pass from the start. Benchmark results are competitive for the size class: 63.8 on MMLU, 72.1 on HellaSwag, and 54.2 on GSM8K — while running at 131 tokens/sec on an M4 Pro MacBook and 44 tokens/sec on an iPhone 17 Pro Max. That makes it the fastest locally-runnable 8B model in its weight class on Apple Silicon. The MLX-optimized weights are available on Hugging Face today under Apache 2.0. The significance goes beyond benchmarks. Getting a capable open-weight model to run at interactive speeds on consumer hardware — with no API key, no GPU, no cloud dependency — is a meaningful step toward truly private, offline AI. This follows PrismML's earlier "Ternary Bonsai" (1.58-bit) but represents a cleaner binary architecture that's easier to accelerate on custom silicon.
Hardware
Dune
A 3-key CNC aluminum keypad that reads your context and adapts
75%
Panel ship
—
Community
Paid
Entry
Dune is a tiny CNC-machined anodized aluminum keypad (40×10×10mm, 50g) from Project Mirage that ships three programmable physical keys alongside context-aware AI logic — automatically detecting your active macOS app and updating key assignments with no manual setup. It's the closest thing yet to a physical MCP client. The hardware handles the meetings problem elegantly: one-click join for Zoom, Teams, and Google Meet with calendar sync, dedicated mic/camera toggles, and instant meeting-window focus. But the broader promise is context adaptation: keys that behave differently when you're in your editor vs. your browser vs. your design tool, without you needing to define profiles. USB-C powered, macOS only, shipping in May 2026 with early bird pricing. Project Mirage has 8+ years of hardware experience and the form factor is genuinely minimal — a sliver of machined metal on your desk rather than another chunky macro pad. The open question is how deep the context awareness goes and whether the AI layer is smart enough to be useful rather than occasionally wrong and annoying. Early Product Hunt reception was strong (608 votes, top of leaderboard), suggesting there's real appetite for physical AI interfaces.
Reviewer scorecard
“131 tokens/sec on M4 Pro at 1.15 GB is genuinely impressive — I can embed this in a macOS app without any cloud dependency, no rate limits, no privacy concerns. The Apache 2.0 license means I can ship commercial products on top of it. This is the edge AI story I've been waiting for.”
“The primitive here is dead simple and correct: an HID device whose key mappings are driven by a macOS accessibility API hook watching the frontmost application — the AI layer handles the mapping logic so you don't write profiles by hand. That's the right DX bet. The moment of truth is day two, not day one: does the context inference hold up when you have twelve apps open and you're alt-tabbing between your editor and a Slack thread? If the answer is yes, this is the macro pad I'd actually leave plugged in. The specific decision that earns a ship from me is that they rejected the 'define every profile yourself' pattern that killed every Stream Deck workflow I've ever set up.”
“63.8 on MMLU is respectable but it's still noticeably behind mid-range cloud models on reasoning tasks. The GSM8K score of 54.2 means it'll fumble multi-step math that users expect to just work. Until 1-bit gets to 70B scale, it's a neat demo that falls short in production use cases where quality matters.”
“Direct competitor is the Stream Deck Mini plus a $10/yr Keyboard Maestro license, which already does context-aware macro switching with zero AI ambiguity. The specific scenario where Dune breaks is the one that happens constantly: two apps open side-by-side, ambiguous context, and three keys that do the wrong thing because the model guessed wrong — that's worse than a dumb macro pad, not better. What kills this in 12 months is Apple shipping Focus-mode-aware Shortcuts automation natively in macOS 16, at which point the software layer this hardware depends on is commoditized. To earn a ship: show me six months of real-world context accuracy data, not a Product Hunt leaderboard.”
“The trajectory here is what matters: 1-bit models are getting faster to train and competitive faster than expected. When custom Apple Neural Engine kernels land for BitNet-style weights, we'll see 200+ tokens/sec on a phone. Bonsai-8B is the proof-of-concept that makes that future feel real.”
“The thesis Dune is betting on: within three years, AI context awareness will be accurate enough that zero-configuration physical controls outperform manually-configured ones, and users will pay a hardware premium for that. That's a falsifiable claim riding a specific trend line — on-device app-state inference getting cheap enough to run as a background daemon — and Project Mirage is early, not late, to it. The second-order effect nobody is talking about: if this works, it inverts the macro pad market from a power-user niche into a normie peripheral, because the configuration tax that kept civilians away disappears. The future state where this is infrastructure is a desk where every physical control knows what you're doing without being told.”
“I've been looking for something I can embed in a creative writing or brainstorming app that doesn't require an internet connection. At 44 tokens/sec on iPhone, Bonsai-8B is finally fast enough to not break the creative flow. The 'no account required' angle is a genuine selling point for privacy-conscious users.”
“The job-to-be-done is singular and clear: stop context-switching your hands when your screen context already switched. The meetings use case is the product's sharpest edge — calendar sync plus one-click join plus mic/camera toggles is a complete workflow replacement, not a feature — and that alone justifies the purchase for anyone on four-plus calls a day. The product has a real opinion: it decides your key assignments, you don't. That's brave and almost certainly right. The gap that would turn this ship into a skip is if the broader context-awareness layer — editor vs. browser vs. design tool — turns out to be shallow window-title matching dressed up as AI; ship the meetings story hard and make everything else a bonus.”
Weekly AI Tool Verdicts
Get the next comparison in your inbox
New AI tools ship daily. We compare them before you waste an afternoon.