cd/entity/Bonsai· home entities Bonsai
grep -l @bonsai /news/*.json | wc -l → 8

Bonsai

mentions 8 type Organization feed RSS

// recent coverage 8 mentions

17:12
2026-08-27
bonsai.io
ai-infrastructure

Float Bloat: vector serialization gone wrong

Bonsai, a vector search company, has identified a pervasive issue it calls 'Float Bloat' where embedding vectors are cast from float32 to float64 during serialization, doubling storage and network cos…

13:55
2026-08-25
thejeshgn.com
artificial-intelligence

Exploring 1-Bit LLMs

Microsoft Research released BitNet b1.58 2B4T, the first open-source native 1-bit LLM at the 2-billion parameter scale, which uses ternary weights (-1, 0, +1) requiring about 1.58 bits per weight, red…

09:10
2026-08-05
sourcefeed.dev
artificial-intelligence

How a 20B Model Hits 120 tok/s on an iPhone

DeepGrove's Maple-Preview, a 20B-parameter mixture-of-experts reasoning model with ternary weights, achieves 120 tokens per second on an iPhone and 218 tok/s on a base M4 Mac mini, according to the co…

19:42
2026-07-22
huggingface.co
artificial-intelligence

Bonsai 27B parameters. 1-bit weights. In your browser

A new AI model called Bonsai with 27 billion parameters and 1-bit weights can now run directly in a web browser, as demonstrated by the WebML Community on Hugging Face. The model's extreme quantizatio…

07:00
2026-06-08
smolhub.com
large-language-models

Bonsai LLM Benchmark: Jetson Orin Nano Super 8GB

NVIDIA Jetson Orin Nano Super 8GB benchmarks show 25W as the energy-efficiency sweet spot for sub-4B Bonsai LLMs, delivering 47-48% more tokens per second than 15W while maintaining or improving outpu…

// co-occurs with top 8 entities
// topics top 6 topics