cd/entity/Cactus Compute· home› entities› Cactus Compute
grep -l @cactus compute /news/*.json | wc -l → 10

Cactus Compute

mentions 10 type Person feed RSS

// recent coverage 10 mentions

05:40
2026-09-24
raspberrypi.com
large-language-models

Turn Text into Actions with Needle, 14Mb

Cactus Compute released Needle 2, a 14MB function-calling LLM that runs entirely on a Raspberry Pi 5's CPU and selected the correct tool for the prompt "Turn the LED on" in 78 milliseconds. The model,…

00:00
2026-09-20
mindstudio.ai
ai-products

Needle 3: The 8-29MB Model Built for On-Device Tool Calling

Cactus Compute released Needle 3, an open-weight foundation model that ships as a single 8-29MB file for on-device tool calling, structured extraction, and embeddings, with a runtime engine under 1MB …

00:00
2026-09-20
mindstudio.ai
ai-products

Needle 3: Running a Tiny On-Device Tool-Calling AI Model

Cactus Compute released Needle 3, a single-file on-device foundation model that ships between 8 and 29 MB depending on layer depth and handles tool calling, structured extraction and text embeddings w…

00:00
2026-09-20
mindstudio.ai
ai-research

Needle 3 Benchmarks: How a Tiny Model Beats 10x Larger LLMs

Cactus Compute's Needle 3, a foundation model shipping as a single 8-29 MB file, beats models ten times its size on mobile tool-calling accuracy and matches models two to three times larger on structu…

07:11
2026-09-19
byteiota.com
ai-tools

Cactus Needle 3: 8MB On-Device AI Without the API Bill

Cactus Compute released Needle 3 on September 17, an Apache 2.0-licensed automation foundation model that ships as an 8 to 29MB binary and runs tool calling, structured extraction, and text embeddings…

05:33
2026-08-26
dev.to
artificial-intelligence

Needle 2: the 14 MB agentic model, tested properly

Cactus Compute released Needle 2, a 14 MB agentic model designed for tool calling, device control, and structured data extraction on edge devices. A developer's hands-on testing revealed several bugs,…

06:12
2026-08-11
byteiota.com
artificial-intelligence

Needle2: A 14MB LLM Runs AI Agents on a Raspberry Pi

Cactus Compute released Needle2, a 45-million-parameter language model compressed to a 14MB binary that runs AI agent tool-calling at 500 tokens per second on a Raspberry Pi 5 and 6,000 tokens per sec…

21:29
2026-08-10
promptcube3.com
artificial-intelligence

Needle 2 fits a functional LLM into just 14MB

Cactus Compute's Needle 2 model fits a functional LLM into just 14MB, using Simple Attention Networks to reduce per-token computation to 70 MFLOPs, compared to 164 MFLOPs for a standard transformer of…

21:19
2026-08-10
runtimewire.com
artificial-intelligence

Cactus ships 14MB Needle 2 for tool calling on cheap devices

Cactus Compute, a Y Combinator Summer 2025 startup, released Needle 2, a 14MB language model for tool calling on low-memory devices, capable of running on phones, wearables, home devices, robots, and …

// co-occurs with top 8 entities
// topics top 6 topics