ESP32 BitNet Cluster Runs a 0.4B LLM for $28
Developer Low-Zi-Hong published a working distributed language model inference engine on GitHub that runs a 0.4-billion-parameter model across seven ESP32-S3 microcontrollers in a SPI daisy-chain pipe…
Developer Low-Zi-Hong published a working distributed language model inference engine on GitHub that runs a 0.4-billion-parameter model across seven ESP32-S3 microcontrollers in a SPI daisy-chain pipe…
A developer built a working FIDO2 hardware security key on an ESP32-S3 running Pico FIDO2 firmware, but OpenAI's Daybreak classified the registered credential as a Passkey rather than a hardware secur…
A developer built PaperMono, an e-ink fridge-magnet shopping list running on an M5Stack PaperMono board (ESP32-S3 with e-ink touchscreen) that syncs with a mobile web page over Wi-Fi and works offline…
A developer detailed how the ESP32-S3 microcontroller paired with Google's TensorFlow Lite Micro framework can replace dedicated voice chips for wake word detection and on-device sound classification.…
GeaStack launched an open-source toolchain that compiles TypeScript, JSX, and CSS into C++ for embedded devices, desktop apps, and consoles, built on the Gea reactive web framework. The stack includes…
CoyoPedal is a standalone guitar amp and effects pedal that runs full-size Neural Amp Modeler A2 captures in real time on a Waveshare ESP32-S3-Touch-AMOLED-2.06 board, with the same firmware also comp…
A developer demonstrates person detection on the ESP32-S3 microcontroller using TinyML, achieving local inference in 80-120 milliseconds without cloud connectivity. The project uses quantized MobileNe…
Researchers introduced FORGE, a forward-only test-time adaptation method for integer-only vision models on microcontrollers, recovering 93% of gradient-based TENT's accuracy gain (+20.9 vs. +24.9 poin…
Deploying LLM-powered vision agents in chaotic school zones remains unreliable due to latency and reasoning gaps, according to a technical analysis. The article proposes a tiered architecture using on…
ShrinkRay, a new open-source command-line tool, compresses and quantizes small neural networks for microcontrollers, providing a fit verdict against a 12-chip database before firmware builds. The tool…
An ESP32-S3 microcontroller now runs a Neural Amp Modeler A2-Lite capture natively in single-precision float, producing bit-identical output to a desktop render, with all 720,000 samples matching by h…
An independent research project demonstrated that three ESP32-S3 microcontrollers can jointly compute Graph Neural Network inference on traffic data using Replicated Secret Sharing, ensuring no single…
An independent researcher demonstrated that timing side channels in multi-agent reinforcement learning (MARL) policies deployed on microcontrollers can leak information about an agent's next action. B…
Makerfabs released the MaTouch ESP32-S3 MaUWB development board, which combines an ESP32-S3 wireless MCU, a 3.95-inch touchscreen, and a UWB module for indoor positioning and ranging. The board uses a…
A developer trained a small language model from scratch on an $8 ESP32-S3 microcontroller, demonstrating that training, not just inference, is possible on ultra-low-cost hardware. The project, inspire…
Slava S. (slvDev) has optimized a 28.9M-parameter LLM to run locally on an ESP32-S3 development board at about 9 tokens/s, generating short stories on an I2C display. The project, called 'esp32-ai', r…
A 27-million-parameter reasoning model that runs entirely offline on two ESP32-S3 microcontrollers, costing about $40 in hardware, achieves 85.5% accuracy on unseen comparative adjectives, up from 63.…
A developer released a shell script that lets users test custom prompts against an already-trained ESP32-S3 model without retraining or rewriting the 15MB model partition. The script tokenizes the pro…
A developer has released a one-shot build script that automates the entire pipeline for running a 28M parameter AI model on an ESP32-S3 microcontroller, from cloning the repository to training, export…
A developer ran a 56M-parameter language model across three ESP32-S3 N16R8 microcontrollers using ESP-NOW wireless communication, achieving distributed inference of a 50.3M-parameter PLE-based transfo…