Jev in 25 lines of Python
NobodyWho published a parody blog post on September 22, 2026 showing a Jev-style decision model implemented in 25 lines of Python using the Qwen3-0.6B-GGUF model via llama-cpp-python, which classifies…
NobodyWho published a parody blog post on September 22, 2026 showing a Jev-style decision model implemented in 25 lines of Python using the Qwen3-0.6B-GGUF model via llama-cpp-python, which classifies…
The open-source Rust inference library behind the NobodyWho Chat test app can register for OS memory-warning signals to automatically unload LLM models before Android and iOS Out Of Memory killers ter…
NobodyWho Edge runs GGUF models through llama.cpp with no conversion step, while Cactus uses its own Cactus Quants (CQ) rotation-and-codebook quantization format from 4-bit down to 1-bit, according to…
NobodyWho, an open-source project providing a Rust core for on-device LLM inference via llama.cpp, announced support for Expo, enabling developers to run large language models entirely on users' phone…
A benchmark of Gemma4-E4B on an AMD Ryzen 7040 CPU shows that using 8 threads instead of 16 improves prompt processing from 94.66 to 105.94 tokens per second and token generation from 11.64 to 16.15 t…
NobodyWho announced speech support in its on-device inference engine, adding Text to Speech (TTS) via Kokoro, Pocket TTS, and Supertonic, and Speech to Text (STT) via Whisper, all running on ONNX Runt…
NobodyWho released NobodyWho Chat, a fully offline AI assistant for iOS and Android that runs open-weight models locally on the device, ensuring no data leaves the phone and no internet connection is …