Meta Releases Muse Glimmer for Local Agentic AI
Meta released Muse Glimmer, a 30-billion-parameter open-weight model for local agentic work, with a 4-bit version fitting under 20 GB and tested within 24 GB or 32 GB memory envelopes on consumer hard…
Meta released Muse Glimmer, a 30-billion-parameter open-weight model for local agentic work, with a 4-bit version fitting under 20 GB and tested within 24 GB or 32 GB memory envelopes on consumer hard…
Newegg is offering a $500 discount on a Skytech Gaming O11 Vision prebuilt PC featuring an NVIDIA RTX 5090 32GB GPU and an AMD Ryzen 7 9800X3D processor, reducing the price to an unspecified amount fr…
RunInfra enables running Kimi-Linear-48B, a distilled version of the full 2.78-trillion-parameter Kimi K3 model, on a single consumer GPU such as the RTX 5090 with 32 GB VRAM, achieving 113.83 tokens …
Jeff Cui and collaborators announced Patch Policy, a robot-control method that beats OpenVLA-OFT with 0.7% of its parameters and trains on a single NVIDIA RTX 5090 GPU. The approach freezes a Vision T…
Thinking Machines' open Inkling model, with 975 billion total parameters and 41 billion active, fits on a single high-memory box at 2-bit or 3-bit quantization, unlike the datacenter-scale Kimi K3. Re…
Unified memory in mini PCs like the AMD Ryzen AI Max+ 395 allows them to run 70-billion-parameter models that exceed the VRAM capacity of high-end GPUs like the NVIDIA RTX 5090, but at significantly s…
A developer migrated a Ryzen 9 9950X3D and RTX 5090 from a full tower to an SFF open-frame custom loop with dual 240mm radiators. The system, named AI-NT-No-Problem, runs local AI inference, remote de…
FastVideo released FastWan-QAD, a family of video generation models that can produce a 5-second 480P video in 1.78 seconds on a single NVIDIA RTX 5090 using quantization-aware distillation. The models…