cd/entity/INT4· home entities INT4
grep -l @int4 /news/*.json | wc -l → 4

INT4

mentions 4 type Organization feed RSS

// recent coverage 4 mentions

20:57
2026-07-31
dev.to
machine-learning

Why INT4 Weight-Only Quantization Doesn't Speed Up Prefill

A developer's analysis shows that INT4 weight-only quantization speeds up decode but not prefill, because prefill is compute-bound while decode is memory-bound. The crossover point where a GEMM become…

08:28
2026-07-01
promptcube3.com
artificial-intelligence

AI\'s Memory Crunch Hits Indian Smartphones

Indian smartphone manufacturers are pushing 12GB or 16GB of RAM as the new standard for mid-range devices to support on-device AI features, which consume significant memory through model weights and K…

05:56
2026-06-12
github.com
large-language-models

LLM for the ESP32-S3

Two ESP32-S3 microcontrollers running a Llama-architecture language model have achieved the first multi-chip pipelined LLM inference on ESP32-class hardware, splitting layers across two boards connect…

// co-occurs with top 8 entities
// topics top 6 topics