cd/entity/BF16· home entities BF16
grep -l @bf16 /news/*.json | wc -l → 8

BF16

mentions 8 type Organization feed RSS

// recent coverage 8 mentions

00:00
2026-08-11
mindstudio.ai
large-language-models

How to Run Nemotron 3.5 Lightning Locally on Your Own GPU

NVIDIA's open-weights Nemotron 3.5 Lightning mixture-of-experts model, with 30 billion total parameters but only 3 billion active, is designed for agentic grunt work and can be run locally on consumer…

20:57
2026-07-31
dev.to
machine-learning

Why INT4 Weight-Only Quantization Doesn't Speed Up Prefill

A developer's analysis shows that INT4 weight-only quantization speeds up decode but not prefill, because prefill is compute-bound while decode is memory-bound. The crossover point where a GEMM become…

14:58
2026-06-06
vettedconsumer.com
large-language-models

GGUF vs. GPTQ vs. AWQ: The Plain-English Guide to LLM Quantization

GGUF, GPTQ, and AWQ are the three dominant formats for running quantized large language models locally, each optimized for different hardware and use cases. GGUF, the format used by llama.cpp and its …

13:50
2026-04-27
x86ecosystem.org
artificial-intelligence

ACE: A Shared Path to Faster Matrix Math on x86

AMD and Intel, through the x86 Ecosystem Advisory Group, are standardizing new matrix math instructions called ACE (AI Computation Extensions) to accelerate AI workloads on x86 processors. ACE introdu…

// co-occurs with top 8 entities
// topics top 6 topics