cd/entity/MXFP4· home› entities› MXFP4
grep -l @mxfp4 /news/*.json | wc -l → 6

MXFP4

mentions 6 type Organization feed RSS

// recent coverage 6 mentions

07:21
2026-10-03
forum.level1techs.com
ai-infrastructure

Quad R9700's on AM4 Shenanigans

A user on the Level1Techs forum reported running four AMD Radeon R9700 GPUs on an AM4 platform to requantize a DeepSeek model to MXFP4 and run it in the vLLM engine, with the first layer of a 48-layer…

02:46
2026-09-02
forgeeks.net
machine-learning

LLM inference now has two ways to get cheaper

Baseten's technical breakdown of LLM inference efficiency identifies two categories of engineering choices: those that trade latency for throughput, such as batch sizing, tensor parallelism, expert pa…

09:35
2026-07-28
twitter.com
artificial-intelligence

Someone runs K3 on 80x 5090s, for 20 tok/s

A team has run the full Kimi K3 model, a 2.8-trillion-parameter mixture-of-experts (MoE) model, on 80 RTX 5090 GPUs achieving 20 tokens per second single-stream inference on day one without tuning, ma…

// co-occurs with top 8 entities
// topics top 6 topics