cd/entity/DeepSeek-V3Β· homeβ€Ί entitiesβ€Ί DeepSeek-V3
grep -l @deepseek-v3 /news/*.json | wc -l β†’ 44

DeepSeek-V3

mentions 44 type Organization page 3/3 feed RSS

// recent coverage 44 mentions

16:00
2026-05-27
dev.to
large-language-models

Why your quantized LLM loses its MTP heads and how to keep them

A developer discovered that standard quantization pipelines for large language models silently discard multi-token prediction (MTP) heads, causing speculative decoding speedups to vanish despite the b…

03:12
2026-05-27
metaworld.me
ai-research

Finding deadlocks in CuTe kernels with SPIN

Researchers at the FlashInfer MLSYS Challenge developed a formal verification method using the SPIN model checker to detect deadlocks in CuTe DSL kernels running on NVIDIA B200 GPUs. The approach, dem…

13:14
2026-05-23
dev.to
large-language-models

Multi-Head Latent Attention (MLA)

**Summary:** Multi-Head Latent Attention (MLA) is an attention mechanism used in DeepSeek-V2/V3 and Kimi K2.x models that compresses the Key-Value (KV) cache by projecting full KV pairs into a shared,…

← prev page 3 / 3
// co-occurs with top 8 entities
// topics top 6 topics