cd/entity/DeepSeek-V4· home entities DeepSeek-V4
grep -l @deepseek-v4 /news/*.json | wc -l → 19

DeepSeek-V4

mentions 19 type Organization feed RSS

// recent coverage 19 mentions

14:00
2026-08-31
kdnuggets.com
large-language-models

Speed Up LLM Inference with DSpark Speculative Decoding

DeepSeek's DSpark speculative decoding technique, which combines parallel drafting with a lightweight sequential component, can improve local LLM generation speed on the same GPU, with DeepSeek report…

08:04
2026-07-31
dataleadsfuture.com
artificial-intelligence

How I Cut Kimi K3 Costs in OpenCode

A developer reports cutting Kimi K3 costs in OpenCode by switching to the k3-256k model, which consumes half the quota of the standard K3, and by choosing the right provider such as Novita.ai. The All…

00:00
2026-07-31
quesma.com
large-language-models

A lesson about retries, hidden in the DeepSeek-V4 paper

An experiment by an unnamed researcher found that retrying failed requests in LLM benchmarks introduces selection bias, shortening average response lengths by 19.2% and reducing long-form outputs by u…

05:09
2026-07-21
sourcefeed.dev
large-language-models

Why Every LLM Vendor Killed the Thinking-Token Budget

Every major LLM vendor has replaced numeric thinking-token budgets with discrete effort levels, converging on a small enum instead of a number. Anthropic now rejects budget_tokens with a 400 error on …

06:29
2026-06-30
venturebeat.com
large-language-models

DeepSeek Open Sources DSpark

Chinese AI firm DeepSeek open-sourced DSpark, a speculative decoding system that accelerates large language model inference by up to 85% without altering output quality, releasing it under the MIT lic…

00:00
2026-06-24
rocm.blogs.amd.com
machine-learning

DP Attention and TBO for DeepSeek-V4 on MI355X

AMD introduces DP Attention and Two-Batch Overlap (TBO) optimizations for DeepSeek-V4 inference on MI355X GPUs, using a coordinated prefill scheduler called PrefillDelayer to reduce padding waste and …

06:19
2026-06-15
dataleadsfuture.com
large-language-models

DeepSeek-V4 Can't Read Images? I Made It Read

A developer created a plugin called 'observer' for OpenCode that enables the DeepSeek-V4 language model to read images by calling a multimodal agent, allowing it to interpret error screenshots, charts…

00:00
2026-04-24
huggingface.co
large-language-models

DeepSeek-V4: a million-token context that agents can actually use

DeepSeek-V4 introduces a new architecture using hybrid attention mechanisms—Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA)—to drastically reduce the computational cost and me…

// co-occurs with top 8 entities
// topics top 6 topics