cd/entity/DeepSeek-V4· home entities DeepSeek-V4
grep -l @deepseek-v4 /news/*.json | wc -l → 15

DeepSeek-V4

mentions 15 type Organization feed RSS

// recent coverage 15 mentions

12:01
2026-07-26
dev.to
artificial-intelligence

Top AI Papers on Hugging Face - 2026-07-26

Hugging Face's top AI papers for July 26, 2026, feature AREX, a recursively self-improving agent for research, and SLAI T-Rex, which demonstrates full-parameter post-training of DeepSeek-V4 on Ascend …

12:01
2026-07-25
dev.to
artificial-intelligence

Top AI Papers on Hugging Face - 2026-07-25

A developer compiled the top 10 most upvoted papers on Hugging Face, covering deep research agents, post-training large models on Ascend SuperPOD, embodied visual tracking, K-12 knowledge graphs, self…

05:09
2026-07-21
sourcefeed.dev
large-language-models

Why Every LLM Vendor Killed the Thinking-Token Budget

Every major LLM vendor has replaced numeric thinking-token budgets with discrete effort levels, converging on a small enum instead of a number. Anthropic now rejects budget_tokens with a 400 error on …

06:29
2026-06-30
venturebeat.com
large-language-models

DeepSeek Open Sources DSpark

Chinese AI firm DeepSeek open-sourced DSpark, a speculative decoding system that accelerates large language model inference by up to 85% without altering output quality, releasing it under the MIT lic…

00:00
2026-06-24
rocm.blogs.amd.com
machine-learning

DP Attention and TBO for DeepSeek-V4 on MI355X

AMD introduces DP Attention and Two-Batch Overlap (TBO) optimizations for DeepSeek-V4 inference on MI355X GPUs, using a coordinated prefill scheduler called PrefillDelayer to reduce padding waste and …

06:19
2026-06-15
dataleadsfuture.com
large-language-models

DeepSeek-V4 Can't Read Images? I Made It Read

A developer created a plugin called 'observer' for OpenCode that enables the DeepSeek-V4 language model to read images by calling a multimodal agent, allowing it to interpret error screenshots, charts…

00:00
2026-04-24
huggingface.co
large-language-models

DeepSeek-V4: a million-token context that agents can actually use

DeepSeek-V4 introduces a new architecture using hybrid attention mechanisms—Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA)—to drastically reduce the computational cost and me…

// co-occurs with top 8 entities
// topics top 6 topics