cd/entity/DeepSeek-V4-Flash-0731· home entities DeepSeek-V4-Flash-0731
grep -l @deepseek-v4-flash-0731 /news/*.json | wc -l → 29

DeepSeek-V4-Flash-0731

mentions 29 type Organization page 2/2 feed RSS

// recent coverage 29 mentions

10:00
2026-08-04
github.com
artificial-intelligence

DeepSeek V4 Flash on a Single AMD MI300X

A production configuration for running DeepSeek-V4-Flash-0731 on a single AMD MI300X GPU achieves 168.6 tok/s median single-stream decode and 542 tok/s aggregate across 8 concurrent streams, with the …

20:34
2026-08-03
twitter.com
artificial-intelligence

DeepSeek-V4-Flash 2.98x faster on 4x B200, lossless

RunInfra AI optimized DeepSeek-V4-Flash-0731, achieving a 2.98x speedup on 4x B200 GPUs, with median latency reduced from 9248 ms to 3095 ms and throughput increased from 113 to 363 tokens per second,…

15:40
2026-08-03
thecoinheadlines.com
artificial-intelligence

Alibaba unveils 2.4-trillion-parameter model as DeepSeek cuts costs

Alibaba Cloud unveiled Qwen3.8-Max on Monday, its largest AI model with 2.4 trillion parameters and a one-million-token context window, while DeepSeek released DeepSeek-V4-Flash-0731 in public beta on…

13:32
2026-08-03
thenewstack.io
artificial-intelligence

DeepSeek’s smaller model just outperformed its own flagship

DeepSeek released DeepSeek-V4-Flash-0731, a smaller model that outperforms its flagship in agent performance without altering the core architecture, according to The New Stack. The model's open weight…

23:59
2026-07-31
simonwillison.net
artificial-intelligence

deepseek-ai/DeepSeek-V4-Flash-0731

DeepSeek released DeepSeek-V4-Flash-0731, a model that Artificial Analysis ranks ahead of MiniMax M3, a 428B model, and at $0.14 per million input tokens and $0.27 per million output tokens it may be …

← prev page 2 / 2
// co-occurs with top 8 entities
// topics top 6 topics