cd/entity/DeepSeek-V4-Flash-0731· home entities DeepSeek-V4-Flash-0731
grep -l @deepseek-v4-flash-0731 /news/*.json | wc -l → 29

DeepSeek-V4-Flash-0731

mentions 29 type Organization page 1/2 feed RSS

// recent coverage 29 mentions

00:45
2026-08-27
runtimewire.com
artificial-intelligence

Z.ai GLM 5.3 Flash tops Editorial Craft benchmark at 0.94

Z.ai's GLM 5.3 Flash ranked first in a 12-task Editorial Craft benchmark with a mean score of 0.94 and an estimated cost of $0.0002 per task, according to RuntimeWire. Step-3.7-flash followed at 0.92,…

08:07
2026-08-26
tokenstead.ai
artificial-intelligence

Ornith-1.5-397B

Ornith AI released Ornith-1.5-397B, a 403B-parameter mixture-of-experts agentic-coding model, on 2026-08-18, claiming coding scores on par with Claude Opus 4.8 and ahead of GLM-5.2 and DeepSeek-V4-Fla…

00:00
2026-08-21
digitalapplied.com
artificial-intelligence

DeepSeek V4 Flash Vision: Images, Same Price, Two Clocks

DeepSeek released deepseek-v4-flash-vision-exp, an experimental multimodal variant of its V4-Flash model, on August 21, 2026, matching text-only capabilities while accepting images at the same per-tok…

14:48
2026-08-19
ornith.ai
artificial-intelligence

Ornith-1.5: From Self-Scaffolding to Self-Improvement

Ornith-1.5, a new family of foundation models from the Ornith project, achieves state-of-the-art performance among open-source models of comparable size, with the flagship Ornith-1.5-397B scoring 86.1…

14:25
2026-08-19
testingcatalog.com
artificial-intelligence

Ornith-1.5 open models launch in 397B, 35B, and 9 B sizes.

DeepReinforce released Ornith-1.5, a family of open models in 397B, 35B, and 9B sizes, extending the self-scaffolding framework into a closed self-improvement loop where the system proposes tasks, gen…

17:18
2026-08-18
promptcube3.com
artificial-intelligence

Four RTX 3060s can actually push 100 tok/s prompt processing on

A developer reports achieving 99.4 tok/s prompt processing and 10.1 tok/s text generation with DeepSeek-V4-Flash-0731 (UD-Q4_K_XL GGUF) on four NVIDIA RTX 3060 12GB cards using llama.cpp build b10181,…

13:26
2026-08-15
runinfra.ai
large-language-models

DeepSeek V4 Flash at 278 tok/s, full precision, no quantization

RunInfra lists DeepSeek V4 Flash, an LLM served as deepseek-ai/DeepSeek-V4-Flash-0731, at $0.13 per 1M input tokens and $0.27 per 1M output tokens, with a 1,048,576-token context window and OpenAI-com…

23:59
2026-08-12
simonwillison.net
artificial-intelligence

DeepSeek V4 Pro 0813 (on OpenRouter)

DeepSeek released DeepSeek V4 Pro 0813 on OpenRouter, following the open-weights releases of DeepSeek-V4-Pro in April and DeepSeek-V4-Flash-0731 in July, though open weights for the new version are no…

02:22
2026-08-10
github.com
artificial-intelligence

DeepSeekV4SSD: DeepSeek-V4-Flash-0731 on an M-series Mac

DeepSeekV4SSD, an experimental app from developer yanun0323, streams routed experts from SSD to run all 284 billion parameters of DeepSeek-V4-Flash-0731 on an M-series Mac with about 30 GB of memory, …

17:01
2026-08-04
pub.towardsai.net
artificial-intelligence

DeepSeek’s $0.14 Model Just Beat Its Own Flagship at Agent Work

DeepSeek released DeepSeek-V4-Flash-0731 on July 31, a $0.14 per million input tokens model that outperforms its own flagship, a 1.6 trillion parameter model, on all nine agent benchmarks published by…

13:10
2026-08-04
sourcefeed.dev
artificial-intelligence

DeepSeek V4 Flash on One AMD GPU Took Nine Patches

A single AMD MI300X GPU with 192 GB of HBM3 now serves DeepSeek's 284B-parameter DeepSeek-V4-Flash-0731 checkpoint in mixed FP4+FP8 format, requiring nine patch overlays against a vLLM ROCm nightly pl…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics