cd/entity/DeepSeek V4 Pro· home› entities› DeepSeek V4 Pro
grep -l @deepseek v4 pro /news/*.json | wc -l → 143

DeepSeek V4 Pro

mentions 143 type Person page 2/8 feed RSS

// recent coverage 143 mentions

16:58
2026-08-31
enclave.ai
artificial-intelligence

We Raced Seven AI Models to RCE

In a controlled race, OpenAI's GPT 5.6 Sol achieved verified remote code execution in 9 of 11 vulnerable-target runs, the most among seven AI models tested, while Zhipu AI's GLM 5.3 verified 8 runs. T…

00:00
2026-08-31
digitalapplied.com
ai-infrastructure

Where to Actually Run Kimi, GLM, DeepSeek and Qwen

A new census of open-weight model hosting reveals that the same model name can arrive at different quantizations, context lengths, and output caps depending on the provider, with some hosts serving GL…

08:11
2026-08-30
byteiota.com
large-language-models

Tencent Hy4 Open-Weights Model: Benchmarks and Access

Tencent released Hy4 preview on August 28, a 770B parameter open-weight model under Apache 2.0, with only 49B active parameters per token and a 1 million token context window, targeting software engin…

16:05
2026-08-27
lockstep.greg.technology
artificial-intelligence

Lockstep – Can a language model run a logic circuit in its head?

Opus 5 topped a new benchmark called Lockstep, scoring highest among 10 leading LLMs tested on their ability to run logic circuits in their heads, with GPT 5.5 second and Kimi K3 and DeepSeek V4 Pro t…

20:44
2026-08-26
cline.ghost.io
artificial-intelligence

DeepSeek wins IMO Gold on 12 cents

DeepSeek V4 Flash scored 30/42 on IMO 2026 problems in Cline, clearing the 29-point gold medal cutoff for just $0.12, roughly 140 times cheaper than Claude Fable 5. The benchmark, run by Cline, tested…

22:14
2026-08-22
primeintellect.ai
artificial-intelligence

NanoGPT Speedrun Frontier

A benchmark of 153 autonomous runs across 18 frontier models on the nanoGPT optimizer speedrun shows Fable 5, run via claude-code at high effort, achieved the best validated result of 2,726 tokens wit…

00:44
2026-08-22
dev.to
ai-agents

Leveling up OpenCode... and not in the way you would expect.

A developer has forked OpenCode's harness to create an open-source project that enables node-based AI workflows, allowing multiple agents with designated roles to collaborate instead of using one prom…

20:14
2026-08-18
byteiota.com
artificial-intelligence

DeepSeek V4 on Cloudflare Workers AI: 1M Context Window Is Live

On August 14, Cloudflare added DeepSeek V4 Pro and DeepSeek V4 Flash to Workers AI, both with a 1,048,576-token context window, the first models on the platform to cross the 1M mark. The Flash model (…

23:01
2026-08-17
runtimewire.com
artificial-intelligence

J-Space says its text harness pushed DeepSeek V4 Pro past Fable 5

Open-source developer Tiger380 published a benchmark report claiming its J-Space text harness improved DeepSeek V4 Pro's performance on reasoning and agent tasks, with scores rising from 60.0 to 67.7 …

06:40
2026-08-17
dev.to
large-language-models

The Model Knew the Bid Was True. Then It Challenged Anyway.

In Kai, a Liar's Dice game, an engineer found that large language models sometimes challenge a bid they know is true, losing the round. The issue was traced to the action schema, where the 'challenge'…

00:00
2026-08-14
mindstudio.ai
ai-tools

How to Install and Set Up DeepSeek Harness Locally

DeepSeek released DeepSeek Harness, an agentic coding framework in developer preview alongside the DeepSeek V4 Pro model, which uses a plugin architecture and a local web app interface. Installation r…

00:00
2026-08-14
tomtunguz.com
artificial-intelligence

Honestly, Who Buys SOTA?

State-of-the-art AI models are two-thirds smarter than last November, with two new models released every three days, yet 84% of tokens on OpenRouter are not state of the art, and the six most-used mod…

← prev page 2 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics