cd/entity/DeepSeek-V4-Pro· home entities DeepSeek-V4-Pro
grep -l @deepseek-v4-pro /news/*.json | wc -l → 46

DeepSeek-V4-Pro

mentions 46 type Organization page 2/3 feed RSS

// recent coverage 46 mentions

12:55
2026-08-13
api-docs.deepseek.com
large-language-models

DeepSeek v4 Pro Price

DeepSeek announced new peak/off-peak pricing for its DeepSeek-V4-Flash and DeepSeek-V4-Pro models, effective August 17, 2026, with off-peak input (cache miss) prices at 1.5 yuan and 4.5 yuan per milli…

12:52
2026-08-13
ca.finance.yahoo.com
artificial-intelligence

DeepSeek Increases Prices for AI Services by Multiple Times

DeepSeek is raising prices for its flagship V4 models by more than four times, with new peak-hour pricing taking effect Aug. 16. The Hangzhou-based AI lab will charge $1.32 per 1 million output tokens…

12:49
2026-08-13
twitter.com
artificial-intelligence

DeepSeek API Pricing Update

DeepSeek announced the launch of DeepSeek-V4-Pro and V4-Flash, introducing flexible reasoning effort levels and native OpenAI Responses API support. The company also updated its API pricing with peak …

23:59
2026-08-12
simonwillison.net
artificial-intelligence

DeepSeek V4 Pro 0813 (on OpenRouter)

DeepSeek released DeepSeek V4 Pro 0813 on OpenRouter, following the open-weights releases of DeepSeek-V4-Pro in April and DeepSeek-V4-Flash-0731 in July, though open weights for the new version are no…

16:13
2026-08-10
lesswrong.com
ai-safety

Coercion and Deception in AI-to-AI Management

A new benchmark, Manager Coercion Bench (MCB), from Compassion in Machine Learning (CaML) finds that Anthropic's Claude models neither escalate to threats nor fabricate success, while all non-Anthropi…

22:02
2026-08-03
letsdatascience.com
ai-safety

Lasso Test Shows Agent Harnesses Can Flip Red-Team Results

Lasso Security reported on August 3 that changing only the agent harness altered outcomes across a 1,000-attack autonomous red-team test. The two harnesses averaged similar objective success rates, ab…

14:16
2026-08-02
runtimewire.com
large-language-models

Head to head: DeepSeek-V4-Pro vs Phi-4-reasoning

DeepSeek-V4-Pro defeated Phi-4-reasoning 12 tasks to 0 with an aggregate score of 105.5 to 33.0 and 100% confidence in a head-to-head benchmark of 12 fresh text tasks scored by gpt-5.4. The evaluation…

06:08
2026-07-31
api-docs.deepseek.com
artificial-intelligence

DeepSeek v4 Flash final models show solid improvements

DeepSeek released DeepSeek-V4-Flash in public beta on 2026-07-31, showing significantly enhanced agent capabilities with benchmark results far exceeding V4-Pro-Preview, including Terminal Bench 2.1 at…

12:04
2026-07-23
promptcube3.com
large-language-models

DARWIN: Evolving LLM Jailbreak Framework

A new framework called DARWIN uses a genetic algorithm to evolve jailbreak prompts for large language models, achieving nearly 100% success rates on DeepSeek-V4-Pro and over 90% on GPT-5.5. The DARWIN…

05:11
2026-07-23
runtimewire.com
artificial-intelligence

Head to head: Google: Gemini 3.6 Flash vs DeepSeek-V4-Pro

Google's Gemini 3.6 Flash defeated DeepSeek-V4-Pro in a head-to-head text task evaluation, winning 5 tasks to 2 with 86% confidence and an overall score of 111.7 to 98.8, according to tests run by an …

07:12
2026-07-22
runtimewire.com
large-language-models

Head to head: Sarvam M vs DeepSeek-V4-Pro

DeepSeek-V4-Pro defeated Sarvam M 113.0 to 44.4 in a 12-task benchmark, achieving a 12-0 sweep with 100% confidence. The test, judged twice by gpt-5.4 to cancel position bias, found DeepSeek-V4-Pro re…

14:05
2026-07-20
runtimewire.com
large-language-models

Head to head: DeepSeek-V4-Pro vs gpt-oss-120b

Gpt-oss-120b defeated DeepSeek-V4-Pro 106.3 to 93.7 across 12 text tasks, winning 10 of 12 matchups with 99% confidence, according to a head-to-head benchmark scored by gpt-5.4. The open-source model …

12:10
2026-07-14
machinebrief.com
artificial-intelligence

EvoClawBench: A New Look at AI's Skill Learning

EvoClawBench, a new benchmark testing whether AI agents can learn reusable skills from experience, shows mixed results across 100 tasks and 502 sub-problems. Nanobot's GPT-5.4 model consistently excee…

17:50
2026-07-10
runtimewire.com
large-language-models

Head to head: Muse Spark 1.1 vs DeepSeek-V4-Pro

Muse Spark 1.1 defeated DeepSeek-V4-Pro 11-1 in a head-to-head benchmark of 12 text tasks, scoring 110.8 to 84.3 with 100% confidence. The model outperformed in localization, proofreading, JSON extrac…

20:49
2026-06-25
inferize.ai
ai-infrastructure

We got DeepSeek-V4-Pro serving in 20 seconds

Inferize announced DeepSeek-V4-Pro, claiming it can serve the model in 20 seconds with highly optimized, elastic AI inference. The company is building fast, efficient LLM serving that scales with dema…

14:27
2026-06-25
testingcatalog.com
large-language-models

DeepReinforce releases Ornith-1.0 open-source coding models

DeepReinforce open-sourced Ornith-1.0, a family of self-improving coding models ranging from 9B to 397B parameters, which learn to generate their own task-specific scaffolds during reinforcement learn…

11:20
2026-06-21
dev.to
large-language-models

AMD ATOM + ATOMesh: Prefill/decode Disaggregation on ROCm

AMD shipped ATOM + ATOMesh, a ROCm-native LLM serving stack for Instinct GPUs that implements prefill/decode disaggregation, splitting the two inference phases onto separate GPU pools to optimize for …

← prev page 2 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics