cd/entity/DeepSeek-V4-Pro· home entities DeepSeek-V4-Pro
grep -l @deepseek-v4-pro /news/*.json | wc -l → 18

DeepSeek-V4-Pro

mentions 18 type Organization feed RSS

// recent coverage 18 mentions

12:04
2026-07-23
promptcube3.com
large-language-models

DARWIN: Evolving LLM Jailbreak Framework

A new framework called DARWIN uses a genetic algorithm to evolve jailbreak prompts for large language models, achieving nearly 100% success rates on DeepSeek-V4-Pro and over 90% on GPT-5.5. The DARWIN…

05:11
2026-07-23
runtimewire.com
artificial-intelligence

Head to head: Google: Gemini 3.6 Flash vs DeepSeek-V4-Pro

Google's Gemini 3.6 Flash defeated DeepSeek-V4-Pro in a head-to-head text task evaluation, winning 5 tasks to 2 with 86% confidence and an overall score of 111.7 to 98.8, according to tests run by an …

07:12
2026-07-22
runtimewire.com
large-language-models

Head to head: Sarvam M vs DeepSeek-V4-Pro

DeepSeek-V4-Pro defeated Sarvam M 113.0 to 44.4 in a 12-task benchmark, achieving a 12-0 sweep with 100% confidence. The test, judged twice by gpt-5.4 to cancel position bias, found DeepSeek-V4-Pro re…

14:05
2026-07-20
runtimewire.com
large-language-models

Head to head: DeepSeek-V4-Pro vs gpt-oss-120b

Gpt-oss-120b defeated DeepSeek-V4-Pro 106.3 to 93.7 across 12 text tasks, winning 10 of 12 matchups with 99% confidence, according to a head-to-head benchmark scored by gpt-5.4. The open-source model …

12:10
2026-07-14
machinebrief.com
artificial-intelligence

EvoClawBench: A New Look at AI's Skill Learning

EvoClawBench, a new benchmark testing whether AI agents can learn reusable skills from experience, shows mixed results across 100 tasks and 502 sub-problems. Nanobot's GPT-5.4 model consistently excee…

17:50
2026-07-10
runtimewire.com
large-language-models

Head to head: Muse Spark 1.1 vs DeepSeek-V4-Pro

Muse Spark 1.1 defeated DeepSeek-V4-Pro 11-1 in a head-to-head benchmark of 12 text tasks, scoring 110.8 to 84.3 with 100% confidence. The model outperformed in localization, proofreading, JSON extrac…

20:49
2026-06-25
inferize.ai
ai-infrastructure

We got DeepSeek-V4-Pro serving in 20 seconds

Inferize announced DeepSeek-V4-Pro, claiming it can serve the model in 20 seconds with highly optimized, elastic AI inference. The company is building fast, efficient LLM serving that scales with dema…

14:27
2026-06-25
testingcatalog.com
large-language-models

DeepReinforce releases Ornith-1.0 open-source coding models

DeepReinforce open-sourced Ornith-1.0, a family of self-improving coding models ranging from 9B to 397B parameters, which learn to generate their own task-specific scaffolds during reinforcement learn…

11:20
2026-06-21
dev.to
large-language-models

AMD ATOM + ATOMesh: Prefill/decode Disaggregation on ROCm

AMD shipped ATOM + ATOMesh, a ROCm-native LLM serving stack for Instinct GPUs that implements prefill/decode disaggregation, splitting the two inference phases onto separate GPU pools to optimize for …

20:18
2026-06-13
runtimewire.com
large-language-models

OpenRouter: Fusion beats DeepSeek-V4-Pro on substance

OpenRouter's Fusion model outperformed DeepSeek-V4-Pro in a head-to-head coding test, delivering a complete, compilable Go rate limit parser with correct edge cases while DeepSeek V4 Pro made basic er…

00:00
2026-04-24
huggingface.co
large-language-models

DeepSeek-V4: a million-token context that agents can actually use

DeepSeek-V4 introduces a new architecture using hybrid attention mechanisms—Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA)—to drastically reduce the computational cost and me…

// co-occurs with top 8 entities
// topics top 6 topics