{"slug": "deepseek-unveils-ai-model-that-approaches-anthropics-performance-at-a-fraction", "title": "DeepSeek unveils AI model that approaches Anthropic’s performance at a fraction of the cost", "summary": "DeepSeek released preview versions of DeepSeek-V4-Pro and V4-Flash on April 24, 2026, with V4-Flash matching Anthropic's Claude Opus 4.8 on reasoning and coding benchmarks while costing approximately $0.28 per output compared to Anthropic's $25 to $30 range, a roughly 99% cost reduction. The multimodal models support up to 1 million tokens and are compatible with OpenAI and Anthropic APIs, potentially lowering compute bills for developers by orders of magnitude.", "body_md": "Via cnet.com\n\n# DeepSeek unveils AI model that approaches Anthropic’s performance at a fraction of the cost\n\nThe Chinese AI lab's V4-Flash model reportedly matches Claude Opus 4.8 on key benchmarks while costing roughly 99% less to run\n\nDeepSeek just dropped what might be the most cost-efficient frontier AI model on the market. The Chinese research lab’s new V4-Flash can process both text and images, performs competitively with Anthropic’s Claude Opus 4.8 on reasoning and coding benchmarks, and does it all for roughly $0.28 per output. Anthropic charges somewhere in the $25 to $30 range for comparable work.\n\n## What DeepSeek actually built\n\nThe company released preview versions of DeepSeek-V4-Pro and V4-Flash on April 24, 2026, the latest entries in its V4 model family. The headline feature is multimodal capability, meaning these models can interpret images alongside text prompts.\n\nThe technical architecture behind V4-Flash is where things get interesting. The model uses approximately 90 KV cache entries to process images, compared to roughly 870 for Claude models. KV cache is essentially the model’s working memory during inference. Fewer entries means less computational overhead, which translates directly into lower costs and faster processing.\n\nDeepSeek’s models also support long-context handling of up to 1 million tokens. The models are compatible with both OpenAI and Anthropic APIs, making them relatively straightforward drop-in replacements for developers already building on those platforms.\n\nSubsequent updates have followed the initial release, including a V4-Flash-0731 version and an experimental vision model dubbed deepseek-v4-flash-vision-exp, suggesting the lab is iterating quickly on its multimodal capabilities.\n\n## The cost gap that keeps widening\n\nDeepSeek has been building toward this moment since January 2025, when its R1 model first rattled the AI industry. That release established a template the company has now refined: match or approach frontier performance while dramatically undercutting on price.\n\nV4-Flash performs on par with Claude Opus 4.8 in benchmarks covering reasoning, coding, and agentic tasks while operating at approximately 99% less cost. Much of this efficiency comes from DeepSeek’s use of Mixture-of-Experts architecture, a design approach where only a subset of the model’s parameters activate for any given task.\n\nA company running thousands of AI inference calls per day could see its compute bill drop by orders of magnitude simply by switching providers. At $0.28 versus $25 to $30 per output, the math doesn’t require a spreadsheet to figure out.\n\n## Why this matters beyond the benchmarks\n\nReports of distillation attacks, where one lab’s model outputs are used to train a competitor’s system, have added tension to the relationship between DeepSeek and Western AI companies. API compatibility with OpenAI and Anthropic means switching costs are relatively low for developers evaluating alternatives.\n\n**Disclosure:** This article was edited by Editorial Team. For more information on how we create and review content, see our\n\n[Editorial Policy](https://cryptobriefing.com/editorial-policy/).", "url": "https://wpnews.pro/news/deepseek-unveils-ai-model-that-approaches-anthropics-performance-at-a-fraction", "canonical_source": "https://cryptobriefing.com/deepseek-v4-flash-rivals-anthropic-claude/", "published_at": "2026-08-21 11:24:46+00:00", "updated_at": "2026-08-21 11:44:06.248133+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "generative-ai", "ai-products"], "entities": ["DeepSeek", "Anthropic", "Claude Opus 4.8", "DeepSeek-V4-Pro", "V4-Flash", "OpenAI"], "alternates": {"html": "https://wpnews.pro/news/deepseek-unveils-ai-model-that-approaches-anthropics-performance-at-a-fraction", "markdown": "https://wpnews.pro/news/deepseek-unveils-ai-model-that-approaches-anthropics-performance-at-a-fraction.md", "text": "https://wpnews.pro/news/deepseek-unveils-ai-model-that-approaches-anthropics-performance-at-a-fraction.txt", "jsonld": "https://wpnews.pro/news/deepseek-unveils-ai-model-that-approaches-anthropics-performance-at-a-fraction.jsonld"}}