deepseek-ai/DeepSeek-V4-Flash-0731 DeepSeek released DeepSeek-V4-Flash-0731, a model that Artificial Analysis ranks ahead of MiniMax M3, a 428B model, and at $0.14 per million input tokens and $0.27 per million output tokens it may be the best value-per-intelligence model available. Simon Willison found that using the default reasoning level via OpenRouter produced disappointing results, but increasing the reasoning effort to high yielded much better output. deepseek-ai/DeepSeek-V4-Flash-0731 https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731 Artificial Analysis rank it https://artificialanalysis.ai/models/deepseek-v4-flash ahead of MiniMax M3 - a 428B model. It's $0.14/million input and $0.27/million output pricing means this may currently be the best value-per-intelligence model out there. It's looking very good on the Intelligence Index vs. Cost per Intelligence Index Task https://artificialanalysis.ai/models/deepseek-v4-flash intelligence-comparison-tabs chart: I got a disappointing pelican https://gist.github.com/simonw/83bfb1171792f1e7a4d8935b5e82317e prompt from it using the default reasoning level via OpenRouter: But when I bumped reasoning level up to high I got something much better https://gist.github.com/simonw/83bfb1171792f1e7a4d8935b5e82317e options : llm -m openrouter/deepseek/deepseek-v4-flash-0731 -t pelican -o reasoning effort high Via Hacker News https://news.ycombinator.com/item?id=49120299 Tags: ai https://simonwillison.net/tags/ai , generative-ai https://simonwillison.net/tags/generative-ai , llms https://simonwillison.net/tags/llms , pelican-riding-a-bicycle https://simonwillison.net/tags/pelican-riding-a-bicycle , deepseek https://simonwillison.net/tags/deepseek , llm-release https://simonwillison.net/tags/llm-release , openrouter https://simonwillison.net/tags/openrouter , ai-in-china https://simonwillison.net/tags/ai-in-china , artificial-analysis https://simonwillison.net/tags/artificial-analysis