cd /news/artificial-intelligence/deepseek-releases-experimental-flash… · home topics artificial-intelligence article
[ARTICLE · art-106382] src=the-decoder.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks

Chinese AI company Deepseek released Deepseek-V4-Flash-Vision-Exp, an experimental multimodal model that adds image understanding to its text capabilities, and on Deepseek's internal benchmarks it nearly matches Opus 4.8 on agent tasks. The model supports JPEG, PNG, GIF, and WebP, works with OpenAI's Chat Completions and Responses APIs and Anthropic's Messages endpoint, and is priced at V4-Flash rates with each image costing at most 384 tokens. Deepseek also released version 0.1.1 of its Harness framework, which supports the new model out of the box.

read2 min views1 publishedAug 21, 2026
Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks
Image: The Decoder

Chinese AI company Deepseek has released V4-Flash-Vision-Exp, an experimental multimodal model that adds image understanding to its text capabilities. On Deepseek's own benchmarks, the model nearly matches Opus 4.8 on agent tasks.

Deepseek-V4-Flash-Vision-Exp extends Deepseek-V4-Flash with image processing while keeping the base model's text performance in reasoning and world knowledge, Deepseek says. On the company's internal multimodal agent benchmarks, the vision variant scores close to Opus 4.8.

Deepseek is targeting visual agent workflows #

Deepseek is positioning the model for agent-based applications. It's designed to work with different agent frameworks and combine visual understanding with tool use. In practice, it can describe images, extract text from screenshots, and analyze diagrams. It handles JPEG, PNG, GIF, and WebP, and determines the format from actual file content rather than the filename or declared MIME type, per the API docs.

The model works with OpenAI's Chat Completions and Responses APIs and Anthropic's Messages endpoint. Deepseek also released version 0.1.1 of its Harness framework, which supports the new model out of the box.

Pricing and image limits #

There are three ways to send images to the model. Developers can embed them directly with Base64 encoding, point to publicly accessible URLs (up to 32 MiB), or use the new, free Files API. The Files API lets you upload a file once and reference it by ID across multiple requests, with a size limit of 64 MiB.

An optional "detail" field downscales images to 512 x 512 pixels, saving tokens when fine visual detail isn't needed. The model automatically normalizes images to roughly 800 x 800 pixels depending on the aspect ratio before processing. Regardless of original resolution, each image costs at most 384 tokens. Pricing follows V4-Flash rates.

A single request can include up to 600 images. Max edge length is 8,192 pixels per side, but that drops to 4,096 pixels once a request contains 15 or more images. Images can only go in user messages.

AI News Without the Hype – Curated by Humans

					Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.				

					Subscribe now

Deepseek

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @deepseek 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/deepseek-releases-ex…] indexed:0 read:2min 2026-08-21 ·