Back to insights

Published on 8/21/2026

Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks

the-decoder.com · ai-productivity-automation · AI Tools & Product Updates

Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks

Insight summary

  • Deepseek released V4-Flash-Vision-Exp, an experimental multimodal model adding image understanding to text capabilities.
  • The model nearly matches Opus 4.8 performance on Deepseek's internal multimodal agent benchmarks.
  • It supports image formats like JPEG, PNG, GIF, and WebP, determining format from file content.
  • Designed for agent-based applications, it integrates visual understanding with tool use and works with OpenAI and Anthropic APIs.
  • Developers can send images via Base64 encoding, public URLs (up to 32 MiB), or a free Files API (up to 64 MiB).
  • Image inputs can be downscaled for token savings, with token cost capped per image; max 600 images per request with size limits based on image count.
  • Pricing aligns with the V4-Flash model rates.

Content details

Industry
ai-productivity-automation
Topic
AI Tools & Product Updates
Source
the-decoder.com
Language
en
View source