DeepSeek V4-Flash Gets Vision — Experimental Image Input Comes to the Budget Model

DeepSeek quietly ships deepseek-v4-flash-vision-exp, an experimental vision variant of its V4-Flash model that accepts image input, alongside V4-Pro and V4-Flash point updates.

Thursday September 3, 2026 Source: DeepSeek
TL;DR — Quick Answer

DeepSeek quietly shipped deepseek-v4-flash-vision-exp, an experimental vision variant of its budget V4-Flash model that accepts image input. Alongside point updates to V4-Flash (0731) and V4-Pro (0813), the move brings multimodality to the token-cheapest tier — meaning image understanding no longer requires premium model pricing.

Key Takeaways

DeepSeek V4-Flash Gets Vision — Experimental Image Input Comes to the Budget Model — AI news article illustration

DeepSeek has quietly shipped an experimental vision variant of its budget flagship: deepseek-v4-flash-vision-exp. The new model accepts image input alongside text, extending the V4-Flash family beyond pure text workloads.

What’s New

According to DeepSeek’s API documentation, the current lineup now includes:

Calling methods remain unchanged for the updated models — existing deepseek-v4-flash and deepseek-v4-pro names automatically route to the latest versions.

Why Vision on Flash Matters

DeepSeek’s Flash models are the workhorses of budget AI — they top global token usage charts precisely because they combine low cost with strong capability. Adding vision to this tier means:

Agent Integration Momentum

The documentation also highlights DeepSeek Harness, now in developer preview for agent harness developers, and notes that the DeepSeek API works directly with tools like Claude Code, GitHub Copilot, and OpenCode — no code changes required. The Anthropic API format is supported natively alongside the OpenAI format.

This positioning reflects DeepSeek’s strategy: be the inexpensive, drop-in backend for the agent ecosystem others are building.

Context

The vision experiment follows a busy stretch for DeepSeek: V4-Flash topped global token usage earlier this summer, and the V4 family’s agent capabilities have made it a favorite for cost-conscious developers. As an “experimental” release, the vision model may change quickly — but it signals that multimodality is no longer a premium feature. It’s table stakes, even at the budget tier.

Frequently Asked Questions

Does DeepSeek V4-Flash support images?

Yes. The experimental deepseek-v4-flash-vision-exp model accepts image input alongside text, bringing vision capabilities to DeepSeek's budget model tier. Set the model name to deepseek-v4-flash-vision-exp to use it.

What is DeepSeek V4-Flash vision?

deepseek-v4-flash-vision-exp is an experimental multimodal variant of DeepSeek's V4-Flash model that additionally accepts image input. It extends the budget-friendly model family beyond pure text workloads.

How much does DeepSeek V4 cost?

DeepSeek V4 models are known for extremely low pricing — the V4-Flash tier topped global token usage charts earlier in 2026 for combining low cost with strong capability. Check DeepSeek's models and pricing page for current rates.

Can DeepSeek be used with Claude Code?

Yes. The DeepSeek API supports both OpenAI and Anthropic API formats natively, so tools like Claude Code, GitHub Copilot, and OpenCode can use DeepSeek as the backend model without code changes.

This article is based on the official announcement from DeepSeek . Read the original for full technical details.

Related Articles

Back to all news