DeepSeek open-sources first vision model in the V4 lineup

DeepSeek open-sources first vision model in the V4 lineup

N
News Editor
2026-08-31 11:03:53
DeepSeek has open-sourced DeepSeek-V4-Flash-Vision-Exp, the first experimental multimodal model released under its DeepSeek-V4 series. The model is built on the V4-Flash architecture and adds a vision module, with image understanding capabilities obtained through continued training. According to the brief, it can interpret screenshots, read charts, and use tools to complete agent tasks. On pure text-based agent tasks, the model is said to maintain performance comparable to V4-Flash. The release comes with 305B parameters and is distributed under the MIT license. The announcement frames the model as an experimental step for the V4 family rather than a general summary of the broader lineup.

DeepSeek has open-sourced DeepSeek-V4-Flash-Vision-Exp, the first experimental multimodal model in the DeepSeek-V4 series.

The model is based on the V4-Flash architecture, adds a vision module, and gained image-understanding capabilities through continued training. According to the brief, it can process screenshots, read charts, and combine those capabilities with tools to complete agent tasks.

For pure text agent tasks, the model still delivers performance comparable to V4-Flash. The model has 305B parameters and is released under the MIT license.

This article was originally published by Bit.Fan. For more cryptocurrency news and market insights, visit www.bit.fan.
500

Disclaimer:

The market information, project data, and third-party content displayed on this platform are for industry information sharing only and do not constitute any form of investment advice or return commitment.

Cryptocurrency trading carries high risks. Users should fully assess their risk tolerance and make independent decisions. All profits, losses, and legal responsibilities are borne by the users themselves.