Caixin

DeepSeek Enters the Multimodal AI Race with Experimental Vision Model

Published: Aug. 22, 2026  12:34 a.m.  GMT+8
00:00
00:00/00:00
Listen to this article 1x
DeepSeek announced the launch of DeepSeek-V4-Flash-Vision-Exp, a multimodal visual-understanding model, and opened its multimodal API service. Photo: VCG
DeepSeek announced the launch of DeepSeek-V4-Flash-Vision-Exp, a multimodal visual-understanding model, and opened its multimodal API service. Photo: VCG

Chinese artificial intelligence startup DeepSeek has launched an experimental multimodal AI model capable of processing visual data and made it available to developers through its application programming interface. 

The release marks a step in the company’s expansion beyond text-based models, as it competes with global rivals to develop more capable, general-purpose AI systems.

loadingImg
You've accessed an article available only to subscribers
VIEW OPTIONS

Unlock exclusive discounts with a Caixin group subscription — ideal for teams and organizations.

Subscribe to both Caixin Global and Bloomberg - for the price of one.

Share this article
Open WeChat and scan the QR code