
The limited release to select users comes just days after the Hangzhou-based company released its new flagship model V4, which was followed by extensive price cuts.
According to DeepSeek multimodal team leader Chen Xiaokang, who made the announcement on Wednesday on social media, the function was initially offered to select users on DeepSeek’s chatbot website and mobile application for beta testing.
Advertisement
On DeepSeek’s chat interface, a new “image recognition mode” had been added alongside the “expert” and “flash” chat modes, which were introduced earlier this month.
Advertisement
As AI continues to rapidly progress, multimodal capabilities are viewed as a necessity to move beyond simple text conversations with users into more complex and economically valuable domains.
While DeepSeek’s breakout moment in January 2025 made it a household name internationally due to its model’s powerful reasoning capabilities and cost-efficiency, the start-up’s lack of a multimodal offering since then has been seen as an Achilles’ heel.

Don't Miss:
-
The US-China tech war has evolved since the last Xi-Trump summit. Who has the edge now?
-
Canada’s British Columbia sues OpenAI in US court over school shooting
-
Top US Senate Democrat Jeanne Shaheen warns Trump not to waver on Taiwan ahead of Xi meet
-
UK PM Burnham agrees to Saudi request for refuelling support
-
Zelensky and US spy chief Ratcliffe meet in Ireland, sources say

Trump administration gave firm with ties to ballroom donor the green light to do business with a sanctioned Russian bank
Claire Sylvia on the Constitutionality of the False Claims Act
School shootings in the Philippines put gun access under scrutiny