DeepSeek Officially Launches Image Recognition Mode on App and Web
DeepSeek has rolled out its image recognition mode on both web and mobile app platforms, enabling advanced visual understanding that goes well beyond simple text extraction.
On June 18, 2026, DeepSeek officially launched its image recognition mode across its web and mobile app interfaces. The announcement came from Xiaokang Chen, a multimodal researcher at DeepSeek, who confirmed the feature is now live for users.
The new mode appears alongside the existing Quick Mode and Expert Mode, giving users a dedicated way to upload images and have DeepSeek "see" and interpret visual content. While the web version is fully available, the mobile app still displays an "image understanding feature in beta" notice during the rollout phase, indicating a gradual stabilization on mobile platforms.
This launch extends DeepSeek’s capabilities far beyond optical character recognition. By enabling image uploads, users can tap into a deeper understanding of objects, scenes, and contextual relationships in photographs, diagrams, and screenshots.
Earlier in April 2026, DeepSeek shed light on the technical foundation of this multimodal leap by publishing details of its core framework: "Thinking with Visual Primitives." The framework underpins how the model processes and reasons about visual information, cementing DeepSeek’s position in the competitive multimodal AI landscape.