DeepSeek "Open Eye" Sets the AI Community Ablaze: I Tested Its Capabilities with 12 Tricky Images to Find Its Limits Five days after DeepSeek delivered a powerful punch with V4, completely detonating the tech circle, Chen Xiaokang, a researcher in charge of multimodality within DeepSeek, posted the following on X and attached the text: Now, we see you. (Image source: Lei Technology) Yes, it means exactly what it says. While everyone was still amazed by the price and coding ability of V4, DeepSeek suddenly started testing the image recognition mode. The multimodal ability that the whole network had been discussing for a whole year has finally been implemented. The speed of this update really makes people wonder if Liang Wenfeng locked the development team in the computer room overnight to avoid being made into meme images of being irresponsible by netizens. It should be noted that this test is not a full - scale test but a small - scale gray - scale test. Only some users can see it in the official DeepSeek App or web version. At this time, in addition to the original quick mode and expert mode above the input field, there will also be a new image recognition mode button, marked with "Image understanding function is in internal testing". (Image source: Lei Technology) Unfortunately, none of my colleagues were selected for the gray - scale test. The number of people selected by the DeepSeek official was actually zero! Fortunately, I actually became the chosen one in ten thousand. Since it's such a coincidence, I'd feel a bit guilty if I didn't test it for everyone. This time, I carefully selected 12 pictures to let everyone see what DeepSeek can actually "see". Strong understanding ability, knowledge base needs to be updated Without further ado, let's start
DeepSeek "Open Eye" Ignites AI Community: I Tested Its Limits with 12 Tricky <b>Images</b>
Read the original article
eu.36kr.com →