3 papers
cs.CV2025
Find Them All: Unveiling MLLMs for Versatile Person Re-identification
Jinhao Li, Zijian Chen, Lirong Deng +2
Person re-identification (ReID) aims to retrieve images of a target person from the gallery set, with wide applications in medical rehabilitation and public security. However, trad…
cs.CV2025
PictOBI-20k: Unveiling Large Multimodal Models in Visual Decipherment for Pictographic Oracle Bone Characters
Zijian Chen, Wenjie Hua, Jinhao Li +4
Deciphering oracle bone characters (OBCs), the oldest attested form of written Chinese, has remained the ultimate, unwavering goal of scholars, offering an irreplaceable key to und…
cs.CV2025
Can Large Models Fool the Eye? A New Turing Test for Biological Animation
Zijian Chen, Lirong Deng, Zhengyu Chen +5
Evaluating the abilities of large models and manifesting their gaps are challenging. Current benchmarks adopt either ground-truth-based score-form evaluation on static datasets or…