4 papers
Visual Bridge: Universal Visual Perception Representations Generating
Yilin Gao, Shuguang Dou, Junzhou Li +4
Recent advances in diffusion models have achieved remarkable success in isolated computer vision tasks such as text-to-image generation, depth estimation, and optical flow. However…
Rethinking the Zigzag Flattening for Image Reading
Qingsong Zhao, Yi Wang, Zhipeng Zhou +4
Sequence ordering of word vector matters a lot to text reading, which has been proven in natural language processing (NLP). However, the rule of different sequence ordering in comp…
Reviving Static Charts into Live Charts
Lu Ying, Yun Wang, Haotian Li +5
Data charts are prevalent across various fields due to their efficacy in conveying complex data relationships. However, static charts may sometimes struggle to engage readers and e…
DROP: Decouple Re-Identification and Human Parsing with Task-specific Features for Occluded Person Re-identification
Shuguang Dou, Xiangyang Jiang, Yuanpeng Tu +4
The paper introduces the Decouple Re-identificatiOn and human Parsing (DROP) method for occluded person re-identification (ReID). Unlike mainstream approaches using global features…