8 papers
Knowledge-guided Disentanglement with Atomic Actions for Action Recognition
Tianci Wu, Siqi Cao, Guangming Zhu +6
Action recognition in complex scenes often involves multiple concurrent fine-grained actions, making it challenging to model internal action structures. Most existing methods rely…
Bridging Vision and Language Concepts through Optimal Transport Semantic Flow
Chenyang Zhang, Anqi Dong, Guangming Zhu +4
Concept Bottleneck Models (CBMs) promise transparent reasoning by predicting through human-interpretable concepts, yet their effectiveness fundamentally depends on how well visual…
Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval
Siyuan Wang, Hanchen Gao, Guangming Zhu +5
Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-drawn sketches capture complex s…
Prompt-guided Disentangled Representation for Action Recognition
Tianci Wu, Guangming Zhu, Jiang Lu +4
Action recognition is a fundamental task in video understanding. Existing methods typically extract unified features to process all actions in one video, which makes it challenging…
Content-Conditioned Generation of Stylized Free hand Sketches
Jiajun Liu, Siyuan Wang, Guangming Zhu +3
In recent years, the recognition of free-hand sketches has remained a popular task. However, in some special fields such as the military field, free-hand sketches are difficult to…
Flowmind2Digital: The First Comprehensive Flowmind Recognition and Conversion Approach
Huanyu Liu, Jianfeng Cai, Tingjia Zhang +6
Flowcharts and mind maps, collectively known as flowmind, are vital in daily activities, with hand-drawn versions facilitating real-time collaboration. However, there's a growing n…