21 citations · 22 across the 7 of their papers we have counts for
7 papers
TSP-Transformer: Task-Specific Prompts Boosted Transformer for Holistic Scene Understanding
Shuo Wang, Jing Li, Zibo Zhao +5
Holistic scene understanding includes semantic segmentation, surface normal estimation, object boundary detection, depth estimation, etc. The key aspect of this problem is to learn…
GraphAdapter: Tuning Vision-Language Models With Dual Knowledge Graph
Xin Li, Dongze Lian, Zhihe Lu +3
Adapter-style efficient transfer learning (ETL) has shown excellent performance in the tuning of vision-language models (VLMs) under the low-data regime, where only a few additiona…
Priority-Centric Human Motion Generation in Discrete Latent Space
Hanyang Kong, Kehong Gong, Dongze Lian +2
Text-to-motion generation is a formidable task, aiming to produce human motions that align with the input text while also adhering to human capabilities and physical laws. While th…
Dataset Quantization
Daquan Zhou, Kai Wang, Jianyang Gu +5
State-of-the-art deep neural networks are trained with large amounts (millions or even billions) of data. The expensive computation and memory costs make it difficult to train them…
Revisiting Event-based Video Frame Interpolation
Jiaben Chen, Yichen Zhu, Dongze Lian +7
Dynamic vision sensors or event cameras provide rich complementary information for video frame interpolation. Existing state-of-the-art methods follow the paradigm of combining bot…
Weakly Supervised Video Representation Learning with Unaligned Text for Sequential Videos
Sixun Dong, Huazhang Hu, Dongze Lian +3
Sequential video understanding, as an emerging video understanding task, has driven lots of researchers' attention because of its goal-oriented nature. This paper studies weakly su…