4 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2024
DQFormer: Towards Unified LiDAR Panoptic Segmentation with Decoupled Queries
Yu Yang, Jianbiao Mei, Liang Liu +6
LiDAR panoptic segmentation, which jointly performs instance and semantic segmentation for things and stuff classes, plays a fundamental role in LiDAR perception tasks. While most…
cs.CV2024★ 4 cited
M2-CLIP: A Multimodal, Multi-task Adapting Framework for Video Action Recognition
Mengmeng Wang, Jiazheng Xing, Boyuan Jiang +6
Recently, the rise of large-scale vision-language pretrained models like CLIP, coupled with the technology of Parameter-Efficient FineTuning (PEFT), has captured substantial attrac…
cs.CV2022★ 3 cited
E-NeRV: Expedite Neural Video Representation with Disentangled Spatial-Temporal Context
Zizhang Li, Mengmeng Wang, Huaijin Pi +3
Recently, the image-wise implicit neural representation of videos, NeRV, has gained popularity for its promising results and swift speed compared to regular pixel-wise implicit rep…