most citedSystem-status-aware Adaptive Network for Online Streaming Video Understanding

3 citations · 5 across the 7 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2024

Action Detection via an Image Diffusion Process

Lin Geng Foo, Tianjiao Li, Hossein Rahmani +1

Action detection aims to localize the starting and ending points of action instances in untrimmed videos, and predict the classes of those instances. In this paper, we make the obs…

cs.CV20243 cited

LLMs are Good Sign Language Translators

Jia Gong, Lin Geng Foo, Yixuan He +2

Sign Language Translation (SLT) is a challenging task that aims to translate sign videos into spoken language. Inspired by the strong translation capabilities of large language mod…

cs.CV2023

Distribution-Aligned Diffusion for Human Mesh Recovery

Lin Geng Foo, Jia Gong, Hossein Rahmani +1

Recovering a 3D human mesh from a single RGB image is a challenging task due to depth ambiguity and self-occlusion, resulting in a high degree of uncertainty. Meanwhile, diffusion…

cs.CV2023

Token Boosting for Robust Self-Supervised Visual Transformer Pre-training

Tianjiao Li, Lin Geng Foo, Ping Hu +4

Learning with large-scale unlabeled data has become a powerful tool for pre-training Visual Transformers (VTs). However, prior works tend to overlook that, in real-world scenarios,…

cs.CV20233 cited

System-status-aware Adaptive Network for Online Streaming Video Understanding

Lin Geng Foo, Jia Gong, Zhipeng Fan +1

Recent years have witnessed great progress in deep neural networks for real-time applications. However, most existing works do not explicitly consider the general case where the de…

cs.CV2023

Progressive Channel-Shrinking Network

Jianhong Pan, Siyuan Yang, Lin Geng Foo +4

Currently, salience-based channel pruning makes continuous breakthroughs in network compression. In the realization, the salience mechanism is used as a metric of channel salience…