49 citations · 135 across the 17 of their papers we have counts for
22 papers
Putting Registers to Work: Task Registers for Token Pruning in Vision Transformers
Hongsen Cao, Mona Jaber, Shanxin Yuan +1
Token-pruning policies are usually designed for a single recognition pipeline, but pretrained Vision Transformers are reused across tasks with different spatial demands. We ask whi…
Exploring Effective Mask Sampling Modeling for Neural Image Compression
Lin Liu, Mingming Zhao, Shanxin Yuan +5
Image compression aims to reduce the information redundancy in images. Most existing neural image compression methods rely on side information from hyperprior or context models to…
NeRFVS: Neural Radiance Fields for Free View Synthesis via Geometry Scaffolds
Chen Yang, Peihao Li, Zanwei Zhou +5
We present NeRFVS, a novel neural radiance fields (NeRF) based method to enable free navigation in a room. NeRF achieves impressive performance in rendering images for novel views…
Graph Neural Networks in Vision-Language Image Understanding: A Survey
Henry Senior, Gregory Slabaugh, Shanxin Yuan +1
2D image understanding is a complex problem within computer vision, but it holds the key to providing human-level scene comprehension. It goes further than identifying the objects…
Low-Light Video Enhancement with Synthetic Event Guidance
Lin Liu, Junfeng An, Jianzhuang Liu +6
Low-light video enhancement (LLVE) is an important yet challenging task with many applications such as photographing and autonomous driving. Unlike single image low-light enhanceme…
Disentangling 3D Attributes from a Single 2D Image: Human Pose, Shape and Garment
Xue Hu, Xinghui Li, Benjamin Busam +3
For visual manipulation tasks, we aim to represent image content with semantically meaningful features. However, learning implicit representations from images often lacks interpret…