activity
20162025
most citedBiFormer: Vision Transformer with Bi-Level Routing Attention

67 citations · 143 across the 16 of their papers we have counts for

collaborators

16 papers

cs.CV2025

Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation

Tianyu Huang, Wangguandong Zheng, Tengfei Wang +8

Real-world applications like video gaming and virtual reality often demand the ability to model 3D scenes that users can explore along custom camera trajectories. While significant…

cs.CV2024

Revisiting the Integration of Convolution and Attention for Vision Backbone

Lei Zhu, Xinjiang Wang, Wayne Zhang +1

Convolutions (Convs) and multi-head self-attentions (MHSAs) are typically considered alternatives to each other for building vision backbones. Although some works try to integrate…

cs.CV2024

LuSh-NeRF: Lighting up and Sharpening NeRFs for Low-light Scenes

Zefan Qu, Ke Xu, Gerhard Petrus Hancke +1

Neural Radiance Fields (NeRFs) have shown remarkable performances in producing novel-view images from high-quality scene images. However, hand-held low-light photography challenges…

cs.CV2024

Inverse Rendering of Glossy Objects via the Neural Plenoptic Function and Radiance Fields

Haoyuan Wang, Wenbo Hu, Lei Zhu +1

Inverse rendering aims at recovering both geometry and materials of objects. It provides a more compatible reconstruction for conventional rendering engines, compared with the neur…

cs.CV2024

Delving into Dark Regions for Robust Shadow Detection

Huankang Guan, Ke Xu, Rynson W. H. Lau

Shadow detection is a challenging task as it requires a comprehensive understanding of shadow characteristics and global/local illumination conditions. We observe from our experime…

cs.CV2024

Recasting Regional Lighting for Shadow Removal

Yuhao Liu, Zhanghan Ke, Ke Xu +3

Removing shadows requires an understanding of both lighting conditions and object textures in a scene. Existing methods typically learn pixel-level color mappings between shadow an…