From the 2 of 23 linked papers with an AI index.
23 papers
ViP-Rig: Visual-Prompted Controllable Rigging
Zihan Qin, Mingze Sun, Yifan Mao +6
ViP-Rig is a framework that lets users control 3D rigging by providing 2D visual prompts, enabling both initial skeleton creation and later editing of skeletons and skinning weight…
RainDancer: RGB-Event Video Deraining with Rain-Oriented Spiking Dynamics
Kui Jiang, Runzhe Li, Zhaocheng Yu +3
RainDancer is a video deraining framework that jointly processes RGB frames and event-camera data, first decomposing rain and background within each modality and then fusing them u…
Reliev3R: Relieving Feed-forward Reconstruction from Multi-View Geometric Annotations
Youyu Chen, Junjun Jiang, Yueru Luo +4
With recent advances, Feed-forward Reconstruction Models (FFRMs) have demonstrated great potential in reconstruction quality and adaptiveness to multiple downstream tasks. However,…
Focus-Scan-Refine: From Human Visual Perception to Efficient Visual Token Pruning
Enwei Tong, Yuanchao Bai, Yao Zhu +2
Vision-language models (VLMs) often generate massive visual tokens that greatly increase inference latency and memory footprint; while training-free token pruning offers a practica…
UDPNet: Unleashing Depth-based Priors for Robust Image Dehazing
Zengyuan Zuo, Junjun Jiang, Gang Wu +1
Image dehazing has witnessed significant advancements with the development of deep learning models. However, most existing methods focus solely on single-modal RGB features, neglec…
Beyond Degradation Redundancy: Contrastive Prompt Learning for All-in-One Image Restoration
Gang Wu, Junjun Jiang, Kui Jiang +2
All-in-One Image Restoration (AiOIR), which addresses diverse degradation types with a unified model, presents significant challenges in designing task-aware prompts that effective…