collaborators

18 papers

cs.CV2026

Seeing What Matters: Lesion-Aware High-Resolution Patch Discovery and Fusion for Chest X-ray Report Generation

Yingshu Li, Yunyi Liu, Zhenghao Chen +5

Despite rapid advances in chest X-ray (CXR) foundation models, most radiology report generation (RRG) systems still rely on heavily downsampled inputs (e.g., 256x256) due to the fi…

cs.CV2026

Neural Stereo Video Compression with Hybrid Disparity Compensation

Shiyin Jiang, Zhenghao Chen, Minghao Han +1

Disparity compensation represents the primary strategy in stereo video compression (SVC) for exploiting cross-view redundancy. These mechanisms can be broadly categorized into two…

cs.CV2026

Differentiable Vector Quantization for Rate-Distortion Optimization of Generative Image Compression

Shiyin Jiang, Wei Long, Minghao Han +3

The rapid growth of visual data under stringent storage and bandwidth constraints makes extremely low-bitrate image compression increasingly important. While Vector Quantization (V…

cs.RO2026

Device-Conditioned Neural Architecture Search for Efficient Robotic Manipulation

Yiming Wu, Huan Wang, Zhenghao Chen +2

The growing complexity of visuomotor policies poses significant challenges for deployment with heterogeneous robotic hardware constraints. However, most existing model-efficient ap…

cs.CV2026

Otter: Mitigating Background Distractions of Wide-Angle Few-Shot Action Recognition with Enhanced RWKV

Wenbo Huang, Jinghui Zhang, Zhenghao Chen +6

Wide-angle videos in few-shot action recognition (FSAR) effectively express actions within specific scenarios. However, without a global understanding of both subjects and backgrou…

cs.GR2026

LoD-Structured 3D Gaussian Splatting for Streaming Video Reconstruction

Xinhui Liu, Can Wang, Lei Liu +4

Free-Viewpoint Video (FVV) reconstruction enables photorealistic and interactive 3D scene visualization; however, real-time streaming is often bottlenecked by sparse-view inputs, p…