6 papers
MIMONet: Multi-scale Input and Multi-scale Output Network for Salient Object Detection
Zhaojian Yao, Wei Gao, Tiesong Zhao +2
The existing methods for saliency detection task focus on the application of multi-level features, aiming to take advantage of the respective strengths of high- and low-level featu…
Flexible Deep Joint Source-Channel Coding: A Vibrotactile Example
Shuijie Li, Kemi Chen, Runjie Wang +2
The increasing demand for real-time tactile communication in multimedia systems has exposed the limitations of existing Joint Source-Channel Coding (JSCC) techniques. While current…
ParaJSCC: A Parameterized Framework for Reusable Multimodal Joint Source-Channel Coding
Kemi Chen, Mingkai Chen, Youjia Chen +3
Multimodal signals, such as visual, audio, and tactile data, are increasingly maintained as persistent digital assets in immersive communication systems and digital twins. In these…
Visual Geometry Foundation-Aware Gaussians for Single-Frame Surround-View Driving Reconstruction
Junhong Lin, Jinlong Wang, Xianda Guo +6
Single-frame surround-view reconstruction faces severe geometric instability and rendering artifacts due to minimal inter-camera overlap. While existing methods rely on complex dec…
SpatialQ: Understanding 3D Gaussian Splatting Scene Quality via Visual-based MLLM
Jingxuan Su, Shenglin Wang, Tiesong Zhao +2
3D Gaussian Splatting (3DGS) has emerged as an effective representation for novel view synthesis and 3D scene reconstruction, creating an increasing demand for reliable quality ass…
Bio-SFT: Asymmetric Cortical Guidance and Retinal Adaptation for Robust HDR Reconstruction
Tingyu Cheng, Ting Zhang, Chongyi Li +2
Recovering high dynamic range (HDR) radiance from a single standard dynamic range (SDR) image is highly ill-posed. Extreme luminance variation and severe quantization in dark regio…