6 papers
Weakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based Diffusion
Wenqi Si, Gongyang Li, Shixiang Shi +1
The paper proposes a weakly‑supervised RGB‑D salient object detection framework that expands sparse scribble labels into dense pseudo annotations using the Segment Anything Model a…
Real-World Scene Recovery for Scattering-Degraded Images Using Spatial and Frequency Priors
Yun Liu, Tao Li, Guanghui Yue +3
Scene recovery from real-world images degraded by scattering effects, such as haze, sandstorm, underwater, and remote sensing conditions, remains a fundamental yet challenging prob…
Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application
Junqi Liu, Yun Zhang, Xiaoxia Huang +2
Just Recognizable Difference (JRD) boosts coding efficiency for machine vision through visibility threshold modeling, but is currently limited to a single-task scenario. To address…
DipGuava: Disentangling Personalized Gaussian Features for 3D Head Avatars from Monocular Video
Jeonghaeng Lee, Seok Keun Choi, Zhixuan Li +2
While recent 3D head avatar creation methods attempt to animate facial dynamics, they often fail to capture personalized details, limiting realism and expressiveness. To fill this…
MUGSQA: Novel Multi-Uncertainty-Based Gaussian Splatting Quality Assessment Method, Dataset, and Benchmarks
Tianang Chen, Jian Jin, Shilv Cai +2
Gaussian Splatting (GS) has recently emerged as a promising technique for 3D object reconstruction, delivering high-quality rendering results with significantly improved reconstruc…
Customizable ROI-Based Deep Image Compression
Jian Jin, Fanxin Xia, Feng Ding +5
Region of Interest (ROI)-based image compression optimizes bit allocation by prioritizing ROI for higher-quality reconstruction. However, as the users (including human clients and…