From the 1 of 13 linked papers with an AI index.
13 papers
Decoder-Guided Lossy Contour Coding Via Anchor Refinement
Ruoyu Yang, Yichi Zhang, Haohong Wang +1
The paper proposes a lossy contour coding scheme that leverages a decoder‑available coarse contour as side information, using anchor extraction and adaptive skipping to achieve a c…
Not Your Stereo-Typical Estimator: Combining Vision and Language for Volume Perception
Gautham Vinod, Bruce Coburn, Siddeshwar Raghavan +1
Accurate volume estimation of objects from visual data is a long-standing challenge in computer vision with significant applications in robotics, logistics, and smart health. Exist…
Can You Hear, Localize, and Segment Continually? An Exemplar-Free Continual Learning Benchmark for Audio-Visual Segmentation
Siddeshwar Raghavan, Gautham Vinod, Bruce Coburn +1
Audio-Visual Segmentation (AVS) aims to produce pixel-level masks of sound producing objects in videos, by jointly learning from audio and visual signals. However, real-world envir…
Uni-LVC: A Unified Method for Intra- and Inter-Mode Learned Video Compression
Yichi Zhang, Ruoyu Yang, Fengqing Zhu
Recent advances in learned video compression (LVC) have led to significant performance gains, with codecs such as DCVC-RT surpassing the H.266/VVC low-delay mode in compression eff…
Temporal Imbalance of Positive and Negative Supervision in Class-Incremental Learning
Jinge Ma, Fengqing Zhu
With the widespread adoption of deep learning in visual tasks, Class-Incremental Learning (CIL) has become an important paradigm for handling dynamically evolving data distribution…
MFP3D: Monocular Food Portion Estimation Leveraging 3D Point Clouds
Jinge Ma, Xiaoyan Zhang, Gautham Vinod +3
Food portion estimation is crucial for monitoring health and tracking dietary intake. Image-based dietary assessment, which involves analyzing eating occasion images using computer…