activity
20152023
most citedMulti-Granularity Denoising and Bidirectional Alignment for Weakly Supervised Semantic Segmentation

45 citations · 70 across the 7 of their papers we have counts for

collaborators

15 papers

cs.CV20242 cited

Real-Time 4K Super-Resolution of Compressed AVIF Images. AIS 2024 Challenge Survey

Marcos V. Conde, Zhijun Lei, Wen Li +72

This paper introduces a novel benchmark as part of the AIS 2024 Real-Time Image Super-Resolution (RTSR) Challenge, which aims to upscale compressed images from 540p to 4K resolutio…

cs.CV20241 cited

DVF: Advancing Robust and Accurate Fine-Grained Image Retrieval with Retrieval Guidelines

Xin Jiang, Hao Tang, Rui Yan +2

Fine-grained image retrieval (FGIR) is to learn visual representations that distinguish visually similar objects while maintaining generalization. Existing methods propose to gener…

cs.CV20241 cited

Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation

Gensheng Pei, Yazhou Yao, Jianbo Jiao +3

Conventional video object segmentation (VOS) methods usually necessitate a substantial volume of pixel-level annotated video data for fully supervised learning. In this paper, we p…

cs.CV2024

ColorMNet: A Memory-based Deep Spatial-Temporal Feature Propagation Network for Video Colorization

Yixin Yang, Jiangxin Dong, Jinhui Tang +1

How to effectively explore spatial-temporal features is important for video colorization. Instead of stacking multiple frames along the temporal dimension or recurrently propagatin…

cs.CV2024

Collaborative Feedback Discriminative Propagation for Video Super-Resolution

Hao Li, Xiang Chen, Jiangxin Dong +2

The key success of existing video super-resolution (VSR) methods stems mainly from exploring spatial and temporal information, which is usually achieved by a recurrent propagation…

cs.CV2023

Learning Comprehensive Representations with Richer Self for Text-to-Image Person Re-Identification

Shuanglin Yan, Neng Dong, Jun Liu +2

Text-to-image person re-identification (TIReID) retrieves pedestrian images of the same identity based on a query text. However, existing methods for TIReID typically treat it as a…