activity
20162026
most citedDual networks based 3D Multi-Person Pose Estimation from Monocular Video

29 citations · 126 across the 26 of their papers we have counts for

collaborators
Showing cs.CVShow all

30 papers · 1 filter

cs.CV20261 cited

Bridging Day and Night: Target-Class Hallucination Suppression in Unpaired Image Translation

Shuwei Li, Lei Tan, Robby T. Tan

Day-to-night unpaired image translation is important to downstream tasks but remains challenging due to large appearance shifts and the lack of direct pixel-level supervision. Exis…

cs.CV2026

Aggregating Diverse Cue Experts for AI-Generated Image Detection

Lei Tan, Shuwei Li, Mohan Kankanhalli +1

The rapid emergence of image synthesis models poses challenges to the generalization of AI-generated image detectors. However, existing methods often rely on model-specific feature…

cs.CV2025

Bridging Annotation Gaps: Transferring Labels to Align Object Detection Datasets

Mikhail Kennerley, Angelica Aviles-Rivero, Carola-Bibiane Schönlieb +1

Combining multiple object detection datasets offers a path to improved generalisation but is hindered by inconsistencies in class semantics and bounding box annotations. Some metho…

cs.CV2025

NTIRE 2025 Challenge on Day and Night Raindrop Removal for Dual-Focused Images: Methods and Results

Xin Li, Yeying Jin, Xin Jin +134

This paper reviews the NTIRE 2025 Challenge on Day and Night Raindrop Removal for Dual-Focused Images. This challenge received a wide range of impressive solutions, which are devel…

cs.CV2024

CAT: Exploiting Inter-Class Dynamics for Domain Adaptive Object Detection

Mikhail Kennerley, Jian-Gang Wang, Bharadwaj Veeravalli +1

Domain adaptive object detection aims to adapt detection models to domains where annotated data is unavailable. Existing methods have been proposed to address the domain gap using…

cs.CV2023

ORTexME: Occlusion-Robust Human Shape and Pose via Temporal Average Texture and Mesh Encoding

Yu Cheng, Bo Wang, Robby T. Tan

In 3D human shape and pose estimation from a monocular video, models trained with limited labeled data cannot generalize well to videos with occlusion, which is common in the wild…