collaborators

8 papers

cs.CV2026

GO-PRE: Goal-Oriented Next-Best-View Selection via Predictive Rendering Entropy for Active 3D Reconstruction

Yan Song, Zhihao Li, Chenglong Li +3

Active 3D reconstruction relies on active view selection to maximize reconstruction fidelity under limited capture budgets. However, most existing methods rely on surrogate signals…

cs.CV2026

Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling

Huanzhen Wang, Ziheng Zhou, Zeng Tao +5

The human brain constructs emotional percepts not by processing facial expressions in isolation, but through a dynamic, hierarchical integration of sensory input with semantic and…

cs.CV2026

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception

Huanzhen Wang, Ziheng Zhou, Jiaqi Song +4

Dynamic facial expression recognition in the wild remains challenging due to data scarcity and long-tail distributions, which hinder models from effectively learning the temporal d…

cs.GR2026

Screen, Cache, and Match: A Training-Free Causality-Consistent Reference Frame Framework for Human Animation

Jianan Wang, Nailei Hei, Li He +8

Human animation aims to generate temporally coherent and visually consistent videos over long sequences, yet modeling long-range dependencies while preserving frame quality remains…

cs.CV2025

MSVCOD:A Large-Scale Multi-Scene Dataset for Video Camouflage Object Detection

Shuyong Gao, Yu'ang Feng, Qishan Wang +5

Video Camouflaged Object Detection (VCOD) is a challenging task which aims to identify objects that seamlessly concealed within the background in videos. The dynamic properties of…

cs.CV2025

AnimatePainter: A Self-Supervised Rendering Framework for Reconstructing Painting Process

Junjie Hu, Shuyong Gao, Qianyu Guo +4

Humans can intuitively decompose an image into a sequence of strokes to create a painting, yet existing methods for generating drawing processes are limited to specific data types…