7 papers · 1 filter
GO-PRE: Goal-Oriented Next-Best-View Selection via Predictive Rendering Entropy for Active 3D Reconstruction
Yan Song, Zhihao Li, Chenglong Li +3
Active 3D reconstruction relies on active view selection to maximize reconstruction fidelity under limited capture budgets. However, most existing methods rely on surrogate signals…
Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling
Huanzhen Wang, Ziheng Zhou, Zeng Tao +5
The human brain constructs emotional percepts not by processing facial expressions in isolation, but through a dynamic, hierarchical integration of sensory input with semantic and…
ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception
Huanzhen Wang, Ziheng Zhou, Jiaqi Song +4
Dynamic facial expression recognition in the wild remains challenging due to data scarcity and long-tail distributions, which hinder models from effectively learning the temporal d…
MSVCOD:A Large-Scale Multi-Scene Dataset for Video Camouflage Object Detection
Shuyong Gao, Yu'ang Feng, Qishan Wang +5
Video Camouflaged Object Detection (VCOD) is a challenging task which aims to identify objects that seamlessly concealed within the background in videos. The dynamic properties of…
AnimatePainter: A Self-Supervised Rendering Framework for Reconstructing Painting Process
Junjie Hu, Shuyong Gao, Qianyu Guo +4
Humans can intuitively decompose an image into a sequence of strokes to create a painting, yet existing methods for generating drawing processes are limited to specific data types…
A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection
Tsui Qin Mok, Shuyong Gao, Haozhe Xing +3
Weakly-Supervised Camouflaged Object Detection (WSCOD) has gained popularity for its promise to train models with weak labels to segment objects that visually blend into their surr…