7 papers · 1 filter
Advancing Complex Wide-Area Scene Understanding with Hierarchical Coresets Selection
Jingyao Wang, Yiming Chen, Lingyu Si +1
Scene understanding is one of the core tasks in computer vision, aiming to extract semantic information from images to identify objects, scene categories, and their interrelationsh…
DiffDesign: Controllable Diffusion with Meta Prior for Efficient Interior Design Generation
Yuxuan Yang, Tao Geng
Interior design is a complex and creative discipline involving aesthetics, functionality, ergonomics, and materials science. Effective solutions must meet diverse requirements, typ…
Spatio-Temporal Fuzzy-oriented Multi-Modal Meta-Learning for Fine-grained Emotion Recognition
Jingyao Wang, Wenwen Qiang, Changwen Zheng +1
Fine-grained emotion recognition (FER) plays a vital role in various fields, such as disease diagnosis, personalized recommendations, and multimedia mining. However, existing FER m…
Causal Prompt Calibration Guided Segment Anything Model for Open-Vocabulary Multi-Entity Segmentation
Jingyao Wang, Jianqi Zhang, Wenwen Qiang +1
Despite the strength of the Segment Anything Model (SAM), it struggles with generalization issues in open-vocabulary multi-entity segmentation (OVMS). Through empirical and causal…
Less yet robust: crucial region selection for scene recognition
Jianqi Zhang, Mengxuan Wang, Jingyao Wang +3
Scene recognition, particularly for aerial and underwater images, often suffers from various types of degradation, such as blurring or overexposure. Previous works that focus on co…
Image-based Freeform Handwriting Authentication with Energy-oriented Self-Supervised Learning
Jingyao Wang, Luntian Mou, Changwen Zheng +1
Freeform handwriting authentication verifies a person's identity from their writing style and habits in messy handwriting data. This technique has gained widespread attention in re…