6 papers
Leader360V: The Large-scale, Real-world 360 Video Dataset for Multi-task Learning in Diverse Environment
Weiming Zhang, Dingwen Xiao, Aobotao Dai +5
360 video captures the complete surrounding scenes with the ultra-large field of view of 360X180. This makes 360 scene understanding tasks, eg, segmentation and tracking, crucial f…
GoodSAM++: Bridging Domain and Capacity Gaps via Segment Anything Model for Panoramic Semantic Segmentation
Weiming Zhang, Yexin Liu, Xu Zheng +1
This paper presents GoodSAM++, a novel framework utilizing the powerful zero-shot instance segmentation capability of SAM (i.e., teacher) to learn a compact panoramic semantic segm…
Learning High-Quality Navigation and Zooming on Omnidirectional Images in Virtual Reality
Zidong Cao, Zhan Wang, Yexin Liu +4
Viewing omnidirectional images (ODIs) in virtual reality (VR) represents a novel form of media that provides immersive experiences for users to navigate and interact with digital c…
MYCloth: Towards Intelligent and Interactive Online T-Shirt Customization based on User's Preference
Yexin Liu, Lin Wang
In conventional online T-shirt customization, consumers, \ie, users, can achieve the intended design only after repeated adjustments of the design prototypes presented by sellers i…
Deep Learning for Event-based Vision: A Comprehensive Survey and Benchmarks
Xu Zheng, Yexin Liu, Yunfan Lu +5
Event cameras are bio-inspired sensors that capture the per-pixel intensity changes asynchronously and produce event streams encoding the time, pixel position, and polarity (sign)…
Unsupervised Visible-Infrared ReID via Pseudo-label Correction and Modality-level Alignment
Yexin Liu, Weiming Zhang, Athanasios V. Vasilakos +1
Unsupervised visible-infrared person re-identification (UVI-ReID) has recently gained great attention due to its potential for enhancing human detection in diverse environments wit…