3 papers
cs.CV2026
More than the Sum: Panorama-Language Models for Adverse Omni-Scenes
Weijia Fan, Ruiping Liu, Jiale Wei +7
Existing vision-language models (VLMs) are tailored for pinhole imagery, stitching multiple narrow field-of-view inputs to piece together a complete omni-scene understanding. Yet,…
cs.CV2025
DAP-MAE: Domain-Adaptive Point Cloud Masked Autoencoder for Effective Cross-Domain Learning
Ziqi Gao, Qiufu Li, Linlin Shen
Compared to 2D data, the scale of point cloud data in different domains available for training, is quite limited. Researchers have been trying to combine these data of different do…
cs.LG2025
BCE vs. CE in Deep Feature Learning
Qiufu Li, Huibin Xiao, Linlin Shen
When training classification models, it expects that the learned features are compact within classes, and can well separate different classes. As the dominant loss function for tra…