3 papers
cs.AI2026
Segment-Aligned Policy Optimization for Multi-Modal Reasoning
Lei Gao, Zhuoming Li, Mengxi Jia +4
Existing reinforcement learning approaches for Large Language Models typically perform policy optimization at the granularity of individual tokens or entire response sequences. How…
cs.CV2026
Attribute Distribution Modeling and Semantic-Visual Alignment for Generative Zero-shot Learning
Haojie Pu, Zhuoming Li, Yongbiao Gao +1
Generative zero-shot learning (ZSL) synthesizes features for unseen classes, leveraging semantic conditions to transfer knowledge from seen classes. However, it also introduces two…
cs.LG2025
DiCaP: Distribution-Calibrated Pseudo-labeling for Semi-Supervised Multi-Label Learning
Bo Han, Zhuoming Li, Xiaoyu Wang +4
Semi-supervised multi-label learning (SSMLL) aims to address the challenge of limited labeled data in multi-label learning (MLL) by leveraging unlabeled data to improve the model's…