2 papers
cs.LG2026
CAMD: Coverage-Aware Multimodal Decoding for Efficient Reasoning of Multimodal Large Language Models
Huijie Guo, Jingyao Wang, Lingyu Si +3
Recent advances in Multimodal Large Language Models (MLLMs) have shown impressive reasoning capabilities across vision-language tasks, yet still face the challenge of compute-diffi…
cs.LG2025
A Generalized Learning Framework for Self-Supervised Contrastive Learning
Lingyu Si, Jingyao Wang, Wenwen Qiang
Self-supervised contrastive learning (SSCL) has recently demonstrated superiority in multiple downstream tasks. In this paper, we generalize the standard SSCL methods to a Generali…