12 citations · 17 across the 5 of their papers we have counts for
5 papers
Wisdom of the Crowd: Reinforcement Learning from Coevolutionary Collective Feedback
Wenzhen Yuan, Shengji Tang, Weihao Lin +8
Reinforcement learning (RL) has significantly enhanced the reasoning capabilities of large language models (LLMs), but its reliance on expensive human-labeled data or complex rewar…
FTII-Bench: A Comprehensive Multimodal Benchmark for Flow Text with Image Insertion
Jiacheng Ruan, Yebin Yang, Zehao Lin +4
Benefiting from the revolutionary advances in large language models (LLMs) and foundational vision models, large vision-language models (LVLMs) have also made significant progress.…
iDAT: inverse Distillation Adapter-Tuning
Jiacheng Ruan, Jingsheng Gao, Mingye Xie +4
Adapter-Tuning (AT) method involves freezing a pre-trained model and introducing trainable adapter modules to acquire downstream knowledge, thereby calibrating the model for better…
EGE-UNet: an Efficient Group Enhanced UNet for skin lesion segmentation
Jiacheng Ruan, Mingye Xie, Jingsheng Gao +2
Transformer and its variants have been widely used for medical image segmentation. However, the large number of parameter and computational load of these models make them unsuitabl…
Learning Robust Visual-Semantic Embedding for Generalizable Person Re-identification
Suncheng Xiang, Jingsheng Gao, Mengyuan Guan +5
Generalizable person re-identification (Re-ID) is a very hot research topic in machine learning and computer vision, which plays a significant role in realistic scenarios due to it…