2 papers
cs.CV2025
Dynamic Token Reduction during Generation for Vision Language Models
Xiaoyu Liang, Chaofeng Guan, Jiaying Lu +3
Vision-Language Models (VLMs) have achieved notable success in multimodal tasks but face practical limitations due to the quadratic complexity of decoder attention mechanisms and a…
cs.CL2024
Minstrel: Structural Prompt Generation with Multi-Agents Coordination for Non-AI Experts
Ming Wang, Yuanzhong Liu, Xiaoyu Liang +8
LLMs have demonstrated commendable performance across diverse domains. Nevertheless, formulating high-quality prompts to assist them in their work poses a challenge for non-AI expe…