2 papers
cs.LG2026
Learning in the Fisher Subspace: A Guided Initialization for LoRA Fine-Tuning
Zhi-Quan Feng, Ying-Jia Lin, Hung-Yu Kao
LoRA adapts large language models (LLMs) by restricting updates to low-rank subspaces of pre-trained weights. While this substantially reduces training cost, the effectiveness of a…
cs.CV2025
VTPerception-R1: Enhancing Multimodal Reasoning via Explicit Visual and Textual Perceptual Grounding
Yizhuo Ding, Mingkang Chen, Zhibang Feng +4
Multimodal large language models (MLLMs) often struggle to ground reasoning in perceptual evidence. We present a systematic study of perception strategies-explicit, implicit, visua…