2 papers
cs.CL2025
MiSS: Revisiting the Trade-off in LoRA with an Efficient Shard-Sharing Structure
Jiale Kang, Qingyu Yin
Low-Rank Adaptation (LoRA) is a widely adopted technique for parameter-efficient fine-tuning, but its slow convergence has spurred the development of numerous variants. Nevertheles…
cs.CL2025
ModRWKV: Transformer Multimodality in Linear Time
Jiale Kang, Ziyin Yue, Qingyu Yin +4
Currently, most multimodal studies are based on large language models (LLMs) with quadratic-complexity Transformer architectures. While linear models like RNNs enjoy low inference…