diffusion models 1instruction-based image editing 1region planning 1reinforcement learning 1vision-language reasoning 1
From the 1 of 13 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
FocuSFT: Bilevel Optimization for Dilution-Aware Long-Context Fine-Tuning
Zehua Pei, Hui-Ling Zhen, Xianzhi Yu +3
Large language models can now process increasingly long inputs, yet their ability to effectively use information spread across long contexts remains limited. We trace this gap to h…
cs.CL2025
Enhancing LLM Knowledge Learning through Generalization
Mingkang Zhu, Xi Chen, Zhongdao Wang +3
As Large language models (LLMs) are increasingly deployed in diverse applications, faithfully integrating evolving factual knowledge into these models remains a critical challenge.…