Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
What Proves You Wrong: Benchmarking Language Models on Falsifiable Research Ideation
Ziyue Wang, Aomufei Yuan, Yiran Yao +10
Large language models are increasingly used to propose research ideas, yet the prevailing ways of judging such ideas supply no shared decision rule: free-form judging sways with st…
cs.CL2026
DiaDem: Advancing Dialogue Descriptions in Audiovisual Video Captioning for Multimodal Large Language Models
Xinlong Chen, Weihong Lin, Jingyun Hua +10
Accurate dialogue description in audiovisual video captioning is crucial for downstream understanding and generation tasks. However, existing models generally struggle to produce f…
cs.CL2025
Mitigating Overthinking through Reasoning Shaping
Feifan Song, Shaohang Wei, Bofei Gao +8
Large reasoning models (LRMs) boosted by Reinforcement Learning from Verifier Reward (RLVR) have shown great power in problem solving, yet they often cause overthinking: excessive,…