Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
CAVE: Competence-Aware Visual Boundary Evidence Alignment for Video Temporal Grounding
Wei Jia, Zhicong Lu, Yu Chen +6
Large vision-language models (LVLMs) have achieved substantial performance gains in Video Temporal Grounding (VTG) through reinforcement learning (RL). However, existing methods pr…
cs.CL2026
Faithful-MR1: Faithful Multimodal Reasoning via Anchoring and Reinforcing Visual Attention
Changyuan Tian, Zhicong Lu, Huaxing Liu +7
Reinforcement learning with verifiable rewards (RLVR) has emerged as a promising paradigm for advancing complex reasoning in large language models, and recent work extends RLVR to…