3 papers
cs.CV2026
RP-OPSD: Resolution-Privileged On-Policy Self-Distillation for Multimodal Large Language Models
Qihui Zhu, Yuchen Wang, Zijian Wen +7
On-Policy Self-Distillation (OPSD) uses privileged information available only to the teacher to provide dense token-level supervision on trajectories generated by the student. Howe…
cs.CL2026
When TableQA Meets Noise: A Dual Denoising Framework for Complex Questions and Large-scale Tables
Shenghao Ye, Yu Guo, Dong Jin +5
Table question answering (TableQA) is a fundamental task in natural language processing (NLP). The strong reasoning capabilities of large language models (LLMs) have brought signif…
cs.AI2026
Bridging Network Fragmentation: A Semantic-Augmented DRL Framework for UAV-aided VANETs
Gaoxiang Cao, Wenke Yuan, Huasen He +4
Urban Vehicular Ad-Hoc Networks (VANETs) can become fragmented because buildings obstruct wireless links and vehicle mobility continuously changes the network topology. Unmanned Ae…