11 papers
CheckRLM: Effective Knowledge-Thought Coherence Checking in Retrieval-Augmented Reasoning
Dingling Xu, Ruobing Wang, Qingfei Zhao +8
Reasoning Language Models (RLMs) have significantly improved performance on complex tasks by extending the reasoning chain. However, these chains are prone to containing factual er…
ECPO: Evidence-Coupled Policy Optimization for Evidence-Certified Candidate Ranking
Miaobo Hu, Shuhao Hu, BoKun Wang +5
Ranking systems used in decision-support settings should not only order candidates but also expose evidence that can be independently checked. We study evidence-certified candidate…
SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text
Miaobo Hu, Xiaobo Guo, Shuhao Hu +5
Schema graphs are an upstream bottleneck of schema-grounded information extraction and knowledge graph construction, yet most extraction systems assume the schema is already availa…
AGPO: Adaptive Group Policy Optimization with Dual Statistical Feedback
Miaobo Hu, Shuhao Hu, Bokun Wang +5
Reinforcement learning improves LLM reasoning, but PPO/GRPO typically use fixed clipping and decoding temperature, which makes training brittle and tuning-heavy. We propose Adaptiv…
SAVER: Selective As-Needed Vision Evidence for Multimodal Information Extraction
Miaobo Hu, Shuhao Hu, Bokun Wang +5
Multimodal IE in social media is difficult because a post may attach multiple images that are weakly related, redundant, or even misleading with respect to the text. In this settin…
ExDR: Explanation-driven Dynamic Retrieval Enhancement for Multimodal Fake News Detection
Guoxuan Ding, Yuqing Li, Ziyan Zhou +3
The rapid spread of multimodal fake news poses a serious societal threat, as its evolving nature and reliance on timely factual details challenge existing detection methods. Dynami…