Showing 2025Show all
2 papers · 1 filter
cs.CL2025
A Question Answering Dataset for Temporal-Sensitive Retrieval-Augmented Generation
Ziyang Chen, Erxue Min, Xiang Zhao +7
We introduce ChronoQA, a large-scale benchmark dataset for Chinese question answering, specifically designed to evaluate temporal reasoning in Retrieval-Augmented Generation (RAG)…
cs.AI2025
PSSD: Making Large Language Models Self-denial via Human Psyche Structure
Jinzhi Liao, Zenghua Liao, Xiang Zhao
The enhance of accuracy in reasoning results of LLMs arouses the community's interests, wherein pioneering studies investigate post-hoc strategies to rectify potential mistakes. De…