benchmark 1cross-document analysis 1event detection 1information extraction 1large language models 1
From the 1 of 3 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
From Single- to Cross-Document: Benchmarking Multi-Granularity Event Analysis of Large Language Models
Tao Wen, Shuai Shao, Pei Ke +7
The paper introduces MiGUE-Bench, a benchmark that evaluates large language models on multi‑granularity event analysis tasks ranging from single‑document event detection to cross‑d…
cs.CL2026
CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models
Qizhi Jiang, Shuo Wang, Pei Ke +2
Large Reasoning Models (LRMs) have achieved remarkable success on complex tasks by leveraging long chain-of-thought (CoT) trajectories, yet they frequently exhibit overthinking on…