2 papers
cs.CL2026
Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery
Tianyun Zhong, Wangyi Jiang, Wei Wang +15
Large language models (LLMs) excel at answering pre-specified questions, yet their ability to navigate the open-ended, pre-conclusion stage of discovery remains largely unmeasured.…
cs.IR2026
HETERQA: Benchmarking Record Retrieval over Multiple Heterogeneous Sources
Yaodong Su, Hanchang Li, Quanqing Xu +2
In emerging systems (e.g., social media and e-commerce platforms), data records are often drawn from heterogeneous sources, such as relational tables, text documents, image reposit…