6 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.CL2026
EvoBrowseComp: Benchmarking Search Agents on Evolving Knowledge
Yunhan Wang, Jiaan Wang, Lianzhe Huang +2
Search Agents -- large language models augmented with search tools -- have intensified the need for future-proof evaluation benchmarks. Existing benchmarks such as BrowseComp rely…
cs.CL2023★ 6 cited
RECALL: A Benchmark for LLMs Robustness against External Counterfactual Knowledge
Yi Liu, Lianzhe Huang, Shicheng Li +5
LLMs and AI chatbots have improved people's efficiency in various fields. However, the necessary knowledge for answering the question may be beyond the models' knowledge boundaries…
cs.CL2022★ 2 cited
Incorporating Hierarchy into Text Encoder: a Contrastive Learning Approach for Hierarchical Text Classification
Zihan Wang, Peiyi Wang, Lianzhe Huang +2
Hierarchical text classification is a challenging subtask of multi-label classification due to its complex label hierarchy. Existing methods encode text and label hierarchy separat…