5 papers
Rethinking Reverse KL as Adaptive Entropy Distillation
Shizhen Li, Zhiyu Shen, Yuyin Lu +4
Knowledge distillation (KD) is widely used to transfer the capabilities of large language models (LLMs) to smaller students, but existing objectives often struggle to balance faith…
Cross-Source Reasoning-based Correction for Author Name Disambiguation
Fanjin Zhang, Yunhe Pang, Bo Chen +4
Author name disambiguation is a critical challenge in academic search systems, often addressed through from-scratch and real-time disambiguation approaches. However, current algori…
RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension
Yelin Chen, Fanjin Zhang, Suping Sun +8
Understanding research papers remains challenging for foundation models due to specialized scientific discourse and complex figures and tables, yet existing benchmarks offer limite…
HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions
Zhiyu Shen, Jiyuan Liu, Yunhe Pang +3
Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creating extensive and high-quality MHQ…
GuARD: Effective Anomaly Detection through a Text-Rich and Graph-Informed Language Model
Yunhe Pang, Bo Chen, Fanjin Zhang +3
Anomaly detection on text-rich graphs is widely prevalent in real life, such as detecting incorrectly assigned academic papers to authors and detecting bots in social networks. The…