11 papers
LexKairos: Benchmarking Legal Temporal Capabilities in LLMs
Chenyang Li, Zejia Feng, Yuqin Huang +2
Large language models (LLMs) have demonstrated strong performance across a wide range of legal tasks. In legal practice, time is a critical concept that governs the validity of sta…
REFACT: Adaptive Fact Restatement for Compact and Faithful Chain-of-Thought Reasoning
Zhensheng Jin, Xin Dai, Zhenghao Liu +5
Large Language Models (LLMs) increasingly leverage long-form reasoning to solve complex tasks, yet their reasoning processes can deviate from the provided context when evidence is…
SHIFT: Gate-Modulated Activation Steering for Knowledge Conflict Mitigation in Retrieval-Augmented Generation
Ruochang Li, Pengcheng Huang, Zhenghao Liu +5
Retrieval-augmented generation (RAG) enhances LLMs by incorporating external knowledge to support response generation. However, conflicts between retrieved context and parametric k…
GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation
Yuxiao Ye, Yiwen Zhang, Huiyuan Xie +2
LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual model capabilities, but also o…
Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling
Shuliang Liu, Zhipeng Xu, Zhenghao Liu +6
Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach plays a vital role in aligning LL…
LexRel: Benchmarking Legal Relation Extraction for Chinese Civil Cases
Yida Cai, Ranjuexiao Hu, Huiyuan Xie +6
Legal relations serve as an important analytical framework for dispute resolution in civil cases. However, legal relations in Chinese civil cases remain underexplored in the field…