6 papers
RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents
Huayi Lai, Shichao Song, Simin Niu +5
Role-playing agents(RPAs) are widely used to steer large language models(LLMs) toward role-consistent behavior, yet existing benchmarks mainly evaluate surface-level fidelity and o…
Dissecting Role Cognition in Medical LLMs via Neuronal Ablation
Xun Liang, Huayi Lai, Hanyu Wang +5
Large language models (LLMs) have gained significant traction in medical decision support systems, particularly in the context of medical question answering and role-playing simula…
Enhancing Spatial Reasoning through Visual and Textual Thinking
Xun Liang, Xin Guo, Zhongming Jin +5
The spatial reasoning task aims to reason about the spatial relationships in 2D and 3D space, which is a fundamental capability for Visual Question Answering (VQA) and robotics. Al…
SEAP: Training-free Sparse Expert Activation Pruning Unlock the Brainpower of Large Language Models
Xun Liang, Hanyu Wang, Huayi Lai +7
Large Language Models have achieved remarkable success across various natural language processing tasks, yet their high computational cost during inference remains a major bottlene…
SurveyX: Academic Survey Automation via Large Language Models
Xun Liang, Jiawei Yang, Yezhaohui Wang +11
Large Language Models (LLMs) have demonstrated exceptional comprehension capabilities and a vast knowledge base, suggesting that LLMs can serve as efficient tools for automated sur…
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model
Xun Liang, Simin Niu, Zhiyu Li +8
The indexing-retrieval-generation paradigm of retrieval-augmented generation (RAG) has been highly successful in solving knowledge-intensive tasks by integrating external knowledge…