30 citations · 134 across the 94 of their papers we have counts for
34 papers · 1 filter
RAdapter: A Routing and Rewriting Adapter for Efficient Hybrid RAG
Yucan Guo, Miao Su, Saiping Guan +6
Retrieval-Augmented Generation (RAG) has become a prevailing paradigm for enhancing Large Language Models (LLMs) with non-parametric knowledge. Vanilla RAG efficiently handles simp…
HiDiffTIR: Hierarchical Difficulty-Aware Policy Optimization for Multi-Turn Tool-Integrated Reasoning
Yucan Guo, Xiaohan Wang, Miao Su +8
Tool-Integrated Reasoning (TIR) is a fundamental capability for LLM agents to solve complex tasks by interacting with external tools iteratively. Reinforcement Learning (RL) has be…
ZenGen: Social Mind for LLMs
ZenGen Team, Ao Xiang, Bi Jingping +56
As large language models move from isolated task solving toward long-term service in human environments, they require social intelligence: the ability to infer mental states, track…
EGAD: Entropy-Guided Adaptive Distillation for Token-Level Knowledge Transfer
Hao Zhang, Zhibin Zhang, Guangxin Wu +3
Large language models (LLMs) have achieved remarkable performance across diverse domains, yet their enormous computational and memory requirements hinder deployment in resource-con…
Detoxification for LLM: From Dataset Itself
Wei Shao, Yihang Wang, Gaoyu Zhu +4
Existing detoxification methods for large language models mainly focus on post-training stage or inference time, while few tackle the source of toxicity, namely, the dataset itself…
PRISM-: Differential Subspace Steering for Prompt Highlighting in Large Language Models
Yuyao Ge, Shenghua Liu, Yiwei Wang +5
Prompt highlighting steers a large language model to prioritize user-specified text spans during generation. A key challenge of existing Key-editing approaches is extracting steeri…