1 citations · 2 across the 6 of their papers we have counts for
9 papers · 1 filter
REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering
Li-Ming Zhan, Bo Liu, Chengqiang Xie +2
Inference-time steering aims to alter a large language model's (LLM's) responses without changing its parameters, but a central challenge is identifying the internal modules that m…
GeoEdit: Geometric Knowledge Editing for Large Language Models
Yujie Feng, Liming Zhan, Zexin Lu +6
Regular updates are essential for maintaining up-to-date knowledge in large language models (LLMs). Consequently, various model editing methods have been developed to update specif…
Understanding Layer Significance in LLM Alignment
Guangyuan Shi, Zexin Lu, Xiaoyu Dong +4
Aligning large language models (LLMs) through supervised fine-tuning is essential for tailoring them to specific applications. Recent studies suggest that alignment primarily adjus…
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection
Bo Liu, Liming Zhan, Yujie Feng +5
In the realm of task-oriented dialogue systems, a robust intent detection mechanism must effectively handle malformed utterances encountered in real-world scenarios. This study pre…
MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models
Lionel Z. Wang, Ka Chung Ng, Yiming Ma +1
Fake news significantly influences decision-making processes by misleading individuals, organizations, and even governments. Large language models (LLMs), as part of generative AI,…
Continual Dialogue State Tracking via Reason-of-Select Distillation
Yujie Feng, Bo Liu, Xiaoyu Dong +4
An ideal dialogue system requires continuous skill acquisition and adaptation to new tasks while retaining prior knowledge. Dialogue State Tracking (DST), vital in these systems, o…