behavioral alignment 1inference-time steering 1large language models 1module selection 1transformer activation analysis 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CL2026
REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering
Li-Ming Zhan, Bo Liu, Chengqiang Xie +2
The paper introduces REAL, a method that trains vector-quantized autoencoders on transformer activations to pinpoint attention heads or layers that most influence a target behavior…
cs.CV2025
GEMeX: A Large-Scale, Groundable, and Explainable Medical VQA Benchmark for Chest X-ray Diagnosis
Bo Liu, Ke Zou, Liming Zhan +7
Medical Visual Question Answering (Med-VQA) combines computer vision and natural language processing to automatically answer clinical inquiries about medical images. However, curre…
cs.CL2024
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection
Bo Liu, Liming Zhan, Yujie Feng +5
In the realm of task-oriented dialogue systems, a robust intent detection mechanism must effectively handle malformed utterances encountered in real-world scenarios. This study pre…