behavioral alignment 1inference-time steering 1large language models 1module selection 1transformer activation analysis 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CL2026
REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering
Li-Ming Zhan, Bo Liu, Chengqiang Xie +2
The paper introduces REAL, a method that trains vector-quantized autoencoders on transformer activations to pinpoint attention heads or layers that most influence a target behavior…
cs.CL2025
GeoEdit: Geometric Knowledge Editing for Large Language Models
Yujie Feng, Liming Zhan, Zexin Lu +6
Regular updates are essential for maintaining up-to-date knowledge in large language models (LLMs). Consequently, various model editing methods have been developed to update specif…
cs.CV2025
GEMeX: A Large-Scale, Groundable, and Explainable Medical VQA Benchmark for Chest X-ray Diagnosis
Bo Liu, Ke Zou, Liming Zhan +7
Medical Visual Question Answering (Med-VQA) combines computer vision and natural language processing to automatically answer clinical inquiries about medical images. However, curre…