Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
What Does Activation Steering Control? Attribution Across Answer Encodings and Output-Sensitive Subspaces
Zhiwei Gao, Shaowen Peng, Shoko Wakamiya +1
Activation steering is often evaluated under the answer encoding used to construct the direction. A reported gain may reflect the intended judgment or compatibility with answer ide…
cs.CL2026
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
Si-Belkacem Yamine Ketir, Lenard Paulo Tamayo, Shohei Hisada +3
Accurate assessment of cognitive decline from spontaneous speech remains challenging due to limited dataset size and class imbalance. In this work, we propose a large language mode…
cs.CL2026
Filling in the Clinical Gaps in Benchmark: Case for HealthBench for the Japanese medical system
Shohei Hisada, Endo Sunao, Himi Yamato +2
This study investigates the applicability of HealthBench, a large-scale, rubric-based medical benchmark, to the Japanese context. Although robust evaluation frameworks are essentia…