4 papers
What Does Activation Steering Control? Attribution Across Answer Encodings and Output-Sensitive Subspaces
Zhiwei Gao, Shaowen Peng, Shoko Wakamiya +1
Activation steering is often evaluated under the answer encoding used to construct the direction. A reported gain may reflect the intended judgment or compatibility with answer ide…
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
Si-Belkacem Yamine Ketir, Lenard Paulo Tamayo, Shohei Hisada +3
Accurate assessment of cognitive decline from spontaneous speech remains challenging due to limited dataset size and class imbalance. In this work, we propose a large language mode…
Filling in the Clinical Gaps in Benchmark: Case for HealthBench for the Japanese medical system
Shohei Hisada, Endo Sunao, Himi Yamato +2
This study investigates the applicability of HealthBench, a large-scale, rubric-based medical benchmark, to the Japanese context. Although robust evaluation frameworks are essentia…
NAIST Academic Travelogue Dataset
Hiroki Ouchi, Hiroyuki Shindo, Shoko Wakamiya +5
We have constructed NAIST Academic Travelogue Dataset (ATD) and released it free of charge for academic research. This dataset is a Japanese text dataset with a total of over 31 mi…