activity
20242026
collaborators

11 papers

cs.CL2026

Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings

Ryo Kishino, Yusuke Takase, Momose Oyama +2

Log-likelihood vectors define a common space for comparing language models as probability distributions, enabling unified comparisons across heterogeneous settings. We extend this…

cs.CL2026

Domain Mixture Design via Log-Likelihood Differences for Aligning Language Models with a Target Model

Ryo Kishino, Riku Shiomi, Hiroaki Yamagiwa +2

Instead of directly distilling a language model, this study addresses the problem of aligning a base model with a target model in distribution by designing the domain mixture of tr…

cs.CL2026

Measuring Affinity between Attention-Head Weight Subspaces via the Projection Kernel

Hiroaki Yamagiwa, Yusuke Takase, Hidetoshi Shimodaira

Understanding relationships between attention heads is essential for interpreting the internal structure of Transformers, yet existing metrics do not capture this structure well. W…

cs.CL2025

Mapping 1,000+ Language Models via the Log-Likelihood Vector

Momose Oyama, Hiroaki Yamagiwa, Yusuke Takase +1

To compare autoregressive language models at scale, we propose using log-likelihood vectors computed on a predefined text set as model features. This approach has a solid theoretic…

cs.CL2025

Quantifying Lexical Semantic Shift via Unbalanced Optimal Transport

Ryo Kishino, Hiroaki Yamagiwa, Ryo Nagata +2

Lexical semantic change detection aims to identify shifts in word meanings over time. While existing methods using embeddings from a diachronic corpus pair estimate the degree of c…

cs.CL2025

Predicting drug-gene relations via analogy tasks with word embeddings

Hiroaki Yamagiwa, Ryoma Hashimoto, Kiwamu Arakane +6

Natural language processing (NLP) is utilized in a wide range of fields, where words in text are typically transformed into feature vectors called embeddings. BioConceptVec is a sp…