From the 1 of 19 linked papers with an AI index.
19 papers
Relation Geometry in Semantic Space of Language Models
Zhihan Cao, Hiroaki Yamada, Simone Teufel +4
The paper investigates how different semantic relations are reflected in the geometric structure of word embedding spaces produced by various language models, examining region clus…
Einstein World Models
Munachiso Samuel Nwadike, Zangir Iklassov, Ali Mekky +2
Does intelligence require the ability to reason about phenomena beyond direct experience? It is natural to suspect that some complex thought cannot be captured through language alo…
Language Models Compare Quantities Using Number-specific and Unit-specific Heuristics
Mutsumi Sasaki, Go kamoda, Ryosuke Takahashi +4
Quantities with measurement units, such as 110 cm and 1.2 m, require language models (LMs) to combine a numeral with a symbolic unit scale. Here, we study how LMs compare such quan…
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
Ying Zhang, Benjamin Heinzerling, Dongyuan Li +1
While fine-tuning is the standard for injecting factual knowledge into large language models (LLMs), the mechanisms enabling reliable fact recall via unseen queries remain poorly u…
Measuring AI Reasoning: A Guide for Researchers
Munachiso Samuel Nwadike, Zangir Iklassov, Kareem Ali +2
In this paper, we offer a guide for researchers on evaluating reasoning in language models, building the case that reasoning should be assessed through evidence of adaptive, multi-…
Cell-Based Representation of Relational Binding in Language Models
Qin Dai, Benjamin Heinzerling, Kentaro Inui
Understanding a discourse requires tracking entities and the relations that hold between them. While Large Language Models (LLMs) perform well on relational reasoning, the mechanis…