most citedProbing Syntax in Large Language Models: Successes and Remaining Challenges

1 citations · 1 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CL20251 cited

Probing Syntax in Large Language Models: Successes and Remaining Challenges

Pablo J. Diego-Simón, Emmanuel Chemla, Jean-Rémi King +1

The syntactic structures of sentences can be readily read-out from the activations of large language models (LLMs). However, the ``structural probes'' that have been developed to r…

cs.CL2025

A Neural Model for Word Repetition

Daniel Dager, Robin Sobczyk, Emmanuel Chemla +1

It takes several years for the developing brain of a baby to fully master word repetition-the task of hearing a word and repeating it aloud. Repeating a new word, such as from a ne…

cs.LG2025

A Minimum Description Length Approach to Regularization in Neural Networks

Matan Abudy, Orr Well, Emmanuel Chemla +2

State-of-the-art neural networks can be trained to become remarkable solutions to many problems. But while these architectures can express symbolic, perfect solutions, trained mode…

cs.CL2025

Large Language Models as Proxies for Theories of Human Linguistic Cognition

Imry Ziv, Nur Lan, Emmanuel Chemla +1

We consider the possible role of current large language models (LLMs) in the study of human linguistic cognition. We focus on the use of such models as proxies for theories of cogn…

cs.CV2024

Disentanglement and Compositionality of Letter Identity and Letter Position in Variational Auto-Encoder Vision Models

Bruno Bianchi, Aakash Agrawal, Stanislas Dehaene +2

Human readers can accurately count how many letters are in a word (e.g., 7 in ``buffalo''), remove a letter from a given position (e.g., ``bufflo'') or add a new one. The human bra…

cs.CL2024

A polar coordinate system represents syntax in large language models

Pablo Diego-Simón, Stéphane D'Ascoli, Emmanuel Chemla +2

Originally formalized with symbolic representations, syntactic trees may also be effectively represented in the activations of large language models (LLMs). Indeed, a 'Structural P…