34 citations · 60 across the 7 of their papers we have counts for
4 papers · 1 filter
On Group Relative Policy Optimization Collapse in Agent Search: The Lazy Likelihood-Displacement
Wenlong Deng, Yushu Li, Boying Gong +3
Tool-integrated (TI) reinforcement learning (RL) enables large language models (LLMs) to perform multi-step reasoning by interacting with external tools such as search engines and…
Inductive Bias and Language Expressivity in Emergent Communication
Shangmin Guo, Yi Ren, Agnieszka Słowik +1
Referential games and reconstruction games are the most common game types for studying emergent languages. We investigate how the type of the language game affects the emergent lan…
Compositional Languages Emerge in a Neural Iterated Learning Model
Yi Ren, Shangmin Guo, Matthieu Labeau +2
The principle of compositionality, which enables natural language to represent complex concepts via a structured combination of simpler ones, allows us to convey an open-ended set…
The Emergence of Compositional Languages for Numeric Concepts Through Iterated Learning in Neural Agents
Shangmin Guo, Yi Ren, Serhii Havrylov +3
Since first introduced, computer simulation has been an increasingly important tool in evolutionary linguistics. Recently, with the development of deep learning techniques, researc…