3 papers
cs.CL2026
Indications of Belief-Guided Agency and Meta-Cognitive Monitoring in Large Language Models
Noam Steinmetz Yalon, Ariel Goldstein, Liad Mudrik +1
Rapid advancements in large language models (LLMs) have sparked the question whether these models possess some form of consciousness. To tackle this challenge, Butlin et al. (2023)…
cs.CL2024
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
Shlomo Neuberger, Niv Eckhaus, Uri Berger +3
Many human interactions, such as political debates, are carried out in group settings, where there are arbitrarily many participants, each with different views and agendas. To expl…
cs.CL2024
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
Daria Lioubashevski, Tomer Schlank, Gabriel Stanovsky +1
Understanding the inner workings of Transformers is crucial for achieving more accurate and efficient predictions. In this work, we analyze the computation performed by Transformer…