From the 1 of 2 linked papers with an AI index.
2 papers
cs.LG2026
Higher Embedding Dimension Creates a Stronger World Model for a Simple Sorting Task
Brady Bhalla, Honglu Fan, Nancy Chen +1
The paper studies how the size of embedding vectors influences the development of internal world models in transformers trained via reinforcement learning to perform bubble‑sort‑st…
cs.LG2026
Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1144
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…