3 papers
cs.LG2026
MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification
Jingwei Song, Xinyu Wang, Hanbin Wang +6
Speculative Decoding (SD) accelerates autoregressive large language model (LLM) inference by decoupling generation and verification. While recent methods improve draft quality by t…
cs.AI2024
Geometry of naturalistic object representations in recurrent neural network models of working memory
Xiaoxuan Lei, Takuya Ito, Pouya Bashivan
Working memory is a central cognitive ability crucial for intelligent decision-making. Recent experimental and computational work studying working memory has primarily used categor…
cs.AI2024
IWISDM: Assessing instruction following in multimodal models at scale
Xiaoxuan Lei, Lucas Gomez, Hao Yuan Bai +1
The ability to perform complex tasks from detailed instructions is a key to many remarkable achievements of our species. As humans, we are not only capable of performing a wide var…