2 papers
cs.LG2026
MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification
Jingwei Song, Xinyu Wang, Hanbin Wang +6
Speculative Decoding (SD) accelerates autoregressive large language model (LLM) inference by decoupling generation and verification. While recent methods improve draft quality by t…
cs.AI2024
Geometry of naturalistic object representations in recurrent neural network models of working memory
Xiaoxuan Lei, Takuya Ito, Pouya Bashivan
Working memory is a central cognitive ability crucial for intelligent decision-making. Recent experimental and computational work studying working memory has primarily used categor…