Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Change of Thought: Adaptive Test-Time Computation
Mrinal Mathur, Mike Doan, Barak Pearlmutter +1
Transformers evaluated in a single, fixed-depth pass are provably limited in expressive power to the constant-depth circuit class TC0. Running a Transformer autoregressively remove…
cs.LG2025
How Overconfidence in Initial Choices and Underconfidence Under Criticism Modulate Change of Mind in Large Language Models
Dharshan Kumaran, Stephen M Fleming, Larisa Markeeva +8
Large language models (LLMs) exhibit strikingly conflicting behaviors: they can appear steadfastly overconfident in their initial answers whilst at the same time being prone to exc…
cs.LG2022
Routing and Placement of Macros using Deep Reinforcement Learning
Mrinal Mathur
Chip placement has been one of the most time consuming task in any semi conductor area, Due to this negligence, many projects are pushed and chips availability in real markets get…