From the 1 of 18 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
OS-Pruner: Pruning Chains-of-Thought of Reasoning Models via Optimal Stopping
Mohammed Ehab, Aymane El Gadarri, Vivek F. Farias +2
The paper proposes OS-Pruner, a lightweight framework that treats chain-of-thought reasoning in large language models as an optimal stopping problem, allowing the model to stop gen…
cs.AI2026
Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning
Richard Dewey, Janos Botyanszki, Ciamac C. Moallemi +1
AI researchers have long focused on poker-like games as a testbed for environments characterized by multi-player dynamics, imperfect information, and reasoning under uncertainty. W…