From the 2 of 16 linked papers with an AI index.
3 papers · 1 filter
Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs
Hamed Damirchi, Ignacio Meza De la Jara, Damith Ranasinghe +2
As language models are increasingly used for tasks that require verifiable reasoning, reliably distinguishing sound reasoning from flawed reasoning has become an important practica…
Certified but Fooled! Breaking Certified Defences with Ghost Certificates
Quoc Viet Vo, Tashreque M. Haq, Paul Montague +3
Certified defenses promise provable robustness guarantees. We study the malicious exploitation of probabilistic certification frameworks to better understand the limits of guarante…
Bayesian Low-Rank LeArning (Bella): A Practical Approach to Bayesian Neural Networks
Bao Gia Doan, Afshar Shamsi, Xiao-Yu Guo +6
Computational complexity of Bayesian learning is impeding its adoption in practical, large-scale tasks. Despite demonstrations of significant merits such as improved robustness and…