3 papers
cs.CL2026
A Theoretical Game of Attacks via Compositional Skills
Xinbo Wu, Huan Zhang, Abhishek Umrawal +1
As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to restrict harmful behavior, the…
cs.CL2026
Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations
Kalyan Cherukuri, Lav R. Varshney
Large language models (LLMs) hallucinate: they produce fluent outputs that are factually incorrect. We present a geometric dynamical systems framework in which hallucinations arise…
cs.LG2026
CAETC: Causal Autoencoding and Treatment Conditioning for Counterfactual Estimation over Time
Nghia D. Nguyen, Pablo Robles-Granda, Lav R. Varshney
Counterfactual estimation over time is important in various applications, such as personalized medicine. However, time-dependent confounding bias in observational data still poses…