Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Manifold-Guided Attention Steering
Ian Li, Kapilesh Guruprasad, Raunak Sengupta +3
Large language models frequently produce errors in reasoning tasks despite possessing the underlying knowledge required for correct reasoning. One possible approach to improve reas…
cs.LG2026
Breaking the Factorization Barrier in Diffusion Language Models
Ian Li, Zilei Shao, Benjie Wang +3
Diffusion language models theoretically allow for efficient parallel generation but are practically hindered by the ``factorization barrier'': the assumption that simultaneously pr…