2 papers
cs.CL2026
TTSR: Test-Time Self-Reflection for Continual Reasoning Improvement
Haoyang He, Zihua Rong, Liangjie Zhao +3
Test-time Training enables model adaptation using only test questions and offers a promising paradigm for improving the reasoning ability of large language models (LLMs). However,…
cs.LG2025
Your Attention Matters: to Improve Model Robustness to Noise and Spurious Correlations
Camilo Tamayo-Rousseau, Yunjia Zhao, Yiqun Zhang +1
Self-attention mechanisms are foundational to Transformer architectures, supporting their impressive success in a wide range of tasks. While there are many self-attention variants,…