Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
How Do Transformers Learn to Associate Tokens: Gradient Leading Terms Bring Mechanistic Interpretability
Shawn Im, Changdae Oh, Zhen Fang +1
Semantic associations such as the link between "bird" and "flew" are foundational for language modeling as they enable models to go beyond memorization and instead generalize and g…
cs.CL2026
Thinking Is Not Telling: Information Disclosure in User-Service LLM Agents
Jiatong Li, Changdae Oh, Hyeong Kyu Choi +2
User-engaged LLM agents increasingly operate in service scenarios where task success depends on coordination between the agent, the user, and a stateful environment. In such intera…