4 papers
ATLAS: Verifier-Guided Adaptive Latent Activation Steering for Efficient LLM Reasoning
Tuc Nguyen, Thai Le
Recent work on activation and latent steering has demonstrated that modifying internal representations can effectively guide large language models (LLMs) toward improved reasoning…
ShareChat: A Dataset of Chatbot Conversations in the Wild
Yueru Yan, Tuc Nguyen, Bo Su +2
By evaluating Large Language Models (LLMs) through uniform, text-only interfaces, current academic benchmarks obscure how the unique designs and affordances of distinct commercial…
Spectral Flattening Is All Muon Needs: How Orthogonalization Controls Learning Rate and Convergence
Tien-Phat Nguyen, Truong Nguyen, Minh-Phuc Truong +3
Muon orthogonalizes the momentum buffer before each update, replacing its singular values with ones via Newton-Schulz iterations. This simple change lets Muon tolerate far larger l…
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
Tuc Nguyen, Yifan Hu, Thai Le
Recent advancements in large language models (LLMs) have been fueled by large scale training corpora drawn from diverse sources such as websites, news articles, and books. These da…