3 papers
cs.LG2026
AREAL-DTA: Dynamic Tree Attention for Efficient Reinforcement Learning of Large Language Models
Jiarui Zhang, Yuchen Yang, Ran Yan +8
Reinforcement learning (RL)-based post-training for large language models (LLMs) is computationally expensive, as it generates many rollout sequences that frequently share long tok…
cs.CL2025
Generative Psycho-Lexical Approach for Constructing Value Systems in Large Language Models
Haoran Ye, Tianze Zhang, Yuhang Xie +4
Values are core drivers of individual and collective perception, cognition, and behavior. Value systems, such as Schwartz's Theory of Basic Human Values, delineate the hierarchy an…
cs.AI2025
SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber World
Jiaqi Zhang, Chen Gao, Liyuan Zhang +2
Recent advances in embodied agents with multimodal perception and reasoning capabilities based on large vision-language models (LVLMs), excel in autonomously interacting either rea…