3 papers
cs.AI2026
Gated-BEPO: Confidence-Gated Bellman Credit Assignment for Large Language Model Agents
Hongxi Yan, Ziyue Huang, Shichao Fan +1
Training large language model agents in long-horizon environments requires assigning credit from sparse terminal outcomes to individual actions. Existing critic-free methods propag…
cs.LG2026
SeeDNorm: Self-Rescaled Dynamic Normalization
Wenrui Cai, Defa Zhu, Qingjie Liu +1
Normalization layer constitutes an essential component in neural networks. In transformers, the predominantly used RMSNorm constrains vectors to a unit hypersphere, followed by dim…
cs.CV2025
A Survey on Remote Sensing Foundation Models: From Vision to Multimodality
Ziyue Huang, Hongxi Yan, Qiqi Zhan +7
The rapid advancement of remote sensing foundation models, particularly vision and multimodal models, has significantly enhanced the capabilities of intelligent geospatial data int…