3 papers
cs.LG2026
DRAN: A Distribution and Relation Adaptive Network for Spatio-temporal Forecasting
Xiaobei Zou, Luolin Xiong, Kexuan Zhang +2
Accurate predictions of spatio-temporal systems are crucial for tasks such as system management, control, and crisis prevention. However, the inherent time variance of many spatio-…
cs.LG2025
HCPO: Hierarchical Conductor-Based Policy Optimization in Multi-Agent Reinforcement Learning
Zejiao Liu, Junqi Tu, Yitian Hong +4
In cooperative Multi-Agent Reinforcement Learning (MARL), efficient exploration is crucial for optimizing the performance of joint policy. However, existing methods often update jo…
cs.AI2025
DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models
Luolin Xiong, Haofen Wang, Xi Chen +7
DeepSeek, a Chinese Artificial Intelligence (AI) startup, has released their V3 and R1 series models, which attracted global attention due to their low cost, high performance, and…