Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
STEMS: Spatial-Temporal Enhanced Safe Multi-Agent Coordination for Building Energy Management
Huiliang Zhang, Di Wu, Arnaud Zinflou +1
Building energy management is essential for achieving carbon reduction goals, improving occupant comfort, and reducing energy costs. Coordinated building energy management faces cr…
cs.AI2025
Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments
Yuran Li, Jama Hussein Mohamud, Chongren Sun +2
Large language models (LLMs) are being widely applied across various fields, but as tasks become more complex, evaluating their responses is increasingly challenging. Compared to h…
cs.AI2024
Robot Policy Learning with Temporal Optimal Transport Reward
Yuwei Fu, Haichao Zhang, Di Wu +2
Reward specification is one of the most tricky problems in Reinforcement Learning, which usually requires tedious hand engineering in practice. One promising approach to tackle thi…