2 papers
cs.LG2026
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL
Erdemt Bao, Xing Lei, Jun Chen
Hierarchical Implicit Q-Learning (HIQL), an offline goal-conditioned RL method, selects subgoals by value-function advantages alone. This rule has two coupled failure modes. Optimi…
cs.CV2026
WaveComm: Lightweight Communication for Collaborative Perception via Wavelet Feature Distillation
Erdemt Bao, Jin Yang
In multi-agent collaborative sensing systems, substantial communication overhead from information exchange significantly limits scalability and real-time performance, especially in…