2 papers
cs.LG2026
Generative Actor-Critic with Soft Bridge Policies
Ke He, Le He, Shunpu Tang +2
Expressive generative policies such as diffusion and flow models are appealing for MaxEnt online reinforcement learning because of their ability to model multimodal and highly non-…
cs.MA2026
AgenticPrecoding: LLM-Empowered Multi-Agent System for Precoding Optimization
Zijiu Yang, Zixiang Zhang, Shunpu Tang +2
Precoding is a key technique for interference management and performance improvement in multi-antenna wireless systems. However, existing precoding methods are typically developed…