activity
20242026
collaborators

12 papers

cs.CL2026

A Training-Free Regeneration Paradigm: Contrastive Reflection Memory Guided Self-Verification and Self-Improvement

Yuran Li, Di Wu, Benoit Boulet

Verification-guided self-improvement has recently emerged as a promising approach to improving the accuracy of large language model (LLM) outputs. However, existing approaches face…

cs.AI2025

STEMS: Spatial-Temporal Enhanced Safe Multi-Agent Coordination for Building Energy Management

Huiliang Zhang, Di Wu, Arnaud Zinflou +1

Building energy management is essential for achieving carbon reduction goals, improving occupant comfort, and reducing energy costs. Coordinated building energy management faces cr…

eess.SY2025

Causal Feature Selection for Weather-Driven Residential Load Forecasting

Elise Zhang, François Mirallès, Stéphane Dellacherie +2

Weather is a dominant external driver of residential electricity demand, but adding many meteorological covariates can inflate model complexity and may even impair accuracy. Select…

cs.LG2025

DRDT3: Diffusion-Refined Decision Test-Time Training Model

Xingshuai Huang, Di Wu, Benoit Boulet

Decision Transformer (DT), a trajectory modelling method, has shown competitive performance compared to traditional offline reinforcement learning (RL) approaches on various classi…

cs.LG2025

Goal-Conditioned Data Augmentation for Offline Reinforcement Learning

Xingshuai Huang, Di Wu, Benoit Boulet

Offline reinforcement learning (RL) enables policy learning from pre-collected offline datasets, relaxing the need to interact directly with the environment. However, limited by th…

cs.LG2025

Leveraging Multivariate Long-Term History Representation for Time Series Forecasting

Huiliang Zhang, Di Wu, Arnaud Zinflou +3

Multivariate Time Series (MTS) forecasting has a wide range of applications in both industry and academia. Recent advances in Spatial-Temporal Graph Neural Network (STGNN) have ach…