Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
AMIGO: Agentic Multi-Image Grounding Oracle Benchmark
Min Wang, Ata Mahjoubfar
Agentic vision-language models increasingly act through extended interactions, but most evaluations still focus on single-image, single-turn correctness. We introduce AMIGO (Agenti…
cs.LG2026
Offline Meta-Reinforcement Learning with Flow-Based Task Inference and Adaptive Correction of Feature Overgeneralization
Min Wang, Xin Li, Mingzhong Wang +1
Offline meta-reinforcement learning (OMRL) combines the strengths of learning from diverse datasets in offline RL with the adaptability to new tasks of meta-RL, promising safe and…
cs.LG2025
Wavelet Predictive Representations for Non-Stationary Reinforcement Learning
Min Wang, Xin Li, Ye He +4
The real world is inherently non-stationary, with ever-changing factors, such as weather conditions and traffic flows, making it challenging for agents to adapt to varying environm…