Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Policy-Conditioned Counterfactual Credit for Verifiable Reinforcement Learning of Long-Horizon Language Agents
Renwei Meng
Reinforcement learning with verifiable rewards improves reasoning and tool use, yet long-horizon language agents still learn unsupported evidence chains, belief drift, and shortcut…
cs.LG2026
Group Resonance Network: Learnable Prototypes and Multi-Subject Resonance for EEG Emotion Recognition
Renwei Meng
Electroencephalography (EEG)-based emotion recognition remains challenging in cross-subject settings due to severe inter-subject variability. Existing methods mainly learn subject-…