3 papers
cs.CL2025
LaSeR: Reinforcement Learning with Last-Token Self-Rewarding
Wenkai Yang, Weijie Liu, Ruobing Xie +4
Reinforcement Learning with Verifiable Rewards (RLVR) has recently emerged as a core paradigm for enhancing the reasoning capabilities of Large Language Models (LLMs). To address t…
cs.CV2024
Spatio-Temporal Multi-Subgraph GCN for 3D Human Motion Prediction
Jiexin Wang, Yiju Guo, Bing Su
Human motion prediction (HMP) involves forecasting future human motion based on historical data. Graph Convolutional Networks (GCNs) have garnered widespread attention in this fiel…
cs.CV2024
Temporal Dynamics Decoupling with Inverse Processing for Enhancing Human Motion Prediction
Jiexin Wang, Yiju Guo, Bing Su
Exploring the bridge between historical and future motion behaviors remains a central challenge in human motion prediction. While most existing methods incorporate a reconstruction…