2 papers
cs.CL2026
TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning
Muyu Pan, Shu Zhao, Nan Zhang +4
This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large language models. This paper exten…
cs.CV2025
Parts-Mamba: Augmenting Joint Context with Part-Level Scanning for Occluded Human Skeleton
Tianyi Shen, Huijuan Xu, Nilesh Ahuja +3
Skeleton action recognition involves recognizing human action from human skeletons. The use of graph convolutional networks (GCNs) has driven major advances in this recognition tas…