3 papers
cs.CL2026
Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning
Kunbin Xu, Xingzuo Li, Xuefeng Bai +1
Test-time reinforcement learning (TTRL) improves the reasoning capabilities of large language models without labeled data by updating the policy with pseudo-labels constructed thro…
cs.CL2025
Generator-Assistant Stepwise Rollback Framework for Large Language Model Agent
Xingzuo Li, Kehai Chen, Yunfei Long +3
Large language model (LLM) agents typically adopt a step-by-step reasoning framework, in which they interleave the processes of thinking and acting to accomplish the given task. Ho…
cs.CL2024
LLM with Relation Classifier for Document-Level Relation Extraction
Xingzuo Li, Kehai Chen, Yunfei Long +1
Large language models (LLMs) have created a new paradigm for natural language processing. Despite their advancement, LLM-based methods still lag behind traditional approaches in do…