12 citations · 25 across the 7 of their papers we have counts for
9 papers
MileGPO: Milestone Inference with Local Evidence for Graph-Based Policy Optimization of Long-Horizon LLM Agents
Bo Qian, Yuting Wu, Shuang Zeng +3
Credit assignment is challenging in long-horizon agentic reinforcement learning, where supervision often comes only from final rewards. Existing methods refine trajectory-level sig…
Incorporating Self-Rewriting into Large Language Model Reasoning Reinforcement
Jiashu Yao, Heyan Huang, Shuang Zeng +6
Through reinforcement learning (RL) with outcome correctness rewards, large reasoning models (LRMs) with scaled inference computation have demonstrated substantial success on compl…
A Two-Stream AMR-enhanced Model for Document-level Event Argument Extraction
Runxin Xu, Peiyi Wang, Tianyu Liu +3
Most previous studies aim at extracting events from a single sentence, while document-level event extraction still remains under-explored. In this paper, we focus on extracting eve…
DISK: Domain-constrained Instance Sketch for Math Word Problem Generation
Tianyang Cao, Shuang Zeng, Xiaodan Xu +2
A math word problem (MWP) is a coherent narrative which reflects the underlying logic of math equations. Successful MWP generation can automate the writing of mathematics questions…
SIRE: Separate Intra- and Inter-sentential Reasoning for Document-level Relation Extraction
Shuang Zeng, Yuting Wu, Baobao Chang
Document-level relation extraction has attracted much attention in recent years. It is usually formulated as a classification problem that predicts relations for all entity pairs i…
Generating Math Word Problems from Equations with Topic Controlling and Commonsense Enforcement
Tianyang Cao, Shuang Zeng, Songge Zhao +2
Recent years have seen significant advancement in text generation tasks with the help of neural language models. However, there exists a challenging task: generating math problem t…