10 papers
Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents
Jianxiang Yu, Jiapeng Zhu, Bochen Lin +3
LLM agents increasingly retrieve externally curated skills-procedural instructions retrieved at decision time-to improve performance on long-horizon interactive tasks. Existing ski…
Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning
Jiapeng Zhu, Jianxiang Yu, Yibo Zhao +5
Equipping large language models with explicit skills has emerged as a promising paradigm for enabling autonomous agents to solve complex tasks. Agent skills can be inherently divid…
GRACE: Reinforcement Learning for Grounded Response and Abstention under Contextual Evidence
Yibo Zhao, Jiapeng Zhu, Zichen Ding +1
Retrieval-Augmented Generation (RAG) integrates external knowledge to enhance Large Language Models (LLMs), yet systems remain susceptible to two critical flaws: providing correct…
Human Cognition Inspired RAG with Knowledge Graph for Complex Problem Solving
Yao Cheng, Yibo Zhao, Jiapeng Zhu +3
Large Language Models (LLMs) have demonstrated significant potential across various domains. However, they often struggle with integrating external knowledge and performing complex…
SEAGraph: Unveiling the Whole Story of Paper Review Comments
Jianxiang Yu, Jiaqi Tan, Zichen Ding +7
Peer review, as a cornerstone of scientific research, ensures the integrity and quality of scholarly work by providing authors with objective feedback for refinement. However, in t…
Text Detoxification: Data Efficiency, Semantic Preservation and Model Generalization
Jing Yu, Yibo Zhao, Jiapeng Zhu +4
The widespread dissemination of toxic content on social media poses a serious threat to both online environments and public discourse, highlighting the urgent need for detoxificati…