collaborators

10 papers

cs.CL2026

Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents

Jianxiang Yu, Jiapeng Zhu, Bochen Lin +3

LLM agents increasingly retrieve externally curated skills-procedural instructions retrieved at decision time-to improve performance on long-horizon interactive tasks. Existing ski…

cs.CL2026

Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning

Jiapeng Zhu, Jianxiang Yu, Yibo Zhao +5

Equipping large language models with explicit skills has emerged as a promising paradigm for enabling autonomous agents to solve complex tasks. Agent skills can be inherently divid…

cs.CL2026

GRACE: Reinforcement Learning for Grounded Response and Abstention under Contextual Evidence

Yibo Zhao, Jiapeng Zhu, Zichen Ding +1

Retrieval-Augmented Generation (RAG) integrates external knowledge to enhance Large Language Models (LLMs), yet systems remain susceptible to two critical flaws: providing correct…

cs.LG2025

Human Cognition Inspired RAG with Knowledge Graph for Complex Problem Solving

Yao Cheng, Yibo Zhao, Jiapeng Zhu +3

Large Language Models (LLMs) have demonstrated significant potential across various domains. However, they often struggle with integrating external knowledge and performing complex…

cs.AI2025

SEAGraph: Unveiling the Whole Story of Paper Review Comments

Jianxiang Yu, Jiaqi Tan, Zichen Ding +7

Peer review, as a cornerstone of scientific research, ensures the integrity and quality of scholarly work by providing authors with objective feedback for refinement. However, in t…

cs.LG2025

Text Detoxification: Data Efficiency, Semantic Preservation and Model Generalization

Jing Yu, Yibo Zhao, Jiapeng Zhu +4

The widespread dissemination of toxic content on social media poses a serious threat to both online environments and public discourse, highlighting the urgent need for detoxificati…