activity
20232026
most citedLadder: A Model-Agnostic Framework Boosting LLM-based Machine Translation to the Next Level

1 citations · 1 across the 15 of their papers we have counts for

collaborators

16 papers

cs.CV2026

SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction

Zhaopeng Feng, Chen Zhi, Xuhong Zhang +2

Reconstructing Computer-Aided Design (CAD) modeling sequences from images is crucial for preserving design intent and supporting parametric editing. However, existing methods typic…

cs.CL2026

SkillComposer: Learning to Evolve Agent Skills for Specification and Generalization

Qi Zhang, Zhaopeng Feng, Xiaonan Shi +8

Agent skills, which consist of reusable strategies that guide agent reasoning and action, have shown strong potential for improving model capability at inference time. However, cur…

cs.CL2026

AgentSwing: Adaptive Parallel Context Management Routing for Long-Horizon Web Agents

Zhaopeng Feng, Liangcai Su, Zhen Zhang +16

As large language models (LLMs) evolve into autonomous agents for long-horizon information-seeking, managing finite context capacity has become a critical bottleneck. Existing cont…

cs.CV2025

Datasets and Recipes for Video Temporal Grounding via Reinforcement Learning

Ruizhe Chen, Zhiting Fan, Tianze Luo +7

Video Temporal Grounding (VTG) aims to localize relevant temporal segments in videos given natural language queries. Despite recent progress with large vision-language models (LVLM…

cs.CL2025

Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning

Xiaotian Zhang, Yuan Wang, Zhaopeng Feng +6

Medical Question-Answering (QA) encompasses a broad spectrum of tasks, including multiple choice questions (MCQ), open-ended text generation, and complex computational reasoning. D…

cs.CL2025

CP-Router: An Uncertainty-Aware Router Between LLM and LRM

Jiayuan Su, Fulin Lin, Zhaopeng Feng +7

Recent advances in Large Reasoning Models (LRMs) have significantly improved long-chain reasoning capabilities over Large Language Models (LLMs). However, LRMs often produce unnece…