works on

From the 1 of 29 linked papers with an AI index.

activity
20242026
collaborators

29 papers

cs.CL2026

EasyOPD: An Easy-to-use On-Policy Distillation Framework for Large Language Models

Jie Sun, Mao Zheng, Mingyang Song +7

The paper introduces EasyOPD, a modular framework that simplifies on-policy distillation for large language models by separating configuration, supervision logic, and distributed e…

cs.LG2026

A Survey of On-Policy Distillation for Large Language Models

Mingyang Song, Mao Zheng

As Large Language Models continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable students has become an important engineering prob…

cs.LG2026

On-Policy Distillation with Curriculum Turn-level Guidance for Multi-turn Agents

Gengsheng Li, Mao Zheng, Mingyang Song +8

Multi-turn agents that plan, invoke tools, and interact with environments offer a promising paradigm for solving complex tasks, yet their capabilities typically rely on very large…

cs.CL2026

Memory Beyond Recall: A Dual-Process Cognitive Memory System for Self-Evolving LLM Agents

Tianxiang Fei, Mingyang Song, Mao Zheng +1

Long-term memory for an LLM agent is more than retrieving the right passage at the right time. Current memory systems collapse belief revision, causal coupling, and cross-domain ab…

cs.CL2026

HardMTBench: Stress-Testing Chinese-English Translation on Knowledge-Intensive Domains

Zheng Li, Mao Zheng, Mingyang Song +1

General-purpose machine translation benchmarks such as FLORES-200 have reached a saturation regime on Chinese-English pairs, where modern large language models cluster within a nar…

cs.CL2026

IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following

Mingrui Sun, Mao Zheng, Zheng Li +1

Modern translation workflows demand more than semantic equivalence. Users routinely require models to preserve JSON or HTML schemas, honor curated glossaries, disambiguate with pro…