1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2026
TED: Training-Free Experience Distillation for Multimodal Reasoning
Shuozhi Yuan, Jinqing Wang, Zihao Liu +5
Knowledge distillation is typically realized by transferring a teacher model's knowledge into a student's parameters through supervised or reinforcement-based optimization. While e…
cs.CL2024
UniCoder: Scaling Code Large Language Model via Universal Code
Tao Sun, Linzheng Chai, Jian Yang +6
Intermediate reasoning or acting steps have successfully improved large language models (LLMs) for handling various downstream natural language processing (NLP) tasks. When applyin…
cs.PL2024★ 1 cited
McEval: Massively Multilingual Code Evaluation
Linzheng Chai, Shukai Liu, Jian Yang +15
Code large language models (LLMs) have shown remarkable advances in code understanding, completion, and generation tasks. Programming benchmarks, comprised of a selection of code c…