1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2025
Walk Before You Run! Concise LLM Reasoning via Reinforcement Learning
Mingyang Song, Mao Zheng
As test-time scaling becomes a pivotal research frontier in Large Language Models (LLMs) development, contemporary and advanced post-training methodologies increasingly focus on ex…
cs.CL2025
TAT-R1: Terminology-Aware Translation with Reinforcement Learning and Word Alignment
Zheng Li, Mao Zheng, Mingyang Song +1
Recently, deep reasoning large language models(LLMs) like DeepSeek-R1 have made significant progress in tasks such as mathematics and coding. Inspired by this, several studies have…
cs.CV2024★ 1 cited
Mitigating Multilingual Hallucination in Large Vision-Language Models
Xiaoye Qu, Mingyang Song, Wei Wei +2
While Large Vision-Language Models (LVLMs) have exhibited remarkable capabilities across a wide range of tasks, they suffer from hallucination problems, where models generate plaus…