7 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.CL2025
Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning
Chengqi Lyu, Songyang Gao, Yuzhe Gu +14
Reasoning abilities, especially those for solving complex math problems, are crucial components of general intelligence. Recent advances by proprietary companies, such as o-series…
cs.CV2024★ 1 cited
MG-LLaVA: Towards Multi-Granularity Visual Instruction Tuning
Xiangyu Zhao, Xiangtai Li, Haodong Duan +4
Multi-modal large language models (MLLMs) have made significant strides in various visual understanding tasks. However, the majority of these models are constrained to process low-…
cs.CV2024★ 7 cited
An Open and Comprehensive Pipeline for Unified Object Grounding and Detection
Xiangyu Zhao, Yicheng Chen, Shilin Xu +4
Grounding-DINO is a state-of-the-art open-set detection model that tackles multiple vision tasks including Open-Vocabulary Detection (OVD), Phrase Grounding (PG), and Referring Exp…