most citedRethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL

1 citations · 1 across the 4 of their papers we have counts for

collaborators

5 papers

cs.RO2025

Seeing to Act, Prompting to Specify: A Bayesian Factorization of Vision Language Action Policy

Kechun Xu, Zhenjie Zhu, Anzhe Chen +7

The pursuit of out-of-distribution generalization in Vision-Language-Action (VLA) models is often hindered by catastrophic forgetting of the Vision-Language Model (VLM) backbone du…

cs.AI20251 cited

Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL

Haoyang He, Zihua Rong, Kun Ji +5

Reinforcement learning (RL) has recently become the dominant paradigm for strengthening the reasoning abilities of large language models (LLMs). Yet the rule-based reward functions…

cs.CV2025

FinMMR: Make Financial Numerical Reasoning More Multimodal, Comprehensive, and Challenging

Zichen Tang, Haihong E, Jiacheng Liu +18

We present FinMMR, a novel bilingual multimodal benchmark tailored to evaluate the reasoning capabilities of multimodal large language models (MLLMs) in financial numerical reasoni…

cs.CL2025

FinanceReasoning: Benchmarking Financial Numerical Reasoning More Credible, Comprehensive and Challenging

Zichen Tang, Haihong E, Ziyan Ma +10

We introduce FinanceReasoning, a novel benchmark designed to evaluate the reasoning capabilities of large reasoning models (LRMs) in financial numerical reasoning problems. Compare…

cs.IR2025

Diffusion-driven SpatioTemporal Graph KANsformer for Medical Examination Recommendation

Jianan Li, Yangtao Zhou, Zhifu Zhao +5

Recommendation systems in AI-based medical diagnostics and treatment constitute a critical component of AI in healthcare. Although some studies have explored this area and made not…