collaborators

5 papers

cs.AI2026

Library Reachability in LSR-Synth: How Anti-Memorization Design Changes the Measurement of Symbolic Discovery

Zhan'ao Yao, Liang Yin, Zhihao Gao +9

Existing benchmarks for scientific equation discovery are largely composed of well-known equations available in the public domain, making it difficult to determine whether a model…

cs.AI2026

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

Shumeng Yang, Yisu Liu, Jiayi Zheng +2

Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, where the policy prematurely con…

cond-mat.mtrl-sci2026

MatMind: A Structure-Activity Knowledge-Driven Generative Foundation Model for Materials Science

Zhan'ao Yao, Boxuan Zhang, Jingyuan Shu +10

Progress in AI-driven crystal materials science has so far been carried by narrow architectures purpose-built for individual tasks -- graph neural networks for property prediction,…

cs.AI2025

Unearthing Gems from Stones: Policy Optimization with Negative Sample Augmentation for LLM Reasoning

Zhaohui Yang, Yuxiao Ye, Shilei Jiang +4

Recent advances in reasoning language models have witnessed a paradigm shift from short to long CoT pattern. Given the substantial computational cost of rollouts in long CoT models…

cs.AI2025

Beyond the First Error: Process Reward Models for Reflective Mathematical Reasoning

Zhaohui Yang, Chenghua He, Xiaowen Shi +4

Many studies focus on data annotation techniques for training effective PRMs. However, current methods encounter a significant issue when applied to long CoT reasoning processes: t…