2 papers
cs.LG2026
Amortized Reasoning Tree Search: Decoupling Proposal and Decision in Large Language Models
Zesheng Hong, Jiadong Yu, Hui Pan
Reinforcement Learning with Verifiable Rewards (RLVR) has established itself as the dominant paradigm for instilling rigorous reasoning capabilities in Large Language Models. While…
cs.CV2024
Out-of-distribution Detection in Medical Image Analysis: A survey
Zesheng Hong, Yubiao Yue, Yubin Chen +10
Computer-aided diagnostics has benefited from the development of deep learning-based computer vision techniques in these years. Traditional supervised deep learning methods assume…