2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CL2026
Distribution-Aware Reward Estimation for Test-Time Reinforcement Learning
Bodong Du, Xuanqi Huang, Xiaomeng Li
Test-time reinforcement learning (TTRL) enables large language models (LLMs) to self-improve on unlabeled inputs, but its effectiveness critically depends on how reward signals are…
cs.AI2025
RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation
Bodong Du, Honglong Yang, Xiaomeng Li
Vision-language models have shown promising results in radiology report generation. However, most existing methods generate reports as flat text and do not explicitly model the sem…
cs.CV2025★ 2 cited
Multi-Modal Explainable Medical AI Assistant for Trustworthy Human-AI Collaboration
Honglong Yang, Shanshan Song, Yi Qin +6
Generalist Medical AI (GMAI) systems have demonstrated expert-level performance in biomedical perception tasks, yet their clinical utility remains limited by inadequate multi-modal…