3 papers
cs.AI2026
SURE-Med: Systematic Uncertainty Reduction for Enhanced Reliability in Medical Report Generation
Yuhang Gu, Xingyu Hu, Yuyu Fan +3
Automated medical report generation (MRG) holds great promise for reducing the heavy workload of radiologists. However, its clinical deployment is hindered by three major sources o…
cs.CL2025
MetaBench: A Multi-task Benchmark for Assessing LLMs in Metabolomics
Yuxing Lu, Xukai Zhao, J. Ben Tamo +6
Large Language Models (LLMs) have demonstrated remarkable capabilities on general text; however, their proficiency in specialized scientific domains that require deep, interconnect…
cs.IR2025
On Negative-aware Preference Optimization for Recommendation
Chenlu Ding, Daoxuan Liu, Jiancan Wu +6
Recommendation systems leverage user interaction data to suggest relevant items while filtering out irrelevant (negative) ones. The rise of large language models (LLMs) has garnere…