3 papers
cs.AI2026
CT Open: An Open-Access, Uncontaminated, Live Platform for the Open Challenge of Clinical Trial Outcome Prediction
Jianyou Wang, Youze Zheng, Longtian Bao +11
Scientists have long sought to accurately predict outcomes of real-world events before they happen. Can AI systems do so more reliably? We study this question through clinical tria…
cs.AI2026
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
Lingfeng Li, Yunlong Lu, Yuefei Zhang +7
Large Language Models (LLMs) are increasingly deployed in interactive environments requiring strategic decision-making, yet systematic evaluation of these capabilities remains chal…
cs.AI2026
Decoupling Return-to-Go for Efficient Decision Transformer
Yongyi Wang, Hanyu Liu, Lingfeng Li +5
The Decision Transformer (DT) has established a powerful sequence modeling approach to offline reinforcement learning. It conditions its action predictions on Return-to-Go (RTG), u…