Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models
Long Lian, Sida Wang, Felix Juefei-Xu +7
Scaling inference-time computation has enabled Large Language Models (LLMs) to achieve strong reasoning performance, but their inherently sequential decoding incurs substantial lat…
cs.LG2025
Data reuse enables cost-efficient randomized trials of medical AI models
Michael Nercessian, Wenxin Zhang, Alexander Schubert +4
Randomized controlled trials (RCTs) are indispensable for establishing the clinical value of medical artificial-intelligence (AI) tools, yet their high cost and long timelines hind…