Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Chain of Operators: An Inference-Time Harness for In-Context Operator Learning
Minghui Yang, Chenghan Wu, Ling Guo +1
While scientific foundation models show immense promise in accelerating physical simulations and numerical forecasting, they remain notoriously brittle when encountering out-of-dis…
cs.LG2026
MLB: A Scenario-Driven Benchmark for Evaluating Large Language Models in Clinical Applications
Qing He, Dongsheng Bi, Jianrong Lu +20
The proliferation of Large Language Models (LLMs) presents transformative potential for healthcare, yet practical deployment is hindered by the absence of frameworks that assess re…