2 papers
cs.LG2026
MLB: A Scenario-Driven Benchmark for Evaluating Large Language Models in Clinical Applications
Qing He, Dongsheng Bi, Jianrong Lu +20
The proliferation of Large Language Models (LLMs) presents transformative potential for healthcare, yet practical deployment is hindered by the absence of frameworks that assess re…
cs.LG2025
Dynamic Backtracking in GFlowNets: Enhancing Decision Steps with Reward-Dependent Adjustment Mechanisms
Shuai Guo, Jielei Chu, Lin Ma +2
Generative Flow Networks (GFlowNets or GFNs) are probabilistic models predicated on Markov flows, and they employ specific amortization algorithms to learn stochastic policies that…