From the 9 of 408 papers with an AI index.
266 citations
- Tsinghua UniversityCN107 papers
- University of Science and Technology of ChinaCN100 papers
- Zhejiang UniversityCN100 papers
- Istituto Nazionale di Fisica Nucleare, Laboratori Nazionali di FrascatiIT99 papers
- Beihang UniversityCN98 papers
- Nanjing Normal UniversityCN98 papers
- Peking UniversityCN98 papers
- National Centre for Nuclear ResearchPL97 papers
- South China Normal UniversityCN97 papers
- University of TarapacáCL97 papers
- University of TurinIT97 papers
- Institute of Modern PhysicsCN96 papers
19 papers · 1 filter
An analysis of machine learning approaches for enhancing decision-making in complex discrete choice tasks
Sheng Lun Christine Cao, Destenie Nock, Alex Davis
Discrete choice modeling is a common tool used for preference elicitation during policy-making, but this is typically done through parametric models. Machine learning can push the…
MacrOData: New Benchmarks of Thousands of Datasets for Tabular Outlier Detection
Xueying Ding, Simon Klüttermann, Haomin Wen +2
Quality benchmarks are essential for fairly and accurately tracking scientific progress and enabling practitioners to make informed methodological choices. Outlier detection (OD) o…
Muscle Synergy Priors Enhance Biomechanical Fidelity in Predictive Musculoskeletal Locomotion Simulation
Ilseung Park, Eunsik Choi, Jangwhan Ahn +1
Human locomotion emerges from high-dimensional neuromuscular control, making predictive musculoskeletal simulation challenging. We present a physiology-informed reinforcement-learn…
Causal methods for LLM development and evaluation
Dennis Frauen, Marie Brockschmidt, Konstantin Hess +10
Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and evaluation pipelines. Here,…
Trading off rewards and errors in multi-armed bandits
Akram Erraqabi, Alessandro Lazaric, Michal Valko +2
In multi-armed bandits, the most-explored arms are the most informative, while reward maximization typically pulls only the best arm. We study the tradeoff between identifying arm…
Unichain and Aperiodicity are Sufficient for Asymptotic Optimality of Average-Reward Restless Bandits
Yige Hong, Qiaomin Xie, Yudong Chen +1
We consider the infinite-horizon, average-reward restless bandit problem in discrete time. We propose a new class of policies that are designed to drive a progressively larger subs…