Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
TextBandit: Evaluating Probabilistic Reasoning in LLMs Through Language-Only Decision Tasks
Jimin Lim, Arjun Damerla, Arthur Jiang +1
Large language models (LLMs) have shown to be increasingly capable of performing reasoning tasks, but their ability to make sequential decisions under uncertainty only using natura…
cs.CL2024
A Korean Legal Judgment Prediction Dataset for Insurance Disputes
Alice Saebom Kwak, Cheonkam Jeong, Ji Weon Lim +1
This paper introduces a Korean legal judgment prediction (LJP) dataset for insurance disputes. Successful LJP models on insurance disputes can benefit insurance companies and their…