2 papers
cs.CL2026
AR-BENCH: Benchmarking Legal Reasoning with Judgment Error Detection, Classification and Correction
Yifei Li, Richong Zhang, Wanyu Tu +4
Legal judgments may contain errors due to the complexity of case circumstances and the abstract nature of legal concepts, while existing appellate review mechanisms face efficiency…
cs.DB2026
ThriftLLM: On Cost-Effective Selection of Large Language Models for Classification Queries
Keke Huang, Yimin Shi, Dujian Ding +4
In recent years, large language models (LLMs) have demonstrated remarkable capabilities in comprehending and generating natural language content, attracting widespread attention in…