Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liability Assessment
Junyu Lu, Qi Wei, Peishuo Zheng +6
Legal Judgment Prediction (LJP) has become a core benchmark for evaluating AI in the criminal legal domain, but it only sees criminal cases that have already passed prosecutorial r…
cs.CL2026
ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring
Deyi Ji, Junyu Lu, Xuanyi Liu +7
Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restrictions on education or aest…
cs.CL2025
L0: Reinforcement Learning to Become General Agents
Junjie Zhang, Jingyi Xi, Zhuoyang Song +7
Training large language models (LLMs) to act as autonomous agents for multi-turn, long-horizon tasks remains significant challenges in scalability and training efficiency. To addre…