2 papers
cs.CL2026
DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain
Hsuvas Borkakoty, Sebastian Pohl, Cheng Wang +2
LLM-based agents are increasingly deployed for routine but consequential tasks in real-world domains, where their behavior is governed by inherently ambiguous domain policies that…
q-fin.TR2025
Advanced simulation paradigm of human behaviour unveils complex financial systemic projection
Cheng Wang, Chuwen Wang, Shirong Zeng +2
The high-order complexity of human behaviour is likely the root cause of extreme difficulty in financial market projections. We consider that behavioural simulation can unveil syst…