1 paper
Manan Roy Choudhury, Adithya Chandramouli, Mannan Anand +1
The rapid integration of large language models (LLMs) into high-stakes legal work has exposed a critical gap: no benchmark exists to systematically stress-test their reliability ag…