3 papers
cs.CL2024
LegalLens Shared Task 2024: Legal Violation Identification in Unstructured Text
Ben Hagag, Liav Harpaz, Gil Semo +5
This paper presents the results of the LegalLens Shared Task, focusing on detecting legal violations within text in the wild across two sub-tasks: LegalLens-NER for identifying leg…
cs.LG2024
Report Cards: Qualitative Evaluation of Language Models Using Natural Language Summaries
Blair Yang, Fuyang Cui, Keiran Paster +4
The rapid development and dynamic nature of large language models (LLMs) make it difficult for conventional quantitative benchmarks to accurately assess their capabilities. We prop…
cs.LG2024
Reward Machines for Deep RL in Noisy and Uncertain Environments
Andrew C. Li, Zizhao Chen, Toryn Q. Klassen +3
Reward Machines provide an automaton-inspired structure for specifying instructions, safety constraints, and other temporally extended reward-worthy behaviour. By exposing the unde…