Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
TOD-ProcBench: Benchmarking Complex Instruction-Following in Task-Oriented Dialogues
Sarik Ghazarian, Abhinav Gullapalli, Swair Shah +4
In real-world task-oriented dialogue (TOD) settings, agents are required to strictly adhere to complex instructions while conducting multi-turn conversations with customers. These…
cs.CL2024
FiNER-ORD: Financial Named Entity Recognition Open Research Dataset
Agam Shah, Abhinav Gullapalli, Ruchit Vithani +2
Over the last two decades, the development of the CoNLL-2003 named entity recognition (NER) dataset has helped enhance the capabilities of deep learning and natural language proces…