10 citations · 10 across the 3 of their papers we have counts for
3 papers · 1 filter
Generating Diverse Q&A Benchmarks for RAG Evaluation with DataMorgana
Simone Filice, Guy Horowitz, David Carmel +3
Evaluating Retrieval-Augmented Generation (RAG) systems, especially in domain-specific contexts, requires benchmarks that address the distinctive requirements of the applicative sc…
Alexa, Let's Work Together: Introducing the First Alexa Prize TaskBot Challenge on Conversational Task Assistance
Anna Gottardi, Osman Ipek, Giuseppe Castellucci +27
Since its inception in 2016, the Alexa Prize program has enabled hundreds of university students to explore and compete to develop conversational agents through the SocialBot Grand…
Answering Product-Questions by Utilizing Questions from Other Contextually Similar Products
Ohad Rozen, David Carmel, Avihai Mejer +2
Predicting the answer to a product-related question is an emerging field of research that recently attracted a lot of attention. Answering subjective and opinion-based questions is…