Publications (12)
SiReRAG: Indexing Similar and Related Information for Multihop Reasoning
Nan Zhang, Prafulla Kumar Choubey, Alexander Fabbri +5
Indexing is an important step towards strong performance in retrieval-augmented generation (RAG) systems. However, existing methods organize data based on either semantic similarit…
Surfer100: Generating Surveys From Web Resources, Wikipedia-style
Irene Li, Alexander Fabbri, Rina Kawamura +7
Fast-developing fields such as Artificial Intelligence (AI) often outpace the efforts of encyclopedic sources such as Wikipedia, which either do not completely cover recently-intro…
A Transfer Learning Pipeline for Educational Resource Discovery with Application in Leading Paragraph Generation
Irene Li, Thomas George, Alexander Fabbri +7
Effective human learning depends on a wide selection of educational materials that align with the learner's current understanding of the topic. While the Internet has revolutionize…
Deep Label-Wise Attentive Temporal Convolutional Networks Improve Medical Coding
Muhammed Yavuz Nuzumlalı, Alexander Fabbri, Irene Li +1
Medical coding is the task of assigning a set of diagnosis and procedure codes for a hospitalization using recorded notes. It requires aggregating information from different parts…
CoSQL: A Conversational Text-to-SQL Challenge Towards Cross-Domain Natural Language Interfaces to Databases
Tao Yu, Rui Zhang, He Yang Er +21
We present CoSQL, a corpus for building cross-domain, general-purpose database (DB) querying dialogue systems. It consists of 30k+ turns plus 10k+ annotated SQL queries, obtained f…
Zero-shot Transfer Learning for Semantic Parsing
Javid Dadashkarimi, Alexander Fabbri, Sekhar Tatikonda +1
While neural networks have shown impressive performance on large datasets, applying these models to tasks where little data is available remains a challenging problem. In this pape…
Fair Abstractive Summarization of Diverse Perspectives
Yusen Zhang, Nan Zhang, Yixin Liu +9
People from different social and demographic groups express diverse perspectives and conflicting opinions on a broad set of topics such as product reviews, healthcare, law, and pol…
Improving Low-Resource Cross-lingual Document Retrieval by Reranking with Deep Bilingual Representations
Rui Zhang, Caitlin Westerfield, Sungrok Shim +5
In this paper, we propose to boost low-resource cross-lingual document retrieval performance with deep bilingual query-document representations. We match queries and documents in b…
From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting
Griffin Adams, Alexander Fabbri, Faisal Ladhak +2
Selecting the ``right'' amount of information to include in a summary is a difficult task. A good summary should be detailed and entity-centric without being overly dense and hard…
Investigating Crowdsourcing Protocols for Evaluating the Factual Consistency of Summaries
Xiangru Tang, Alexander Fabbri, Haoran Li +6
Current pre-trained models applied to summarization are prone to factual inconsistencies which either misrepresent the source text or introduce extraneous information. Thus, compar…
CLICKER: A Computational LInguistics Classification Scheme for Educational Resources
Swapnil Hingmire, Irene Li, Rena Kawamura +11
A classification scheme of a scientific subject gives an overview of its body of knowledge. It can also be used to facilitate access to research articles and other materials relate…
R-VGAE: Relational-variational Graph Autoencoder for Unsupervised Prerequisite Chain Learning
Irene Li, Alexander Fabbri, Swapnil Hingmire +1
The task of concept prerequisite chain learning is to automatically determine the existence of prerequisite relationships among concept pairs. In this paper, we frame learning prer…