5 papers
Evaluating the Robustness of Dense Retrievers in Interdisciplinary Domains
Sarthak Chaturvedi, Anurag Acharya, Rounak Meyur +3
Evaluation benchmark characteristics may distort the true benefits of domain adaptation in retrieval models. This creates misleading assessments that influence deployment decisions…
Benchmarking LLMs for Environmental Review and Permitting
Rounak Meyur, Hung Phan, Koby Hayashi +12
The National Environment Policy Act (NEPA) stands as a foundational piece of environmental legislation in the United States, requiring federal agencies to consider the environmenta…
WeQA: A Benchmark for Retrieval Augmented Generation in Wind Energy Domain
Rounak Meyur, Hung Phan, Sridevi Wagle +5
Wind energy project assessments present significant challenges for decision-makers, who must navigate and synthesize hundreds of pages of environmental and scientific documentation…
DISHONEST: Dissecting misInformation Spread using Homogeneous sOcial NEtworks and Semantic Topic classification
Caleb Stam, Emily Saldanha, Mahantesh Halappanavar +1
The emergence of the COVID-19 pandemic resulted in a significant rise in the spread of misinformation on online platforms such as Twitter. Oftentimes this growth is blamed on the i…
Exploring the Benefits of Domain-Pretraining of Generative Large Language Models for Chemistry
Anurag Acharya, Shivam Sharma, Robin Cosbey +3
A proliferation of Large Language Models (the GPT series, BLOOM, LLaMA, and more) are driving forward novel development of multipurpose AI for a variety of tasks, particularly natu…