2 papers
cs.CL2024
Large Language Models as Evaluators for Scientific Synthesis
Julia Evans, Jennifer D'Souza, Sören Auer
Our study explores how well the state-of-the-art Large Language Models (LLMs), like GPT-4 and Mistral, can assess the quality of scientific summaries or, more fittingly, scientific…
cs.CL2024
Astro-NER -- Astronomy Named Entity Recognition: Is GPT a Good Domain Expert Annotator?
Julia Evans, Sameer Sadruddin, Jennifer D'Souza
In this study, we address one of the challenges of developing NER models for scholarly domains, namely the scarcity of suitable labeled data. We experiment with an approach using p…