3 papers
cs.CL2026
Generative Active Testing: Efficient LLM Evaluation via Proxy Task Adaptation
Aashish Anantha Ramakrishnan, Ardavan Saeedi, Hamid Reza Hassanzadeh +2
With the widespread adoption of pre-trained Large Language Models (LLM), there exists a high demand for task-specific test sets to benchmark their performance in domains such as he…
cs.CL2025
LLMs are Better Than You Think: Label-Guided In-Context Learning for Named Entity Recognition
Fan Bai, Hamid Hassanzadeh, Ardavan Saeedi +1
In-context learning (ICL) enables large language models (LLMs) to perform new tasks using only a few demonstrations. However, in Named Entity Recognition (NER), existing ICL method…
cs.CL2024
Give me Some Hard Questions: Synthetic Data Generation for Clinical QA
Fan Bai, Keith Harrigian, Joel Stremmel +3
Clinical Question Answering (QA) systems enable doctors to quickly access patient information from electronic health records (EHRs). However, training these systems requires signif…