5 papers
Expert Preference-based Evaluation of Automated Related Work Generation
Furkan Åahinuç, Subhabrata Dutta, Iryna Gurevych
Expert domain writing, such as scientific writing, typically demands extensive domain knowledge. Although large language models (LLMs) show promising potential in this task, evalua…
Reward Modeling for Scientific Writing Evaluation
Furkan Åahinuç, Subhabrata Dutta, Iryna Gurevych
Scientific writing is an expert-domain task that demands deep domain knowledge, task-specific requirements and reasoning capabilities that leverage the domain knowledge to satisfy…
Efficient Performance Tracking: Leveraging Large Language Models for Automated Construction of Scientific Leaderboards
Furkan Åahinuç, Thy Thy Tran, Yulia Grishina +3
Scientific leaderboards are standardized ranking systems that facilitate evaluating and comparing competitive methods. Typically, a leaderboard is defined by a task, dataset, and e…
MiDe22: An Annotated Multi-Event Tweet Dataset for Misinformation Detection
Cagri Toraman, Oguzhan Ozcelik, Furkan Åahinuç +1
The rapid dissemination of misinformation through online social networks poses a pressing issue with harmful consequences jeopardizing human health, public safety, democracy, and t…
Systematic Task Exploration with LLMs: A Study in Citation Text Generation
Furkan Åahinuç, Ilia Kuznetsov, Yufang Hou +1
Large language models (LLMs) bring unprecedented flexibility in defining and executing complex, creative natural language generation (NLG) tasks. Yet, this flexibility brings new c…