17 papers
Uncertainty-Aware Generation and Decision-Making Under Ambiguity
Nico Daheim, Iryna Gurevych
With rapidly improving capabilities, Large Language Models (LLMs) are increasingly used in many complex real-world tasks. Beyond requiring in-depth knowledge and reasoning skills,…
Variational Model Merging for Pareto Front Estimation in Multitask Finetuning
Hugo Monzón Maldonado, Nico Daheim, Thomas Möllenhoff +2
Pareto fronts are useful to find good task-mixing strategies for multitask finetuning, but they are also costly to compute. To reduce costs, recent works have used existing model m…
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
Andreas Waldis, Yotam Perlitz, Leshem Choshen +2
We introduce Holmes, a new benchmark designed to assess language models (LMs) linguistic competence - their unconscious understanding of linguistic phenomena. Specifically, we use…
In-depth Research Impact Summarization through Fine-Grained Temporal Citation Analysis
Hiba Arnaout, Noy Sternlicht, Tom Hope +1
Understanding the impact of scientific publications is crucial for identifying breakthroughs and guiding future research. Traditional metrics based on citation counts often miss th…
GRITHopper: Decomposition-Free Multi-Hop Dense Retrieval
Justus-Jonas Erker, Nils Reimers, Iryna Gurevych
Decomposition-based multi-hop retrieval methods rely on many autoregressive steps to break down complex queries, which breaks end-to-end differentiability and is computationally ex…
Decision-Making with Deliberation: Meta-reviewing as a Document-grounded Dialogue
Sukannya Purkayastha, Nils Dycke, Anne Lauscher +1
Meta-reviewing is a pivotal stage in the peer-review process, serving as the final step in determining whether a paper is recommended for acceptance. Prior research on meta-reviewi…