4 papers
Do LLMs Use Cultural Knowledge Without Being Told? A Multilingual Evaluation of Implicit Pragmatic Adaptation
Mehwish Nasim, Sanjeevan Selvaganapathy, Neel Ganapathi Sabhahit +6
Many benchmarks show that large language models can answer direct questions about culture. We study a different question: do they also change how they speak when culture is only im…
BONO-Bench: A Comprehensive Test Suite for Bi-objective Numerical Optimization with Traceable Pareto Sets
Lennart Schäpermeier, Pascal Kerschke
The evaluation of heuristic optimizers on test problems, better known as \emph{benchmarking}, is a cornerstone of research in multi-objective optimization. However, most test probl…
R2 v2: The Pareto-compliant R2 Indicator for Better Benchmarking in Bi-objective Optimization
Lennart Schäpermeier, Pascal Kerschke
In multi-objective optimization, set-based quality indicators are a cornerstone of benchmarking and performance assessment. They capture the quality of a set of trade-off solutions…
Greedy Restart Schedules: A Baseline for Dynamic Algorithm Selection on Numerical Black-box Optimization Problems
Lennart Schäpermeier
In many optimization domains, there are multiple different solvers that contribute to the overall state-of-the-art, each performing better on some, and worse on other types of prob…