2 papers
cs.CL2025
LLM-as-a-Judge is Bad, Based on AI Attempting the Exam Qualifying for the Member of the Polish National Board of Appeal
MichaŠKarp, Anna Kubaszewska, Magdalena Król +4
This study provides an empirical assessment of whether current large language models (LLMs) can pass the official qualifying examination for membership in Poland's National Appeal…
cs.LG2025
FeNeC: Enhancing Continual Learning via Feature Clustering with Neighbor- or Logit-Based Classification
Kamil KsiÄ Å¼ek, Hubert JastrzÄbski, Bartosz Trojan +3
The ability of deep learning models to learn continuously is essential for adapting to new data categories and evolving data distributions. In recent years, approaches leveraging f…