2 papers
cs.CV2026
MGRegBench: A Novel Benchmark Dataset with Anatomical Landmarks for Mammography Image Registration
Svetlana Krasnova, Emiliya Starikova, Ilia Naletov +2
Robust mammography registration is essential for clinically relevant applications like tracking disease progression in breast tissue. However, progress has been limited by the abse…
cs.CL2024
BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack
Yuri Kuratov, Aydar Bulatov, Petr Anokhin +4
In recent years, the input context sizes of large language models (LLMs) have increased dramatically. However, existing evaluation methods have not kept pace, failing to comprehens…