2 papers
cs.CL2025
DARE: Diverse Visual Question Answering with Robustness Evaluation
Hannah Sterz, Jonas Pfeiffer, Ivan VuliÄ
Vision Language Models (VLMs) extend remarkable capabilities of text-only large language models and vision-only models, and are able to learn from and process multi-modal vision-te…
cs.CL2025
Retrofitting Large Language Models with Dynamic Tokenization
Darius Feher, Ivan VuliÄ, Benjamin Minixhofer
Current language models (LMs) use a fixed, static subword tokenizer. This default choice typically results in degraded efficiency and language capabilities, especially in languages…