5 papers
Locations of Characters in Narratives: Andersen and Persuasion Datasets
Batuhan Ozyurt, Roya Arkhmammadova, Deniz Yuret
The ability of machines to grasp spatial understanding within narrative contexts is an intriguing aspect of reading comprehension that continues to be studied. Motivated by the goa…
How much do LLMs learn from negative examples?
Shadi Hamdan, Deniz Yuret
Large language models (LLMs) undergo a three-phase training process: unsupervised pre-training, supervised fine-tuning (SFT), and learning from human feedback (RLHF/DPO). Notably,…
Neurocache: Efficient Vector Retrieval for Long-range Language Modeling
Ali Safaya, Deniz Yuret
This paper introduces Neurocache, an approach to extend the effective context size of large language models (LLMs) using an external vector cache to store its past states. Like rec…
Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking
Emre Can Acikgoz, Mete Erdogan, Deniz Yuret
Large Language Models (LLMs) are becoming crucial across various fields, emphasizing the urgency for high-quality models in underrepresented languages. This study explores the uniq…
Sequential Compositional Generalization in Multimodal Models
Semih Yagcioglu, Osman Batur İnce, Aykut Erdem +3
The rise of large-scale multimodal models has paved the pathway for groundbreaking advances in generative modeling and reasoning, unlocking transformative applications in a variety…