activity
20242026
collaborators

5 papers

cs.CL2026

From Raw Corpora to Domain Benchmarks: Automated Evaluation of LLM Domain Expertise

Nitin Sharma, Thomas Wolfers, Çağatay Yıldız

Accurate domain-specific benchmarking of LLMs is essential, specifically in domains with direct implications for humans, such as law, healthcare, and education. However, existing b…

cs.CV2025

Object-level Self-Distillation for Vision Pretraining

Çağlar Hızlı, Çağatay Yıldız, Pekka Marttinen

State-of-the-art vision pretraining methods rely on image-level self-distillation from object-centric datasets such as ImageNet, implicitly assuming each image contains a single ob…

cs.CL2025

Investigating Continual Pretraining in Large Language Models: Insights and Implications

Çağatay Yıldız, Nishaanth Kanna Ravichandran, Nitin Sharma +2

Continual learning (CL) in large language models (LLMs) is an evolving domain that focuses on developing efficient and sustainable training strategies to adapt models to emerging k…

cs.LG2024

Infinite dSprites for Disentangled Continual Learning: Separating Memory Edits from Generalization

Sebastian Dziadzio, Çağatay Yıldız, Gido M. van de Ven +3

The ability of machine learning systems to learn continually is hindered by catastrophic forgetting, the tendency of neural networks to overwrite previously acquired knowledge when…

cs.LG2024

Identifying latent state transition in non-linear dynamical systems

Çağlar Hızlı, Çağatay Yıldız, Matthias Bethge +2

This work aims to improve generalization and interpretability of dynamical systems by recovering the underlying lower-dimensional latent states and their time evolutions. Previous…