activity
20162024
most citedLearning to Generate Images of Outdoor Scenes from Attributes and Semantic Layouts

116 citations · 175 across the 10 of their papers we have counts for

collaborators

10 papers

cs.LG20243 cited

Hippocrates: An Open-Source Framework for Advancing Large Language Models in Healthcare

Emre Can Acikgoz, Osman Batur İnce, Rayene Bench +4

The integration of Large Language Models (LLMs) into healthcare promises to transform medical diagnostics, research, and patient care. Yet, the progression of medical LLMs faces ob…

cs.CL2024

Sequential Compositional Generalization in Multimodal Models

Semih Yagcioglu, Osman Batur İnce, Aykut Erdem +3

The rise of large-scale multimodal models has paved the pathway for groundbreaking advances in generative modeling and reasoning, unlocking transformative applications in a variety…

cs.CL2023

ViLMA: A Zero-Shot Benchmark for Linguistic and Temporal Grounding in Video-Language Models

Ilker Kesen, Andrea Pedrotti, Mustafa Dogan +8

With the ever-increasing popularity of pretrained Video-Language Models (VidLMs), there is a pressing need to develop robust evaluation methodologies that delve deeper into their v…

cs.CL2023

Harnessing Dataset Cartography for Improved Compositional Generalization in Transformers

Osman Batur İnce, Tanin Zeraati, Semih Yagcioglu +3

Neural networks have revolutionized language modeling and excelled in various downstream tasks. However, the extent to which these models achieve compositional generalization compa…

eess.IV202340 cited

Hyperspectral Image Denoising via Self-Modulating Convolutional Neural Networks

Orhan Torun, Seniha Esen Yuksel, Erkut Erdem +2

Compared to natural images, hyperspectral images (HSIs) consist of a large number of bands, with each band capturing different spectral information from a certain wavelength, even…

cs.CV2023

Spherical Vision Transformer for 360-degree Video Saliency Prediction

Mert Cokelek, Nevrez Imamoglu, Cagri Ozcinar +2

The growing interest in omnidirectional videos (ODVs) that capture the full field-of-view (FOV) has gained 360-degree saliency prediction importance in computer vision. However, pr…