3 papers
eess.IV2026
Prompt-Guided Prefiltering for VLM Image Compression
Bardia Azizian, Ivan V. Bajic
The rapid progress of large Vision-Language Models (VLMs) has enabled a wide range of applications, such as image understanding and Visual Question Answering (VQA). Query images ar…
cs.CV2025
How Universal Are SAM2 Features?
Masoud Khairi Atani, Alon Harell, Hyomin Choi +3
The trade-off between general-purpose foundation vision models and their specialized counterparts is critical for efficient feature coding design and is not yet fully understood. W…
eess.IV2024
Learned Scalable Video Coding For Humans and Machines
Hadi Hadizadeh, Ivan V. BajiÄ
Video coding has traditionally been developed to support services such as video streaming, videoconferencing, digital TV, and so on. The main intent was to enable human viewing of…