activity
20172024
most citedUnderstanding Visual Saliency in Mobile User Interfaces

81 citations · 180 across the 10 of their papers we have counts for

collaborators

17 papers

cs.CV20248 cited

AIM 2024 Challenge on Video Saliency Prediction: Methods and Results

Andrey Moskalenko, Alexey Bryncev, Dmitry Vatolin +30

This paper reviews the Challenge on Video Saliency Prediction at AIM 2024. The goal of the participants was to develop a method for predicting accurate saliency maps for the provid…

cs.CV20223 cited

Leveraging progressive model and overfitting for efficient learned image compression

Honglei Zhang, Francesco Cricri, Hamed Rezazadegan Tavakoli +2

Deep learning is overwhelmingly dominant in the field of computer vision and image/video processing for the last decade. However, for image and video compression, it lags behind th…

eess.IV2021

Lossless Image Compression Using a Multi-Scale Progressive Statistical Model

Honglei Zhang, Francesco Cricri, Hamed R. Tavakoli +3

Lossless image compression is an important technique for image storage and transmission when information loss is not allowed. With the fast development of deep learning techniques,…

eess.IV202159 cited

Learned Image Coding for Machines: A Content-Adaptive Approach

Nam Le, Honglei Zhang, Francesco Cricri +3

Today, according to the Cisco Annual Internet Report (2018-2023), the fastest-growing category of Internet traffic is machine-to-machine communication. In particular, machine-to-ma…

cs.HC202181 cited

Understanding Visual Saliency in Mobile User Interfaces

Luis A. Leiva, Yunfei Xue, Avya Bansal +4

For graphical user interface (UI) design, it is important to understand what attracts visual attention. While previous work on saliency has focused on desktop and web-based UIs, mo…

eess.IV20202 cited

End-to-End Learning for Video Frame Compression with Self-Attention

Nannan Zou, Honglei Zhang, Francesco Cricri +5

One of the core components of conventional (i.e., non-learned) video codecs consists of predicting a frame from a previously-decoded frame, by leveraging temporal correlations. In…