28 citations · 35 across the 7 of their papers we have counts for
7 papers
Explainable Interfaces for Rapid Gaze-Based Interactions in Mixed Reality
Mengjie Yu, Dustin Harris, Ian Jones +12
Gaze-based interactions offer a potential way for users to naturally engage with mixed reality (XR) interfaces. Black-box machine learning models enabled higher accuracy for gaze-b…
NTIRE 2024 Challenge on Short-form UGC Video Quality Assessment: Methods and Results
Xin Li, Kun Yuan, Yajing Pei +65
This paper reviews the NTIRE 2024 Challenge on Shortform UGC Video Quality Assessment (S-UGC VQA), where various excellent solutions are submitted and evaluated on the collected da…
AnyMAL: An Efficient and Scalable Any-Modality Augmented Language Model
Seungwhan Moon, Andrea Madotto, Zhaojiang Lin +10
We present Any-Modality Augmented Language Model (AnyMAL), a unified model that reasons over diverse input modality signals (i.e. text, image, video, audio, IMU motion sensor), and…
Sparse-based Domain Adaptation Network for OCTA Image Super-Resolution Reconstruction
Huaying Hao, Cong Xu, Dan Zhang +4
Retinal Optical Coherence Tomography Angiography (OCTA) with high-resolution is important for the quantification and analysis of retinal vasculature. However, the resolution of OCT…
Multiple Kernel Clustering with Dual Noise Minimization
Junpu Zhang, Liang Li, Siwei Wang +4
Clustering is a representative unsupervised method widely applied in multi-modal and multi-view scenarios. Multiple kernel clustering (MKC) aims to group data by integrating comple…
Spatial Transformation for Image Composition via Correspondence Learning
Bo Zhang, Yue Liu, Kaixin Lu +2
When using cut-and-paste to acquire a composite image, the geometry inconsistency between foreground and background may severely harm its fidelity. To address the geometry inconsis…