activity
20242026
collaborators

7 papers

cs.CV2026

iDiff: Interpretable Difference-aware Framework for Pairwise Image Quality Assessment

Xinli Yue, JianHui Sun, Tao Shao +3

Pairwise image quality assessment (IQA) in professional photography requires a model not only to identify the preferred image between two candidates, but also to provide convincing…

cs.SE2025

From User Interface to Agent Interface: Efficiency Optimization of UI Representations for LLM Agents

Dezhi Ran, Zhi Gong, Yuzhe Guo +10

While Large Language Model (LLM) agents show great potential for automated UI navigation such as automated UI testing and AI assistants, their efficiency has been largely overlooke…

cs.CV2025

iDETEX: Empowering MLLMs for Intelligent DETailed EXplainable IQA

Zhaoran Zhao, Xinli Yue, Jianhui Sun +5

Image Quality Assessment (IQA) has progressed from scalar quality prediction to more interpretable, human-aligned evaluation paradigms. In this work, we address the emerging challe…

cs.CV2025

Instruction-augmented Multimodal Alignment for Image-Text and Element Matching

Xinli Yue, JianHui Sun, Junda Lu +6

With the rapid advancement of text-to-image (T2I) generation models, assessing the semantic alignment between generated images and text descriptions has become a significant resear…

cs.SE2025

Beyond Pass or Fail: Multi-Dimensional Benchmarking of Foundation Models for Goal-based Mobile UI Navigation

Dezhi Ran, Mengzhou Wu, Hao Yu +15

Recent advances of foundation models (FMs) have made navigating mobile applications (apps) based on high-level goal instructions within reach, with significant industrial applicati…

cs.CV2024

Advancing Video Quality Assessment for AIGC

Xinli Yue, Jianhui Sun, Han Kong +9

In recent years, AI generative models have made remarkable progress across various domains, including text generation, image generation, and video generation. However, assessing th…