5 papers
MRSE: An Efficient Multi-modality Retrieval System for Large Scale E-commerce
Hao Jiang, Haoxiang Zhang, Qingshan Hou +4
Providing high-quality item recall for text queries is crucial in large-scale e-commerce search systems. Current Embedding-based Retrieval Systems (ERS) embed queries and items int…
LMM-PCQA: Assisting Point Cloud Quality Assessment with LMM
Zicheng Zhang, Haoning Wu, Yingjie Zhou +7
Although large multi-modality models (LMMs) have seen extensive exploration and application in various quality assessment studies, their integration into Point Cloud Quality Assess…
Q-Ground: Image Quality Grounding with Large Multi-modality Models
Chaofeng Chen, Sensen Yang, Haoning Wu +6
Recent advances of large multi-modality models (LMM) have greatly improved the ability of image quality assessment (IQA) method to evaluate and explain the quality of visual conten…
Enhancing Diffusion Models with Text-Encoder Reinforcement Learning
Chaofeng Chen, Annan Wang, Haoning Wu +4
Text-to-image diffusion models are typically trained to optimize the log-likelihood objective, which presents challenges in meeting specific requirements for downstream tasks, such…
G-Refine: A General Quality Refiner for Text-to-Image Generation
Chunyi Li, Haoning Wu, Hongkun Hao +7
With the evolution of Text-to-Image (T2I) models, the quality defects of AI-Generated Images (AIGIs) pose a significant barrier to their widespread adoption. In terms of both perce…