6 papers
GT-PCQA: Geometry-Texture Decoupled Point Cloud Quality Assessment with MLLM
Guohua Zhang, Jian Jin, Meiqin Liu +3
With the rapid advancement of Multi-modal Large Language Models (MLLMs), MLLM-based Image Quality Assessment (IQA) methods have shown promising generalization. However, directly ex…
High-Fidelity 3D Facial Avatar Synthesis with Controllable Fine-Grained Expressions
Yikang He, Jichao Zhang, Wei Wang +2
Facial expression editing methods can be mainly categorized into two types based on their architectures: 2D-based and 3D-based methods. The former lacks 3D face modeling capabiliti…
Enhancing Blind Face Restoration through Online Reinforcement Learning
Bin Wu, Yahui Liu, Chi Zhang +2
Blind Face Restoration (BFR) encounters inherent challenges in exploring its large solution space, leading to common artifacts like missing details and identity ambiguity in the re…
Stable-Hair v2: Real-World Hair Transfer via Multiple-View Diffusion Model
Kuiyuan Sun, Yuxuan Zhang, Jichao Zhang +4
While diffusion-based methods have shown impressive capabilities in capturing diverse and complex hairstyles, their ability to generate consistent and high-quality multi-view outpu…
Guiding the Experts: Semantic Priors for Efficient and Focused MoE Routing
Chengxi Min, Wei Wang, Yahui Liu +4
Mixture-of-Experts (MoE) models have emerged as a promising direction for scaling vision architectures efficiently. Among them, Soft MoE improves training stability by assigning ea…
DiffusionReward: Enhancing Blind Face Restoration through Reward Feedback Learning
Bin Wu, Wei Wang, Yahui Liu +2
Reward Feedback Learning (ReFL) has recently shown great potential in aligning model outputs with human preferences across various generative tasks. In this work, we introduce a Re…