12 papers
GS-RealBlur: A Flexible Data Acquisition Framework for Real-World Image Deblurring
Mingyang Chen, Zhilu Zhang, Honglei Xu +3
High-quality, large-scale paired data is essential for training learning-based image deblurring models. However, synthetic blurry images generally lack realism, while real-world ca…
Illuminating Unified Multimodal Model for Free-form Interleaved Text-Image Generation
Chonghuinan Wang, Zhikai Chen, Chunwei Wang +9
The advancement of generative AI models capable of producing text and image marks a critical step forward in the realm of multimodal intelligence, particularly for tasks involving…
S2AM3D: Scale-controllable Part Segmentation of 3D Point Clouds
Han Su, Tianyu Huang, Zichen Wan +2
Part-level point cloud segmentation has recently attracted significant attention in 3D computer vision. Nevertheless, existing research is constrained by two major challenges: nati…
SelfHVD: Self-Supervised Handheld Video Deblurring
Honglei Xu, Zhilu Zhang, Junjie Fan +2
Shooting video with handheld shooting devices often results in blurry frames due to shaking hands and other instability factors. Although previous video deblurring methods have ach…
CREval: An Automated Interpretable Evaluation for Creative Image Manipulation under Complex Instructions
Chonghuinan Wang, Zihan Chen, Yuxiang Wei +5
Instruction-based multimodal image manipulation has recently made rapid progress. However, existing evaluation methods lack a systematic and human-aligned framework for assessing m…
MiM-DiT: MoE in MoE with Diffusion Transformers for All-in-One Image Restoration
Lingshun Kong, Jiawei Zhang, Zhengpeng Duan +6
All-in-one image restoration is challenging because different degradation types, such as haze, blur, noise, and low-light, impose diverse requirements on restoration strategies, ma…