Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Enhancing Image Aesthetics with Dual-Conditioned Diffusion Models Guided by Multimodal Perception
Xinyu Nan, Ning Wang, Yuyao Zhai +1
Image aesthetic enhancement aims to perceive aesthetic deficiencies in images and perform corresponding editing operations, which is highly challenging and requires the model to po…
cs.CV2025
UniDGF: A Unified Detection-to-Generation Framework for Hierarchical Object Visual Recognition
Xinyu Nan, Lingtao Mao, Huangyu Dai +8
Achieving visual semantic understanding requires a unified framework that simultaneously handles object detection, category prediction, and attribute recognition. However, current…
cs.CV2024
PAM: A Propagation-Based Model for Segmenting Any 3D Objects across Multi-Modal Medical Images
Zifan Chen, Xinyu Nan, Jiazheng Li +9
Volumetric segmentation is important in medical imaging, but current methods face challenges like requiring lots of manual annotations and being tailored to specific tasks, which l…