1 citations · 3 across the 8 of their papers we have counts for
4 papers · 1 filter
ProCrop: Learning Aesthetic Image Cropping from Professional Compositions
Ke Zhang, Tianyu Ding, Jiachen Jiang +4
Image cropping is crucial for enhancing the visual appeal and narrative impact of photographs, yet existing rule-based and data-driven approaches often lack diversity or require an…
Analyzing and Mitigating Model Collapse in Rectified Flow Models
Huminhao Zhu, Fangyikang Wang, Tianyu Ding +2
Training with synthetic data is becoming increasingly inevitable as synthetic content proliferates across the web, driven by the remarkable performance of recent deep generative mo…
OFER: Occluded Face Expression Reconstruction
Pratheba Selvaraju, Victoria Fernandez Abrevaya, Timo Bolkart +4
Reconstructing 3D face models from a single image is an inherently ill-posed problem, which becomes even more challenging in the presence of occlusions. In addition to fewer availa…
Probing the Robustness of Vision-Language Pretrained Models: A Multimodal Adversarial Attack Approach
Jiwei Guan, Tianyu Ding, Longbing Cao +3
Vision-language pretraining (VLP) with transformers has demonstrated exceptional performance across numerous multimodal tasks. However, the adversarial robustness of these models h…