2 papers
cs.MM2024
Towards Unified Representation of Multi-Modal Pre-training for 3D Understanding via Differentiable Rendering
Ben Fei, Yixuan Li, Weidong Yang +2
State-of-the-art 3D models, which excel in recognition tasks, typically depend on large-scale datasets and well-defined category sets. Recent advances in multi-modal pre-training h…
cs.CV2024
Deep Shape-Texture Statistics for Completely Blind Image Quality Evaluation
Yixuan Li, Peilin Chen, Hanwei Zhu +3
Opinion-Unaware Blind Image Quality Assessment (OU-BIQA) models aim to predict image quality without training on reference images and subjective quality scores. Thereinto, image st…