Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Learning to Hear by Seeing: It's Time for Vision Language Models to Understand Artistic Emotion from Sight and Sound
Dengming Zhang, Weitao You, Jingxiong Li +6
Emotion understanding is critical for making Large Language Models (LLMs) more general, reliable, and aligned with humans. Art conveys emotion through the joint design of visual an…
cs.CV2024
Efficient and Scalable Chinese Vector Font Generation via Component Composition
Jinyu Song, Weitao You, Shuhui Shi +3
Chinese vector font generation is challenging due to the complex structure and huge amount of Chinese characters. Recent advances remain limited to generating a small set of charac…
cs.CV2023
Reducing Spatial Fitting Error in Distillation of Denoising Diffusion Models
Shengzhe Zhou, Zejian Lee, Shengyuan Zhang +5
Denoising Diffusion models have exhibited remarkable capabilities in image generation. However, generating high-quality samples requires a large number of iterations. Knowledge dis…