2 papers
cs.CV2024
Bridging the Intent Gap: Knowledge-Enhanced Visual Generation
Yi Cheng, Ziwei Xu, Dongyun Lin +5
For visual content generation, discrepancies between user intentions and the generated content have been a longstanding problem. This discrepancy arises from two main factors. Firs…
cs.CV2024
PEVA-Net: Prompt-Enhanced View Aggregation Network for Zero/Few-Shot Multi-View 3D Shape Recognition
Dongyun Lin, Yi Cheng, Shangbo Mao +2
Large vision-language models have impressively promote the performance of 2D visual recognition under zero/few-shot scenarios. In this paper, we focus on exploiting the large visio…