343 citations · 416 across the 10 of their papers we have counts for
20 papers · 1 filter
Alleviating Hallucinations in Large Vision-Language Models through Hallucination-Induced Optimization
Xinyu Lyu, Beitao Chen, Lianli Gao +2
Although Large Visual Language Models (LVLMs) have demonstrated exceptional abilities in understanding multimodal data, they invariably suffer from hallucinations, leading to a dis…
AICL: Action In-Context Learning for Video Diffusion Model
Jianzhi Liu, Junchen Zhu, Lianli Gao +2
The open-domain video generation models are constrained by the scale of the training video datasets, and some less common actions still cannot be generated. Some researchers explor…
Make-A-Storyboard: A General Framework for Storyboard with Disentangled and Merged Control
Sitong Su, Litao Guo, Lianli Gao +2
Story Visualization aims to generate images aligned with story prompts, reflecting the coherence of storybooks through visual consistency among characters and scenes.Whereas curren…
ALF: Adaptive Label Finetuning for Scene Graph Generation
Qishen Chen, Jianzhi Liu, Xinyu Lyu +3
Scene Graph Generation (SGG) endeavors to predict the relationships between subjects and objects in a given image. Nevertheless, the long-tail distribution of relations often leads…
JoReS-Diff: Joint Retinex and Semantic Priors in Diffusion Model for Low-light Image Enhancement
Yuhui Wu, Guoqing Wang, Zhiwen Wang +5
Low-light image enhancement (LLIE) has achieved promising performance by employing conditional diffusion models. Despite the success of some conditional methods, previous methods m…
CUCL: Codebook for Unsupervised Continual Learning
Chen Cheng, Jingkuan Song, Xiaosu Zhu +3
The focus of this study is on Unsupervised Continual Learning (UCL), as it presents an alternative to Supervised Continual Learning which needs high-quality manual labeled data. Th…