14 citations · 32 across the 15 of their papers we have counts for
15 papers
SC-Tune: Unleashing Self-Consistent Referential Comprehension in Large Vision Language Models
Tongtian Yue, Jie Cheng, Longteng Guo +6
Recent trends in Large Vision Language Models (LVLMs) research have been increasingly focusing on advancing beyond general image understanding towards more nuanced, object-level re…
Repetitive nonoverlapping sequential pattern mining
Meng Geng, Youxi Wu, Yan Li +4
Sequential pattern mining (SPM) is an important branch of knowledge discovery that aims to mine frequent sub-sequences (patterns) in a sequential database. Various SPM methods have…
GLOBER: Coherent Non-autoregressive Video Generation via GLOBal Guided Video DecodER
Mingzhen Sun, Weining Wang, Zihan Qin +3
Video generation necessitates both global coherence and local realism. This work presents a novel non-autoregressive method GLOBER, which first generates global features to obtain…
PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models
Li Chen, Mengyi Zhao, Yiheng Liu +8
Personalized text-to-image generation has emerged as a powerful and sought-after tool, empowering users to create customized images based on their specific concepts and prompts. Ho…
Perceptual Quality Assessment of Omnidirectional Audio-visual Signals
Xilei Zhu, Huiyu Duan, Yuqin Cao +6
Omnidirectional videos (ODVs) play an increasingly important role in the application fields of medical, education, advertising, tourism, etc. Assessing the quality of ODVs is signi…
AIGCIQA2023: A Large-scale Image Quality Assessment Database for AI Generated Images: from the Perspectives of Quality, Authenticity and Correspondence
Jiarui Wang, Huiyu Duan, Jing Liu +3
In this paper, in order to get a better understanding of the human visual preferences for AIGIs, a large-scale IQA database for AIGC is established, which is named as AIGCIQA2023.…