3 papers
cs.LG2024
Scalable Weibull Graph Attention Autoencoder for Modeling Document Networks
Chaojie Wang, Xinyang Liu, Dongsheng Wang +3
Although existing variational graph autoencoders (VGAEs) have been widely used for modeling and generating graph-structured data, most of them are still not flexible enough to appr…
cs.CV2024
Instruction Tuning-free Visual Token Complement for Multimodal LLMs
Dongsheng Wang, Jiequan Cui, Miaoge Li +3
As the open community of large language models (LLMs) matures, multimodal LLMs (MLLMs) have promised an elegant bridge between vision and language. However, current research is inh…
cs.CV2024
Patch-Prompt Aligned Bayesian Prompt Tuning for Vision-Language Models
Xinyang Liu, Dongsheng Wang, Bowei Fang +5
For downstream applications of vision-language pre-trained models, there has been significant interest in constructing effective prompts. Existing works on prompt engineering, whic…