3 citations · 4 across the 8 of their papers we have counts for
4 papers · 2 filters
Autoregressive Video Generation without Vector Quantization
Haoge Deng, Ting Pan, Haiwen Diao +6
This paper presents a novel approach that enables autoregressive video generation with high efficiency. We propose to reformulate the video generation problem as a non-quantized au…
GSSF: Generalized Structural Sparse Function for Deep Cross-modal Metric Learning
Haiwen Diao, Ying Zhang, Shang Gao +3
Cross-modal metric learning is a prominent research topic that bridges the semantic heterogeneity between vision and language. Existing methods frequently utilize simple cosine or…
SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning
Haiwen Diao, Bo Wan, Xu Jia +4
Parameter-efficient transfer learning (PETL) has emerged as a flourishing research field for adapting large pre-trained models to downstream tasks, greatly reducing trainable param…
Unveiling Encoder-Free Vision-Language Models
Haiwen Diao, Yufeng Cui, Xiaotong Li +3
Existing vision-language models (VLMs) mostly rely on vision encoders to extract visual features followed by large language models (LLMs) for visual-language tasks. However, the vi…