228 citations · 315 across the 22 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.CV2024
TG-LLaVA: Text Guided LLaVA via Learnable Latent Embeddings
Dawei Yan, Pengcheng Li, Yang Li +7
Currently, inspired by the success of vision-language models (VLMs), an increasing number of researchers are focusing on improving VLMs and have achieved promising results. However…
cs.CV2024★ 2 cited
3D-RCNet: Learning from Transformer to Build a 3D Relational ConvNet for Hyperspectral Image Classification
Haizhao Jing, Liuwei Wan, Xizhe Xue +2
Recently, the Vision Transformer (ViT) model has replaced the classical Convolutional Neural Network (ConvNet) in various computer vision tasks due to its superior performance. Eve…