20 citations · 20 across the 2 of their papers we have counts for
4 papers
LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Peng Xu, Wenqi Shao, Kaipeng Zhang +7
Large Vision-Language Models (LVLMs) have recently played a dominant role in multimodal vision-language learning. Despite the great success, it lacks a holistic evaluation of their…
FarSee-Net: Real-Time Semantic Segmentation by Efficient Multi-scale Context Aggregation and Feature Space Super-resolution
Zhanpeng Zhang, Kaipeng Zhang
Real-time semantic segmentation is desirable in many robotic applications with limited computation resources. One challenge of semantic segmentation is to deal with the object scal…
Bootstrap Model Ensemble and Rank Loss for Engagement Intensity Regression
Kai Wang, Jianfei Yang, Da Guo +3
This paper presents our approach for the engagement intensity regression task of EmotiW 2019. The task is to predict the engagement intensity value of a student when he or she is w…
Super-Identity Convolutional Neural Network for Face Hallucination
Kaipeng Zhang, Zhanpeng Zhang, Chia-Wen Cheng +4
Face hallucination is a generative task to super-resolve the facial image with low resolution while human perception of face heavily relies on identity information. However, previo…