58 citations · 64 across the 16 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Beyond Emotion Recognition: A Multi-Turn Multimodal Emotion Understanding and Reasoning Benchmark
Jinpeng Hu, Hongchang Shi, Chongyuan Dai +3
Multimodal large language models (MLLMs) have been widely applied across various fields due to their powerful perceptual and reasoning capabilities. In the realm of psychology, the…
cs.CV2022
Improving Radiology Summarization with Radiograph and Anatomy Prompts
Jinpeng Hu, Zhihong Chen, Yang Liu +2
The impression is crucial for the referring physicians to grasp key information since it is concluded from the findings and reasoning of radiologists. To alleviate the workload of…
cs.CV2022
Multi-Modal Masked Autoencoders for Medical Vision-and-Language Pre-Training
Zhihong Chen, Yuhao Du, Jinpeng Hu +4
Medical vision-and-language pre-training provides a feasible solution to extract effective vision-and-language representations from medical images and texts. However, few studies h…