10 citations · 47 across the 23 of their papers we have counts for
4 papers · 1 filter
JMed48k: A Multi-Profession Japanese Medical Licensing Benchmark for Vision-Language Model Evaluation
Yue Xun, Junyu Liu, Qian Niu +10
We introduce JMed48k, a multi-profession Japanese healthcare licensing benchmark for evaluating vision-language models. Built from official PDF materials released by the Japanese M…
Deep Learning, Machine Learning -- Digital Signal and Image Processing: From Theory to Application
Weiche Hsieh, Ziqian Bi, Junyu Liu +16
Digital Signal Processing (DSP) and Digital Image Processing (DIP) with Machine Learning (ML) and Deep Learning (DL) are popular research areas in Computer Vision and related field…
From Pixels to Prose: Advancing Multi-Modal Language Models for Remote Sensing
Xintian Sun, Benji Peng, Charles Zhang +10
Remote sensing has evolved from simple image acquisition to complex systems capable of integrating and processing visual and textual data. This review examines the development and…
Deep Learning and Machine Learning -- Object Detection and Semantic Segmentation: From Theory to Applications
Jintao Ren, Ziqian Bi, Qian Niu +16
An in-depth exploration of object detection and semantic segmentation is provided, combining theoretical foundations with practical applications. State-of-the-art advancements in m…