6 papers
RefDrone: A Challenging Benchmark for Referring Expression Comprehension in Drone Scenes
Zhichao Sun, Yepeng Liu, Zhiling Su +7
Drones have become prevalent robotic platforms with diverse applications, showing significant potential in Embodied Artificial Intelligence (Embodied AI). Referring Expression Comp…
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching
Yepeng Liu, Zhichao Sun, Baosheng Yu +4
Many keypoint detection and description methods have been proposed for image matching or registration. While these methods demonstrate promising performance for single-modality ima…
Leveraging Textual Anatomical Knowledge for Class-Imbalanced Semi-Supervised Multi-Organ Segmentation
Yuliang Gu, Weilun Tsao, Bo Du +2
Annotating 3D medical images demands substantial time and expertise, driving the adoption of semi-supervised learning (SSL) for segmentation tasks. However, the complex anatomical…
Consistent-Point: Consistent Pseudo-Points for Semi-Supervised Crowd Counting and Localization
Yuda Zou, Zelong Liu, Yuliang Gu +2
Crowd counting and localization are important in applications such as public security and traffic management. Existing methods have achieved impressive results thanks to extensive…
CellSeg1: Robust Cell Segmentation with One Training Image
Peilin Zhou, Bo Du, Yongchao Xu
Recent trends in cell segmentation have shifted towards universal models to handle diverse cell morphologies and imaging modalities. However, for continuously emerging cell types a…
Progressive Retinal Image Registration via Global and Local Deformable Transformations
Yepeng Liu, Baosheng Yu, Tian Chen +4
Retinal image registration plays an important role in the ophthalmological diagnosis process. Since there exist variances in viewing angles and anatomical structures across differe…