15 citations · 19 across the 5 of their papers we have counts for
5 papers
ZALM3: Zero-Shot Enhancement of Vision-Language Alignment via In-Context Information in Multi-Turn Multimodal Medical Dialogue
Zhangpu Li, Changhong Zou, Suxue Ma +13
The rocketing prosperity of large language models (LLMs) in recent years has boosted the prevalence of vision-language models (VLMs) in the medical sector. In our online medical co…
DiffMOT: A Real-time Diffusion-based Multiple Object Tracker with Non-linear Prediction
Weiyi Lv, Yuhang Huang, Ning Zhang +3
In Multiple Object Tracking, objects often exhibit non-linear motion of acceleration and deceleration, with irregular direction changes. Tacking-by-detection (TBD) trackers with Ka…
HSTFormer: Hierarchical Spatial-Temporal Transformers for 3D Human Pose Estimation
Xiaoye Qian, Youbao Tang, Ning Zhang +4
Transformer-based approaches have been successfully proposed for 3D human pose estimation (HPE) from 2D pose sequence and achieved state-of-the-art (SOTA) performance. However, cur…
Accurate and Robust Lesion RECIST Diameter Prediction and Segmentation with Transformers
Youbao Tang, Ning Zhang, Yirui Wang +4
Automatically measuring lesion/tumor size with RECIST (Response Evaluation Criteria In Solid Tumors) diameters and segmentation is important for computer-aided diagnosis. Although…
PieTrack: An MOT solution based on synthetic data training and self-supervised domain adaptation
Yirui Wang, Shenghua He, Youbao Tang +8
In order to cope with the increasing demand for labeling data and privacy issues with human detection, synthetic data has been used as a substitute and showing promising results in…