activity
20222024
most citedHSTFormer: Hierarchical Spatial-Temporal Transformers for 3D Human Pose Estimation

15 citations · 19 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CL2024

ZALM3: Zero-Shot Enhancement of Vision-Language Alignment via In-Context Information in Multi-Turn Multimodal Medical Dialogue

Zhangpu Li, Changhong Zou, Suxue Ma +13

The rocketing prosperity of large language models (LLMs) in recent years has boosted the prevalence of vision-language models (VLMs) in the medical sector. In our online medical co…

cs.CV20241 cited

DiffMOT: A Real-time Diffusion-based Multiple Object Tracker with Non-linear Prediction

Weiyi Lv, Yuhang Huang, Ning Zhang +3

In Multiple Object Tracking, objects often exhibit non-linear motion of acceleration and deceleration, with irregular direction changes. Tacking-by-detection (TBD) trackers with Ka…

cs.CV202315 cited

HSTFormer: Hierarchical Spatial-Temporal Transformers for 3D Human Pose Estimation

Xiaoye Qian, Youbao Tang, Ning Zhang +4

Transformer-based approaches have been successfully proposed for 3D human pose estimation (HPE) from 2D pose sequence and achieved state-of-the-art (SOTA) performance. However, cur…

eess.IV20222 cited

Accurate and Robust Lesion RECIST Diameter Prediction and Segmentation with Transformers

Youbao Tang, Ning Zhang, Yirui Wang +4

Automatically measuring lesion/tumor size with RECIST (Response Evaluation Criteria In Solid Tumors) diameters and segmentation is important for computer-aided diagnosis. Although…

cs.CV20221 cited

PieTrack: An MOT solution based on synthetic data training and self-supervised domain adaptation

Yirui Wang, Shenghua He, Youbao Tang +8

In order to cope with the increasing demand for labeling data and privacy issues with human detection, synthetic data has been used as a substitute and showing promising results in…