1 citations · 1 across the 2 of their papers we have counts for
4 papers
MedFact-R1: Towards Factual Medical Reasoning via Pseudo-Label Augmentation
Gengliang Li, Rongyu Chen, Bin Li +2
Ensuring factual consistency and reliable reasoning remains a critical challenge for medical vision-language models. We introduce MEDFACT-R1, a two-stage framework that integrates…
A Temporal Modeling Framework for Video Pre-Training on Video Instance Segmentation
Qing Zhong, Peng-Tao Jiang, Wen Wang +3
Contemporary Video Instance Segmentation (VIS) methods typically adhere to a pre-train then fine-tune regime, where a segmentation model trained on images is fine-tuned on videos.…
Condensing Action Segmentation Datasets via Generative Network Inversion
Guodong Ding, Rongyu Chen, Angela Yao
This work presents the first condensation approach for procedural video datasets used in temporal action segmentation. We propose a condensation framework that leverages generative…
OnlineTAS: An Online Baseline for Temporal Action Segmentation
Qing Zhong, Guodong Ding, Angela Yao
Temporal context plays a significant role in temporal action segmentation. In an offline setting, the context is typically captured by the segmentation network after observing the…