Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024
Let Video Teaches You More: Video-to-Image Knowledge Distillation using DEtection TRansformer for Medical Video Lesion Detection
Yuncheng Jiang, Zixun Zhang, Jun Wei +5
AI-assisted lesion detection models play a crucial role in the early screening of cancer. However, previous image-based models ignore the inter-frame contextual information present…
cs.CV2024
ToDER: Towards Colonoscopy Depth Estimation and Reconstruction with Geometry Constraint Adaptation
Zhenhua Wu, Yanlin Jin, Liangdong Qiu +3
Visualizing colonoscopy is crucial for medical auxiliary diagnosis to prevent undetected polyps in areas that are not fully observed. Traditional feature-based and depth-based reco…
cs.CV2024
Large Multimodal Agents: A Survey
Junlin Xie, Zhihong Chen, Ruifei Zhang +2
Large language models (LLMs) have achieved superior performance in powering text-based AI agents, endowing them with decision-making and reasoning abilities akin to humans. Concurr…