5 papers
SutureFormer: Learning Surgical Trajectories via Goal-conditioned Offline RL in Pixel Space
Huanrong Liu, Chunlin Tian, Tongyu Jia +8
Predicting surgical needle trajectories from endoscopic video is critical for robot-assisted suturing, enabling anticipatory planning, real-time guidance, and safer motion executio…
ESAM++: Efficient Online 3D Perception on the Edge
Qin Liu, Lavisha Aggarwal, Saptarashmi Bandyopadhyay +4
Online 3D scene perception in real time is essential for robotics, AR/VR, and autonomous systems, particularly in edge computing scenarios where computational resources are limited…
Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy
Jiaheng Dai, Huanrong Liu, Tailai Zhou +7
Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing gestures with variable durat…
PPBoost: Progressive Prompt Boosting for Text-Driven Medical Image Segmentation
Xuchen Li, Hengrui Gu, Mohan Zhang +6
Text-prompted foundation models for medical image segmentation offer an intuitive way to delineate anatomical structures from natural language queries, but their predictions often…
LiVOS: Light Video Object Segmentation with Gated Linear Matching
Qin Liu, Jianfeng Wang, Zhengyuan Yang +4
Semi-supervised video object segmentation (VOS) has been largely driven by space-time memory (STM) networks, which store past frame features in a spatiotemporal memory to segment t…