activity
20242026
collaborators

5 papers

cs.RO2026

SutureFormer: Learning Surgical Trajectories via Goal-conditioned Offline RL in Pixel Space

Huanrong Liu, Chunlin Tian, Tongyu Jia +8

Predicting surgical needle trajectories from endoscopic video is critical for robot-assisted suturing, enabling anticipatory planning, real-time guidance, and safer motion executio…

cs.CV2026

ESAM++: Efficient Online 3D Perception on the Edge

Qin Liu, Lavisha Aggarwal, Saptarashmi Bandyopadhyay +4

Online 3D scene perception in real time is essential for robotics, AR/VR, and autonomous systems, particularly in edge computing scenarios where computational resources are limited…

cs.CV2026

Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy

Jiaheng Dai, Huanrong Liu, Tailai Zhou +7

Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing gestures with variable durat…

cs.CV2025

PPBoost: Progressive Prompt Boosting for Text-Driven Medical Image Segmentation

Xuchen Li, Hengrui Gu, Mohan Zhang +6

Text-prompted foundation models for medical image segmentation offer an intuitive way to delineate anatomical structures from natural language queries, but their predictions often…

cs.CV2024

LiVOS: Light Video Object Segmentation with Gated Linear Matching

Qin Liu, Jianfeng Wang, Zhengyuan Yang +4

Semi-supervised video object segmentation (VOS) has been largely driven by space-time memory (STM) networks, which store past frame features in a spatiotemporal memory to segment t…