4 papers
SurgVLA-Bench: Towards Evaluating Vision-Language-Action Models for Laparoscopic Surgical Robotics
Jiashuo Sun, Yue He, Wenxuan Liu +4
Vision-Language-Action (VLA) models represent a promising direction for embodied intelligence in surgical robotics. Despite the prevalence of VLA benchmarks for general robotics, s…
A Step to Decouple Optimization in 3DGS
Renjie Ding, Yaonan Wang, Min Liu +6
3D Gaussian Splatting (3DGS) has emerged as a powerful technique for real-time novel view synthesis. As an explicit representation optimized through gradient propagation among prim…
Reg-TTR, Test-Time Refinement for Fast, Robust and Accurate Image Registration
Lin Chen, Yue He, Fengting Zhang +4
Traditional image registration methods are robust but slow due to their iterative nature. While deep learning has accelerated inference, it often struggles with domain shifts. Emer…
Reconsider the Template Mesh in Deep Learning-based Mesh Reconstruction
Fengting Zhang, Boxu Liang, Qinghao Liu +3
Mesh reconstruction is a cornerstone process across various applications, including in-silico trials, digital twins, surgical planning, and navigation. Recent advancements in deep…