3 papers
cs.CV2025
See&Trek: Training-Free Spatial Prompting for Multimodal Large Language Model
Pengteng Li, Pinhao Song, Wuyang Li +5
We introduce SEE&TREK, the first training-free prompting framework tailored to enhance the spatial understanding of Multimodal Large Language Models (MLLMS) under vision-only const…
cs.CV2025
From Events to Enhancement: A Survey on Event-Based Imaging Technologies
Yunfan Lu, Xiaogang Xu, Pengteng Li +4
Event cameras offering high dynamic range and low latency have emerged as disruptive technologies in imaging. Despite growing research on leveraging these benefits for different im…
cs.RO2024
Robot Trajectron: Trajectory Prediction-based Shared Control for Robot Manipulation
Pinhao Song, Pengteng Li, Erwin Aertbelien +1
We address the problem of (a) predicting the trajectory of an arm reaching motion, based on a few seconds of the motion's onset, and (b) leveraging this predictor to facilitate sha…