3 papers
cs.CV2026
LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel
Zhe Feng, Sen Lian, Changwei Wang +5
The quadratic complexity of softmax attention presents a major obstacle for scaling Transformers to high-resolution vision tasks. Existing linear attention variants often replace t…
cs.RO2026
HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare
Rongtao Xu, Mingming Yu, Xiaofeng Han +7
The rapid advancement of Embodied Intelligence has opened transformative opportunities in healthcare, particularly in physical therapy and rehabilitation. However, critical challen…
cs.RO2025
Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
Xiaofeng Han, Shunpeng Chen, Zenghuang Fu +9
Robot vision has greatly benefited from advancements in multimodal fusion techniques and vision-language models (VLMs). We adopt a task-oriented perspective to systematically revie…