From the 1 of 4 linked papers with an AI index.
4 papers
AgenticFocus: Object-Preserving Mixed Reality Synthesis from Human FPV Video for Dexterous Humanoid Learning
Iaroslav Kolomiets, Miguel Altamirano Cabrera, Artem Lykov +6
The paper presents AgenticFocus, a mixed-reality pipeline that turns ordinary first-person human videos into robot-ready demonstrations by reconstructing hidden object geometry, co…
VLH: Vision-Language-Haptics Foundation Model
Luis Francisco Moreno Fuentes, Muhammad Haris Khan, Miguel Altamirano Cabrera +5
We present VLH, a novel Visual-Language-Haptic Foundation Model that unifies perception, language, and tactile feedback in aerial robotics and virtual reality. Unlike prior work th…
HapticVLM: VLM-Driven Texture Recognition Aimed at Intelligent Haptic Interaction
Muhammad Haris Khan, Miguel Altamirano Cabrera, Dmitrii Iarchuk +4
This paper introduces HapticVLM, a novel multimodal system that integrates vision-language reasoning with deep convolutional networks to enable real-time haptic feedback. HapticVLM…
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
Muhamamd Haris Khan, Selamawit Asfaw, Dmitrii Iarchuk +4
This paper introduces Shake-VLA, a Vision-Language-Action (VLA) model-based system designed to enable bimanual robotic manipulation for automated cocktail preparation. The system i…