works on

From the 1 of 5 linked papers with an AI index.

collaborators

5 papers

cs.HC2026

SpaceVLA: Spatially Grounded VLA for Robotic Manipulation with User-Authored Grasp and Place Anchors

Daniia Zinniatullina, Iaroslav Kolomiets, Mikhail Konenkov +2

Vision-language-action (VLA) models follow language commands but often lack explicit spatial intent for manipulation. We present Visual Intent Anchors, an XR pipeline that lets use…

cs.RO2026

ORCESTRA: VLM-driven Visual Robot programming in Mixed Reality

Ivan Snegirev, Elizaveta Semenyakina, Mikhail Konenkov +3

ORCESTRA is a mixed-reality system for programming robot digital twins through no-code waypoint teaching and language-guided control. In a passthrough mixed-reality workspace, user…

cs.RO2026

AgenticFocus: Object-Preserving Mixed Reality Synthesis from Human FPV Video for Dexterous Humanoid Learning

Iaroslav Kolomiets, Miguel Altamirano Cabrera, Artem Lykov +6

The paper presents AgenticFocus, a mixed-reality pipeline that turns ordinary first-person human videos into robot-ready demonstrations by reconstructing hidden object geometry, co…

cs.RO2026

VersualRL: Closed-Loop Verbal Reinforcement Learning with Visual Execution Feedback for Task-Level Robot Planning

Dmitrii Plotnikov, Iaroslav Kolomiets, Dmitrii Maliukov +9

We introduce VersualRL, a closed-loop framework for task-level robot planning that uses visual execution feedback to iteratively refine executable Behavior Trees through structured…

cs.RO2025

PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models

Artem Lykov, Jeffrin Sam, Hung Khang Nguyen +6

We introduce PhysicalAgent, an agentic framework for robotic manipulation that integrates iterative reasoning, diffusion-based video generation, and closed-loop execution. Given a…