17 papers
GeminiPainter's sequence-formed pipeline comprised of perception, cognition, planning, and action stages
Miguel Altamirano Cabrera, Aleksey Fedoseev, Iana Zhura +1
We present an autonomous robotic portrait-generation system combining real-time face detection, AI-based sketch generation, and robotic drawing. The system captures video frames, e…
GoalVLM: VLM-driven Object Goal Navigation for Multi-Agent System
MoniJesu James, Amir Atef Habel, Aleksey Fedoseev +1
Object-goal navigation has traditionally been limited to ground robots with closed-set object vocabularies. Existing multi-agent approaches depend on precomputed probabilistic grap…
GoalSwarm: Multi-UAV Semantic Coordination for Open-Vocabulary Object Navigation
MoniJesu Wonders James, Amir Atef Habel, Aleksey Fedoseev +1
Cooperative visual semantic navigation is a foundational capability for aerial robot teams operating in unknown environments. However, achieving robust open-vocabulary object-goal…
ImpedanceDiffusion: Diffusion-Based Global Path Planning for UAV Swarm Navigation with Generative Impedance Control
Faryal Batool, Yasheerah Yaqoot, Muhammad Ahsan Mustafa +3
Safe swarm navigation in cluttered indoor environment requires long-horizon planning, reactive obstacle avoidance, and adaptive compliance. We propose ImpedanceDiffusion, a hierarc…
Hybrid F' and ROS2 Architecture for Vision-Based Autonomous Flight: Design and Experimental Validation
Abdelrahman Metwally, Monijesu James, Aleksey Fedoseev +3
Autonomous aerospace systems require architectures that balance deterministic real-time control with advanced perception capabilities. This paper presents an integrated system comb…
DiffusionCinema: Text-to-Aerial Cinematography
Valerii Serpiva, Artem Lykov, Jeffrin Sam +2
We propose a novel Unmanned Aerial Vehicles (UAV) assisted creative capture system that leverages diffusion models to interpret high-level natural language prompts and automaticall…