collaborators

7 papers

cs.CV2026

UniTrack: Differentiable Graph Representation Learning for Multi-Object Tracking

Bishoy Galoaa, Xiangyu Bai, Utsav Nandi +3

We present UniTrack, a plug-and-play graph-theoretic loss function designed to significantly enhance multi-object tracking (MOT) performance by directly optimizing tracking-specifi…

cs.CV2025

Broadening View Synthesis of Dynamic Scenes from Constrained Monocular Videos

Le Jiang, Shaotong Zhu, Yedi Luo +2

In dynamic Neural Radiance Fields (NeRF) systems, state-of-the-art novel view synthesis methods often fail under significant viewpoint deviations, producing unstable and unrealisti…

cs.CV2025

K-Track: Kalman-Enhanced Tracking for Accelerating Deep Point Trackers on Edge Devices

Bishoy Galoaa, Pau Closas, Sarah Ostadabbas

Point tracking in video sequences is a foundational capability for real-world computer vision applications, including robotics, autonomous systems, augmented reality, and video ana…

cs.CV2025

Lang2Motion: Bridging Language and Motion through Joint Embedding Spaces

Bishoy Galoaa, Xiangyu Bai, Sarah Ostadabbas

We present Lang2Motion, a framework for language-guided point trajectory generation by aligning motion manifolds with joint embedding spaces. Unlike prior work focusing on human mo…

cs.CV2025

MoReGen: Multi-Agent Motion-Reasoning Engine for Code-based Text-to-Video Synthesis

Xiangyu Bai, He Liang, Bishoy Galoaa +4

While text-to-video (T2V) generation has achieved remarkable progress in photorealism, generating intent-aligned videos that faithfully obey physics principles remains a core chall…

cs.CV2025

Look Around and Pay Attention: Multi-camera Point Tracking Reimagined with Transformers

Bishoy Galoaa, Xiangyu Bai, Shayda Moezzi +4

This paper presents LAPA (Look Around and Pay Attention), a novel end-to-end transformer-based architecture for multi-camera point tracking that integrates appearance-based matchin…