activity
20242026
collaborators

6 papers

cs.CV2026

Unified Video Dense Prediction from Disjoint Data

Yihong Sun, Seoung Wug Oh, Jiahui Huang +2

Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragmented across incompatible, doma…

cs.CV2026

Efficient Tracking and Understanding Object Transformations

Yihong Sun, Bharath Hariharan

Tracking objects through state transformations is essential for understanding real-world dynamics. However, existing methods are computationally expensive. TubeletGraph recently sh…

cs.CV2026

Live Interactive Training for Video Segmentation

Xinyu Yang, Haozheng Yu, Yihong Sun +2

Interactive video segmentation often requires many user interventions for robust performance in challenging scenarios (e.g., occlusions, object separations, camouflage, etc.). Yet,…

cs.CV2026

Tracking and Understanding Object Transformations

Yihong Sun, Xinyu Yang, Jennifer J. Sun +1

Real-world objects frequently undergo state transformations. From an apple being cut into pieces to a butterfly emerging from its cocoon, tracking through these changes is importan…

cs.CV2025

Learning 3D Perception from Others' Predictions

Jinsu Yoo, Zhenyang Feng, Tai-Yu Pan +7

Accurate 3D object detection in real-world environments requires a huge amount of annotated data with high quality. Acquiring such data is tedious and expensive, and often needs re…

cs.CV2024

Video Creation by Demonstration

Yihong Sun, Hao Zhou, Liangzhe Yuan +7

We explore a novel video creation experience, namely Video Creation by Demonstration. Given a demonstration video and a context image from a different scene, we generate a physical…