5 papers
Flying in Clutter on Monocular RGB by Learning in 3D Radiance Fields with Domain Adaptation
Xijie Huang, Jinhan Li, Tianyue Wu +3
Modern autonomous navigation systems predominantly rely on lidar and depth cameras. However, a fundamental question remains: Can flying robots navigate in clutter using solely mono…
EfficientIML: Efficient High-Resolution Image Manipulation Localization
Jinhan Li, Haoyang He, Lei Xie +1
With imaging devices delivering ever-higher resolutions and the emerging diffusion-based forgery methods, current detectors trained only on traditional datasets (with splicing, cop…
Sing it, Narrate it: Quality Musical Lyrics Translation
Zhuorui Ye, Jinhan Li, Rongwu Xu
Translating lyrics for musicals presents unique challenges due to the need to ensure high translation quality while adhering to singability requirements such as length and rhyme. E…
Harmon: Whole-Body Motion Generation of Humanoid Robots from Language Descriptions
Zhenyu Jiang, Yuqi Xie, Jinhan Li +3
Humanoid robots, with their human-like embodiment, have the potential to integrate seamlessly into human environments. Critical to their coexistence and cooperation with humans is…
OKAMI: Teaching Humanoid Robots Manipulation Skills through Single Video Imitation
Jinhan Li, Yifeng Zhu, Yuqi Xie +4
We study the problem of teaching humanoid robots manipulation skills by imitating from single video demonstrations. We introduce OKAMI, a method that generates a manipulation plan…