5 papers
AVI-Edit: Audio-sync Video Instance Editing with Granularity-Aware Mask Refiner
Haojie Zheng, Shuchen Weng, Jingqi Liu +3
Recent advancements in video generation highlight that realistic audio-visual synchronization is crucial for engaging content creation. However, existing video editing methods larg…
MeniMV: A Multi-view Benchmark for Meniscus Injury Severity Grading
Shurui Xu, Siqi Yang, Jiapin Ren +5
Precise grading of meniscal horn tears is critical in knee injury diagnosis but remains underexplored in automated MRI analysis. Existing methods often rely on coarse study-level l…
PanoWan: Lifting Diffusion Video Generation Models to 360° with Latitude/Longitude-aware Mechanisms
Yifei Xia, Shuchen Weng, Siqi Yang +6
Panoramic video generation enables immersive 360° content creation, valuable in applications that demand scene-consistent world exploration. However, existing panoramic video gene…
Unsupervised Cross-Domain Regression for Fine-grained 3D Game Character Reconstruction
Qi Wen, Xiang Wen, Hao Jiang +5
With the rise of the ``metaverse'' and the rapid development of games, it has become more and more critical to reconstruct characters in the virtual world faithfully. The immersive…
RiEMann: Near Real-Time SE(3)-Equivariant Robot Manipulation without Point Cloud Segmentation
Chongkai Gao, Zhengrong Xue, Shuying Deng +4
We present RiEMann, an end-to-end near Real-time SE(3)-Equivariant Robot Manipulation imitation learning framework from scene point cloud input. Compared to previous methods that r…