activity
20242026
collaborators

5 papers

cs.CV2026

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators

Jintao Rong, Xin Xie, Xinyi Yu +4

Training-free motion customization imposes motion patterns from reference videos onto video generators through test-time computation. Most existing methods target full diffusion mo…

cs.CV2026

Exploring Spatial Intelligence from a Generative Perspective

Muzhi Zhu, Shunyao Jiang, Huanyi Zheng +9

Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective. We ask whether modern gener…

cs.CV2026

RealHD: A High-Quality Dataset for Robust Detection of State-of-the-Art AI-Generated Images

Hanzhe Yu, Yun Ye, Jintao Rong +2

The rapid advancement of generative AI has raised concerns about the authenticity of digital images, as highly realistic fake images can now be generated at low cost, potentially i…

cs.RO2025

GSORB-SLAM: Gaussian Splatting SLAM benefits from ORB features and Transmittance information

Wancai Zheng, Xinyi Yu, Jintao Rong +3

The emergence of 3D Gaussian Splatting (3DGS) has recently ignited a renewed wave of research in dense visual SLAM. However, existing approaches encounter challenges, including sen…

cs.CV2024

Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification

Jintao Rong, Hao Chen, Linlin Ou +3

The Contrastive Language-Image Pretraining (CLIP) model has been widely used in various downstream vision tasks. The few-shot learning paradigm has been widely adopted to augment i…